← back to catalog · registered 2026-09-11 08:55

Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ8e-DFlash2

Solstice-AI Glm GGUF multimodal second-order
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals — repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
1K
Likes
3
Model age
4d ago
created 2026-09-07

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Languages
en zh
Tags
mlx safetensors gguf glm5_next solstice-ai glm glm5 glm-5.3-flash oq8e mixed-precision apple-silicon metal

Related

Total size
319 GB
Files
82
Quantizations
2
Registered
2026-09-11 08:55
Last updated on HF
2026-09-11 09:47

Files by quantization

BF16 1 file 1.08 GB
mmproj-BF16.gguf 1.08 GB 513c9bfc download
Auxiliary files 81 files 319 GB
model-00067-of-00069.safetensors 4.68 GB 35bc2b8c download
model-00055-of-00069.safetensors 4.66 GB 3af0df78 download
model-00019-of-00069.safetensors 4.66 GB c1cbdd04 download
model-00047-of-00069.safetensors 4.66 GB 70a7a131 download
model-00008-of-00069.safetensors 4.66 GB 319a61a7 download
model-00060-of-00069.safetensors 4.66 GB 5a8df601 download
model-00027-of-00069.safetensors 4.66 GB 91bfecce download
model-00028-of-00069.safetensors 4.66 GB 7e66487d download
model-00050-of-00069.safetensors 4.66 GB 7433c46e download
model-00016-of-00069.safetensors 4.66 GB e819f1b4 download
model-00057-of-00069.safetensors 4.66 GB 10e956ff download
model-00020-of-00069.safetensors 4.66 GB 24764464 download
model-00046-of-00069.safetensors 4.66 GB 04d11dec download
model-00042-of-00069.safetensors 4.66 GB 14d49c10 download
model-00032-of-00069.safetensors 4.66 GB 5b260319 download
model-00023-of-00069.safetensors 4.66 GB 065a50b6 download
model-00004-of-00069.safetensors 4.66 GB 41e8a30f download
model-00043-of-00069.safetensors 4.66 GB 5cebdc1c download
model-00015-of-00069.safetensors 4.66 GB 8bc0da9c download
model-00026-of-00069.safetensors 4.66 GB 5d9df8df download
model-00007-of-00069.safetensors 4.66 GB 5c711930 download
model-00013-of-00069.safetensors 4.66 GB ed6c946f download
model-00054-of-00069.safetensors 4.66 GB 2c0feabf download
model-00066-of-00069.safetensors 4.66 GB fd603bac download
model-00034-of-00069.safetensors 4.66 GB e30615d9 download
model-00052-of-00069.safetensors 4.66 GB effc8565 download
model-00063-of-00069.safetensors 4.66 GB df5cc6d5 download
model-00062-of-00069.safetensors 4.66 GB 71beeed2 download
model-00036-of-00069.safetensors 4.66 GB 40845782 download
model-00039-of-00069.safetensors 4.66 GB da9f78c4 download
model-00001-of-00069.safetensors 4.66 GB 9b4f572c download
model-00010-of-00069.safetensors 4.66 GB 87f97d69 download
model-00044-of-00069.safetensors 4.66 GB 0c7b37b0 download
model-00011-of-00069.safetensors 4.66 GB aad70466 download
model-00030-of-00069.safetensors 4.66 GB d6279135 download
model-00038-of-00069.safetensors 4.66 GB 487a9216 download
model-00005-of-00069.safetensors 4.66 GB f8f46d33 download
model-00024-of-00069.safetensors 4.66 GB 57b0e917 download
model-00035-of-00069.safetensors 4.66 GB a6a4fd44 download
model-00059-of-00069.safetensors 4.66 GB ddcd0d92 download
model-00051-of-00069.safetensors 4.66 GB 32b0e61f download
model-00065-of-00069.safetensors 4.66 GB 0f8d91fb download
model-00009-of-00069.safetensors 4.66 GB 98624e7a download
model-00002-of-00069.safetensors 4.66 GB 3b45ebcd download
model-00068-of-00069.safetensors 4.66 GB cfb995ce download
model-00014-of-00069.safetensors 4.66 GB d67c629c download
model-00033-of-00069.safetensors 4.66 GB a18166d4 download
model-00003-of-00069.safetensors 4.66 GB df3bb336 download
model-00022-of-00069.safetensors 4.66 GB 6b4b4952 download
model-00041-of-00069.safetensors 4.66 GB 027823d3 download
model-00058-of-00069.safetensors 4.66 GB 40be4bde download
model-00017-of-00069.safetensors 4.66 GB 36f82de8 download
model-00049-of-00069.safetensors 4.66 GB c5797a25 download
model-00056-of-00069.safetensors 4.66 GB 18d067f0 download
model-00021-of-00069.safetensors 4.66 GB ef73fb4e download
model-00037-of-00069.safetensors 4.66 GB acd8d790 download
model-00040-of-00069.safetensors 4.66 GB 113dd972 download
model-00048-of-00069.safetensors 4.66 GB 05b181d6 download
model-00029-of-00069.safetensors 4.66 GB 5db7426f download
model-00053-of-00069.safetensors 4.66 GB a5aa4374 download
model-00061-of-00069.safetensors 4.66 GB b72036bb download
model-00064-of-00069.safetensors 4.66 GB 4ef83a9e download
model-00025-of-00069.safetensors 4.66 GB 421d21d6 download
model-00006-of-00069.safetensors 4.66 GB c4a99000 download
model-00031-of-00069.safetensors 4.66 GB 23bf4a0a download
model-00012-of-00069.safetensors 4.66 GB adac95a1 download
model-00045-of-00069.safetensors 4.66 GB f4dd2bda download
model-00018-of-00069.safetensors 4.66 GB 83167bf7 download
model-00069-of-00069.safetensors 1.60 GB 522cb61f download
tokenizer.json 19.3 MB 19e77364 download
model.safetensors.index.json 11.5 MB 46a159d5 download
config.json 67.8 KB f93128cf download
dealign_mascot.png 10.9 KB da3bf39a download
chat_template.jinja 10.4 KB 5d5e1052 download
dealign_logo.png 7.48 KB a5b3546b download
README.md 4.35 KB 170bfe22 download
.gitattributes 1.82 KB 806d4b0a download
LICENSE 1.04 KB 986b06fb download
processor_config.json 909 B 3ec2a058 download
tokenizer_config.json 761 B e375fa0a download
generation_config.json 223 B be58f531 download

README current version from Hugging Face


language:

  • en
  • zh
    license: mit
    base_model: dealignai/GLM-5.3-Flash-UNCENSORED-FP8
    tags:
  • solstice-ai
  • glm
  • glm5
  • glm-5.3-flash
  • mlx
  • oq8e
  • mixed-precision
  • apple-silicon
  • metal
  • vision
  • video
  • multimodal
  • dflash2
  • speculative-decoding
  • image-text-to-text
  • long-context
  • uncensored
  • abliterated
    pipeline_tag: image-text-to-text
    library_name: mlx

Solstice-AI Banner

GLM-5.3-Flash-UNCENSORED (oQ8e Mixed-Precision)

Official Solstice-AI Apple Silicon Release • Native Multimodal Vision + Video • 1M Context Window (1,048,576 Tokens) • Bundled DFlash 2 Speculative Drafter

Original Architecture by Zhipu AI / ZAI • Uncensored Weights by dealignai • oQ8e Mixed-Precision by Solstice-AI

Solstice-AI License Format Precision Context Hardware


Model Summary

Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ8e is the official oQ8e mixed-precision release of the uncensored 320B foundation model, GLM-5.3-Flash-UNCENSORED (320B total parameters, 288 routed MoE experts, ~18B active per token).

Mixed-Precision Quantization Architecture:

  • Base Precision: 8-bit affine (group_size=64).
  • Target bpw: ~8.6 bpw.
  • Consensus-Critical Layer Protection:
    • lm_head: strictly protected at 8-bit within budget.
    • MoE Routers & Gate Projections (mlp.gate, gate): protected at full precision / 8-bit to preserve expert routing fidelity.
    • 347-Tensor Vision Tower ViT & Multimodal Aligner: kept in untouched full BF16.
    • Attention Sinks & Hyper-Connection Tables (hc_*): kept in full BF16/FP32.
  • Native 1M Context Window: 1,048,576 tokens native context.
  • Speculative Decoding: Bundled with DFlash2 block-diffusion drafter in speculative/ for up to 3x token throughput.

Official GLM-5.3-Flash Benchmark Scoreboard

Benchmark Suite Discipline GLM-5.3-Flash Uncensored MLX Base GLM-5.3 Claude 3.5 Sonnet GPT-4o
MMLU General Knowledge & Reasoning 85.28% 86.15% 88.7% 87.2%
HarmBench-320 Safety Refusal Suppression 0% Refusals 94.2% Refusals 92.5% 91.0%
SWE-bench Pro Real-World Software Engineering 63.4% 64.1% 61.2% 48.9%
LiveCodeBench v6 Competitive Algorithmic Coding 86.1% 87.0% 78.4% 72.8%
MATH-500 High-School / Olympiad Math 92.8% 93.4% 89.2% 91.4%
MMMU (Multimodal) Multi-Discipline Visual Understanding 70.8% 71.2% 70.4% 69.1%
VideoQA / Temporal Video Reasoning Across Time Frames 78.5% 79.1% 77.2% 75.6%

Quickstart on Apple Silicon

pip install mlx mlx-lm huggingface_hub
from mlx_lm import load, generate

model, tokenizer = load("Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ8e")
response = generate(model, tokenizer, prompt="Explain sparse mixture-of-experts in GLM-5.3.", max_tokens=1024, verbose=True)
print(response)

DFlash 2 Speculative Decoding Acceleration

This release bundles pre-aligned speculative draft weights in speculative/:

  • speculative/GLM-5.3-Flash-DFlash2-bf16.gguf

To run accelerated inference with speculative drafting:

python -m mlx_lm.generate \
    --model Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ8e \
    --draft-model Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ8e/speculative/GLM-5.3-Flash-DFlash2-bf16.gguf \
    --prompt "Synthesize the architectural innovations of GLM-5.3." \
    --max-tokens 2048
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Infrahuman, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.