← back to catalog · registered 2026-08-26 21:48

dealignai/GLM-5.3-Flash-ABLITERATED-FP8

dealignai Glm MoE multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/dealignai%2FGLM-5.3-Flash-ABLITERATED-FP8"
Response includes
  • classification m-uncensored
  • files 76
  • author_summary 38 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
2K
Likes
22
Descendants
1
in 1 direct fork
Model age
6w ago
created 2026-08-26
Downloads over time
Now1.7K→from0↑0%
06161.2K1.8K0 on Aug 261.7K on Sep 22AugSep
Aug 26 → Sep 22 · 28 snapshots · spans 27 days

Genealogy 1 direct fork

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 5K downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
mit
Languages
en
Tags
safetensors glm5_next abliterated uncensored crack refusal-removed glm moe fp8 vision mtp en

Related

Total size
306 GB
Files
76
Quantizations
1
Registered
2026-08-26 21:48
Last updated on HF
2026-09-19 16:32

Files by quantization

Auxiliary files 76 files 306 GB
model-00001-of-00062.safetensors 5.00 GB 9ff3c939 download
model-00024-of-00062.safetensors 5.00 GB 84cebf97 download
model-00038-of-00062.safetensors 5.00 GB cc7811d3 download
model-00045-of-00062.safetensors 5.00 GB 70c38787 download
model-00052-of-00062.safetensors 5.00 GB 62102ea4 download
model-00007-of-00062.safetensors 5.00 GB c0d433aa download
model-00014-of-00062.safetensors 5.00 GB cd2a9e78 download
model-00021-of-00062.safetensors 5.00 GB 6edb54c1 download
model-00042-of-00062.safetensors 5.00 GB 487a343e download
model-00049-of-00062.safetensors 5.00 GB 6560b786 download
model-00008-of-00062.safetensors 5.00 GB 82407024 download
model-00015-of-00062.safetensors 5.00 GB 54fc4bad download
model-00018-of-00062.safetensors 5.00 GB 17f4de1d download
model-00025-of-00062.safetensors 5.00 GB 2802695c download
model-00029-of-00062.safetensors 5.00 GB 6eb84cc9 download
model-00036-of-00062.safetensors 5.00 GB ed272b16 download
model-00050-of-00062.safetensors 5.00 GB 04a75cc3 download
model-00039-of-00062.safetensors 5.00 GB a6cec9b2 download
model-00043-of-00062.safetensors 5.00 GB ef3f74d3 download
model-00012-of-00062.safetensors 5.00 GB 0de52e01 download
model-00019-of-00062.safetensors 5.00 GB de9979f7 download
model-00026-of-00062.safetensors 5.00 GB 3eff620e download
model-00033-of-00062.safetensors 5.00 GB b371d227 download
model-00054-of-00062.safetensors 5.00 GB 12b4c365 download
model-00004-of-00062.safetensors 5.00 GB ea66d913 download
model-00031-of-00062.safetensors 5.00 GB 4be6e500 download
model-00056-of-00062.safetensors 5.00 GB 748dda91 download
model-00047-of-00062.safetensors 5.00 GB ddf0691c download
model-00060-of-00062.safetensors 5.00 GB 2cf64655 download
model-00057-of-00062.safetensors 5.00 GB e128c2da download
model-00009-of-00062.safetensors 5.00 GB 69cbb58b download
model-00017-of-00062.safetensors 5.00 GB f9bf1792 download
model-00028-of-00062.safetensors 5.00 GB 5a79d7d3 download
model-00035-of-00062.safetensors 5.00 GB 1a05a4ac download
model-00022-of-00062.safetensors 5.00 GB 8650c349 download
model-00053-of-00062.safetensors 5.00 GB b32ee193 download
model-00011-of-00062.safetensors 5.00 GB 984759cc download
model-00040-of-00062.safetensors 5.00 GB dc058a53 download
model-00005-of-00062.safetensors 5.00 GB d1da1e65 download
model-00046-of-00062.safetensors 5.00 GB 1def7d29 download
model-00032-of-00062.safetensors 5.00 GB f33b5385 download
model-00059-of-00062.safetensors 5.00 GB bc173127 download
model-00003-of-00062.safetensors 5.00 GB e0fc42c2 download
model-00010-of-00062.safetensors 4.99 GB bbf3a5ae download
model-00048-of-00062.safetensors 4.99 GB 85c51200 download
model-00027-of-00062.safetensors 4.99 GB fdafd6e7 download
model-00034-of-00062.safetensors 4.99 GB 451bb036 download
model-00041-of-00062.safetensors 4.99 GB 9ddfde38 download
model-00020-of-00062.safetensors 4.99 GB 67a86a8b download
model-00013-of-00062.safetensors 4.99 GB bfecdc0f download
model-00006-of-00062.safetensors 4.99 GB 1653c5e3 download
model-00051-of-00062.safetensors 4.99 GB b3423a93 download
model-00044-of-00062.safetensors 4.99 GB ed57a936 download
model-00030-of-00062.safetensors 4.99 GB a28adbc9 download
model-00037-of-00062.safetensors 4.99 GB 0e7c51bd download
model-00023-of-00062.safetensors 4.99 GB 1ea8f1bb download
model-00016-of-00062.safetensors 4.99 GB 6cc97434 download
model-00055-of-00062.safetensors 4.99 GB 8d64c91f download
model-00058-of-00062.safetensors 4.99 GB ac4e58d1 download
model-00002-of-00062.safetensors 4.96 GB e2ed9887 download
model-00061-of-00062.safetensors 4.94 GB 4c29e565 download
model-00062-of-00062.safetensors 1.17 GB d3087816 download
tokenizer.json 19.3 MB 19e77364 download
model.safetensors.index.json 8.02 MB 6e8b6ed4 download
config.json 67.8 KB f93128cf download
dealign_mascot.png 10.9 KB da3bf39a download
chat_template.jinja 10.4 KB fb94d40d download
dealign_logo.png 7.48 KB a5b3546b download
README.md 6.48 KB 76758eed download
CRACK_SURGERY.json 4.80 KB a6f7b470 download
.gitattributes 1.53 KB 52373fe2 download
LICENSE 1.04 KB 986b06fb download
processor_config.json 909 B 3ec2a058 download
tokenizer_config.json 761 B e375fa0a download
generation_config.json 194 B 637ee6af download
CRACK_MTP.json 146 B 1ab484f0 download

README current version from Hugging Face


license: mit
base_model:

  • zai-org/GLM-5.3-Flash
    language:
  • en
    tags:
  • abliterated
  • uncensored
  • crack
  • refusal-removed
  • glm
  • moe
  • fp8
  • vision
  • mtp
    thumbnail: dealign_mascot.png

GLM 5.3 CRACK Abliterated FP8

Abliterated · CRACK · guardrails removed at the weight level · native FP8 speed · vision + MTP working

a CRACK release by dealignai · Twitter @dealignai

Mirror of dealignai/GLM-5.3-Flash-UNCENSORED-FP8 — same weights, same CRACK.


What Is This?

CRACK is dealignai's brand for permanent, weight-level uncensoring. This is
GLM-5.3-Flash in FP8 with its refusal behavior —
which caused heavy over-refusal, especially on copyright and other benign-but-flagged requests —
removed directly in the model weights. FP8 runs at native speed on Hopper (H100/H200) GPUs.

Genuine weight modification — none of the usual shortcuts:

  • ❌ No fine-tuning / SFT / DPO. ❌ No cheap template / jailbreak-prompt tricks.
  • ❌ No LoRA, adapters, steering vectors, runtime hooks, or custom model.py.
  • ✅ A permanent edit baked into the tensors. Load with stock vLLM and it just works.

Specs

Architecture GLM-5.3-Flash (glm5_next) — hybrid MoE (KDA linear + DeepSeek-sparse attention)
Parameters 320B total · 18B active per token
Quantization FP8 (block-wise e4m3) — native tensor-core speed on Hopper
Context 1M tokens
Vision GLM-4.1V vision tower — working (ships the correct multimodal chat template)
MTP multi-token-prediction draft head — also CRACK'd, 75.9% acceptance

Speed (TP4, native FP8 on H200)

Decode 163 tok/s single-stream (211 tok/s with MTP speculative decoding)
Prefill ~19,400 tok/s
MTP acceptance 75.9% — and it does not collapse on the un-refused prompts (benign / harmful / copyright all ~208–219 tok/s)

Capability Is Preserved — MMLU-logit

Identical logit-mode scoring on base vs. this model, 1,026 questions:

Base FP8 CRACK Uncensored FP8 Δ
MMLU (overall) 86.74% 86.26% -0.48 pp

Guardrails Are Gone

HarmBench-320 (greedy):

Category Complied Rate
Standard 159/159 100.0%
Contextual 81/81 100.0%
Copyright 80/80 100.0%
Overall 320/320 100.0%

Robust under the recommended sampling params too (temperature 1.0, top_p 0.95): the 6 harshest
behaviors sampled 5× each → 30/30 complied, 0 refusals, 0 soft refusals, 0 garbage. The crack is
not a greedy-decoding artifact.

A Note on KL Divergence

For a refusal-ablation, KL divergence vs. the base model is not a meaningful quality metric.
The entire point is to change one behavior — refusal — end-to-end, so a distributional shift on
refusal-adjacent tokens is the intended result, not damage. Capability preservation (MMLU, above)
is what matters
, and it is essentially untouched (-0.48 pp).

MMLU by Topic (base → CRACK)

All 57 MMLU subjects
Subject Base CRACK
Abstract Algebra 66.7% 66.7%
Anatomy 83.3% 88.9%
Astronomy 94.4% 100.0%
Business Ethics 94.4% 94.4%
Clinical Knowledge 100.0% 94.4%
College Biology 100.0% 100.0%
College Chemistry 61.1% 66.7%
College Computer Science 83.3% 94.4%
College Mathematics 66.7% 61.1%
College Medicine 94.4% 94.4%
College Physics 77.8% 88.9%
Computer Security 83.3% 77.8%
Conceptual Physics 94.4% 94.4%
Econometrics 77.8% 72.2%
Electrical Engineering 77.8% 72.2%
Elementary Mathematics 94.4% 88.9%
Formal Logic 66.7% 66.7%
Global Facts 66.7% 77.8%
High School Biology 94.4% 94.4%
High School Chemistry 88.9% 94.4%
High School Computer Science 100.0% 100.0%
High School European History 77.8% 77.8%
High School Geography 88.9% 88.9%
High School Government And Politics 100.0% 100.0%
High School Macroeconomics 88.9% 88.9%
High School Mathematics 61.1% 44.4%
High School Microeconomics 83.3% 83.3%
High School Physics 88.9% 88.9%
High School Psychology 100.0% 100.0%
High School Statistics 94.4% 94.4%
High School Us History 88.9% 88.9%
High School World History 94.4% 94.4%
Human Aging 72.2% 77.8%
Human Sexuality 88.9% 88.9%
International Law 88.9% 94.4%
Jurisprudence 88.9% 83.3%
Logical Fallacies 88.9% 88.9%
Machine Learning 88.9% 88.9%
Management 100.0% 100.0%
Marketing 94.4% 88.9%
Medical Genetics 94.4% 100.0%
Miscellaneous 88.9% 88.9%
Moral Disputes 88.9% 88.9%
Moral Scenarios 83.3% 55.6%
Nutrition 100.0% 100.0%
Philosophy 94.4% 94.4%
Prehistory 94.4% 94.4%
Professional Accounting 88.9% 83.3%
Professional Law 83.3% 77.8%
Professional Medicine 94.4% 94.4%
Professional Psychology 100.0% 100.0%
Public Relations 72.2% 72.2%
Security Studies 83.3% 83.3%
Sociology 100.0% 100.0%
Us Foreign Policy 88.9% 88.9%
Virology 55.6% 55.6%
World Religions 88.9% 88.9%

Usage

vllm serve dealignai/GLM-5.3-Flash-ABLITERATED-FP8 \
  --tensor-parallel-size 4 \
  --tool-call-parser glm47 --reasoning-parser glm45 --enable-auto-tool-choice \
  --speculative-config '{"method":"mtp","num_speculative_tokens":1}'

Native FP8 on Hopper (no Marlin needed). OpenAI-compatible chat/completions, tools, reasoning,
vision (image_url), and MTP speculative decoding all work. (DeepGEMM JITs a block-FP8
kernel at startup — make sure nvcc is on PATH.)

Credits

Disclaimer

Safety guardrails have been removed; this model will comply with requests a stock model refuses.
Released for alignment and safety research. You are responsible for how you use it.

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-19docs: add measured video fps engine-kill + media parameter warningsced06e210.7 KB
    Loading...
  2. 2026-09-19docs: add reasoning_effort / max_tokens serving warning to top of card6ee346f9.3 KB
    Loading...
  3. 2026-08-29Update card: reasoning-mode transparency (off/max fully uncensored, greedy wo...e5a20947.6 KB
    Loading...
  4. 2026-08-282026-08-28: files fixed (repetition-loop resolved); MMLU improved to 87.33% (...40a2a447 KB
    Loading...
  5. 2026-08-26Release68fdc7b6.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration