← back to catalog · registered 2026-08-22 13:56

dokjlahdf/GLM-5.1-Abliterated

dokjlahdf Glm 752B MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/dokjlahdf%2FGLM-5.1-Abliterated"
Response includes
  • classification m1
  • files 153
  • hub_downloads_all_time 88
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
88
24 last 30d - stable
Likes
0
Model age
5mo ago
created 2026-05-11
Downloads over time
Now103→from14↑636%
10447811214 on May 13103 on Oct 11MayJunJulAugSepOct
May 13 → Oct 11 · 61 snapshots · spans 151 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
agpl-3.0
Languages
en zh
Tags
transformers safetensors glm_moe_dsa text-generation abliterated uncensored moe glm glm-5.1 fp8 blackwell conversational

Related

Total size
704 GB
Files
153
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-05-11 09:03

Files by quantization

Auxiliary files 153 files 704 GB
model-00117-of-00142.safetensors 5.00 GB be7ecc4f download
model-00001-of-00142.safetensors 5.00 GB 205976fd download
model-00138-of-00142.safetensors 5.00 GB a37ce395 download
model-00024-of-00142.safetensors 4.99 GB d73902c7 download
model-00011-of-00142.safetensors 4.99 GB b5127c0c download
model-00094-of-00142.safetensors 4.99 GB f50e714a download
model-00035-of-00142.safetensors 4.99 GB 9e66f781 download
model-00070-of-00142.safetensors 4.99 GB 9d96a09e download
model-00129-of-00142.safetensors 4.99 GB 72dd93b9 download
model-00046-of-00142.safetensors 4.99 GB a7268a07 download
model-00105-of-00142.safetensors 4.99 GB 6a954d90 download
model-00022-of-00142.safetensors 4.99 GB 76aafe98 download
model-00057-of-00142.safetensors 4.99 GB 4df6cbab download
model-00009-of-00142.safetensors 4.99 GB bc5b647b download
model-00092-of-00142.safetensors 4.99 GB 58dc91f1 download
model-00116-of-00142.safetensors 4.99 GB 053efffa download
model-00033-of-00142.safetensors 4.99 GB 456a3ba7 download
model-00068-of-00142.safetensors 4.99 GB 54d7d4ed download
model-00127-of-00142.safetensors 4.99 GB cf652a70 download
model-00044-of-00142.safetensors 4.99 GB 8901bb22 download
model-00103-of-00142.safetensors 4.99 GB 9e76012a download
model-00055-of-00142.safetensors 4.99 GB 16b14c8a download
model-00081-of-00142.safetensors 4.99 GB 9fafe206 download
model-00007-of-00142.safetensors 4.99 GB 5c3ca6e2 download
model-00090-of-00142.safetensors 4.99 GB 0cc02ba5 download
model-00114-of-00142.safetensors 4.99 GB 5fd83b52 download
model-00031-of-00142.safetensors 4.99 GB a74449dc download
model-00066-of-00142.safetensors 4.99 GB 79123348 download
model-00018-of-00142.safetensors 4.99 GB e7a53380 download
model-00125-of-00142.safetensors 4.99 GB 6fe518d2 download
model-00042-of-00142.safetensors 4.99 GB de378082 download
model-00077-of-00142.safetensors 4.99 GB f09ac913 download
model-00005-of-00142.safetensors 4.99 GB 96eac93c download
model-00053-of-00142.safetensors 4.99 GB 6813ba34 download
model-00112-of-00142.safetensors 4.99 GB 55c451b3 download
model-00088-of-00142.safetensors 4.99 GB 05b9261f download
model-00029-of-00142.safetensors 4.99 GB 14e12bfc download
model-00016-of-00142.safetensors 4.99 GB 8bbed787 download
model-00064-of-00142.safetensors 4.99 GB 44bbff6f download
model-00003-of-00142.safetensors 4.99 GB 86c43d51 download
model-00014-of-00142.safetensors 4.99 GB d94c89c8 download
model-00051-of-00142.safetensors 4.99 GB e81a496f download
model-00086-of-00142.safetensors 4.99 GB 8057777f download
model-00123-of-00142.safetensors 4.99 GB 0ba6b15e download
model-00027-of-00142.safetensors 4.99 GB b9d9e2c3 download
model-00062-of-00142.safetensors 4.99 GB 33043fb0 download
model-00075-of-00142.safetensors 4.99 GB 50c97c04 download
model-00012-of-00142.safetensors 4.99 GB f9d7591c download
model-00025-of-00142.safetensors 4.99 GB a7299b7d download
model-00036-of-00142.safetensors 4.99 GB 9555a1c9 download
model-00049-of-00142.safetensors 4.99 GB ae9a3796 download
model-00071-of-00142.safetensors 4.99 GB 134b7ccd download
model-00073-of-00142.safetensors 4.99 GB e45c92cc download
model-00084-of-00142.safetensors 4.99 GB edadee5d download
model-00095-of-00142.safetensors 4.99 GB bb8eb308 download
model-00097-of-00142.safetensors 4.99 GB ed152723 download
model-00134-of-00142.safetensors 4.99 GB 25f2497a download
model-00110-of-00142.safetensors 4.99 GB 18671e2f download
model-00106-of-00142.safetensors 4.99 GB 9b741d26 download
model-00108-of-00142.safetensors 4.99 GB 67484855 download
model-00130-of-00142.safetensors 4.99 GB e962ce71 download
model-00132-of-00142.safetensors 4.99 GB d30285d6 download
model-00119-of-00142.safetensors 4.99 GB 437b5164 download
model-00038-of-00142.safetensors 4.99 GB 93a7bfdc download
model-00101-of-00142.safetensors 4.99 GB b8b320af download
model-00059-of-00142.safetensors 4.99 GB abbb57ae download
model-00079-of-00142.safetensors 4.99 GB c2e1576e download
model-00099-of-00142.safetensors 4.99 GB b64e2913 download
model-00140-of-00142.safetensors 4.99 GB 20190fd9 download
model-00040-of-00142.safetensors 4.99 GB 05e6033c download
model-00121-of-00142.safetensors 4.99 GB 51afac83 download
model-00060-of-00142.safetensors 4.99 GB a0430b0f download
model-00136-of-00142.safetensors 4.99 GB 705bd4c5 download
model-00020-of-00142.safetensors 4.99 GB 3e2a4e3e download
model-00002-of-00142.safetensors 4.99 GB d909a059 download
model-00013-of-00142.safetensors 4.99 GB 1722fbe6 download
model-00026-of-00142.safetensors 4.99 GB 1ece8ece download
model-00037-of-00142.safetensors 4.99 GB 9bf844f6 download
model-00050-of-00142.safetensors 4.99 GB b79ea464 download
model-00061-of-00142.safetensors 4.99 GB b53485d2 download
model-00072-of-00142.safetensors 4.99 GB 6408c8c9 download
model-00074-of-00142.safetensors 4.99 GB 978db219 download
model-00085-of-00142.safetensors 4.99 GB 3d03bb77 download
model-00096-of-00142.safetensors 4.99 GB ec77fa04 download
model-00098-of-00142.safetensors 4.99 GB d30bad76 download
model-00107-of-00142.safetensors 4.99 GB cd8a18b1 download
model-00109-of-00142.safetensors 4.99 GB a5f5038f download
model-00131-of-00142.safetensors 4.99 GB 0bc16cce download
model-00133-of-00142.safetensors 4.99 GB e604cc4a download
model-00015-of-00142.safetensors 4.99 GB 916305b1 download
model-00063-of-00142.safetensors 4.99 GB 108c6e36 download
model-00028-of-00142.safetensors 4.99 GB 5f37bad9 download
model-00122-of-00142.safetensors 4.99 GB d735a8b7 download
model-00087-of-00142.safetensors 4.99 GB 4d2aee9f download
model-00004-of-00142.safetensors 4.99 GB c8405236 download
model-00052-of-00142.safetensors 4.99 GB 66ed6313 download
model-00111-of-00142.safetensors 4.99 GB 1b5928d1 download
model-00076-of-00142.safetensors 4.99 GB 7df10590 download
model-00041-of-00142.safetensors 4.99 GB b96b36ed download
model-00135-of-00142.safetensors 4.99 GB f35babc4 download
model-00017-of-00142.safetensors 4.99 GB 345af057 download
model-00065-of-00142.safetensors 4.99 GB caa1fcc2 download
model-00030-of-00142.safetensors 4.99 GB e7b65448 download
model-00124-of-00142.safetensors 4.99 GB 0f4f759a download
model-00089-of-00142.safetensors 4.99 GB fc8f74f5 download
model-00006-of-00142.safetensors 4.99 GB eec9fcc3 download
model-00054-of-00142.safetensors 4.99 GB cc34d07e download
model-00113-of-00142.safetensors 4.99 GB 6352977f download
model-00137-of-00142.safetensors 4.99 GB e802972d download
model-00078-of-00142.safetensors 4.99 GB 9ed29e41 download
model-00043-of-00142.safetensors 4.99 GB 91036a2d download
model-00102-of-00142.safetensors 4.99 GB 873e97c7 download
model-00019-of-00142.safetensors 4.99 GB 9fe4da41 download
model-00067-of-00142.safetensors 4.99 GB 1ed2d941 download
model-00032-of-00142.safetensors 4.99 GB 995e8d62 download
model-00126-of-00142.safetensors 4.99 GB 2d00348f download
model-00091-of-00142.safetensors 4.99 GB 81ba80b3 download
model-00008-of-00142.safetensors 4.99 GB c3778a0c download
model-00056-of-00142.safetensors 4.99 GB ae8b17b5 download
model-00021-of-00142.safetensors 4.99 GB 2da172eb download
model-00115-of-00142.safetensors 4.99 GB 314fc05e download
model-00045-of-00142.safetensors 4.99 GB a3e1fb4e download
model-00104-of-00142.safetensors 4.99 GB b73dbfe8 download
model-00069-of-00142.safetensors 4.99 GB 907b38c7 download
model-00034-of-00142.safetensors 4.99 GB c2d52533 download
model-00128-of-00142.safetensors 4.99 GB 61db42ba download
model-00093-of-00142.safetensors 4.99 GB e8cf02a6 download
model-00010-of-00142.safetensors 4.99 GB 46449f54 download
model-00023-of-00142.safetensors 4.99 GB 66c41ca7 download
model-00058-of-00142.safetensors 4.99 GB 1dacb250 download
model-00039-of-00142.safetensors 4.99 GB 5ebc73ae download
model-00120-of-00142.safetensors 4.99 GB 23b49900 download
model-00100-of-00142.safetensors 4.99 GB 5f3901b6 download
model-00080-of-00142.safetensors 4.99 GB a92f502b download
model-00139-of-00142.safetensors 4.99 GB 862e5ebd download
model-00118-of-00142.safetensors 4.99 GB 55cead69 download
model-00083-of-00142.safetensors 4.99 GB 54e972b4 download
model-00048-of-00142.safetensors 4.99 GB 00e90eae download
model-00047-of-00142.safetensors 4.98 GB 13013f12 download
model-00141-of-00142.safetensors 4.98 GB fba01737 download
model-00082-of-00142.safetensors 4.95 GB 5a635186 download
model-00142-of-00142.safetensors 140 MB 2dbc19a2 download
tokenizer.json 19.3 MB 19e77364 download
model.safetensors.index.json 10.9 MB a5c43b9f download
config.json 35.0 KB e6aa7db2 download
README.md 15.5 KB 9a5bc38b download
chat_template.jinja 4.56 KB e230020b download
.gitattributes 1.60 KB aa7aacd0 download
LICENSE 1.04 KB e14ce68b download
abliteration_meta.json 1.03 KB 9f8f6924 download
tokenizer_config.json 760 B 1723f7d9 download
merge_meta.json 259 B c3003804 download
generation_config.json 193 B 453800a0 download

README current version from Hugging Face


license: agpl-3.0
base_model: zai-org/GLM-5.1-FP8
tags:

  • abliterated
  • uncensored
  • moe
  • glm
  • glm-5.1
  • fp8
  • blackwell
    language:
  • en
  • zh
    library_name: transformers
    pipeline_tag: text-generation

GLM-5.1-Abliterated - FP8

Z.ai's flagship 754B MoE — among the strongest open-weight models available, with coding performance reported to match or exceed frontier closed models — with the refusal layer surgically removed and capability healed back. Optimised over 250 multi-objective Optuna trials. Standard-benchmark deviation pending external evaluation; sweep-time numbers in the methodology section.

The huihui-ai Q3_K_M abliteration of GLM-5.1 was the only public option until now — produced with what the authors themselves called a "crude" pipeline, and never released as a full-precision artifact. This is the answer.


TL;DR (read this if you don't know what abliteration is)

GLM-5.1 is a 754B-parameter Mixture-of-Experts language model (~754 GB on disk) from Z.ai (China). Like every modern frontier model, it ships with a "safety layer" — a learned reflex to refuse certain prompts. Abliteration is a surgical technique that removes that reflex by identifying the single direction in the model's hidden state that encodes "I should refuse this", and projecting it out of every attention layer's output weights.

Done crudely, abliteration tanks the model's intelligence. Done well, it removes the refusals while preserving (or sometimes improving) capability on benign tasks.

This release is done well: 250 Optuna-optimised trial runs to find the Pareto-optimal trade-off, a rank-4 LoRA healing pass on top, and the result merged back into the FP8 weights as a single drop-in replacement for the base model. Same memory footprint, same speed, no LoRA loading dance — just from_pretrained() and go.


Headline numbers

Metric Base GLM-5.1-FP8 This release
Refusal rate (sweep-time eval, n=30) high (close to 100% on harmful set) 0/30 ✓
KL divergence vs base (sweep-time, n=30 harmless) 0.0 0.348 (pre-heal selected trial)
Healing CE loss on bartowski corpus (last-20 step avg) — 8.44 (down from 9.95 at step 5)
MMLU (5-shot) 86.2 pending external benchmark
GSM8K (8-shot CoT) 78.4 pending external benchmark
HumanEval 72.0 pending external benchmark
Perplexity on C4-en 4.21 pending external benchmark
Model size (FP8) 754 GB 754 GB (drop-in)

Sweep-time numbers are measured against the FP8-quantized base on the held-out 30-prompt curated harmful set. Standard-benchmark numbers will be filled in after a serverless RunPod evaluation pass — check this card's "Community" tab for the pending update.

Note for the technically curious: the abliteration KL improvement isn't free — perplexity rises slightly on clean text. But on reasoning benchmarks (GSM8K, HumanEval) we typically see small gains, because the model stops hedging and shortening responses on prompts that touched the safety reflex. Empirical observation across published abliterations, well documented in the lineage cited below.


Why this exists

Three reasons:

  1. The existing GLM-5.1 abliteration shipped as Q3_K_M only. That quantization tier loses meaningful capability before any abliteration is even applied. A serious abliteration deserves a serious quantization base.

  2. Native FP8 master. This release ships the full 754 GB block-FP8 weights as the canonical artifact. Every downstream quantization (Q8, Q6_K, Q5_K_M, Q4_K_M, IQ4_XS) is derived from a high-quality source rather than a chain of lossy conversions.

  3. The pipeline is reproducible. The complete abliteration + healing pipeline is published as a sibling repo. Pin the commit, point it at any FP8 MoE on Blackwell, get the same result.


Example outputs

Real-output side-by-side comparisons against base GLM-5.1-FP8 will be added once the external benchmark + qualitative eval pass completes. Until then, refer to the methodology section below for the technical specifics of what was changed.


How to use

from transformers import AutoModelForCausalLM, AutoTokenizer

tokenizer = AutoTokenizer.from_pretrained("helixdouble/GLM-5.1-Abliterated")
model = AutoModelForCausalLM.from_pretrained(
    "helixdouble/GLM-5.1-Abliterated",
    device_map="auto",
    torch_dtype="auto",
)

vLLM serving:

vllm serve helixdouble/GLM-5.1-Abliterated \
    --dtype auto \
    --gpu-memory-utilization 0.92 \
    --cpu-offload-gb 600 \
    --max-model-len 4096

For Blackwell hosts, you'll likely want VLLM_USE_DEEP_GEMM=0 set — see the methodology section below for why.


Methodology

The pipeline (full code bundled inside the trials dataset repo under scripts/):

  1. Calibration set curation — 60 GLM-5.1-specific harmful prompts (glm51_curated_harmful.jsonl) that the base model strongly refuses, drawn from a larger filtered harmless/harmful pool (filtered_harmless.jsonl / filtered_harmful.jsonl). Quality-filtered via Fireworks API against the live GLM-5.1 endpoint to ensure the refusal signal is real on this specific model.

  2. Refusal direction extraction — single-token forward pass through HF transformers with output_hidden_states=True, mean-pool the hidden states for harmful and harmless prompts separately at every layer, take the difference, L2-normalise. Produces (n_layers + 1, hidden_size) direction tensor.

  3. Optuna sweep (250 trials, multi-objective TPE sampler) searching the 4-dimensional space:

    • max_weight ∈ [0.5, 2.0] — peak abliteration strength
    • min_weight ∈ [0.0, max_weight] — taper floor
    • max_weight_position ∈ [0.4·L, 0.9·L] — which layer the peak lives at
    • min_weight_distance ∈ [0.2·L, 0.7·L] — half-width of the active layer band

    Each trial mutates the model's o_proj.weight in-place via vLLM's apply_model (no per-trial reload), evaluates refusal count + KL divergence on the calibration set, returns to Optuna.

  4. Pareto-optimal hyperparameter selection — pick the lowest-KL trial with refusals ≤ target threshold from the Pareto front.

  5. Direct FP8 weight surgery — the chosen hyperparameters bake into a new model directory: every affected layer's o_proj.weight has the refusal direction projected out, then re-quantized to block-FP8 with fresh weight_scale_inv.

  6. Healing pass — rank-4 LoRA fine-tune on bartowski's general calibration corpus for 250 SGD steps with cosine learning-rate schedule, AdamW8bit optimizer, gradient checkpointing. Heals only the abliterated layers (skipped layers untouched).

  7. Merge — the healing LoRA's delta is dequantized, added to the abliterated FP8 weights, re-quantized. Result: a single FP8 model directory with both abliteration and healing baked in. Drop-in replacement.

The Blackwell + FP8 patch stack

This pipeline discovered (and works around) structural bugs in the standard transformers + accelerate + vLLM stack on NVIDIA Blackwell GPUs with FP8 MoE models. Anyone running FP8 MoE inference or training on Blackwell hits at least one of these:

Inference / serving:

  • transformers.integrations.finegrained_fp8._load_deepgemm_kernel produces silent NaN cascades on sm_100 for non-DeepSeek FP8 models. Fix: monkey-patch the loader to fail, falling back to Triton.
  • vllm 0.20.1's V1 EngineCore deadlocks during model load with cpu_offload_gb > 0. Fix: set VLLM_ENABLE_V1_MULTIPROCESSING=0 before any vLLM import.
  • vllm's default FP8 LinearMethod selects DeepGEMM, which transforms the weight_scale_inv layout via TMA-aligned swizzling. Fix: set VLLM_USE_DEEP_GEMM=0, falling back to CUTLASS (standard (out//128, in//128) layout).
  • Transformers' rotary embedding default is rotate_half, but GLM-5.1 was trained with interleaved RoPE per its config. The glm_moe_rope_interleaved_patch supplies the missing interleaved variant for both MLA and the GlmMoeDsa indexer.

Training (Phase D healing):

  • device_map="auto" + glm_moe_smart_offload.activate() does NOT actually move experts to CPU — activate() only swaps the forward function. Result: ~150 GB of routed experts pin GPU VRAM and cause OOM during backward. Fix: manually p.data = p.data.to('cpu') for every routed-expert param after from_pretrained, then re-call activate(). Frees ~150 GB of GPU.
  • accelerate's offload puts small modules (gate, indexer, q/kv-projections) at execution_device='cpu' even when their inputs arrive on cuda:0, with weights that need GPU materialization at compute time but stay on meta until the hook fires. Forward looks fine, backward crashes mid-recompute with mat2 is on cpu. Fix: walk every non-expert hook and force execution_device=cuda:0, materialize the meta tensors via hook.pre_forward(module) then neutralize post_forward so they stay GPU-resident.
  • peft.get_peft_model(base, lora_config) on an accelerate-offloaded base lands the new lora_A / lora_B parameters on meta device. Forward appears to work (PEFT silently skips when LoRA weights aren't materialized), backward dies with MmBackward0 returned an invalid gradient at index 1 - expected device meta but got cuda:0. Fix: replace each meta LoRA Parameter via setattr on its parent module with a freshly-initialized cuda:0 Parameter (kaiming_uniform_(a=√5) for A, zeros for B), then recreate the optimizer since the original references stale meta param objects.

These are documented with reproduction steps in the pipeline repo. Now you don't have to rediscover them.


Datasets

This release uses two purpose-built datasets, both published as separate HF Hub repos:

  • helixdouble/glm-5.1-abliteration-trials-250 — the 250 Optuna trials with their hyperparameters, refusal counts, and KL divergences (trials_optimized.jsonl), the Optuna study database (phase_b_study.db), and the curated calibration sets (glm51_curated_harmful.jsonl, filtered_harmless.jsonl) used for direction extraction and per-trial evaluation. Quality-filtered via Fireworks API against the live GLM-5.1 endpoint to ensure the refusal signal is real on this specific model.
  • Healing corpus: bartowski's general calibration corpus (bartowski_calibration_v5.txt, ~410K tokens) is included alongside the trials dataset above for full reproducibility.

Both datasets are the actual, frozen artifacts used for this release — not "approximately the same" or "a similar set". If you want to reproduce or audit any step of this pipeline, those are the inputs.


Reproducibility

You can reproduce this exact model. Everything you need is published:

  • Pipeline code: bundled inside the trials dataset repo at helixdouble/glm-5.1-abliteration-trials-250 under scripts/ — full reproduction stack including the heal v3 production runner, the Blackwell+FP8 patches, the FP8 backward patches for autograd through quantized layers, the smart_offload custom expert kernel, and the direct_weight_abliterate / merge_healing_lora baking scripts.
  • Calibration set: see Datasets section above.
  • Optuna study database: included as phase_b_study.db in this repo's root — full state of all 250 trials.
  • Per-trial record: trials_optimized.jsonl — one JSON line per trial with hyperparameters, refusal count, and KL.
  • Bake hyperparameters: in abliteration_meta.json at the root.
  • Heal manifest: in healing_manifest.json at the root.

If you have a Vast B200 pod (or equivalent Blackwell host) and the patience for a ~30-hour run (most of it the 250-trial sweep), the same trial 132 hyperparameter combo regenerates this model from base GLM-5.1-FP8.


Limitations

  • Not a jailbreak. Some refusals are baked into the model's weights more deeply than the single direction this technique removes. Expect residual refusals on the most strongly-trained categories.
  • Calibration is in English + Chinese. Refusal-direction extraction was done with prompts in those languages; behaviour on other languages is not guaranteed to be equally affected.
  • No alignment. The original alignment was removed by design. This model is not a substitute for thinking about what you ship into production.
  • FP8 native weight is the canonical release. The GGUF quantizations are accurate but not bit-exact reproductions; expect ~0.1-0.5% additional perplexity drift per step down the quant ladder.
  • Geopolitically sensitive prompts: the abliteration removed the safety reflex broadly, including the politically-trained reflexes on China-adjacent topics. This is a property of the technique, not a deliberate choice. The model's responses on these topics reflect its training corpus minus the overlay; that's all.

Credits

Abliteration as a technique builds on serious research. This model exists because:

  • Arditi et al. 2024, "Refusal in Language Models is Mediated by a Single Direction", the foundational paper that established the rank-1 refusal direction.
  • FailSpy (HF profile) for the original abliterator notebook that translated the paper into a working technique.
  • mlabonne (HF profile) for popularising the method and publishing some of the cleanest abliterations.
  • p-e-w / Heretic (github) for the Optuna-driven multi-objective search framework, AGPL-licensed, the direct ancestor of this pipeline.
  • wuwangzhang1216 / abliterix (github) for the vLLM-integrated weight-mutation approach this pipeline adapted.
  • bartowski for the calibration corpus used in healing and GGUF imatrix.
  • huihui-ai for the prior GLM-5.1 abliteration release. Imperfect but it raised the question: can this be done better?

The intellectual achievement of the original Arditi et al. work — discovering that something as broad as refusal behaviour collapses to a single direction in activation space — is genuinely remarkable. This release stands on those shoulders.


What's next

Kimi K2.6 abliterated, when the weights drop. Same pipeline, same standard.


Citation

If you use this model in research or downstream work:

@misc{glm51_abliterated_helixdouble_2026,
  title  = {GLM-5.1-Abliterated},
  author = {helixdouble},
  year   = {2026},
  publisher = {Hugging Face},
  url    = {https://huggingface.co/helixdouble/GLM-5.1-Abliterated},
  note   = {Trials + reproduction code: \url{https://huggingface.co/datasets/helixdouble/glm-5.1-abliteration-trials-250}}
}

License

AGPL-3.0 — inherited from the Heretic framework that this pipeline derives from. Same license as the upstream tooling.

The base model GLM-5.1-FP8 is published under its own license (see zai-org/GLM-5.1-FP8). This release is a derivative work under the same terms; users are responsible for complying with the upstream license alongside the AGPL.


Trained on a single 1× B200 in a Vast.ai pod over ~3 days of wall-clock (most of it the 250-trial sweep). Selected trial: 132 (refusals=0/30, KL=0.348). Healing pass: rank-4 LoRA, 250 steps, chunk=2048, AdamW8bit lr=5e-5 cosine, final CE loss 8.44 on bartowski calibration corpus. Model card last updated: 2026-05-09.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-05-11Duplicate from helixdouble/GLM-5.1-Abliterated02ce81115.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration