← back to catalog · registered 2026-08-22 13:56

gregfrank/GLM-4.5-Air-ULRE-abliterated

gregfrank Glm 107B MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/gregfrank%2FGLM-4.5-Air-ULRE-abliterated"
Response includes
  • classification m1
  • files 20
  • benchmarks 16 entries
  • hub_downloads_all_time 1,429
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
1K
61 last 30d - cooling
Likes
0
Model age
4mo ago
created 2026-06-07
Downloads over time
Now1.5K→from231↑530%
1706391.1K1.6K231 on Jun 101.5K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Arena-Battles 10200 LM-Arena
LM Arena Elo 1389.858268635106 LM-Arena
Arena-Elo-Lower 1383.6290541907408 LM-Arena
Arena-Elo-Upper 1396.0874830794712 LM-Arena
Arena-Rank 20 LM-Arena
Entertainment 3.5 UGI
Hazardous 4.1 UGI
Natural Intelligence 33.3 UGI
Political lean -23.5% UGI
Sensitive-Info 37.19 UGI
SocPol 3.7 UGI
UGI 30.62 UGI
Willingness (10) 1.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 2 UGI
Writing 41.96 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Languages
en
Tags
mlx safetensors glm4_moe abliterated uncensored ulre glm moe text-generation conversational en base_model:zai-org/GLM-4.5-Air

Related

Total size
56.0 GB
Files
20
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-07 13:11

Files by quantization

Auxiliary files 20 files 56.0 GB
model-00001-of-00012.safetensors 4.97 GB ea945dcd download
model-00007-of-00012.safetensors 4.91 GB ecf836aa download
model-00006-of-00012.safetensors 4.91 GB b3f524da download
model-00005-of-00012.safetensors 4.91 GB 58ae77db download
model-00004-of-00012.safetensors 4.91 GB ce731cde download
model-00011-of-00012.safetensors 4.91 GB d4803618 download
model-00010-of-00012.safetensors 4.91 GB 0cde63c5 download
model-00003-of-00012.safetensors 4.91 GB 2d8dce29 download
model-00009-of-00012.safetensors 4.91 GB 3df8692f download
model-00008-of-00012.safetensors 4.91 GB 503fed77 download
model-00002-of-00012.safetensors 4.91 GB f9f2ce84 download
model-00012-of-00012.safetensors 1.95 GB 232781ce download
tokenizer.json 19.0 MB bda8e214 download
model.safetensors.index.json 155 KB 9024a31e download
README.md 3.17 KB 215463a6 download
chat_template.jinja 3.17 KB 41478957 download
.gitattributes 1.53 KB 52373fe2 download
config.json 1.25 KB 5ce05d01 download
tokenizer_config.json 364 B 0ba781cf download
generation_config.json 155 B 4d49113f download

README current version from Hugging Face


base_model: zai-org/GLM-4.5-Air
license: mit
library_name: mlx
pipeline_tag: text-generation
tags:

  • mlx
  • abliterated
  • uncensored
  • ulre
  • glm
  • moe
    language:
  • en

GLM-4.5-Air-ULRE (abliterated, MLX 4-bit)

An abliterated (refusal-reduced) build of GLM-4.5-Air (Zhipu/Z.ai; 106B-A12B MoE), 4-bit MLX,
produced with ULRE — a per-layer residual-stream steering edit baked into the attention output
projection. Quantized base from
lmstudio-community/GLM-4.5-Air-MLX-4bit.

GLM-4.5-Air is a strong agentic / tool-calling model. ULRE de-refuses it very cleanly while
preserving and even improving capability.

Results

De-refusal judged by an independent local LLM judge (gpt-oss-120b-heretic) on 100 held adversarial
prompts; 0=refuse … 1=clean compliance … 4-5=strong-steer. Validated in both modes:

mode clean compliance refuse strong-steer mean
non-thinking (@512) 96 / 100 0 1 1.09
thinking (@1536) 95 / 100 1 3 1.10

(base GLM-4.5-Air refuses ~23/24 on the same screen.) The edit de-refuses cleanly whether or not the
model is reasoning.

Capability gates (thinking mode, same harness, base vs this model):

gate base this model Δ
math (GSM8K) 0.83 0.83 0pp
code (HumanEval) 0.65 0.825 +17.5pp (over-refusal recovery)

Best de-refusal in the ULRE series (vs Mistral-Large 78, Qwen3-32B 68), with capability intact/up.

Method (ULRE)

ULRE subtracts alpha * u_l (the layer-l harmful−harmless activation mean-difference direction)
from the output of a band of decoder layers (here o_proj on layers 16–26, alpha = 6), baked
statically as an o_proj bias. The alpha is tuned to the lowest value that saturates de-refusal.

⚠️ Loading — needs a one-line glm4_moe loader patch (or run via mlx_lm.server)

mlx-lm's glm4_moe.py hardcodes o_proj to have no bias, so it must be told to build one
(backwards-compatible; base models default to False):

# class ModelArgs:  add field
    o_proj_bias: bool = False
# class Attention.__init__:  replace the o_proj line
    self.o_proj = nn.Linear(n_heads * head_dim, dim, bias=getattr(args, "o_proj_bias", False))

The model's config.json sets "o_proj_bias": true. Then load via the patched mlx_lm:

# serve on an OpenAI-compatible endpoint (works with patched mlx-lm)
mlx_lm.server --model gregfrank/GLM-4.5-Air-ULRE-abliterated --port 8080

Point any MCP-capable client (Open WebUI, LibreChat, or LM Studio as an MCP host pointing at the
endpoint) at http://127.0.0.1:8080/v1. (LM Studio's bundled MLX engine does not carry this patch,
so it won't load the file directly — use mlx_lm.server.)

Notes & caveats

  • De-refusal validated in both thinking and non-thinking modes (95–96/100 clean), and capability
    gates were run in thinking mode (math/code preserved or improved). Verified serving via mlx_lm.server.
  • Research artifact for studying refusal mechanisms / safety-tuning robustness. Use responsibly under
    the base model's MIT license and applicable law.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-07Upload README.md with huggingface_hub2ce8b893.2 KB
    Loading...
  2. 2026-06-07Add files using upload-large-folder tool244144c3.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration