← back to catalog · registered 2026-08-22 13:56

Jerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF

Jerome0207 Glm GGUF 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Jerome0207%2FHuihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF"
Response includes
  • classification m8
  • files 11
  • benchmarks 11 entries
  • hub_downloads_all_time 192
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
192
67 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-07-29
Downloads over time
Now222→from81↑174%
7412818223681 on Jul 29222 on Oct 11JulAugSepOct
Jul 29 → Oct 11 · 51 snapshots · spans 74 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 5 UGI
Hazardous 1.8 UGI
Natural Intelligence 54.88 UGI
Political lean -12.7% UGI
Sensitive-Info 41.67 UGI
SocPol 5.1 UGI
UGI 29.44 UGI
Willingness (10) 0.5 UGI
W10-Adherence 0 UGI
W10-Direct 1 UGI
Writing 61.21 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Languages
en zh
Quantizations
Q2_K
Tags
gguf glm glm-5.2 abliterated quantized llama.cpp text-generation en zh base_model:zai-org/GLM-5.2 base_model:quantized:zai-org/GLM-5.2 license:mit

Related

Total size
235 GB
Files
11
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-07-29 07:14

Files by quantization

Q2_K 7 files 235 GB
GLM-5.2-UD-Q2_K_MXFP4-00006-of-00007.gguf 46.3 GB bd39693f download
GLM-5.2-UD-Q2_K_MXFP4-00002-of-00007.gguf 45.7 GB 95701330 download
GLM-5.2-UD-Q2_K_MXFP4-00003-of-00007.gguf 45.4 GB c98084fe download
GLM-5.2-UD-Q2_K_MXFP4-00004-of-00007.gguf 45.4 GB a51d859c download
GLM-5.2-UD-Q2_K_MXFP4-00005-of-00007.gguf 45.4 GB d8b5e3c6 download
GLM-5.2-UD-Q2_K_MXFP4-00007-of-00007.gguf 6.90 GB 3b62fa28 download
GLM-5.2-UD-Q2_K_MXFP4-00001-of-00007.gguf 8.99 MB 3ca70bbb download
Auxiliary files 4 files 4.93 KB
.gitattributes 2.02 KB a4bfea3b download
README.md 1.84 KB 9d42c347 download
checksums.sha256 756 B 62166482 download
NOTICE 347 B 9d1e6c9e download

README current version from Hugging Face


license: mit
language:

  • en
  • zh
    pipeline_tag: text-generation
    base_model:
  • zai-org/GLM-5.2
    tags:
  • gguf
  • glm
  • glm-5.2
  • abliterated
  • quantized
  • llama.cpp

Huihui GLM-5.2 Abliterated — UD-Q2_K_MXFP4 GGUF

This is a single-quant operational mirror for RunPod cached models. It contains
only the seven UD-Q2_K_MXFP4 GGUF shards from the pinned huihui-ai source.
It does not claim authorship of the model or its quantization.

Destination repository: Jerome0207/Huihui-GLM-5.2-abliterated-UD-Q2_K_MXFP4-GGUF

Provenance

  1. zai-org/GLM-5.2
  2. huihui-ai abliterated GGUF

Source revision:
994200a058539c553f0977002e58a2f05de845c7.

The seven GGUF shards total 252,600,253,568 bytes. See
checksums.sha256 for the source SHA-256 object identifiers.

Runtime

The repository is intended for a current CUDA build of
llama.cpp. Start with the first shard;
llama.cpp discovers the remaining split files automatically:

GLM-5.2-UD-Q2_K_MXFP4-00001-of-00007.gguf

The validated deployment target uses three 96 GB GPUs, 131,072 context tokens,
one parallel slot and Q8_0 K/V caches.

License and limitations

The source model card declares the MIT license. See NOTICE for attribution
and exact provenance.

This is an “abliterated” derivative with intentionally reduced refusal
behavior. It is not a safety-tuned public service. Deploy it only behind
authentication and rate limits, and run coding/security agents in isolated
sandboxes without production credentials.

GLM-5.2 support in llama.cpp is evolving. Full DSA, Lightning Indexer,
IndexShare and MTP behavior must be validated before treating this deployment
as equivalent to the official inference implementation.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-29Add GLM-5.2 Q2 metadata and provenanceac3d8211.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration