← back to catalog · registered 2026-08-22 13:56

QuantFactory/glm-4-9b-chat-abliterated-GGUF

QuantFactory Glm 9B GGUF 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/QuantFactory%2Fglm-4-9b-chat-abliterated-GGUF"
Response includes
  • classification m8
  • files 16
  • hub_downloads_all_time 8,000
  • author_summary 48 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

No other method signals detected in this model.
Confidence
HIGH
Why this label 3 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=quantfactory (M8 quantization producer)
  • is_gguf=1
  • base_model='THUDM/glm-4-9b-chat' (source unknown method)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
8K
987 last 30d - stable
Likes
5
Model age
2.1y ago
created 2024-08-31
Downloads over time
Now8.4K→from377↑2,122%
03.1K6.1K9.2K377 on Aug 28, 20248.4K on Oct 11Aug '24Dec '24Apr '25Aug '25Dec '25AprAug
Aug 28, 2024 → Oct 11 · 150 snapshots · spans 774 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
zh en
Quantizations
Q2_K Q3_K Q4 Q4_K Q5 Q5_K Q6_K Q8_0
Tags
transformers gguf glm chatglm thudm chat abliterated text-generation zh en base_model:zai-org/glm-4-9b-chat base_model:quantized:zai-org/glm-4-9b-chat

Related

Total size
82.1 GB
Files
16
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2024-08-31 08:11

Files by quantization

Q8_0 1 file 9.31 GB
glm-4-9b-chat-abliterated.Q8_0.gguf 9.31 GB c374d8bb download
Q6_K 1 file 7.69 GB
glm-4-9b-chat-abliterated.Q6_K.gguf 7.69 GB 5f3f8749 download
Q5_K 2 files 12.9 GB
glm-4-9b-chat-abliterated.Q5_K_M.gguf 6.65 GB a4dff678 download
glm-4-9b-chat-abliterated.Q5_K_S.gguf 6.23 GB 3fc0ad86 download
Q5 2 files 12.7 GB
glm-4-9b-chat-abliterated.Q5_1.gguf 6.61 GB b2a89a1c download
glm-4-9b-chat-abliterated.Q5_0.gguf 6.10 GB d8b134c7 download
Q4_K 2 files 11.2 GB
glm-4-9b-chat-abliterated.Q4_K_M.gguf 5.82 GB a1c1dd1a download
glm-4-9b-chat-abliterated.Q4_K_S.gguf 5.36 GB 53965549 download
Q4 2 files 10.7 GB
glm-4-9b-chat-abliterated.Q4_1.gguf 5.59 GB 8f2074b1 download
glm-4-9b-chat-abliterated.Q4_0.gguf 5.08 GB c5f465a3 download
Q3_K 3 files 13.9 GB
glm-4-9b-chat-abliterated.Q3_K_L.gguf 4.92 GB ea112b0d download
glm-4-9b-chat-abliterated.Q3_K_M.gguf 4.72 GB c4d087e3 download
glm-4-9b-chat-abliterated.Q3_K_S.gguf 4.27 GB 767ff436 download
Q2_K 1 file 3.72 GB
glm-4-9b-chat-abliterated.Q2_K.gguf 3.72 GB 300ae680 download
Auxiliary files 2 files 3.77 KB
.gitattributes 2.48 KB e67160b1 download
README.md 1.28 KB 076aa560 download

README current version from Hugging Face


base_model: THUDM/glm-4-9b-chat
pipeline_tag: text-generation
license: other
license_name: glm-4
license_link: https://huggingface.co/THUDM/glm-4-9b-chat/blob/main/LICENSE
language:

  • zh
  • en
    tags:
  • glm
  • chatglm
  • thudm
  • chat
  • abliterated
    library_name: transformers

QuantFactory/glm-4-9b-chat-abliterated-GGUF

This is quantized version of byroneverson/glm-4-9b-chat-abliterated created using llama.cpp

Original Model Card

GLM 4 9B Chat - Abliterated

Check out the jupyter notebook for details of how this model was abliterated from glm-4-9b-chat.

The python package "tiktoken" is required to quantize the model into gguf format. So I had to create a fork of GGUF My Repo (+tiktoken).

Logo

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-08-31Upload README.md with huggingface_hubade834f1.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration