← back to catalog · registered 2026-08-22 13:56

justfrfn/GLM-4.7-Flash-Uncensored-HauhauCS-Balanced

justfrfn Glm GGUF MoE 203K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/justfrfn%2FGLM-4.7-Flash-Uncensored-HauhauCS-Balanced"
Response includes
  • classification m-uncensored
  • files 6
  • hub_downloads_all_time 237
  • author_summary 22 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
237
172 last 30d - active
Likes
0
Model age
4mo ago
created 2026-06-08
Downloads over time
Now334→from51↑555%
3714525436251 on Jun 10334 on Oct 11334 on Oct 10JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Metadata

License
mit
Languages
en zh
Quantizations
F16 Q4_K Q6_K Q8_0
Tags
gguf uncensored glm4 moe en zh license:mit endpoints_compatible region:us conversational

Related

Total size
125 GB
Files
6
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2026-06-08 15:44

Files by quantization

F16 1 file 55.8 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-FP16.gguf 55.8 GB 064533e4 download
Q8_0 1 file 29.7 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q8_0.gguf 29.7 GB ffb16bb2 download
Q6_K 1 file 22.9 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q6_K.gguf 22.9 GB d397ce6d download
Q4_K 1 file 16.9 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf 16.9 GB 8c1ac6d1 download
Auxiliary files 2 files 3.44 KB
.gitattributes 1.83 KB bdd34fec download
README.md 1.61 KB 3316f23e download

README current version from Hugging Face


license: mit
tags:

  • uncensored
  • glm4
  • moe
    language:
  • en
  • zh

GLM-4.7-Flash-Uncensored-HauhauCS-Balanced

Join the Discord for updates, roadmaps, projects, or just to chat.

GLM-4.7 Flash uncensored by HauhauCS.

About

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.

These are meant to be the best lossless uncensored models out there.

Agentic Coding

If you're doing agentic coding, use the Balanced variants. Good balance between capability and not refusing everything.

Downloads

File Quant Size
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-FP16.gguf FP16 56 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q8_0.gguf Q8_0 30 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q6_K.gguf Q6_K 23 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf Q4_K_M 17 GB

Specs

Recommended Settings

From the official Z.ai authors:

General use:

  • --temp 1.0 --top-p 0.95

Tool-calling / agentic:

  • --temp 0.7 --top-p 1.0

Important:

  • Disable repeat penalty (or --repeat-penalty 1.0)
  • For llama.cpp: use --min-p 0.01 (default 0.05 is too high)
  • Use --jinja flag for llama.cpp

Note: Not recommended for Ollama due to chat template issues. Works well with llama.cpp, LM Studio, Jan.

Usage

Works with llama.cpp, LM Studio, Jan, koboldcpp, etc.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-08Duplicate from HauhauCS/GLM-4.7-Flash-Uncensored-HauhauCS-Balancedf9f67e61.6 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration