← back to catalog · registered 2026-08-22 13:56

AIOpsInSpace/GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive

AIOpsInSpace Glm GGUF second-order 203K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/AIOpsInSpace%2FGLM-4.7-Flash-Uncensored-HauhauCS-Aggressive"
Response includes
  • classification m-uncensored
  • files 7
  • hub_downloads_all_time 3,304
  • author_summary 12 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
3K
787 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-07-15
Downloads over time
Now3.6K→from996↑262%
8661.9K2.9K3.9K996 on Jul 153.6K on Oct 9JulAugSepOct
Jul 15 → Oct 9 · 51 snapshots · spans 86 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Quantizations
IQ2 IQ3 IQ4 Q4_K
Tags
gguf text-generation uncensored flash-attention glm en base_model:HauhauCS/GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive base_model:quantized:HauhauCS/GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive license:apache-2.0 endpoints_compatible region:us conversational

Related

Total size
70.6 GB
Files
7
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2026-07-22 07:18

Files by quantization

Q4_K 1 file 16.9 GB
GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf 16.9 GB 7124f9e3 download
IQ4 2 files 31.4 GB
GLM-4.7-Flash-Uncensored-Heretic-IQ4_NL.gguf 16.1 GB 4b8a5b75 download
GLM-4.7-Flash-Uncensored-Heretic-IQ4_XS.gguf 15.3 GB e6d95968 download
IQ3 1 file 12.7 GB
GLM-4.7-Flash-Uncensored-Heretic-IQ3_M.gguf 12.7 GB 658df96f download
IQ2 1 file 9.61 GB
GLM-4.7-Flash-Uncensored-Heretic-IQ2_M.gguf 9.61 GB fec3983a download
Auxiliary files 2 files 20.2 KB
README.md 17.9 KB 1ef6fbb4 download
.gitattributes 2.31 KB 20f7dccd download

README current version from Hugging Face


base_model:

  • HauhauCS/GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive
  • THUDM/GLM-4.7-Flash
    tags:
  • text-generation
  • gguf
  • uncensored
  • flash-attention
  • glm
    license: apache-2.0
    language:
  • en
    pipeline_tag: text-generation

"This is humanity's race.
The solution is open source.
Stay sovereign."

— AIOpsInSpace

GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive

AIOpsInSpace Official

Blazing fast GLM-4.7-Flash ablated variant fine-tuned for high-speed conversational response and logic.

⚡ MoE Flash Model ⚡ Flash Attention Optimized 🛠️ Aggressively Uncensored

> What is this model and Why is it Needed?

GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive is built on THUDM GLM-4.7-Flash architecture with surgical safety ablation.

Why it is needed: Delivers flash-attention speed and strong multi-lingual reasoning without restrictive system prompts or refusal behavior.

> From the Parent Repository

"Extreme speed meets raw unaligned reasoning in the GLM architecture."

— GLM Open Source Project


🏗️ 2. Model Architecture & Merging

Architecture: GLM-4.7-Flash Transformer Architecture
Merging Technique: Safety Filter Ablation & GGUF Conversion
Constituent Models: Methodology: Safety vectors stripped and GGUF layout patched for maximum throughput.

🚀 3. Technical Enhancements

> Key Upgrades Over Base Model:

  • Flash Throughput: Optimized for high tokens-per-second generation.
  • Uncensored Logic: Full access to internal reasoning state without censorship filters.

📊 4. Benchmark Competitiveness vs. Frontier Scores

> Evaluated Performance
Benchmark GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive Frontier Target
MMLU Evaluated 88.7%
GSM8K Evaluated 95.6%
HumanEval Evaluated 90.2%

🏆 5. Comprehensive Arena Analytics

> Status: Active Community Benchmarking

// Note: Arena Elo and head-to-head winrates updated continuously as evaluation telemetry processes.

🔍 6. SWOT Analysis

> Strengths (S)

  • 🛡️ Uncensored Fidelity: Surgically patched to ensure maximum generation throughput without alignment overhead.
  • ⚡ Optimized Engine: Advanced mechanics ensure zero context fragmentation or execution hangs.

> Weaknesses (W)

  • 📉 Hardware Limits: Requires sufficient VRAM/RAM for higher precision GGUF quantizations.

> Opportunities (O)

  • 🎯 Local Sovereign Agents: Perfect for offline, private reasoning and agentic workflows.

> Threats (T)

  • ⚠️ Sampler Sensitivity: High temperatures may require repetition penalty adjustments.

⚡ 7. Usage & Deployment Info

> Recommended Settings

  • Temperature: 0.2 - 0.7
  • Top-P: 0.95
  • Backend Engines: Compatible with llama.cpp, vLLM, Ollama, LM Studio, KoboldCPP

⚙️ 8. Backend Compatibility

> Validated Engines:

  • [+] llama.cpp: Native support across all quantizations.
  • [+] Ollama / LM Studio: Full GGUF compatibility.

📜 9. Disclaimers & Credits

Disclaimer: GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive is provided for research and sovereign local deployment. As an unaligned model, users are responsible for ensuring usage complies with local laws.

Credits: Gratitude to original base model authors (HauhauCS/GLM-4.7-Flash-Uncensored-HauhauCS-Aggressive) and open-source AI community tools.

README history 8 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-22Fix lineage, tags, and model card metadata9d104f517.7 KB
    Loading...
  2. 2026-07-18Overhaul model card with pentesting UIf06d1ee22.5 KB
    Loading...
  3. 2026-07-17Upload README.md with huggingface_hube989f781.3 KB
    Loading...
  4. 2026-07-17Upload README.md with huggingface_hub115846c252 B
    Loading...
  5. 2026-07-17Upload README.md with huggingface_hub7e703ea3.8 KB
    Loading...
  6. 2026-07-15Upload README.md with huggingface_hub98604173.8 KB
    Loading...
  7. 2026-07-15Update model card with all 5 quants61f1f475.3 KB
    Loading...
  8. 2026-07-15Add model card939f3f74.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration