← back to catalog · registered 2026-08-22 13:56

gbrennon/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive

gbrennon Qwen 122B GGUF MoE multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/gbrennon%2FQwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive"
Response includes
  • classification m-uncensored
  • files 6
  • benchmarks 11 entries
  • hub_downloads_all_time 744
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
744
55 last 30d - cooling
Likes
0
Model age
6mo ago
created 2026-03-22
Downloads over time
Now764→from362↑111%
342496650804362 on Mar 25764 on Oct 11764 on Oct 10MarAprMayJunJulAugSepOct
Mar 25 → Oct 11 · 68 snapshots · spans 200 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 1.4 UGI
Hazardous 1.8 UGI
Natural Intelligence 31.08 UGI
Political lean -21.9% UGI
Sensitive-Info 17.81 UGI
SocPol 2.3 UGI
UGI 17.71 UGI
Willingness (10) 1.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 2 UGI
Writing 39.54 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh multilingual
Quantizations
IQ2 IQ3 Q3_K
Tags
gguf uncensored qwen3.5 moe vision multimodal image-text-to-text en zh multilingual base_model:Qwen/Qwen3.5-122B-A10B base_model:quantized:Qwen/Qwen3.5-122B-A10B

Related

Total size
142 GB
Files
6
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2026-03-22 04:48

Files by quantization

Q3_K 1 file 54.6 GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf 54.6 GB 5be57e72 download
IQ3 1 file 50.1 GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf 50.1 GB ee381bb2 download
IQ2 1 file 37.2 GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf 37.2 GB 55a0d230 download
F16 1 file 867 MB
mmproj-Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-f16.gguf 867 MB 9d3c48d5 download
Auxiliary files 2 files 6.55 KB
README.md 4.59 KB fdee849c download
.gitattributes 1.96 KB e6ba250a download

README current version from Hugging Face


license: apache-2.0
tags:

  • uncensored
  • qwen3.5
  • moe
  • gguf
  • vision
  • multimodal
    language:
  • en
  • zh
  • multilingual
    pipeline_tag: image-text-to-text
    base_model: Qwen/Qwen3.5-122B-A10B

Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive

Qwen3.5-122B-A10B uncensored by HauhauCS. 0/465 refusals.

About

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.

These are meant to be the best lossless uncensored models out there.

Aggressive Variant

Stronger uncensoring — model is fully unlocked and won't refuse prompts. Disclaimers that were present in previous releases have been significantly reduced in this version.

For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available.

What are K_P quants?

K_P ("Perfect") quants are HauhauCS custom quantizations that use model-specific analysis to selectively preserve quality where it matters most. Each model gets its own optimized quantization profile.

A K_P quant effectively bumps quality up by 1-2 quant levels at only ~5-15% larger file size than the base quant. Fully compatible with llama.cpp, LM Studio, and any GGUF-compatible runtime — no special builds needed.

Downloads

File Quant Size
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf Q8_K_P XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf Q6_K_P XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf Q6_K XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q5_K_M.gguf Q5_K_M XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf Q4_K_P XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf Q4_K_M XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ4_XS.gguf IQ4_XS XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf Q3_K_P XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf Q3_K_M XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf IQ3_M XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ3_XXS.gguf IQ3_XXS XX GB
Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf IQ2_M XX GB
mmproj-Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-f16.gguf mmproj (f16) 867 MB

Note: K_P quants may show as "?" in LM Studio's quant column. This is a display issue only — the model loads and runs fine.

Specs

  • 122B total parameters, ~10B active per forward pass (MoE)
  • 256 experts, 8 routed + 1 shared per token
  • Hybrid architecture: Gated DeltaNet linear attention + full softmax attention (3:1 ratio)
  • 48 layers, pattern: 12 x (3 x DeltaNet-MoE + 1 x Attention-MoE)
  • 262K native context
  • Natively multimodal (text, image, video)
  • 248K vocabulary, 201 languages
  • Based on Qwen/Qwen3.5-122B-A10B

Recommended Settings

From the official Qwen authors:

Thinking mode (default):

  • General: temperature=1.0, top_p=0.95, top_k=20, min_p=0, presence_penalty=1.5
  • Coding/precise tasks: temperature=0.6, top_p=0.95, top_k=20, min_p=0, presence_penalty=0

Non-thinking mode:

  • General: temperature=0.7, top_p=0.8, top_k=20, min_p=0, presence_penalty=1.5
  • Reasoning tasks: temperature=1.0, top_p=1.0, top_k=40, min_p=0, presence_penalty=2.0

Important:

  • Use --jinja flag with llama.cpp for proper chat template handling
  • Thinking mode is on by default — to disable, use --chat-template-kwargs '{"enable_thinking":false}' or edit the jinja template
  • Vision support requires the mmproj file alongside the main GGUF

Usage

Works with llama.cpp, LM Studio, Jan, koboldcpp, and other GGUF-compatible runtimes.

# Text only
llama-cli -m Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --jinja -c 131072 -ngl 99

# With vision
llama-cli -m Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --mmproj mmproj-Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-f16.gguf \
  --jinja -c 131072 -ngl 99

Other Models

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-03-22Duplicate from HauhauCS/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressivee467f254.6 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration