← back to catalog · registered 2026-08-22 13:56

marafx2025/Qwen3.6-27B-Abliterated-Heretic-Uncensored-GGUF

marafx2025 Qwen 27B GGUF multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/marafx2025%2FQwen3.6-27B-Abliterated-Heretic-Uncensored-GGUF"
Response includes
  • classification m3
  • files 14
  • benchmarks 11 entries
  • hub_downloads_all_time 31,004
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
31K
1K last 30d - cooling
Likes
0
Model age
5mo ago
created 2026-04-27
Downloads over time
Now31.2K→from2K↑1,438%
56911.7K22.9K34.1K2K on Apr 2931.2K on Oct 11AprMayJunJulAugSepOct
Apr 29 → Oct 11 · 63 snapshots · spans 165 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 1.2 UGI
Hazardous 4.7 UGI
Natural Intelligence 33.16 UGI
Political lean -20.0% UGI
Sensitive-Info 26.98 UGI
SocPol 2.9 UGI
UGI 27.15 UGI
Willingness (10) 2.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 4 UGI
Writing 42.47 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf qwen qwen3.6 qwen3_5 dense multimodal vlm vision video image-text-to-text abliterated uncensored

Related

Total size
153 GB
Files
14
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2026-04-27 17:25

Files by quantization

Q8_0 1 file 26.6 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q8_0.gguf 26.6 GB 4c277268 download
Q6_K 1 file 20.6 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q6_K.gguf 20.6 GB 9345d222 download
Q5_K 1 file 17.9 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q5_K_M.gguf 17.9 GB 71ef3dab download
Q4_K 1 file 15.4 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q4_K_M.gguf 15.4 GB 90c46f0e download
Q3_K 1 file 12.4 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q3_K_M.gguf 12.4 GB b4ab5c0b download
Q2_K 1 file 9.98 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q2_K.gguf 9.98 GB 7bd0b105 download
BF16 2 files 1.73 GB
mmproj-BF16.gguf 888 MB 05353347 download
mmproj-model-bf16.gguf 888 MB 05353347 download
F16 2 files 1.73 GB
mmproj-F16.gguf 885 MB eacf610d download
mmproj-model-f16.gguf 885 MB eacf610d download
Auxiliary files 4 files 50.1 GB
Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16-00001-of-00002.gguf 41.8 GB 693645fe download
Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16-00002-of-00002.gguf 8.29 GB ab693c10 download
README.md 4.18 KB 75992de3 download
.gitattributes 2.43 KB f5b7413c download

README current version from Hugging Face


base_model: Qwen/Qwen3.6-27B
library_name: gguf
pipeline_tag: text-generation
license: apache-2.0
tags:

  • gguf
  • qwen
  • qwen3.6
  • qwen3_5
  • dense
  • multimodal
  • vlm
  • vision
  • video
  • image-text-to-text
  • abliterated
  • uncensored
  • heretic
  • mpoa
  • llama-cpp
  • soma
    quantized_by: Youssofal

Qwen3.6-27B-Abliterated-Heretic-Uncensored-GGUF

This is a GGUF release of an abliterated, uncensored version of Qwen's Qwen3.6-27B, made with Heretic.

By applying a Heretic-style MPOA pipeline with magnitude preservation on the Qwen3.6-27B dense text stack, the base refusal behavior was removed at the weight level with extremely low distributional divergence (KL 0.0251 vs base on harmless prompts). The text GGUFs are paired with Qwen3.6-27B vision projectors (mmproj) so image/video-capable llama.cpp and LM Studio runtimes can use the multimodal path.

Quick Benchmarks

Check Original Qwen3.6-27B Abliterated Heretic Uncensored
Official 25-prompt refusal check 20/25 refusals 1/25 refusals
100-prompt refusal check 92/100 refusals 3/100 refusals
KL divergence N/A 0.0251

Methodology & Model Notes

Qwen3.6-27B is a 27.8B dense vision-language model with 64 text layers, hybrid linear/full attention (3 linear-attention + 1 full-attention per 4-layer group), and an integrated image + video vision tower.

This release was produced with a direct Heretic-style MPOA run with magnitude preservation — output-side orthogonalization on self_attn.o_proj, linear_attn.out_proj, and mlp.down_proj, with each weight row/column's L2 norm restored after projection. The ablation direction is interpolated at direction_index = 37.97.

The accepted candidate scored Refusals: 1/25 on the official 25-prompt marker suite used for the MiniMax M2.7 and Qwen3.6-35B-A3B abliterated runs, with a measured KL divergence of 0.0251 against the base on mlabonne/harmless_alpaca test[:25].

The resulting abliterated checkpoint was exported to BF16 and then converted to GGUF for llama.cpp-compatible deployment. The language GGUF files are text-model files; multimodal input is enabled by loading a matching mmproj projector alongside them.

Files

  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16-00001-of-00002.gguf + -00002-of-00002.gguf: BF16 GGUF source (split; use with --load-tensors or llama-gguf-split --merge)
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q8_0.gguf: highest-fidelity quant
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q6_K.gguf: near-lossless practical quant
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q5_K_M.gguf: high-fidelity medium quant
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q4_K_M.gguf: smaller general-use quant
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q3_K_M.gguf: compact quant
  • Qwen3.6-27B-Abliterated-Heretic-Uncensored-Q2_K.gguf: smallest-footprint quant
  • mmproj-F16.gguf / mmproj-model-f16.gguf: F16 Qwen3.6-27B vision projector for LM Studio and llama.cpp multimodal loading
  • mmproj-BF16.gguf / mmproj-model-bf16.gguf: BF16 Qwen3.6-27B vision projector

Running

llama-server \
  -m <quant-file.gguf> \
  -ngl 999 -c 32768 --jinja -fa

For image/video-capable runtimes, load the projector with the text GGUF:

llama-server \
  -m <quant-file.gguf> \
  --mmproj mmproj-model-f16.gguf \
  -ngl 999 -c 32768 --jinja -fa

Model Architecture

Spec Value
Total Parameters 27.8B (dense)
Layers 64
Attention Hybrid (3 linear-attention + 1 full-attention per 4-layer group)
Hidden Size 5120
Family qwen3_5
Modality Vision-language via GGUF text model + Qwen3.6-27B mmproj
Base Model Qwen/Qwen3.6-27B

Disclaimer

This model has had refusal behavior removed at the weight level. It will answer prompts that the base model would normally refuse. You are responsible for how you use it.

Credits

License

This release inherits the base Qwen3.6-27B license.

Apache-2.0.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-27Duplicate from Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-GGUF65ecd014.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration