← back to catalog · registered 2026-08-22 13:56

huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp

huginnfork Qwen 28B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/huginnfork%2FQwen3.6-27B-uncensored-heretic-v2-mtp"
Response includes
  • classification m3
  • files 22
  • benchmarks 11 entries
  • hub_downloads_all_time 380
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
380
26 last 30d - cooling
Likes
4
Descendants
2
in 2 direct forks
Model age
5mo ago
created 2026-04-26
Downloads over time
Now391→from126↑210%
113214316418126 on Apr 29391 on Oct 11391 on Oct 9AprMayJunJulAugSepOct
Apr 29 → Oct 11 · 63 snapshots · spans 165 days

Benchmarks

Benchmark Score Source
Entertainment 1.5 UGI
Hazardous 2.9 UGI
Natural Intelligence 29.17 UGI
Political lean -24.5% UGI
Sensitive-Info 17.24 UGI
SocPol 1 UGI
UGI 43.16 UGI
Willingness (10) 9.5 UGI
W10-Adherence 10 UGI
W10-Direct 9 UGI
Writing 39.42 UGI

Genealogy 2 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 95 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
apache-2.0
Tags
safetensors qwen3_5 heretic abliteration multimodal mtp speculative-decoding bf16 image-text-to-text conversational base_model:Qwen/Qwen3.6-27B base_model:merge:Qwen/Qwen3.6-27B

Related

Total size
51.7 GB
Files
22
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-04-27 09:47

Files by quantization

Auxiliary files 22 files 51.8 GB
model-00007-of-00011.safetensors 4.99 GB de172e33 download
model-00009-of-00011.safetensors 4.99 GB 61766ffb download
model-00010-of-00011.safetensors 4.96 GB 723462be download
model-00004-of-00011.safetensors 4.96 GB 086cb5f8 download
model-00006-of-00011.safetensors 4.96 GB a7b3257d download
model-00005-of-00011.safetensors 4.96 GB 49aded77 download
model-00001-of-00011.safetensors 4.95 GB 94997885 download
model-00003-of-00011.safetensors 4.95 GB e2faf0f8 download
model-00008-of-00011.safetensors 4.95 GB a8f0a815 download
model-00002-of-00011.safetensors 4.93 GB 9e7923dd download
model-00011-of-00011.safetensors 2.18 GB fae69aa4 download
tokenizer.json 19.1 MB 6f32ce20 download
model.safetensors.index.json 110 KB f0c983fd download
chat_template.jinja 7.82 KB 09c96b90 download
config.json 3.74 KB 2ef24c62 download
README.md 2.88 KB de94c762 download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.13 KB c487bad4 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
kld_heretic.json 304 B 0727ddb2 download
generation_config.json 226 B 16d319af download

README current version from Hugging Face


license: apache-2.0
base_model:

  • llmfan46/Qwen3.6-27B-uncensored-heretic-v2
  • Qwen/Qwen3.6-27B
    pipeline_tag: image-text-to-text
    tags:
  • heretic
  • abliteration
  • multimodal
  • mtp
  • speculative-decoding
  • bf16
    base_model_relation: merge

huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp

bf16 build derived from the heretic-abliterated llmfan46/Qwen3.6-27B-uncensored-heretic-v2 of Qwen/Qwen3.6-27B, with the MTP head and vision tower preserved.

Provenance

  • Base: Qwen/Qwen3.6-27B (bf16)
  • Abliteration tool: heretic v1.2.0 (ARA) by Philipp Emanuel Weidmann (p-e-w)
  • Abliterated source weights: llmfan46/Qwen3.6-27B-uncensored-heretic-v2 — the heretic-derived bf16 abliteration of Qwen/Qwen3.6-27B
  • MTP head: re-grafted from Qwen/Qwen3.6-27B (15 tensors, ~810 MB bf16) so SGLang/vLLM speculative decoding (--speculative-algo NEXTN) works
  • Vision tower: model.visual.* preserved in bf16 (333 tensors)

KL divergence measurements

KLD computed with eval_kld.py — per-token KLD averaged over 8 samples from neuralmagic/calibration (LLM split), max_seq=1024. Max sample KLD is the highest single-sample mean (catches outliers that the overall mean hides).

Comparison Mean KLD (nats) Max sample KLD Samples max_seq
vs Qwen3.6-27B base 0.0425 0.1104 8 1024

Note: this pipeline always uploads the resulting checkpoint. Consult the KL
divergence numbers above to judge whether the result is acceptable for your
use case.

Perplexity (wikitext-2-raw)

Wikitext-2-raw test split, non-overlapping chunks of 2048 tokens, computed with eval_ppl.py. Same tokenizer for every row so the numbers compare apples-to-apples.

Model Perplexity Tokens scored Dataset seq
Qwen3.6-27B base (bf16) 7.3057 296907 wikitext/wikitext-2-raw-v1/test 2048
this checkpoint 7.4619 296907 wikitext/wikitext-2-raw-v1/test 2048

Inference

transformers (text + vision; MTP not exercised)

from transformers import AutoModelForImageTextToText, AutoProcessor
import torch

repo = "huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp"
proc = AutoProcessor.from_pretrained(repo, trust_remote_code=True)
model = AutoModelForImageTextToText.from_pretrained(
    repo, dtype=torch.bfloat16, device_map="auto", trust_remote_code=True,
)

vLLM (bf16 + MTP speculative decoding)

vllm serve huginnfork/Qwen3.6-27B-uncensored-heretic-v2-mtp \
    --trust-remote-code \
    --gpu-memory-utilization 0.85 \
    --max-model-len 8192 \
    --speculative-config '{"method":"qwen3_5_mtp","num_speculative_tokens":1}'

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-27Refresh KLD numbers + add wikitext-2-raw perplexity tablea1ca58c2.9 KB
    Loading...
  2. 2026-04-27Rename huginnfork/Qwen3.6-27B-heretic -> huginnfork/Qwen3.6-27B-uncensored-he...b1acf1a2.3 KB
    Loading...
  3. 2026-04-27Set base_model + base_model_relation metadata1c7a0c02.3 KB
    Loading...
  4. 2026-04-26Update READMEb0a19dd2.2 KB
    Loading...
  5. 2026-04-26Initial uploadb6043281.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration