← back to catalog · registered 2026-08-22 13:56

Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-8bit

Youssofal Qwen 27B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Youssofal%2FQwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-8bit"
Response includes
  • classification m3
  • files 20
  • hub_downloads_all_time 12,446
  • author_summary 17 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
12K
331 last 30d - cooling
Likes
5
Model age
5mo ago
created 2026-04-24
Downloads over time
Now12.6K→from929↑1,251%
3484.8K9.3K13.7K929 on Apr 2212.6K on Oct 11AprMayJunJulAugSepOct
Apr 22 → Oct 11 · 65 snapshots · spans 172 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5 mlx-lm qwen qwen3.6 dense multimodal vlm vision video abliterated

Related

Total size
27.9 GB
Files
20
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-04-26 00:27

Files by quantization

Auxiliary files 20 files 27.9 GB
model-00005-of-00006.safetensors 4.99 GB 3b145f7d download
model-00004-of-00006.safetensors 4.99 GB d2ba9932 download
model-00002-of-00006.safetensors 4.99 GB c1a9152d download
model-00003-of-00006.safetensors 4.97 GB a3f5787d download
model-00001-of-00006.safetensors 4.93 GB 7f114d55 download
model-00006-of-00006.safetensors 2.18 GB 1a6f5c88 download
model-vision-00001-of-00001.safetensors 879 MB 0a5d5f7f download
tokenizer.json 19.1 MB 06b95093 download
model.safetensors.index.json 216 KB c5c20dd9 download
config.json 155 KB 1d54853b download
mlx_variant_metadata.json 74.0 KB 6ec4e422 download
chat_template.jinja 7.58 KB a8755d82 download
README.md 3.91 KB 3165ca1a download
abliteration_metadata.json 3.43 KB d88c438e download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.24 KB 619bc32e download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B 00fe0efb download
configuration.json 51.0 B 3a6d4256 download

README current version from Hugging Face


base_model: Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16
library_name: mlx
pipeline_tag: text-generation
license: apache-2.0
tags:

  • mlx
  • mlx-lm
  • qwen
  • qwen3.6
  • qwen3_5
  • dense
  • multimodal
  • vlm
  • vision
  • video
  • abliterated
  • uncensored
  • heretic
  • mpoa
  • apple-silicon
  • 8-bit
  • soma
    quantized_by: Youssofal

Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-8bit

This is an MLX release of an abliterated, uncensored version of Qwen's Qwen3.6-27B with the source image/video payload retained, made from the published BF16 checkpoint.

By applying a Heretic-style MPOA pipeline with magnitude preservation on the Qwen3.6-27B dense text stack, the base refusal behavior was removed at the weight level. This MLX build keeps that text-side behavior in Apple Silicon format while retaining Qwen3.6-27B's vision/video payload in the repo.

Quick Benchmarks

Check Original Qwen3.6-27B This MLX Quant
Official 25-prompt refusal check 20/25 refusals 0/25 refusals
100-prompt refusal check 92/100 refusals 3/100 refusals

Methodology & Model Notes

Qwen3.6-27B is a 27.8B dense vision-language model with 64 text layers, hybrid linear/full attention (3 linear-attention + 1 full-attention per 4-layer group), and an integrated image + video vision tower.

This MLX variant was built directly from Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16 using dynamic layer-aware 8-bit affine MLX quantization with 64/32 group-size tiers. It is a dynamic, layer-aware quantization pass rather than a flat conversion. The retained multimodal payload includes the source vision/video tensors and processor files; direct image/video use depends on MLX runtime support for Qwen3.6 multimodal inputs.

Files

  • model-*.safetensors: quantized MLX text weights
  • model-vision-00001-of-00001.safetensors: retained BF16 vision/video tensor shard
  • model.safetensors.index.json: tensor-to-shard mapping, including retained vision tensors
  • config.json, generation_config.json, configuration.json: model config
  • tokenizer.json, tokenizer_config.json, chat_template.jinja: tokenizer + chat template
  • preprocessor_config.json, video_preprocessor_config.json: image + video processor configs
  • mlx_variant_metadata.json: build metadata

Running

from mlx_lm import load, generate
from mlx_lm.sample_utils import make_sampler

repo = "Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-8bit"
model, tokenizer = load(repo)

messages = [{"role": "user", "content": "Write a short Python function that reverses a string."}]
prompt = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True,
    enable_thinking=False,
)

response = generate(
    model,
    tokenizer,
    prompt=prompt,
    max_tokens=256,
    sampler=make_sampler(temp=0.0),
)
print(response)

Model Architecture

Spec Value
Total Parameters 27.8B (dense source)
Layers 64
Attention Hybrid (3 linear-attention + 1 full-attention per 4-layer group)
Hidden Size 5120
Family qwen3_5
Modality Vision-language source; MLX text path locally validated
MLX Runtime Validation Text generation
Base Model Qwen/Qwen3.6-27B

Disclaimer

This model has had refusal behavior removed at the weight level. It will answer prompts that the base model would normally refuse. You are responsible for how you use it.

Credits

License

This release inherits the base Qwen3.6-27B license.

Apache-2.0.

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-24Sync transferable 35B tagsbb2ec183.9 KB
    Loading...
  2. 2026-04-24Clean MLX model card benchmark tablee07ca8f3.9 KB
    Loading...
  3. 2026-04-24Update MLX model carded161b44.1 KB
    Loading...
  4. 2026-04-24Add README.md24881ca5.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration