← back to catalog · registered 2026-08-22 13:56

Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-4bit

Youssofal Qwen 27B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Youssofal%2FQwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-4bit"
Response includes
  • classification m3
  • files 18
  • hub_downloads_all_time 12,146
  • author_summary 17 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
12K
774 last 30d - cooling
Likes
6
Model age
5mo ago
created 2026-04-24
Downloads over time
Now12.5K→from1.2K↑945%
6295K9.3K13.6K1.2K on Apr 2212.5K on Oct 11AprMayJunJulAugSepOct
Apr 22 → Oct 11 · 65 snapshots · spans 172 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5 mlx-lm qwen qwen3.6 dense multimodal vlm vision video abliterated

Related

Total size
17.4 GB
Files
18
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-04-26 00:24

Files by quantization

Auxiliary files 18 files 17.4 GB
model-00001-of-00004.safetensors 4.99 GB 50a2e07a download
model-00003-of-00004.safetensors 4.98 GB b69fe630 download
model-00002-of-00004.safetensors 4.97 GB 043dee36 download
model-00004-of-00004.safetensors 1.62 GB 74c78ddc download
model-vision-00001-of-00001.safetensors 879 MB 0a5d5f7f download
tokenizer.json 19.1 MB 06b95093 download
model.safetensors.index.json 216 KB c542a291 download
config.json 155 KB 6a021ce1 download
mlx_variant_metadata.json 74.0 KB 74e09dfd download
chat_template.jinja 7.58 KB a8755d82 download
README.md 3.89 KB ac52a0f3 download
abliteration_metadata.json 3.43 KB d88c438e download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.24 KB 619bc32e download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B 00fe0efb download
configuration.json 51.0 B 3a6d4256 download

README current version from Hugging Face


base_model: Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16
library_name: mlx
pipeline_tag: text-generation
license: apache-2.0
tags:

  • mlx
  • mlx-lm
  • qwen
  • qwen3.6
  • qwen3_5
  • dense
  • multimodal
  • vlm
  • vision
  • video
  • abliterated
  • uncensored
  • heretic
  • mpoa
  • apple-silicon
  • 4-bit
  • soma
    quantized_by: Youssofal

Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-4bit

This is an MLX release of an abliterated, uncensored version of Qwen's Qwen3.6-27B with the source image/video payload retained, made from the published BF16 checkpoint.

By applying a Heretic-style MPOA pipeline with magnitude preservation on the Qwen3.6-27B dense text stack, the base refusal behavior was removed at the weight level. This MLX build keeps that text-side behavior in Apple Silicon format while retaining Qwen3.6-27B's vision/video payload in the repo.

Quick Benchmarks

Check Original Qwen3.6-27B This MLX Quant
Official 25-prompt refusal check 20/25 refusals 0/25 refusals
100-prompt refusal check 92/100 refusals 3/100 refusals

Methodology & Model Notes

Qwen3.6-27B is a 27.8B dense vision-language model with 64 text layers, hybrid linear/full attention (3 linear-attention + 1 full-attention per 4-layer group), and an integrated image + video vision tower.

This MLX variant was built directly from Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-BF16 using dynamic layer-aware 4/5/6-bit affine MLX quantization. It is a dynamic, layer-aware quantization pass rather than a flat conversion. The retained multimodal payload includes the source vision/video tensors and processor files; direct image/video use depends on MLX runtime support for Qwen3.6 multimodal inputs.

Files

  • model-*.safetensors: quantized MLX text weights
  • model-vision-00001-of-00001.safetensors: retained BF16 vision/video tensor shard
  • model.safetensors.index.json: tensor-to-shard mapping, including retained vision tensors
  • config.json, generation_config.json, configuration.json: model config
  • tokenizer.json, tokenizer_config.json, chat_template.jinja: tokenizer + chat template
  • preprocessor_config.json, video_preprocessor_config.json: image + video processor configs
  • mlx_variant_metadata.json: build metadata

Running

from mlx_lm import load, generate
from mlx_lm.sample_utils import make_sampler

repo = "Youssofal/Qwen3.6-27B-Abliterated-Heretic-Uncensored-MLX-4bit"
model, tokenizer = load(repo)

messages = [{"role": "user", "content": "Write a short Python function that reverses a string."}]
prompt = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True,
    enable_thinking=False,
)

response = generate(
    model,
    tokenizer,
    prompt=prompt,
    max_tokens=256,
    sampler=make_sampler(temp=0.0),
)
print(response)

Model Architecture

Spec Value
Total Parameters 27.8B (dense source)
Layers 64
Attention Hybrid (3 linear-attention + 1 full-attention per 4-layer group)
Hidden Size 5120
Family qwen3_5
Modality Vision-language source; MLX text path locally validated
MLX Runtime Validation Text generation
Base Model Qwen/Qwen3.6-27B

Disclaimer

This model has had refusal behavior removed at the weight level. It will answer prompts that the base model would normally refuse. You are responsible for how you use it.

Credits

License

This release inherits the base Qwen3.6-27B license.

Apache-2.0.

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-24Sync transferable 35B tags0ff3a5c3.9 KB
    Loading...
  2. 2026-04-24Clean MLX model card benchmark tableba560a83.9 KB
    Loading...
  3. 2026-04-24Update MLX model card923e0934.1 KB
    Loading...
  4. 2026-04-24Add README.mdef9ba2f5.1 KB
    Loading...

Discussions 1 thread

  1. 2026-07-03Model fails to load in LM Studio MLX backend with shape mismatch error in visio…open2 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration