← back to catalog · registered 2026-08-22 13:56

zecanard/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-6bit-int6-affine

zecanard Qwen 39B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/zecanard%2FQwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-6bit-int6-affine"
Response includes
  • classification m3
  • files 20
  • hub_downloads_all_time 1,108
  • author_summary 95 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
1K
38 last 30d - cooling
Likes
1
Model age
4mo ago
created 2026-06-11

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now1.1K→from106↑955%
554438311.2K106 on Jun 101.1K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh
Tags
mlx safetensors qwen3_5 decensored unsloth fine tune heretic uncensored abliterated multi-stage tuned. all use cases coder

Related

Total size
32.7 GB
Files
20
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-12 08:41

Files by quantization

Auxiliary files 20 files 32.7 GB
model-00003-of-00007.safetensors 5.00 GB f306bf60 download
model-00001-of-00007.safetensors 4.99 GB a1613696 download
model-00004-of-00007.safetensors 4.99 GB 2ec2902b download
model-00005-of-00007.safetensors 4.99 GB e9002046 download
model-00002-of-00007.safetensors 4.97 GB 59ed0c6f download
model-00006-of-00007.safetensors 4.96 GB 1c820c80 download
model-00007-of-00007.safetensors 2.80 GB 19c4fc2c download
tokenizer.json 19.1 MB 06b95093 download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 306 KB b75a7a68 download
chat_template.jinja 7.97 KB 56865aba download
config.json 5.92 KB 8e0616a1 download
README.md 2.15 KB 579e2191 download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.14 KB 1d134cd2 download
processor_config.json 991 B 8f29fe38 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 202 B 023756cf download
configuration.json 51.0 B 3a6d4256 download

README current version from Hugging Face


base_model: DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking
library_name: mlx
language:

  • en
  • zh
    license: apache-2.0
    tags:
  • mlx
  • decensored
  • unsloth
  • fine tune
  • heretic
  • uncensored
  • abliterated
  • multi-stage tuned.
  • all use cases
  • coder
  • creative
  • creative writing
  • fiction writing
  • plot generation
  • sub-plot generation
  • fiction writing
  • story generation
  • scene continue
  • storytelling
  • fiction story
  • science fiction
  • romance
  • all genres
  • story
  • writing
  • vivid prosing
  • vivid writing
  • fiction
  • roleplaying
  • bfloat16
  • all use cases
    datasets:
  • TeichAI/claude-4.5-opus-high-reasoning-250x
  • DavidAU/PkDick-Deckard-5-Datasets
    pipeline_tag: image-text-to-text

🦆 zecanard/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-6bit-int6-affine

This model was converted to MLX from DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking using mlx-vlm version 0.6.3.
Please refer to the original model card for more details.

🌟 Quality

Quantized vision language model with an effective 7.106 bits per weight.

mlx_vlm.convert --quantize --q-group-size 32 --q-bits 6 --q-mode affine

🛠️ Customizations

This quant is aware of the current date, and also enables thinking (if available). You may disable this behavior by deleting the following line from the chat template, or changing true to false:

{%- set enable_thinking = true %}

A fix is also included for a thinking-related performance issue in Qwen 3.6.

🖥️ Use with mlx

pip install -U mlx-vlm
mlx_vlm.generate --model zecanard/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-6bit-int6-affine --max-tokens 100 --temperature 0 --prompt "Describe this image." --image <path_to_image>

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-11Add files using upload-large-folder toold14ca122.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration