← back to catalog · registered 2026-08-22 13:56

zecanard/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-4bit-mixed_4_6

zecanard Qwen 39B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/zecanard%2FQwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-4bit-mixed_4_6"
Response includes
  • classification m3
  • files 16
  • hub_downloads_all_time 1,282
  • author_summary 95 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
1K
58 last 30d - cooling
Likes
1
Model age
5mo ago
created 2026-05-08

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now1.3K→from65↑1,898%
34769491.4K65 on May 61.3K on Oct 111.3K on Oct 9MayJunJulAugSepOct
May 6 → Oct 11 · 62 snapshots · spans 158 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh
Tags
mlx safetensors qwen3_5 decensored unsloth fine tune heretic uncensored abliterated multi-stage tuned. all use cases coder

Related

Total size
22.9 GB
Files
16
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-12 10:45

Files by quantization

Auxiliary files 16 files 23.0 GB
model-00004-of-00005.safetensors 5.00 GB f27668a3 download
model-00001-of-00005.safetensors 4.97 GB 45d06585 download
model-00003-of-00005.safetensors 4.95 GB 37f00ab5 download
model-00002-of-00005.safetensors 4.95 GB bdf27fa0 download
model-00005-of-00005.safetensors 3.06 GB d065b02b download
tokenizer.json 19.1 MB 06b95093 download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 306 KB 7cb1c526 download
config.json 187 KB 35bca84c download
chat_template.jinja 8.17 KB 65328ad8 download
README.md 2.19 KB 2ba11ced download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.14 KB 1d134cd2 download
processor_config.json 991 B 8f29fe38 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download

README current version from Hugging Face


language:

  • en
  • zh
    license: apache-2.0
    tags:
  • mlx
  • decensored
  • unsloth
  • fine tune
  • heretic
  • uncensored
  • abliterated
  • multi-stage tuned.
  • all use cases
  • coder
  • creative
  • creative writing
  • fiction writing
  • plot generation
  • sub-plot generation
  • fiction writing
  • story generation
  • scene continue
  • storytelling
  • fiction story
  • science fiction
  • romance
  • all genres
  • story
  • writing
  • vivid prosing
  • vivid writing
  • fiction
  • roleplaying
  • bfloat16
  • all use cases
    datasets:
  • TeichAI/claude-4.5-opus-high-reasoning-250x
  • DavidAU/PkDick-Deckard-5-Datasets
    base_model: DavidAU/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking
    library_name: mlx
    pipeline_tag: image-text-to-text

🦆 zecanard/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-4bit-mixed_4_6

This model was converted to MLX from DavidAU/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking using mlx-vlm version 0.6.3.
Please refer to the original model card for more details.

🌟 Quality

Mixed-precision quantized vision language model with an effective 4.983 bits per weight. Combines the size and speed benefits of a 4-bit quant with higher precision where it matters most.

mlx_vlm.convert --quantize --q-group-size 32 --quant-predicate mixed_4_6

🛠️ Customizations

This quant is aware of the current date, and also enables thinking (if available). You may disable this behavior by deleting the following line from the chat template, or changing true to false:

{%- set enable_thinking = true %}

🖥️ Use with mlx

pip install -U mlx-vlm
mlx_vlm.generate --model zecanard/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-4bit-mixed_4_6 --max-tokens 100 --temperature 0 --prompt "Describe this image." --image <path_to_image>

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-11Add files using upload-large-folder tool18e02022.2 KB
    Loading...
  2. 2026-05-08Add files using upload-large-folder tool21bbabb2.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration