language:
- en
- zh
license: apache-2.0
tags: - mlx
- unsloth
- fine tune
- heretic
- uncensored
- abliterated
- multi-stage tuned.
- all use cases
- coder
- creative
- creative writing
- fiction writing
- plot generation
- sub-plot generation
- fiction writing
- story generation
- scene continue
- storytelling
- fiction story
- science fiction
- romance
- all genres
- story
- writing
- vivid prosing
- vivid writing
- fiction
- roleplaying
- bfloat16
- all use cases
datasets: - TeichAI/claude-4.5-opus-high-reasoning-250x
- DavidAU/PkDick-Deckard-5-Datasets
base_model: DavidAU/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking
library_name: mlx - DavidAU/Qwen3.5-27B-Deckard-PKD-Heretic-Uncensored-Thinking
pipeline_tag: image-text-to-text
🦆 zecanard/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-2bit-affine
This model was converted to MLX from DavidAU/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking using mlx-vlm version 0.4.4.
Please refer to the original model card for more details.
🌟 Quality
Quantized vision language model with 3.153 bits per weight.
mlx_vlm.convert --quantize --q-bits 2 --q-group-size 32 --q-mode affine
🛠️ Customizations
This quant is aware of the current date, and also enables thinking (if available). You may disable this behavior by deleting the following line from the chat template:
{%- set enable_thinking = true %}
🖥️ Use with mlx
pip install -U mlx-vlm
mlx_vlm.generate --model zecanard/Qwen3.5-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-MLX-2bit-affine --max-tokens 100 --temperature 0 --prompt "Describe this image." --image <path_to_image>