language:
- en
- zh
license: apache-2.0
tags: - mlx
- 6-bit
- quantized
- uncensored
- heretic
- roleplaying
- creative-writing
- storytelling
- fiction
datasets: - TeichAI/claude-4.5-opus-high-reasoning-250x
- DavidAU/PkDick-Deckard-5-Datasets
pipeline_tag: text-generation
library_name: mlx
base_model: DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking
Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking — 6-bit MLX
This is a 6-bit MLX quantization of DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking, converted with Apple's mlx-lm for fast local inference on Apple Silicon.
Credit
All credit for the underlying model, fine-tuning, and merge work goes to DavidAU. This repository only contains a quantized MLX conversion. See the original model card for full details on training data, intended use, and capabilities.
Released under the same license as the base model: Apache 2.0.
Quantization details
- Method:
mlx_lm.convertwith-q --q-bits 6(affine quantization, group size 64) - Actual bits per weight after quantization: 6.501 (group/scale overhead)
- Source dtype: bfloat16
- Converted on an Apple M1 Max, 64 GB unified memory
Use with mlx-lm
pip install mlx-lm
mlx_lm.generate --model amit21186/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-6bit-MLX --prompt "Your prompt here"
Use with LM Studio
Search this repo name in LM Studio's model search, or download and place the folder under LM Studio's local models directory.