license: apache-2.0
base_model: huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated
base_model_relation: quantized
library_name: mlx
pipeline_tag: image-text-to-text
tags:
- mlx
- mtplx
- mtp
- qwen3_6
- 4-bit
- abliterated
Huihui ThinkingCap Qwen3.6 27B Abliterated — MLX 4-bit MTP
MTPLX-compatible conversion of huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated, pinned to revision 44f63da.
Runtime-specific artifact: do not load this repository in LM Studio. LM Studio flattens the MTPLX sidecar into its target directory, causing the target loader to reject the 15
mtp.*tensors. For the recommended MTP experience, use the oMLX Native-MTP model; oMLX is the faster, more mature integrated path for this model.
- Trunk: MLX affine 4-bit, group size 64.
- MTP head: 15-tensor BF16 sidecar at
mtp/weights.safetensors. - Runtime: MTPLX 2.1.0, Apple Silicon only.
pip install -U mtplx
mtplx run --model pixelkaiser/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-MTP --depth 3 "Hello"
MTPLX Forge classified this artifact as verified-native. A bounded 32-token M4 Max verification measured 24.94 tok/s AR and 46.02 tok/s at MTP depth 3 (1.85x); treat this as a packaging smoke, not a general benchmark.
This is an abliterated model with reduced safety behavior. Review the upstream model card before use.