← back to catalog · registered 2026-08-22 13:56

pixelkaiser/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-oMLX-MTP

pixelkaiser Qwen 27B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/pixelkaiser%2FHuihui-ThinkingCap-Qwen3.6-27B-abliterated-MLX-4bit-oMLX-MTP"
Response includes
  • classification m1
  • files 20
  • hub_downloads_all_time 3,192
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
3K
676 last 30d - stable
Likes
4
Model age
2mo ago
created 2026-07-17
Downloads over time
Now3.5K→from140↑2,379%
01.3K2.5K3.8K140 on Jul 153.5K on Oct 11JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
BF16
Tags
mlx safetensors qwen3_5 omlx qwen3.6 mtp speculative-decoding 4-bit image-text-to-text conversational base_model:huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated base_model:quantized:huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated

Related

Total size
15.7 GB
Files
20
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-07-17 18:16

Files by quantization

BF16 1 file 810 MB
model-mtp-bf16.safetensors 810 MB 25ac9ba2 download
Auxiliary files 19 files 15.0 GB
model-00003-of-00003.safetensors 4.99 GB 6eedfe7a download
model-00002-of-00003.safetensors 4.99 GB 2b2e16c9 download
model-00001-of-00003.safetensors 4.98 GB 44be9aa6 download
tokenizer.json 19.1 MB 06b95093 download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 206 KB 75057fd8 download
merge_verify_report.json 35.5 KB cd0fc3ee download
chat_template.jinja 7.58 KB a8755d82 download
OMLX_MTP_BUILD.json 4.17 KB 3243be27 download
config.json 4.14 KB 0b2001c3 download
README.md 2.60 KB 636bd096 download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.14 KB 1d134cd2 download
weight_check_report.json 1.05 KB d02433a6 download
processor_config.json 991 B 8f29fe38 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B c20033f9 download
configuration.json 51.0 B 3a6d4256 download

README current version from Hugging Face


license: apache-2.0
library_name: mlx
base_model: huihui-ai/Huihui-ThinkingCap-Qwen3.6-27B-abliterated
pipeline_tag: image-text-to-text
tags:

  • mlx
  • omlx
  • qwen3.6
  • mtp
  • speculative-decoding
  • 4-bit

Huihui ThinkingCap Qwen3.6 27B — oMLX 4-bit with MTP

This repository is packaged specifically for oMLX native MTP. The target checkpoint's main safetensors index includes all 15 language_model.mtp.* tensors, which bind directly to oMLX's Qwen3.5/3.6 VLM MTP model tree.

Recommended MTP runtime: use this oMLX artifact for MTP. oMLX provides the faster, more mature integrated path for this model and requires no separate drafter.

oMLX

Add this model repository to oMLX and enable Native MTP in the model settings. No separate draft-model repository is required for this oMLX artifact.

Compatibility

Runtime Recommended artifact
oMLX This repository: embedded/indexed language_model.mtp.* target
Direct mlx-vlm Use the normal 4-bit target plus the standalone direct mlx-vlm MTP drafter
LM Studio MLX Use the normal 4-bit target without MTP; runtime 1.10.1 does not support draft models for this batched VLM
MTPLX Use the MTPLX sidecar target

Do not use this embedded-MTP target as an LM Studio target: LM Studio's current MLX target loader and oMLX use different MTP packaging contracts.

Verified runtime

Verified end-to-end on Apple Silicon with oMLX 0.4.4rc1: the model loaded directly through VLMBatchedEngine, oMLX reported its native MTP patch active, and a bounded chat generation completed with the MTP path active. That smoke accepted 2 of 5 drafted tokens (40%); acceptance depends on the prompt.

Technical details

  • Target trunk: MLX affine 4-bit, group size 64
  • MTP tensors: 15 language_model.mtp.* entries in the main model.safetensors.index.json
  • MTP precision: BF16
  • Source revision: 44f63da8141407af529405c1e4b83fa39b70abe0
  • The three previously validated target trunk shards are unchanged; the MTP payload is an additional indexed shard.

Upstream

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-17Recommend oMLX Native MTP75446b32.6 KB
    Loading...
  2. 2026-07-17Clarify oMLX Native MTP usageb0c502a2.3 KB
    Loading...
  3. 2026-07-17Publish oMLX indexed native-MTP target3bab1ee2.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration