← back to catalog · registered 2026-08-22 13:56

AITRADER/Huihui-Qwen3.5-27B-Claude-4.6-Opus-abliterated-mlx-8bit

AITRADER Qwen 24B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/AITRADER%2FHuihui-Qwen3.5-27B-Claude-4.6-Opus-abliterated-mlx-8bit"
Response includes
  • classification m1
  • files 22
  • hub_downloads_all_time 1,566
  • author_summary 30 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
2K
68 last 30d - cooling
Likes
1
Model age
6mo ago
created 2026-03-25
Downloads over time
Now1.6K→from412↑286%
3538051.3K1.7K412 on Mar 251.6K on Oct 11MarAprMayJunJulAugSepOct
Mar 25 → Oct 11 · 68 snapshots · spans 200 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5 vision qwen3.5 abliterated tool-use function-calling image-text-to-text conversational license:apache-2.0 8-bit

Related

Total size
29.7 GB
Files
22
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-03-26 13:04

Files by quantization

Auxiliary files 22 files 29.7 GB
model-00001-of-00011.safetensors 4.10 GB 37217d51 download
model-00005-of-00011.safetensors 2.64 GB 0776510d download
model-00009-of-00011.safetensors 2.64 GB 7833ccf5 download
model-00003-of-00011.safetensors 2.63 GB a0198394 download
model-00007-of-00011.safetensors 2.63 GB afef0ffc download
model-00004-of-00011.safetensors 2.63 GB 13115b02 download
model-00008-of-00011.safetensors 2.63 GB 76549a56 download
model-00006-of-00011.safetensors 2.63 GB eaf1c96e download
model-00002-of-00011.safetensors 2.63 GB 4c39c8fb download
model-00011-of-00011.safetensors 2.37 GB 9b77b5f8 download
model-00010-of-00011.safetensors 2.14 GB 29765f5d download
tokenizer.json 9.61 MB 92b4bc16 download
model.safetensors.index.json 204 KB c061a1c2 download
chat_template.jinja 7.57 KB a585dec8 download
config.json 3.88 KB 2e88dd05 download
.gitattributes 1.48 KB a6344aac download
README.md 1.37 KB 020941b4 download
processor_config.json 1.27 KB 7ad6acdf download
tokenizer_config.json 1.24 KB 855dd6e0 download
preprocessor_config.json 337 B e501264c download
video_preprocessor_config.json 332 B 4d1c4049 download
generation_config.json 178 B bb6b4de8 download

README current version from Hugging Face


base_model: huihui-ai/Qwen3.5-27B-abliterated
tags:

  • mlx
  • vision
  • qwen3.5
  • abliterated
  • tool-use
  • function-calling
    license: apache-2.0
    pipeline_tag: image-text-to-text
    library_name: mlx

Huihui-Qwen3.5-27B-Claude-4.6-Opus-abliterated MLX 8-bit

8-bit quantized MLX version of Huihui-Qwen3.5-27B-Claude-4.6-Opus-abliterated.

Model Details

  • Architecture: Qwen 3.5 27B (hybrid linear attention + full attention)
  • Quantization: 8-bit affine, group_size=64
  • Size: ~32 GB (down from ~55 GB bf16)
  • Context Length: 262,144 tokens
  • Vision: Full image and video understanding via integrated vision tower (27 ViT blocks)
  • Tool Use: Native function calling support
  • Thinking: Chain-of-thought reasoning mode

Quantization Strategy

  • Quantized (8-bit): All large projection weights in 64 language model layers (MLP, attention, linear attention)
  • Kept in bf16: Vision tower, embeddings, LM head, layer norms, SSM parameters (A_log, dt_bias, conv1d)
  • This preserves vision quality and model stability while significantly reducing memory usage

Capabilities

  • Image understanding and description
  • Video understanding
  • Tool use / function calling
  • Multi-step agent reasoning
  • Thinking/reasoning mode
  • Multilingual support
  • Long context (262K tokens)

Usage

Works with LM Studio, MLX, and other MLX-compatible frameworks.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-03-25Add README.md for vision support72de4311.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration