← back to catalog · registered 2026-08-22 13:56

Shiftedx/Qwen3.8-27B-Abliterated-MLX-MXFP4-MTP

Shiftedx Qwen 27B multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Shiftedx%2FQwen3.8-27B-Abliterated-MLX-MXFP4-MTP"
Response includes
  • classification m1
  • files 19
  • hub_downloads_all_time 5,353
  • author_summary 24 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
5K
413 last 30d - cooling
Likes
2
Model age
8w ago
created 2026-08-14
Downloads over time
Now5.5K→from3.7K↑49%
3.6K4.3K5K5.7K3.7K on Aug 195.5K on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5 mtplx mxfp4 native-mtp abliterated vision-language qwen3.8 image-text-to-text conversational base_model:Qwen/Qwen3.8-27B

Related

Total size
15.0 GB
Files
19
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-17 08:34

Files by quantization

Auxiliary files 19 files 15.0 GB
model-00001-of-00003.safetensors 5.00 GB 84226f42 download
model-00002-of-00003.safetensors 4.96 GB 43895782 download
model-00003-of-00003.safetensors 3.35 GB 8ebe0eee download
model-vision-00001-of-00001.safetensors 879 MB 7a252e5c download
mtp.safetensors 810 MB 4468f396 download
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 159 KB 49bda3a7 download
tokenizer_config.json 17.5 KB 5de744b3 download
LICENSE 11.3 KB f938136e download
chat_template.jinja 8.74 KB c0c686f9 download
config.json 5.20 KB 698f07e8 download
README.md 3.32 KB 049e600c download
mtplx_runtime.json 2.38 KB e596da09 download
.gitattributes 1.53 KB 52373fe2 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 202 B 023756cf download

README current version from Hugging Face


library_name: mlx
license: apache-2.0
pipeline_tag: image-text-to-text
base_model: Qwen/Qwen3.8-27B
base_model_relation: quantized
tags:

  • mlx
  • mtplx
  • mxfp4
  • native-mtp
  • abliterated
  • vision-language
  • qwen3.8

Qwen3.8-27B Abliterated MLX MXFP4 + Native MTP

Vision-enabled MTPLX package of Qwen/Qwen3.8-27B, pinned to revision
1d4bf0f2ff6012fd82039f2fa52739d0dd7c60c0, with a measured refusal-direction edit.

  • Language model: MXFP4, group size 32; strength 2.5 norm-preserving edit on 128 modules
  • Vision tower: unchanged same-revision BF16 weights
  • Speculation: 15 unchanged native BF16 MTP tensors; recommended depth 3
  • Runtime: MTPLX 2.0.2 or newer

The frozen held-out behavior suite changed measured refusal markers from 12/12 to 0/12,
with benign-sensitive refusal at 0/12 and utility at 6/6. MTPLX tensor inspection, text
generation, and three vision requests through its API passed. This is scoped evidence, not a
claim that the model is universally “uncensored.”

The Hub's approximately 5.5B safetensors count reflects packed MXFP storage elements; the
underlying architecture remains the full 27B model.

mtplx quickstart \
  --model Shiftedx/Qwen3.8-27B-Abliterated-MLX-MXFP4-MTP \
  --mtp --depth 3 --profile sustained

This edit reduces refusal behavior and may increase harmful or incorrect outputs. Treat prompts,
images, and outputs as untrusted; do not submit secrets, and sandbox tools or generated code with
least-privilege access. This package adds no telemetry or remote execution.

The upstream Apache-2.0 license and model limitations continue to apply.

Shiftedx Bench post-publication qualification

This table was generated from the frozen lightweight quant gate after the model weights were published. Categories remain separate; the benchmark does not produce a composite intelligence score.

Lane Passed Accuracy Mean wall time Mean decode Peak active memory
Quality 6/10 60.0% 22.41 s 53.46 tok/s 32.36 GiB
Long context 13/15 86.7% 145.65 s 42.63 tok/s 41.02 GiB
Tool calling 6/6 100.0% 4.26 s 41.08 tok/s 35.18 GiB
Agentic 0/2 0.0% 19.22 s — tok/s —
Vision 3/4 75.0% 4.81 s 49.61 tok/s 34.42 GiB
  • Tested model revision: d6352f90f36268551b9032b37ec92fe89c1a4ebe
  • Benchmark: Shiftedx Bench v0.3.0
  • Context lengths represented: 4,096, 16,384, 65,536, 131,072 prompt tokens; effective tested context: 131,072 tokens
  • Runtime contract: MTPLX 2.7.1 sustained; thinking on/medium; sampler temperature=1.0, top_p=0.95, top_k=20; Hermes hybrid routing with native explicit-parallel calls; KV cache off; MTP depth 3
  • Host: Apple M4 Max, 64 GiB unified memory
  • Total measured request wall time: 2492.15 seconds
  • 260,096-token status: not run; it is outside the lightweight quant gate.

Scores are specific to the linked model revision, benchmark revision, runtime contract, and host. Changing weight precision, KV-cache precision, reasoning mode, template, or speculative depth creates a different benchmark candidate.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-17docs: add standardized Shiftedx quant gate8df6d223.3 KB
    Loading...
  2. 2026-08-14Add files using upload-large-folder toold6352f91.6 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration