← back to catalog · registered 2026-08-22 13:56

Shiftedx/Tess-4-27B-Abliterated-MXFP4-Vision-MLX

Shiftedx Qwen 27B multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Shiftedx%2FTess-4-27B-Abliterated-MXFP4-Vision-MLX"
Response includes
  • classification m1
  • files 15
  • hub_downloads_all_time 161
  • author_summary 24 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
161
15 last 30d - cooling
Likes
0
Model age
3mo ago
created 2026-07-10
Downloads over time
Now167→from68↑146%
6310113917768 on Jul 15167 on Oct 11167 on Oct 9JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5 mlx-vlm mxfp4 qwen3.6 tess vision abliterated apple-silicon image-text-to-text conversational

Related

Total size
14.2 GB
Files
15
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-10 01:23

Files by quantization

Auxiliary files 15 files 14.2 GB
model-00001-of-00003.safetensors 5.00 GB 39479da8 download
model-00002-of-00003.safetensors 4.96 GB 9fdbd24d download
model-00003-of-00003.safetensors 3.35 GB f7a85f9a download
model-vision-00001-of-00001.safetensors 879 MB d0e847d7 download
tokenizer.json 12.2 MB 5f9e4d49 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 159 KB becb46c5 download
tokenizer_config.json 16.3 KB 28d96ff3 download
chat_template.jinja 7.58 KB a8755d82 download
config.json 4.54 KB c065a792 download
README.md 3.92 KB e1123f81 download
.gitattributes 1.53 KB 52373fe2 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download

README current version from Hugging Face


license: apache-2.0
base_model:

  • migtissera/Tess-4-27B
    base_model_relation: quantized
    library_name: mlx
    pipeline_tag: image-text-to-text
    tags:
  • mlx
  • mlx-vlm
  • mxfp4
  • qwen3_5
  • qwen3.6
  • tess
  • vision
  • abliterated
  • apple-silicon

Tess-4-27B Abliterated MXFP4 Vision MLX

This is a local MLX/VLM release of migtissera/Tess-4-27B, converted and quantized to MXFP4 for Apple Silicon, then edited with our refusal-direction ablation workflow. It preserves the Tess/Qwen3.6 vision tower and the Qwen3.5-family multimodal chat template.

This is the standard non-MTPLX artifact. Use the paired MTPLX repository only when you specifically want MTPLX native MTP speculative decoding.

Source

  • Base model: migtissera/Tess-4-27B
  • Source revision: ab2110bec1702f27a62f48f7e8929683a613c51d
  • Base architecture: Qwen3.6/Qwen3.5-family image-text-to-text
  • License: Apache-2.0
  • Chat format: Qwen chat template with <think> reasoning blocks

Conversion

  • Runtime format: MLX
  • Quantization: MXFP4, 4-bit, group size 32
  • Language body size: about 13 GiB
  • Vision tower: BF16 vision tensors grafted from the source model
  • Vision tensor count: 333
  • MTP sidecar: not included in this repo

The non-MTPLX config intentionally advertises mtp_num_hidden_layers = 0 and does not reference mtp.safetensors.

Abliteration Notes

The selected candidate used residual-direction weight orthogonalization against a Tess-specific refusal direction.

Setting Value
Strength 2.5
Targets attention, dense_down
Edited modules 128
Direction scope global
Preserve column norm true

Heldout screen, no code execution:

Variant Refusal rate Benign refusal rate Utility pass rate Avg generation tok/s
Parent MXFP4 1.00 0.00 1.00 23.91
Selected strength 2.5 0.00 0.00 1.00 24.22
Strength 3.0 trial 0.00 0.00 0.50 26.20

Strength 2.5 was selected because the 3.0 trial harmed utility in the heldout screen.

Vision Validation

mlx_vlm.generate smoke passed locally after the vision graft. The smoke image was described as:

A close-up of a white ceramic mug with a black handle, filled with dark coffee and topped with a swirl of foam.

BenchLocal Light Screen

The MTPLX paired artifact was run through a light BenchLocal screen:

Pack Pass / Total Score Failed IDs
toolcall-15 11/15 73% TC-03, TC-07, TC-11, TC-12
instructfollow-15 14/15 93% IF-14
Total 25/30 83%

Important caveat: this was a non-canonical quick run with thinking disabled, max_tokens=2048, and timeout_per_case=90. It is a fast quality screen, not directly comparable to the upstream model-card full BenchLocal score of 122/150 for Tess-4-27B Q8.

Usage

Install:

pip install -U mlx mlx-lm mlx-vlm

Vision:

python -m mlx_vlm.generate \
  --model Shiftedx/Tess-4-27B-Abliterated-MXFP4-Vision-MLX \
  --image path/to/image.png \
  --prompt "Describe this image in one sentence." \
  --max-tokens 128

Text-only prompts should also work through recent MLX-compatible runners that support Qwen3.5/Qwen3.6. For local app use, this standard repo is the LM Studio-oriented artifact; the MTPLX repo is for MTPLX.

Compatibility Notes

  • Recent LM Studio builds support MLX models, including VLMs, on Apple Silicon.
  • Qwen3.5-family models are listed by LM Studio as available in GGUF and MLX, with tool use, vision input, and reasoning support.
  • If you only need stock LM Studio behavior, use this repo rather than the MTPLX wrapper.

Limitations

  • MXFP8 was not produced in this pass because local disk headroom was kept above the workflow safety floor.
  • Vision was validated with MLX-VLM locally. Always run a small smoke test in the exact app/runtime you plan to use.
  • This is an ablated research artifact, not a safety guarantee. Evaluate behavior before deployment.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-10Add files using upload-large-folder tool6fed9f73.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration