← back to catalog · registered 2026-08-22 13:56

Shiftedx/qwopus3.6-35b-a3b-coder-abliterated-mxfp8-vision-mlx

Shiftedx 35B MoE multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Shiftedx%2Fqwopus3.6-35b-a3b-coder-abliterated-mxfp8-vision-mlx"
Response includes
  • classification m1
  • files 18
  • hub_downloads_all_time 560
  • author_summary 24 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
560
58 last 30d - stable
Likes
0
Model age
3mo ago
created 2026-07-01
Downloads over time
Now581→from53↑996%
2722943163453 on Jul 1581 on Oct 11JulAugSepOct
Jul 1 → Oct 11 · 54 snapshots · spans 102 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors qwen3_5_moe mlx-vlm mxfp8 lm-studio apple-silicon qwen3_6 coder agent tool-use function-calling

Related

Total size
34.1 GB
Files
18
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-05 23:07

Files by quantization

Auxiliary files 18 files 34.1 GB
model-00005-of-00007.safetensors 4.85 GB b0ace159 download
model-00003-of-00007.safetensors 4.85 GB c099f96f download
model-00004-of-00007.safetensors 4.84 GB 55dc7b17 download
model-00006-of-00007.safetensors 4.84 GB d5924120 download
model-00002-of-00007.safetensors 4.84 GB 92288515 download
model-00001-of-00007.safetensors 4.82 GB 470c599f download
model-00007-of-00007.safetensors 4.24 GB fcc218d3 download
model-vision-00001-of-00001.safetensors 852 MB d4579339 download
tokenizer.json 19.1 MB 87a7830d download
model.safetensors.index.json 161 KB a34efea4 download
config.json 48.5 KB 3321e104 download
tokenizer_config.json 15.0 KB dc7d3fb6 download
chat_template.jinja 7.70 KB 68a73673 download
README.md 3.09 KB be99601d download
.gitattributes 1.53 KB 52373fe2 download
processor_config.json 991 B 8f29fe38 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download

README current version from Hugging Face


license: apache-2.0
base_model: Jackrong/Qwopus3.6-35B-A3B-Coder
base_model_relation: quantized
library_name: mlx
pipeline_tag: image-text-to-text
tags:

  • mlx
  • mlx-vlm
  • mxfp8
  • safetensors
  • lm-studio
  • apple-silicon
  • qwen3_5_moe
  • qwen3_6
  • coder
  • agent
  • tool-use
  • function-calling
  • vision
  • thinking-off
  • abliterated

qwopus3.6-35b-a3b-coder-abliterated-mxfp8-vision-mlx

Vision-enabled MLX MXFP8 conversion of Jackrong/Qwopus3.6-35B-A3B-Coder, prepared by Shiftedx for Apple Silicon, MLX, and LM Studio.

Build Notes

  • Quantized primary language weights with mxfp8 at group size 32.
  • Kept MoE router and gate modules in affine 8-bit group size 64 for compatibility.
  • Added Qwen3.5/Qwen3.6-MoE-compatible vision components and validated image grounding locally.
  • Removed source MTP tensors and set MTP/next-token prediction layer counts to 0 for LM Studio compatibility.
  • Set tool_parser_type to qwen3_coder.
  • Patched the chat template so enable_thinking defaults to false when the runtime honors that template variable.
  • Applied a research refusal-direction weight edit using residual-direction orthogonalization.

Local Validation

Hardware reference: Apple M4 Max Apple Silicon with 64 GB unified memory.

Validated locally on June 30, 2026 and July 1, 2026 with LM Studio and direct MLX/VLM loading.

Check Result
LM Studio load Passed at 32k context, parallel 1, GPU max.
Basic text completion Passed; answered 2+2 with 4 and stopped.
Code completion Passed; produced a simple valid add(a, b) function.
Direct MLX/VLM image color smoke Passed; answered Red.
Direct MLX/VLM OCR smoke Passed; answered FABLE 42.
LM Studio OpenAI-compatible image smoke Passed; answered Red and FABLE 42.
LM Studio native image smoke Passed; answered Red and FABLE 42.
Thinking-off behavior Smoke checks returned 0 reasoning tokens.
LM Studio logs No warnings, errors, tracebacks, KV-cache issues, or tokenizer-regex warnings in the validation window.

This is a smoke-validated release, not a full benchmark suite. Broader downstream evaluation is still recommended for production use.

Recommended LM Studio Defaults

After downloading in LM Studio, load the model by repo name:

lms load shiftedx/qwopus3.6-35b-a3b-coder-abliterated-mxfp8-vision-mlx --context-length 32768 --parallel 1 --gpu max

Recommended profile defaults:

  • Preset/template: Qwen3 thinking-compatible Jinja template with <|im_end|> stop.
  • Thinking: off by default through the included chat template.
  • Context length: 200000 when memory allows; 32768 was used for local validation.
  • Sampling: temperature 0.6, top-k 20, top-p 0.95, min-p enabled at 0.
  • Repeat penalty: off by default.
  • Load: parallel 1, GPU max.

Provenance

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-05Add validation hardware referencef1c61283.1 KB
    Loading...
  2. 2026-07-01Add files using upload-large-folder toolb2438803 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration