← back to catalog · registered 2026-08-22 13:56

OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-MLX-4bit

OpenYourMind Qwen 122B MoE multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/OpenYourMind%2FQwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-MLX-4bit"
Response includes
  • classification m4
  • files 26
  • hub_downloads_all_time 4,264
  • author_summary 11 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M4
Primary method

Abliterate + heal

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 2 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'healed'/'orpo'/'dpo' in name suggests heal step after abliteration
  • M4 = abliterate + heal pipeline
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
4K
226 last 30d - cooling
Likes
1
Model age
4mo ago
created 2026-05-20
Downloads over time
Now4.3K→from408↑965%
2111.7K3.2K4.7K408 on May 204.3K on Oct 11MayJunJulAugSepOct
May 20 → Oct 11 · 61 snapshots · spans 144 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Tags
mlx safetensors qwen3_5_moe qwen qwen3 qwen3.5 moe abliterated uncensored dpo opus qwopus

Related

Total size
64.8 GB
Files
26
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-16 09:28

Files by quantization

Auxiliary files 26 files 64.9 GB
model-00011-of-00014.safetensors 4.84 GB fdec5259 download
model-00008-of-00014.safetensors 4.84 GB 11b1fbd7 download
model-00005-of-00014.safetensors 4.84 GB f527899e download
model-00012-of-00014.safetensors 4.84 GB 0adcf79a download
model-00006-of-00014.safetensors 4.84 GB 5e9222d3 download
model-00009-of-00014.safetensors 4.84 GB bad0a01f download
model-00003-of-00014.safetensors 4.84 GB 06f685c2 download
model-00002-of-00014.safetensors 4.84 GB 571a4927 download
model-00013-of-00014.safetensors 4.80 GB e92cb3a4 download
model-00004-of-00014.safetensors 4.79 GB 685e89af download
model-00007-of-00014.safetensors 4.79 GB 46aa2ff8 download
model-00010-of-00014.safetensors 4.79 GB 6f3b2563 download
model-00001-of-00014.safetensors 4.77 GB 8b85826a download
model-00014-of-00014.safetensors 2.14 GB e68bd36d download
tokenizer.json 19.1 MB 639e352c download
OYM_banner.png 1.58 MB a714b89b download
model.safetensors.index.json 247 KB 00324b36 download
config.json 27.2 KB 52dfef37 download
chat_template.jinja 7.57 KB a585dec8 download
README.md 4.02 KB 1946f729 download
.gitattributes 1.58 KB b6da95e2 download
processor_config.json 1.27 KB 7ad6acdf download
tokenizer_config.json 1.24 KB 9de16b5b download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B 318011ae download

README current version from Hugging Face


license: other
library_name: mlx
base_model: OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated
tags:

  • mlx
  • qwen
  • qwen3
  • qwen3.5
  • moe
  • abliterated
  • uncensored
  • dpo
  • opus
  • qwopus
  • kimi
  • kimi-k2
  • multimodal
  • vision
  • 4-bit
    pipeline_tag: image-text-to-text

OpenYourMind

Support & Community

☕ If these models are useful to you, consider supporting my work — it funds compute for more & larger abliterations.

Buy Me A Coffee

buymeacoffee.com/oym.kuato

💬 Discord: discord.gg/rhUZY5GEZr  ·  ₿ Bitcoin: bc1qsvfduzj9fjs9fugpc52yver3f2g8fp7xjxecdv


Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated — MLX 4-bit

Overview

MLX 4-bit quant of OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated for Apple Silicon (LM Studio / mlx-vlm). See the parent repo for the full pipeline: refusal ablation → constrained-LoRA Opus reasoning SFT → unconstrained chosen-completion SFT → Kimi K2.6 reasoning DPO (≈3,000 distilled samples + synthetic data, improving reasoning verbosity on ~12% of requests and removing looping on 2–6% of long-tail conversations).

  • Multimodal: the vision tower is included and kept in full precision (BF16) — image input works.
  • Language: 4-bit, group size 64 (MoE routing gates kept at higher precision via the model's quant predicate), ≈ 4.5 bits/weight overall.

MTP

The MTP head is not included in this MLX build: no current MLX runtime (mlx-lm / mlx-vlm, including LM Studio's bundled engine) consumes the qwen3.5 MTP head — they drop mtp.* on load — so it was omitted to keep the model loading cleanly. If/when an MLX runtime adds qwen3.5 MTP support, an MTP-enabled build can be produced from the full weights. (Note: MTP gave no measurable gain in our testing and would need retraining to be useful — see the parent card. For MTP today, use the GGUF build with llama.cpp.)

Files

File Description Size
model-*-of-00014.safetensors 4-bit language weights + BF16 vision tower ~65 GB total
config.json Qwen3_5MoeForConditionalGeneration + quantization (4-bit, g64) —
preprocessor_config.json, video_preprocessor_config.json, processor_config.json Qwen3-VL image/video processor —
tokenizer*, chat_template.jinja, generation_config.json Standard —

Total on disk: ~65 GB.

Usage

pip install mlx-vlm
python -m mlx_vlm.generate \
  --model OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-MLX-4bit \
  --image path/to/image.jpg \
  --prompt "Describe this image." --max-tokens 256

In LM Studio: drop the folder under your models directory (publisher OpenYourMind) and load it with the MLX runtime.

Hardware

~65 GB on disk; needs roughly ≥ 72 GB unified memory to load with usable context. Runs on M-series Macs with 96 GB+.

Notes

  • License: Other (inherits from the Qwen3.5 base license)
  • Parent (full weights): Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated
  • Modality: Text + Vision (image / video). MTP not included in this build.
  • Architecture: Qwen3 MoE (~10B active / 122B total) + Qwen3-VL vision tower

Disclaimer

Use is the responsibility of the user. Ensure your usage complies with applicable laws, platform rules, and deployment requirements.

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-16Move Support & Community section to top (below banner); unify across models51aeda94 KB
    Loading...
  2. 2026-06-08Add OYM banner to top of model card28a21bb4 KB
    Loading...
  3. 2026-06-08Add highlighted Buy Me a Coffee support section38c8b7e3.9 KB
    Loading...
  4. 2026-05-20Add model card31a390b3.5 KB
    Loading...

Discussions 1 thread

  1. 2026-05-25Native MTP Speculative Decoding on Apple Siliconopen5 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration