← back to catalog · registered 2026-08-22 13:56

Jonyses/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-GGUF

Jonyses Qwen 122B GGUF MoE multimodal second-order 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Jonyses%2FQwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-GGUF"
Response includes
  • classification m4
  • files 5
  • hub_downloads_all_time 786
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M4
Primary method

Abliterate + heal

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 2 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'healed'/'orpo'/'dpo' in name suggests heal step after abliteration
  • M4 = abliterate + heal pipeline
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
786
224 last 30d - stable
Likes
2
Model age
7w ago
created 2026-08-21
Downloads over time
Now891→from91↑879%
5135866497191 on Aug 19891 on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Quantizations
Q4_K
Tags
gguf qwen qwen3 qwen3.5 moe abliterated uncensored dpo opus qwopus kimi kimi-k2

Related

Total size
70.6 GB
Files
5
Quantizations
3
Registered
2026-08-22 13:56
Last updated on HF
2026-08-21 19:52

Files by quantization

Q4_K 1 file 70.6 GB
Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-Q4_K_M.gguf 70.6 GB a504c684 download
F16 1 file 867 MB
mmproj-Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-F16.gguf 867 MB e77aff84 download
Auxiliary files 3 files 1.59 MB
OYM_banner.png 1.58 MB a714b89b download
README.md 4.41 KB 75aad2e6 download
.gitattributes 1.82 KB 0d33850c download

README current version from Hugging Face


license: other
library_name: gguf
base_model: OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated
tags:

  • gguf
  • qwen
  • qwen3
  • qwen3.5
  • moe
  • abliterated
  • uncensored
  • dpo
  • opus
  • qwopus
  • kimi
  • kimi-k2
  • multimodal
  • vision
  • mmproj
  • mtp
  • q4_k_m
    pipeline_tag: image-text-to-text

OpenYourMind

Support & Community

☕ If these models are useful to you, consider supporting my work — it funds compute for more & larger abliterations.

Buy Me A Coffee

buymeacoffee.com/oym.kuato

💬 Discord: discord.gg/rhUZY5GEZr  ·  ₿ Bitcoin: bc1qsvfduzj9fjs9fugpc52yver3f2g8fp7xjxecdv


Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated — GGUF

Overview

GGUF build of OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated. See the parent repo for the full pipeline: refusal ablation → constrained-LoRA Opus reasoning SFT → unconstrained chosen-completion SFT → Kimi K2.6 reasoning DPO (≈3,000 distilled samples + synthetic data, improving reasoning verbosity on ~12% of requests and removing looping on 2–6% of long-tail conversations).

This repo ships both the language model and the vision projector (mmproj), so it runs as a full multimodal (image + text) model in llama.cpp / LM Studio.

Files

File Bits/weight Size Notes
Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-Q4_K_M.gguf ~4.6 ~76 GB Language model. Q4_K_M keeps output.weight at higher precision. MTP head included.
mmproj-Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-F16.gguf F16 ~0.9 GB Vision projector (qwen3vl_merger, Qwen3.5 vision tower). Load alongside the model for image input.

Vision (mmproj)

Pass the mmproj file to enable image input. The vision tower is the standard Qwen3.5-122B-A10B Qwen3-VL encoder (carried over unchanged from the base model), F16.

# llama.cpp multimodal CLI
llama-mtmd-cli \
  -m Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-Q4_K_M.gguf \
  --mmproj mmproj-Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-F16.gguf \
  --image path/to/image.jpg \
  -p "Describe this image." -ngl 99

In LM Studio: keep the mmproj-*.gguf in the same folder as the model — it is detected automatically and the image-attachment button becomes available.

MTP (multi-token prediction)

This build keeps the MTP head (blk.48.nextn.*, qwen35moe.nextn_predict_layers). Recent llama.cpp with qwen35moe MTP support (e.g. LM Studio's llama.cpp 2.15.0) can load it and expose "MTP Speculative Decoding" in the advanced load settings.

⚠️ Caveat: in our testing the MTP head gave no measurable performance gain on this checkpoint. It is shipped for completeness and would need to be retrained to be genuinely useful — happy to do so if there is interest in the model. The model runs fine with MTP speculative decoding off.

Usage (text-only)

Requires a recent llama.cpp build that supports the qwen35moe architecture (Gated DeltaNet linear-attn + MoE).

llama-cli -m Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated-Q4_K_M.gguf \
  -p "Explain why the sky is blue." -ngl 99 -c 8192

Hardware

Q4_K_M (~76 GB) + mmproj (~0.9 GB) fits on a single 96 GB GPU, an Apple Silicon machine with ≥ 96 GB unified memory, or CPU + RAM. Leave headroom for KV cache / context.

Notes

Disclaimer

Use is the responsibility of the user. Ensure your usage complies with applicable laws, platform rules, and deployment requirements.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-21Duplicate from OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abli...e69ab5a4.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration