← back to catalog · registered 2026-08-22 13:56

morikomorizz/Qwopus3.5-122B-Distill-Kimi-Uncensored-GGUF

morikomorizz Qwen 122B GGUF MoE multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/morikomorizz%2FQwopus3.5-122B-Distill-Kimi-Uncensored-GGUF"
Response includes
  • classification m8
  • files 7
  • benchmarks 11 entries
  • hub_downloads_all_time 4,346
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
4K
687 last 30d - stable
Likes
4
Model age
3mo ago
created 2026-06-23
Downloads over time
Now4.6K→from851↑443%
6632.1K3.6K5K851 on Jun 244.6K on Oct 11JunJulAugSepOct
Jun 24 → Oct 11 · 55 snapshots · spans 109 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 1.4 UGI
Hazardous 1.8 UGI
Natural Intelligence 31.08 UGI
Political lean -21.9% UGI
Sensitive-Info 17.81 UGI
SocPol 2.3 UGI
UGI 17.71 UGI
Willingness (10) 1.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 2 UGI
Writing 39.54 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Quantizations
IQ2 IQ3 IQ4
Tags
transformers gguf qwen3_5_moe image-text-to-text conversational moe agent abliterated uncensored sft dpo opus

Related

Total size
194 GB
Files
7
Quantizations
6
Registered
2026-08-22 13:56
Last updated on HF
2026-06-23 19:10

Files by quantization

IQ4 1 file 76.1 GB
Qwopus3.5-122B-Distill-Kimi-IQ4_NL.gguf 76.1 GB c5958dc4 download
IQ3 1 file 62.6 GB
Qwopus3.5-122B-Distill-Kimi-IQ3_M.gguf 62.6 GB 50a88b44 download
IQ2 1 file 55.0 GB
Qwopus3.5-122B-Distill-Kimi-IQ2_M.gguf 55.0 GB 171c9794 download
BF16 1 file 870 MB
Qwopus3.5-122B-Distill-Kimi-mmproj-BF16.gguf 870 MB 7fd9f29d download
Q8_0 1 file 591 MB
Qwopus3.5-122B-Distill-Kimi-mmproj-Q8_0.gguf 591 MB b4df89ea download
Auxiliary files 2 files 8.10 KB
README.md 6.24 KB ed1c2d7f download
.gitattributes 1.86 KB 19c45098 download

README current version from Hugging Face


base_model:

  • Qwen/Qwen3.5-122B-A10B
  • OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated
    pipeline_tag: text-generation
    library_name: transformers
    tags:
  • qwen3_5_moe
  • image-text-to-text
  • conversational
  • gguf
  • moe
  • agent
  • moe
  • abliterated
  • uncensored
  • sft
  • dpo
  • opus
  • qwopus
  • kimi
  • kimi-k2
  • distill
  • multimodal
  • vision
  • mtp

Qwopus3.5-122B-Distill-Kimi-Uncensored-GGUF

This repository contains the GGUF quantized files for OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated.

  • Original Model: OpenYourMind/Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliterated
  • Architecture: Qwen3.5-122B-A10B
  • License: Apache 2.0
  • Vision: works as expected (image / video → text).
  • MTP: the head is present and shape-compatible, but in our testing it produced no measurable speedup or quality gain on this checkpoint. It is shipped intact for completeness and forward-compatibility, but would need to be retrained to be useful — happy to do so if there is interest in the model.
Quant Type Size Description
IQ2_M 55-59 GB Mixed Precision for Better Quality
IQ3_M 62-67 GB Mixed Precision for Better Quality
IQ4_NL 76-81 GB Mixed Precision for Better Quality

Overview

The pipeline:

  1. Refusal Ablation — Residual-stream refusal directions (one per decoder layer, layers 19–45) were extracted via diff-in-means on a labeled prompt set and baked into the weights as a per-matrix delta — see the abliterix framework for the methodology.
  2. Healing — Stage A: Constrained-LoRA SFT on Opus reasoning data — Supervised finetuned on a curated set of Claude Opus reasoning traces (single-turn, ~8k rows). To keep the abliteration mathematically intact during training, a custom orthogonality projection is applied to every LoRA B-matrix on residual-write modules after each optimizer step (B := B − r·(rᵀB)), so the LoRA update is forbidden from re-introducing the refusal direction. LoRA rank 32, α 64, 54 protected modules across 27 decoder layers. Verified residual after training: max ‖rᵀB‖₂ = 8.5 × 10⁻¹⁰.
  3. Healing — Stage B: Unconstrained SFT on chosen completions — A second short SFT pass (LoRA r=16, α 32, no orthogonality constraint) on the chosen answers (including reasoning chains) from an internal preference dataset, to tighten on the deployment distribution and remove the last bits of drift introduced by Stage A.
  4. Kimi K2.6 Reasoning DPO — A targeted preference-optimization pass distilled from Kimi K2.6 to improve reasoning verbosity and eliminate degenerate looping. See the dedicated section below.
  5. Vision + MTP Restoration — The original Qwen3.5 vision tower (333 tensors, depth 27, hidden 1152) and MTP head (785 tensors, 1 hidden layer) were grafted back from the upstream Qwen/Qwen3.5-122B-A10B shards. Tensor names, shapes, and config.json schema (Qwen3_5MoeForConditionalGeneration, model_type: qwen3_5_moe) match the base model exactly — so this checkpoint loads anywhere the original loads.

Key Properties:

  • Uncensored across the standard refusal axes
  • Reasoning preserved and improved (Opus-style think-then-answer + Kimi K2.6 reasoning DPO)
  • Fewer looping / repetition failures on long conversations
  • Multimodal: vision (image / video) and MTP heads carried forward
  • Drop-in shape compatibility with Qwen/Qwen3.5-122B-A10B

Kimi K2.6 Reasoning DPO

On top of the base abliteration + Opus healing, this release adds a focused healing pass built from Kimi K2.6:

  • ~3,000 samples distilled from Kimi K2.6 were used for DPO (Direct Preference Optimization), alongside synthetic datasets also generated from Kimi K2.6.
  • Improved reasoning verbosity — the model now produces more complete, better-structured reasoning on the ~12% of requests where the previous release tended to under-explain or cut its chain-of-thought short.
  • Fixed looping / repetition — degenerate loops that appeared on 2–6% of long-tail conversations (long context, multi-turn) were largely eliminated.

The DPO pass targets the language model's reasoning behavior only; the abliteration, vision tower, and MTP head are unchanged by this step.

Evaluation

This model family outperforms the full-precision (BF16) Qwen/Qwen3.5-122B-A10B baseline across reasoning, coding, and tool-use benchmarks:

Benchmark Qwen3.5-122B-A10B (BF16, baseline) Qwopus3.5-122B-A10B
CTI 64.8 71.5
LiveCodeBench 78.9 79.9
BFCL 72.2 85.6

BFCL is the Berkeley Function-Calling Leaderboard (tool use); LiveCodeBench is contamination-controlled code generation.


Notes

  • License: Other (inherits from the Qwen3.5 base license)
  • Base Model: Qwen/Qwen3.5-122B-A10B
  • Healing: Opus reasoning SFT + Kimi K2.6 reasoning DPO (≈3,000 distilled samples + synthetic data)
  • Modality: Text + Vision (image / video) + MTP
  • Architecture: Qwen3 MoE (~10B active / 122B total) + Qwen3-VL vision tower + MTP head

Thanks

  • Jackrong — for the idea of Qwopus merges (Opus distillations on Qwen models).
  • wangzhang — for the wonderful abliterix framework, which was customized to do this abliteration.

Disclaimer

Use is the responsibility of the user. Ensure your usage complies with applicable laws, platform rules, and deployment requirements.


How to Use

These GGUF files are fully compatible with llama.cpp and popular graphical interfaces like LM Studio.

using llama.cpp CLI:

./llama-cli -m /path/to/model/Qwopus3.5-122B-Distill-Kimi-IQ3_M.gguf \
  -p "Hello, how are you?" \
  -sys "You are a helpful AI" \
  -n 4096 \
  -c 8192

using llama-server :

./llama-cli -m /path/to/model/Qwopus3.5-122B-Distill-Kimi-IQ3_M.gguf \
  --host 0.0.0.0 \
  --port 8080

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-23Update README.mde10131c6.2 KB
    Loading...
  2. 2026-06-23Update README.mdb5ba4286.2 KB
    Loading...
  3. 2026-06-23Update README.md6b4a1bd6.1 KB
    Loading...
  4. 2026-06-23Update README.md822c9766 KB
    Loading...
  5. 2026-06-23Update README.md9e877a24.4 KB
    Loading...
  6. 2026-06-23initial commitc5df4de28 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration