← back to catalog · registered 2026-08-22 13:56

Flakily6416/Qwen3.5-27B-Opus-Reasoning-v2-Abliterated-EvoPress-GGUF

Flakily6416 Qwen 27B GGUF multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Flakily6416%2FQwen3.5-27B-Opus-Reasoning-v2-Abliterated-EvoPress-GGUF"
Response includes
  • classification m8
  • files 10
  • hub_downloads_all_time 9,145
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
9K
520 last 30d - cooling
Likes
3
Model age
6mo ago
created 2026-04-02

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now9.5K→from7.6K↑25%
7.5K8.2K8.9K9.7K7.6K on Apr 159.5K on Oct 11AprMayJunJulAugSepOct
Apr 15 → Oct 11 · 65 snapshots · spans 179 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh ko
Quantizations
Q2_K Q3_K Q4_K Q5_K Q6_K
Tags
gguf mamba uncensored reasoning chain-of-thought qwen3.5 image-text-to-text en zh ko dataset:nohurry/Opus-4.6-Reasoning-3000x-filtered dataset:Jackrong/Qwen3.5-reasoning-700x
Total size
119 GB
Files
10
Quantizations
6
Registered
2026-08-22 13:56
Last updated on HF
2026-04-06 00:34

Files by quantization

Q6_K 1 file 20.6 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q6_K.gguf 20.6 GB 9039742b download
Q5_K 1 file 17.9 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q5_K.gguf 17.9 GB 943ee36f download
Q4_K 1 file 15.4 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q4_K.gguf 15.4 GB f3789a6d download
Q3_K 1 file 12.4 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q3_K.gguf 12.4 GB 3d9202b3 download
Q2_K 1 file 9.98 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q2_K.gguf 9.98 GB cd76b134 download
Auxiliary files 5 files 42.9 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP5.0_bpw.gguf 17.0 GB 2dd3ee26 download
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP4.25_bpw.gguf 14.4 GB 36280098 download
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP3.5_bpw.gguf 11.6 GB 9ffbe1e4 download
README.md 3.50 KB ee9afd7d download
.gitattributes 3.15 KB a9a2c801 download

README current version from Hugging Face


license: apache-2.0
language:

  • en
  • zh
  • ko
    tags:
  • gguf
  • mamba
  • uncensored
  • reasoning
  • chain-of-thought
  • qwen3.5
    base_model: Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2
    pipeline_tag: image-text-to-text
    datasets:
  • nohurry/Opus-4.6-Reasoning-3000x-filtered
  • Jackrong/Qwen3.5-reasoning-700x
  • Roman1111111/claude-opus-4.6-10000x

Note: This release currently does not have vision capabilities due to an oversight.

I'll get this fixed as soon as my free lightning ai credits reset (or please get in touch if you would like to sponsor some A100 hours).

Qwen 3.5 27B Opus-Reasoning v2 (Abliterated) - Mixed Precision GGUFs

This repository features traditional and EvoPress GGUF quants of an abliterated reasoning model. Built upon Jackrong's Claude-4.6-Opus Distillation v2, the model was uncensored via the Orion-Zhen pipeline before being quantized with the EvoPress mixed-precision strategy.

🔥 Model Lineage & Highlights

  1. Base: Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2
    • Distilled using 14,000+ Claude 4.6 Opus-style samples to drastically improve Chain-of-Thought (CoT) efficiency.
    • Reduces unnecessarily long internal reasoning chains while maintaining top-tier benchmark scores (e.g., 96.91% pass@1 on HumanEval).
  2. Abliteration: Processed via Orion-Zhen's open-source pipeline.

⚡ EvoPress Mixed-Precision Quantization

This release utilizes the EvoPress methodology to maximize intelligence-per-gigabyte in Hybrid Mamba architectures. Standard quantization often degrades the sensitive State Space Model (SSM) components; these quants solve that by using tiered-precision mapping.

Key Methodology: EvoPress (GitHub/HF)

File Name Target BPW VRAM Fit Optimization Strategy
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP3.5_bpw.gguf 3.5 12GB - 16GB Q3_K Base + F32 Mamba/Norms
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP4.25_bpw.gguf 4.25 16GB - 24GB Q4_K Base + F32 Mamba/Norms
Qwen-3.5-27B-Opus-Reasoning-Abliterated-EP5.0_bpw.gguf 5.0 24GB+ Q5_K Base + F32 Mamba/Norms

📊 Standard/Traditional Quantizations

Included for comparison and compatibility with older hardware.

File Type Size
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q2_K.gguf Q2_K ~10 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q3_K.gguf Q3_K ~13 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q4_K.gguf Q4_K ~16 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q5_K.gguf Q5_K ~19 GB
Qwen-3.5-27B-Opus-Reasoning-Abliterated-Q6_K.gguf Q6_K ~23 GB

🛠️ Technical Details & Setup

  • Architecture: Hybrid Mamba-Transformer (Qwen 3.5)
  • Quantization: Performed using a modified gptq-gguf-toolkit for Mamba-aware layer mapping.
  • Requirements: Use a recent build of llama.cpp (March 2026+) for full Hybrid Mamba support.

🤝 Credits & Acknowledgements

  • Jackrong: For the Claude-4.6-Opus reasoning distillation methodology and published model.
  • Orion-Zhen: For the abliteration and refusal-removal pipeline.
  • Alibaba Qwen Team: For the base Qwen 3.5 architecture.
  • Flakily6416: Quantization, layer-mapping, and Mixed-Precision optimization.

README history 9 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-06Update README.mddc00fca3.5 KB
    Loading...
  2. 2026-04-06Update README.mdf9b359f3.6 KB
    Loading...
  3. 2026-04-05Update README.md0c8deae3.6 KB
    Loading...
  4. 2026-04-02Update README.md434edfb3.4 KB
    Loading...
  5. 2026-04-02Update README.md18c6c8f3.2 KB
    Loading...
  6. 2026-04-02Update README.md22f57cd3.2 KB
    Loading...
  7. 2026-04-02Update README.md371fb3c3.2 KB
    Loading...
  8. 2026-04-02Initial release: HauhauCS Aggressive Uncensored (EvoPress GGUFs)5b5671c2.3 KB
    Loading...
  9. 2026-04-02Initial release: HauhauCS Aggressive Uncensored (EvoPress GGUFs)ab03ed12.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration