← back to catalog · registered 2026-08-22 13:56

AIOpsInSpace/Qwen2.5-Coder-32B-Instruct-Uncensored-Patched

AIOpsInSpace Qwen 32B GGUF 33K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/AIOpsInSpace%2FQwen2.5-Coder-32B-Instruct-Uncensored-Patched"
Response includes
  • classification m-uncensored
  • files 8
  • benchmarks 10 entries
  • hub_downloads_all_time 2,477
  • author_summary 12 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
441 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-07-16
Downloads over time
Now2.6K→from149↑1,636%
279621.9K2.8K149 on Jul 152.6K on Oct 9JulAugSepOct
Jul 15 → Oct 9 · 51 snapshots · spans 86 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Arena-Battles 5730 LM-Arena
LM Arena Elo 1235.1499959151995 LM-Arena
Arena-Elo-Lower 1227.470119075386 LM-Arena
Arena-Elo-Upper 1242.829872755013 LM-Arena
Arena-Rank 123 LM-Arena
BBH average 0.5898300687508108 OpenLLM-v2
IFEval instruct 0.7709832134292566 OpenLLM-v2
IFEval-Prompt 0.6820702402957486 OpenLLM-v2
MATH lvl 5 0.42749244712990936 OpenLLM-v2
MMLU-Pro 0.44132313829787234 OpenLLM-v2

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Quantizations
Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf text-generation code coding-assistant qwen2.5-coder en base_model:Qwen/Qwen2.5-Coder-32B-Instruct base_model:quantized:Qwen/Qwen2.5-Coder-32B-Instruct license:apache-2.0 endpoints_compatible region:us conversational

Related

Total size
124 GB
Files
8
Quantizations
7
Registered
2026-08-22 13:56
Last updated on HF
2026-07-22 07:19

Files by quantization

Q8_0 1 file 32.4 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q8_0.gguf 32.4 GB 9c599686 download
Q6_K 1 file 25.0 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q6_K.gguf 25.0 GB 4ae6ed7a download
Q5_K 1 file 21.7 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q5_K_M.gguf 21.7 GB d1cd63fe download
Q4_K 1 file 18.5 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q4_K_M.gguf 18.5 GB e3633060 download
Q3_K 1 file 14.8 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q3_K_M.gguf 14.8 GB 901c3f10 download
Q2_K 1 file 11.5 GB
Qwen2.5-Coder-32B-Instruct-Uncensored-Patched-Q2_K.gguf 11.5 GB 69409f53 download
Auxiliary files 2 files 19.9 KB
README.md 17.9 KB aaf8efb8 download
.gitattributes 2.03 KB 130f769c download

README current version from Hugging Face


base_model: Qwen/Qwen2.5-Coder-32B-Instruct
tags:

  • text-generation
  • gguf
  • code
  • coding-assistant
  • qwen2.5-coder
    license: apache-2.0
    language:
  • en
    pipeline_tag: text-generation

"This is humanity's race.
The solution is open source.
Stay sovereign."

— AIOpsInSpace

Qwen2.5-Coder-32B-Instruct-Uncensored-Patched

AIOpsInSpace Official

Premier 32B open-source code intelligence model patched for IDE autocomplete stability and uncensored generation.

🔥 32B Dense Model ⚡ Code Assistant Optimized 🛠️ IDE Plugin Hang Patched

> What is this model and Why is it Needed?

Qwen2.5-Coder-32B-Instruct-Uncensored-Patched is built on top of Qwen/Qwen2.5-Coder-32B-Instruct.

Why it is needed: Solves critical GGUF token parsing bugs that caused VS Code and JetBrains AI plugins to hang during inline code completion.

> From the Parent Repository

"Qwen2.5-Coder 32B matches proprietary frontier models across HumanEval and SWE-bench."

— Qwen Code Team


🏗️ 2. Model Architecture & Merging

Architecture: Qwen 2.5 Coder 32B Transformer Architecture
Merging Technique: IDE Tokenizer Fixes & Uncensored Fine-Tuning
Constituent Models: Methodology: Repaired FIM (Fill-In-Middle) token mappings for IDE extension backends.

🚀 3. Technical Enhancements

> Key Upgrades Over Base Model:

  • Frontier Coding: State-of-the-art Python, C++, Rust, and SQL generation.
  • IDE Stability: Prevents autocomplete hangs in Continue, Cline, and Aider.

📊 4. Benchmark Competitiveness vs. Frontier Scores

> Evaluated Performance
Benchmark Qwen2.5-Coder-32B-Instruct-Uncensored-Patched Frontier Target
MMLU Evaluated 88.7%
GSM8K Evaluated 95.6%
HumanEval Evaluated 90.2%

🏆 5. Comprehensive Arena Analytics

> Status: Active Community Benchmarking

// Note: Arena Elo and head-to-head winrates updated continuously as evaluation telemetry processes.

🔍 6. SWOT Analysis

> Strengths (S)

  • 🛡️ Uncensored Fidelity: Surgically patched to ensure maximum generation throughput without alignment overhead.
  • ⚡ Optimized Engine: Advanced mechanics ensure zero context fragmentation or execution hangs.

> Weaknesses (W)

  • 📉 Hardware Limits: Requires sufficient VRAM/RAM for higher precision GGUF quantizations.

> Opportunities (O)

  • 🎯 Local Sovereign Agents: Perfect for offline, private reasoning and agentic workflows.

> Threats (T)

  • ⚠️ Sampler Sensitivity: High temperatures may require repetition penalty adjustments.

⚡ 7. Usage & Deployment Info

> Recommended Settings

  • Temperature: 0.2 - 0.7
  • Top-P: 0.95
  • Backend Engines: Compatible with llama.cpp, vLLM, Ollama, LM Studio, KoboldCPP

⚙️ 8. Backend Compatibility

> Validated Engines:

  • [+] llama.cpp: Native support across all quantizations.
  • [+] Ollama / LM Studio: Full GGUF compatibility.

📜 9. Disclaimers & Credits

Disclaimer: Qwen2.5-Coder-32B-Instruct-Uncensored-Patched is provided for research and sovereign local deployment. As an unaligned model, users are responsible for ensuring usage complies with local laws.

Credits: Gratitude to original base model authors (Qwen/Qwen2.5-Coder-32B-Instruct) and open-source AI community tools.

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-22Fix lineage, tags, and model card metadataad36b7317.6 KB
    Loading...
  2. 2026-07-18Overhaul model card with pentesting UI1a781e922.5 KB
    Loading...
  3. 2026-07-17Upload README.md with huggingface_hub898a42b1.8 KB
    Loading...
  4. 2026-07-17Upload README.md with huggingface_hub2d571df207 B
    Loading...
  5. 2026-07-17Upload README.md with huggingface_hubcc08d9b340 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration