← back to catalog · registered 2026-08-22 13:56

Andycurrent/Qwen2.5-7B-Instruct-Uncensored_GGUF

Andycurrent Qwen 7B GGUF second-order 33K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Andycurrent%2FQwen2.5-7B-Instruct-Uncensored_GGUF"
Response includes
  • classification m-uncensored
  • files 9
  • benchmarks 5 entries
  • hub_downloads_all_time 10,425
  • author_summary 13 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
10K
1K last 30d - stable
Likes
5
Model age
8mo ago
created 2026-01-21
Downloads over time
Now11.1K→from1.6K↑579%
1.2K4.8K8.4K12K1.6K on Jan 2111.1K on Oct 11JanMarMayJulSep
Jan 21 → Oct 11 · 79 snapshots · spans 263 days

Benchmarks

Benchmark Score Source
BBH average 0.49643274095213386 OpenLLM-v2
IFEval instruct 0.7661870503597122 OpenLLM-v2
IFEval-Prompt 0.6746765249537893 OpenLLM-v2
MATH lvl 5 0.013595166163141994 OpenLLM-v2
MMLU-Pro 0.4426529255319149 OpenLLM-v2

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Quantizations
F16 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf conversational chat agent instruction-tunned text-generation en base_model:Orion-zhen/Qwen2.5-7B-Instruct-Uncensored base_model:quantized:Orion-zhen/Qwen2.5-7B-Instruct-Uncensored license:apache-2.0 endpoints_compatible region:us

Related

Total size
43.3 GB
Files
9
Quantizations
8
Registered
2026-08-22 13:56
Last updated on HF
2026-01-22 05:15

Files by quantization

F16 1 file 14.2 GB
Qwen2.5-7B-Instruct-Uncensored_F16.gguf 14.2 GB 6a353c70 download
Q8_0 1 file 7.54 GB
Qwen2.5-7B-Instruct-Uncensored_Q8_0.gguf 7.54 GB bb5090fb download
Q6_K 1 file 5.82 GB
Qwen2.5-7B-Instruct-Uncensored_Q6_K.gguf 5.82 GB 1369ccc3 download
Q5_K 1 file 5.07 GB
Qwen2.5-7B-Instruct-Uncensored_Q5_K_M.gguf 5.07 GB a4095a2a download
Q4_K 1 file 4.36 GB
Qwen2.5-7B-Instruct-Uncensored_Q4_K_M.gguf 4.36 GB 11b11835 download
Q3_K 1 file 3.55 GB
Qwen2.5-7B-Instruct-Uncensored_Q3_K_M.gguf 3.55 GB 1dda0b55 download
Q2_K 1 file 2.81 GB
Qwen2.5-7B-Instruct-Uncensored_Q2_K.gguf 2.81 GB 7fc7c045 download
Auxiliary files 2 files 5.44 KB
README.md 3.42 KB 0b3222fd download
.gitattributes 2.01 KB 7f920c93 download

README current version from Hugging Face


license: apache-2.0
language:

  • en
    base_model:
  • Orion-zhen/Qwen2.5-7B-Instruct-Uncensored
    tags:
  • conversational
  • chat
  • agent
  • instruction-tunned
    pipeline_tag: text-generation

Qwen2.5-7B-Instruct-Uncensored

Qwen2.5-7B-Instruct-Uncensored is a 7-billion-parameter instruction-following language model designed for open-ended interaction, research experimentation, and local deployment scenarios where users require minimal alignment constraints and maximal behavioral flexibility.

This model is intended for technically proficient users who want direct control over prompting, alignment, and downstream usage without heavy built-in moderation layers.


Model Summary

  • Model Name: Qwen2.5-7B-Instruct-Uncensored
  • Base Architecture: Qwen2.5-7B
  • Maintainer: Orion-zhen
  • Parameter Count: 7B
  • Model Type: Decoder-only transformer, instruction-tuned
  • License: Inherits the license terms of the original Qwen2.5 base model
  • Primary Focus: Open instruction following with reduced safety filtering

Design Philosophy

This release emphasizes instruction fidelity and conversational openness over restrictive alignment.
The uncensored variant is designed to:

  • Respond directly to user instructions without excessive refusal patterns
  • Support experimentation with prompt engineering and alignment research
  • Enable private, offline, or air-gapped deployments
  • Serve as a flexible base for further fine-tuning or specialization

Instruction Format

For best results, interactions should follow a structured chat format compatible with Qwen-style instruction tuning:

<|system|>
Optional system-level guidance or role definition
<|user|>
User input or task description
<|assistant|>
Model response

Clear role separation improves consistency, especially in multi-turn conversations and complex reasoning tasks.


Core Capabilities

  • Strong adherence to explicit user instructions
  • Capable of multi-step reasoning and long-form responses
  • Performs well in coding, analysis, writing, and ideation tasks
  • Suitable for creative generation, simulations, and role-based interactions
  • Stable in extended dialogues without excessive context loss
  • Compatible with local inference stacks and quantized runtimes

Suggested Applications

  • Local AI assistants for private workflows
  • Research environments studying model behavior and alignment
  • Developer tooling such as code explanation and generation
  • Creative projects including storytelling and world-building
  • Prompt engineering experimentation
  • Offline or privacy-sensitive deployments

Responsible Usage Notice

This model intentionally minimizes automated content restrictions.
Users are responsible for ensuring that their usage complies with applicable laws, regulations, and ethical standards.

It is recommended only for users who understand the implications of operating uncensored language models.


Deployment Notes

  • Best suited for self-hosted or research environments
  • Not recommended for unattended public-facing services
  • Works well with standard transformer inference frameworks
  • Supports further fine-tuning and alignment layering if desired

Acknowledgements

Thanks to the Qwen development team for releasing the base architecture and to the open-source community for providing tools, evaluations, and infrastructure that make experimentation with large language models accessible.


README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-01-22Update README.md39d7a133.4 KB
    Loading...
  2. 2026-01-22Update README.mddb838143.4 KB
    Loading...
  3. 2026-01-22Update README.md88ee0933.4 KB
    Loading...
  4. 2026-01-22Update README.md36b93173.3 KB
    Loading...
  5. 2026-01-22Update README.md5c952663.5 KB
    Loading...
  6. 2026-01-21initial commitbdd7a1128 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration