← back to catalog · registered 2026-08-22 13:56

Andycurrent/DeepSeek-R1-Distill-Qwen-7B-Uncensored_GGUF

Andycurrent Qwen 7B GGUF 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Andycurrent%2FDeepSeek-R1-Distill-Qwen-7B-Uncensored_GGUF"
Response includes
  • classification m-uncensored
  • files 9
  • benchmarks 5 entries
  • hub_downloads_all_time 9,765
  • author_summary 13 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
10K
960 last 30d - cooling
Likes
5
Model age
8mo ago
created 2026-01-30
Downloads over time
Now10.1K→from0↑0%
03.7K7.4K11.1K0 on Jan 2810.1K on Oct 11JanMarMayJulSep
Jan 28 → Oct 11 · 78 snapshots · spans 256 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
BBH average 0.33389544688026984 OpenLLM-v2
IFEval instruct 0.4748201438848921 OpenLLM-v2
IFEval-Prompt 0.33271719038817005 OpenLLM-v2
MATH lvl 5 0 OpenLLM-v2
MMLU-Pro 0.2321309840425532 OpenLLM-v2

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh
Quantizations
F16 Q8_0
Tags
gguf Reasoning Instruct Uncensored Distilled GGUF Quantized en zh base_model:deepseek-ai/DeepSeek-R1-Distill-Qwen-7B base_model:quantized:deepseek-ai/DeepSeek-R1-Distill-Qwen-7B license:apache-2.0

Related

Total size
10.2 GB
Files
9
Quantizations
3
Registered
2026-08-22 13:56
Last updated on HF
2026-01-30 11:45

Files by quantization

F16 1 file 3.32 GB
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_F16.gguf 3.32 GB 7b969ee4 download
Q8_0 1 file 1.76 GB
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q8_0.gguf 1.76 GB a83031e4 download
Auxiliary files 7 files 5.16 GB
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q6_k.gguf 1.36 GB 06e2a8c4 download
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q5_k_m.gguf 1.20 GB 0de398f9 download
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q4_k_m.gguf 1.04 GB e23a4aab download
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q3_k_m.gguf 882 MB bc421eb3 download
DeepSeek-R1-Distill-Qwen-1.5B-uncensored_Q2_k.gguf 718 MB 3bac3653 download
README.md 4.48 KB e51ba374 download
.gitattributes 2.08 KB 65a4bbfd download

README current version from Hugging Face


license: apache-2.0
language:

  • en
  • zh
    base_model:
  • deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
    tags:
  • Reasoning
  • Instruct
  • Uncensored
  • Distilled
  • GGUF
  • Quantized

DeepSeek-R1-Distill-Qwen-7B-Uncensored

This repository hosts uncensored and efficiency-focused builds of DeepSeek-R1-Distill-Qwen-7B, intended for users who require direct model behavior, strong reasoning, and full local control without aggressive automated filtering.

The model is suitable for advanced experimentation, private deployments, and research scenarios where transparency and flexibility are prioritized.


Model Overview

  • Model Name: DeepSeek-R1-Distill-Qwen-7B-Uncensored
  • Base Model: DeepSeek-R1-Distill-Qwen-7B
  • Architecture: Decoder-only Transformer
  • Parameter Count: ~7B
  • Modalities: Text
  • Context Length: Up to 32K tokens (runtime dependent)
  • Developer (Base): DeepSeek AI
  • Distillation Target: Qwen-based reasoning model
  • License: Apache-2.0 (inherits base model license)
  • Languages: Multilingual (English, Chinese, others)

Project Intent

This release is designed for users who want minimal behavioral constraints while preserving the structured reasoning and instruction-following strengths of the DeepSeek-R1 distillation.

Key objectives include:

  • Predictable, direct responses without heavy content suppression
  • Strong multi-step reasoning and analytical depth
  • Compatibility with local and offline inference setups
  • A solid foundation for further alignment, fine-tuning, or research

This is not a consumer-safety-aligned assistant and is intended for controlled environments.


Quantized Variants (GGUF)

To support a wide range of hardware, multiple GGUF quantization levels are provided.

Q2_K (2-bit)

  • Extremely small memory footprint
  • Intended for experimentation or extreme hardware constraints
  • Severe degradation in reasoning and instruction accuracy

Q3_K_M (3-bit)

  • Slight improvement over 2-bit
  • Lightweight and fast
  • Limited suitability for complex reasoning tasks

Q4_K_M (4-bit)

  • Strong efficiency-to-quality tradeoff
  • Works well on CPUs and low-VRAM GPUs
  • Suitable for general chat and exploratory reasoning

Q5_K_M (5-bit)

  • Recommended default for most users
  • Retains most reasoning and instruction-following ability
  • Balanced memory usage and output quality

Q6_K (6-bit)

  • Higher reasoning fidelity
  • Increased memory requirements
  • Better performance on long or complex prompts

Q8_0 (8-bit)

  • Near full-precision behavior
  • Highest quality quantized variant
  • Best choice when memory is not a limiting factor

Output quality depends heavily on context length, sampling parameters, and inference backend.


Prompting Format

The model performs best with a structured chat format:


<|system|>
High-level instructions or behavioral guidance
<|user|>
User prompt
<|assistant|>

Clear system messages are recommended to guide tone, verbosity, and task focus.


Suggested Settings

  • Temperature: 0.6 – 0.8 for analytical tasks
  • Use Q5_K_M or higher for reasoning-heavy prompts
  • Avoid ultra-low-bit quantizations for long-context analysis

Capabilities

  • Strong logical and mathematical reasoning
  • Effective multi-step analysis and planning
  • Clear instruction-following behavior
  • Suitable for research into reasoning and alignment
  • Performs well in uncensored local deployments
  • Maintains coherence over extended conversations

Recommended Use Cases

  • Local reasoning assistants
  • Research and alignment studies
  • Offline analysis and experimentation
  • Advanced prompt engineering workflows
  • Private deployments requiring full user control

Important Notes

  • This model intentionally avoids strong automated moderation
  • Users are responsible for ensuring lawful and ethical usage
  • Not recommended for unsupervised or public-facing applications
  • Quantized variants may hallucinate more than full-precision models

Always evaluate outputs in the context of your intended application.


Acknowledgements

  • DeepSeek AI for releasing the DeepSeek-R1 model family
  • Qwen team for the underlying architecture contributions
  • The llama.cpp and GGUF ecosystem for enabling efficient local inference
  • Open-source contributors supporting transparent LLM research

Contact

For issues related to quantization files or repository content, please open an issue in this repository.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-01-30Update README.mdd2de22e4.5 KB
    Loading...
  2. 2026-01-30Update README.mda244e784.7 KB
    Loading...
  3. 2026-01-30initial commit7465cda21 B
    Loading...

Discussions 1 thread

  1. 2026-03-16Not a 7B modelopen1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration