← back to catalog · registered 2026-08-22 13:56

XythicK/Mag-Mell-R1-Uncensored-21B-GGUF

XythicK 21B GGUF second-order 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/XythicK%2FMag-Mell-R1-Uncensored-21B-GGUF"
Response includes
  • classification m-uncensored
  • files 14
  • hub_downloads_all_time 5,522
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
6K
971 last 30d - stable
Likes
0
Model age
9mo ago
created 2025-12-28
Downloads over time
Now5.8K→from515↑1,036%
2482.3K4.3K6.4K515 on Dec 31, 20255.8K on Oct 11Dec '25FebAprJunAugOct
Dec 31, 2025 → Oct 11 · 80 snapshots · spans 284 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Languages
en
Tags
transformers gguf quantization llm inference efficient-llm text-generation en endpoints_compatible region:us conversational

Related

Total size
144 GB
Files
14
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-12-28 17:53

Files by quantization

Auxiliary files 14 files 144 GB
mag-mell-r1-uncensored-21b-q8_0.gguf 20.2 GB cdc3ced2 download
mag-mell-r1-uncensored-21b-q6_k.gguf 15.6 GB 28464a15 download
mag-mell-r1-uncensored-21b-q5_k_m.gguf 13.5 GB 89a506e7 download
mag-mell-r1-uncensored-21b-q5_0.gguf 13.2 GB e898a58a download
mag-mell-r1-uncensored-21b-q5_k_s.gguf 13.2 GB 752b7c88 download
mag-mell-r1-uncensored-21b-q4_k_m.gguf 11.5 GB 8b512a15 download
mag-mell-r1-uncensored-21b-q4_k_s.gguf 10.9 GB 8f2a72f4 download
mag-mell-r1-uncensored-21b-q4_0.gguf 10.9 GB 5fbb3d18 download
mag-mell-r1-uncensored-21b-q3_k_l.gguf 10.1 GB 05c6d331 download
mag-mell-r1-uncensored-21b-q3_k_m.gguf 9.33 GB 89d26735 download
mag-mell-r1-uncensored-21b-q3_k_s.gguf 8.43 GB d7cdbd19 download
mag-mell-r1-uncensored-21b-q2_k.gguf 7.26 GB 1389932d download
README.md 3.48 KB 65bec321 download
.gitattributes 2.35 KB 17f2453d download

README current version from Hugging Face


base_model:

  • JustOnion/Mag-Mell-R1-Uncensored-21B
    language:
  • en
    library_name: transformers
    pipeline_tag: text-generation
    tags:
  • quantization
  • gguf
  • llm
  • inference
  • efficient-llm

Mag-Mell-R1-Uncensored-21B-GGUF

🧠 Model Overview

Mag-Mell-R1-Uncensored-21B-GGUF is a quantized version of Mag-Mell-R1-Uncensored-21B, optimized for efficient inference with reduced memory usage and faster runtime while preserving as much of the original model quality as possible.

This repository provides multiple quantized variants suitable for:

  • Local inference
  • Low-VRAM GPUs
  • CPU-only environments

🔗 Original Model


📦 Quantization Details

  • Quantization method: GGUF
  • Quantization tool: llama.cpp
  • Precision: Mixed (2-8,bit depands in variant)
  • Activation aware: No (weight-only quantinization)
  • Group size: 256 (K-quant variants)

📦 Available Quantized Files

Quant Format File Name Approx. Size VRAM / RAM Needed Notes
Q2_K mag-mell-r1-uncensored-21b-q2_k.gguf ~7.8 GB ~8 GB Extreme compression; noticeable quality loss
Q3_K_S mag-mell-r1-uncensored-21b-q3_k_s.gguf ~9 GB ~10 GB Smaller, faster, lower quality
Q3_K_M mag-mell-r1-uncensored-21b-q3_k_m.gguf ~10 GB ~11 GB Better balance than Q3_K_S
Q3_K_L mag-mell-r1-uncensored-21b-q3_k_l.gguf ~10.8 GB ~11.5 GB Highest-quality 3-bit variant
Q4_0 mag-mell-r1-uncensored-21b-q4_0.gguf ~11.7 GB ~12.9 GB Legacy format; simpler quantization
Q4_K_S mag-mell-r1-uncensored-21b-q4_k_s.gguf ~11.7 GB ~13 GB Smaller grouped 4-bit
Q4_K_M mag-mell-r1-uncensored-21b-q4_k_m.gguf ~12.4 GB ~14 GB Recommended default
Q5_0 mag-mell-r1-uncensored-21b-q5_0.gguf ~14.1 GB ~16 GB Higher quality, larger size
Q5_K_S mag-mell-r1-uncensored-21b-q5_k_s.gguf ~14 GB ~15.1 GB Efficient high-quality variant
Q5_K_M mag-mell-r1-uncensored-21b-q5_K_M.gguf ~14.5 GB ~16 GB Near-FP16 quality
Q6_K mag-mell-r1-uncensored-21b-q6_k.gguf ~16.8 GB ~18 GB Minimal quantization loss
Q8_0 mag-mell-r1-uncensored-21b-q8_0.gguf ~21.6 GB ~23 GB Maximum quality; large memory

💡 Recommendation: Start with Q4_K_M for the best quality-to-performance ratio.


🚀 Usage Example

llama.cpp

./main -m mag-mell-r1-uncensored-21b-q5_0.gguf -p "The World is beautiful isn't it?" -n 256

Python (llama-cpp-python)

from llama_cpp import Llama

llm = Llama(
    model_path="<MODEL_FILE>.gguf",
    n_ctx=4096,
    n_threads=8
)

print(llm("Your prompt here"))

🙋 Contact

Maintainer: M Mashhudur Rahim [XythicK]

Role:
Independent Machine Learning Researcher & Model Infrastructure Maintainer

(Focused on model quantization, optimization, and efficient deployment)

For issues, improvement requests, or additional quantization formats, please use the Hugging Face Discussions or Issues tab.

❤️ Acknowledgements

Thanks to the original model authors for their ongoing contributions to open AI research, and to Hugging Face and the open-source machine learning community for providing the tools and platforms that make efficient model sharing and deployment possible.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-12-28Update README.mde9b25ac3.5 KB
    Loading...
  2. 2025-12-28Update README.md1f5c41b3.5 KB
    Loading...
  3. 2025-12-28Create README.md0bf8c7658 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration