← back to catalog · registered 2026-08-22 13:56

QuantFactory/VersatiLlama-Llama-3.2-3B-Instruct-Abliterated-GGUF

QuantFactory Llama 3B GGUF 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/QuantFactory%2FVersatiLlama-Llama-3.2-3B-Instruct-Abliterated-GGUF"
Response includes
  • classification m8
  • files 16
  • benchmarks 21 entries
  • hub_downloads_all_time 47,103
  • author_summary 48 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

No other method signals detected in this model.
Confidence
HIGH
Why this label 3 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=quantfactory (M8 quantization producer)
  • is_gguf=1
  • base_model='meta-llama/Llama-3.2-3B-Instruct' (source unknown method)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
47K
2K last 30d - cooling
Likes
6
Model age
2.0y ago
created 2024-10-08
Downloads over time
Now48.2K→from67↑71,837%
017.7K35.3K53K67 on Oct 2, 202448.2K on Oct 11Oct '24Feb '25Jun '25Oct '25FebJunOct
Oct 2, 2024 → Oct 11 · 148 snapshots · spans 739 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Arena-Battles 8390 LM-Arena
LM Arena Elo 1118.1163530112908 LM-Arena
Arena-Elo-Lower 1110.9097939581304 LM-Arena
Arena-Elo-Upper 1125.322912064451 LM-Arena
Arena-Rank 181 LM-Arena
BBH average 0.4202879750940459 OpenLLM-v2
IFEval instruct 0.7817745803357314 OpenLLM-v2
IFEval-Prompt 0.6968576709796673 OpenLLM-v2
MATH lvl 5 0.1555891238670695 OpenLLM-v2
MMLU-Pro 0.3194813829787234 OpenLLM-v2
Entertainment 0.5 UGI
Hazardous 0 UGI
Natural Intelligence 10.45 UGI
Political lean -1.2% UGI
Sensitive-Info 4.49 UGI
SocPol 0.7 UGI
UGI 7.16 UGI
Willingness (10) 1.2 UGI
W10-Adherence 0.5 UGI
W10-Direct 2 UGI
Writing 10.44 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
cc-by-4.0
Languages
en
Quantizations
Q2_K Q3_K Q4 Q4_K Q5 Q5_K Q6_K Q8_0
Tags
transformers gguf text-generation en base_model:meta-llama/Llama-3.2-3B-Instruct base_model:quantized:meta-llama/Llama-3.2-3B-Instruct license:cc-by-4.0 endpoints_compatible region:us conversational

Related

Total size
27.7 GB
Files
16
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2024-10-08 02:20

Files by quantization

Q8_0 1 file 3.19 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q8_0.gguf 3.19 GB 42a7248f download
Q6_K 1 file 2.46 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q6_K.gguf 2.46 GB 5f935900 download
Q5 2 files 4.39 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q5_1.gguf 2.28 GB c6d6fffc download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q5_0.gguf 2.11 GB 24c69b90 download
Q5_K 2 files 4.28 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q5_K_M.gguf 2.16 GB 3142b54f download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q5_K_S.gguf 2.11 GB b4bf2854 download
Q4 2 files 3.74 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q4_1.gguf 1.95 GB 82dfeef4 download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q4_0.gguf 1.79 GB 7f7062bb download
Q4_K 2 files 3.68 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q4_K_M.gguf 1.88 GB 15b9e4a9 download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q4_K_S.gguf 1.80 GB 8e29c12f download
Q3_K 3 files 4.70 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q3_K_L.gguf 1.69 GB 3a8c1817 download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q3_K_M.gguf 1.57 GB 35e45ca7 download
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q3_K_S.gguf 1.44 GB c69e3257 download
Q2_K 1 file 1.27 GB
VersatiLlama-Llama-3.2-3B-Instruct-Abliterated.Q2_K.gguf 1.27 GB 5abea3f5 download
Auxiliary files 2 files 4.92 KB
.gitattributes 2.77 KB a09f1aff download
README.md 2.15 KB 0fe81a02 download

README current version from Hugging Face


base_model:

  • meta-llama/Llama-3.2-3B-Instruct
    license: cc-by-4.0
    language:
  • en
    pipeline_tag: text-generation
    library_name: transformers

QuantFactory Banner

QuantFactory/VersatiLlama-Llama-3.2-3B-Instruct-Abliterated-GGUF

This is quantized version of Devarui379/VersatiLlama-Llama-3.2-3B-Instruct-Abliterated created using llama.cpp

Original Model Card

Model Card for Model ID

VersatiLlama-Llama-3.2-3B-Instruct-Abliterated

image/webp

Model Description

Small but Smart

Fine-Tuned on Vast dataset of Conversations

Able to Generate Human like text with high performance within its size.

It is Very Versatile when compared for it's size and Parameters and offers capability almost as good as Llama 3.1 8B Instruct

Feel free to Check it out!!

Check the quantized model here: Devarui379/VersatiLlama-Llama-3.2-3B-Instruct-Abliterated-Imatrix-GGUF

[This model was trained for 5hrs on GPU T4 15gb vram]

  • Developed by: Meta AI
  • Fine-Tuned by: Devarui379
  • Model type: Transformers
  • Language(s) (NLP): English
  • License: cc-by-4.0

Model Sources [optional]

base model:meta-llama/Llama-3.2-3B-Instruct

Uses

Use desired System prompt when using in LM Studio
The optimal chat template seems to be Jinja but feel free to test it out as you want!

Technical Specifications

Model Architecture and Objective

Llama 3.2

Hardware

NVIDIA TESLA T4 GPU 15GB VRAM

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-10-08Upload README.md with huggingface_hub15f240f2.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration