← back to catalog · registered 2026-08-22 13:56

QuantFactory/llama2_7b_chat_uncensored-GGUF

QuantFactory Llama GGUF 2K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/QuantFactory%2Fllama2_7b_chat_uncensored-GGUF"
Response includes
  • classification m8
  • files 16
  • hub_downloads_all_time 6,694
  • author_summary 48 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 2 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • author=quantfactory (M8 quantization producer)
  • is_gguf=1
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
7K
276 last 30d - cooling
Likes
4
Model age
23mo ago
created 2024-11-07

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now6.8K→from598↑1,034%
02.5K5K7.4K598 on Nov 6, 20246.8K on Oct 11Nov '24Feb '25May '25Aug '25Nov '25FebMayAug
Nov 6, 2024 → Oct 11 · 140 snapshots · spans 704 days

Metadata

License
other
Quantizations
Q2_K Q3_K Q4 Q4_K Q5 Q5_K Q6_K Q8_0
Tags
gguf dataset:georgesung/wizard_vicuna_70k_unfiltered license:other endpoints_compatible region:us

Related

Total size
56.1 GB
Files
16
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2024-11-07 06:21

Files by quantization

Q8_0 1 file 6.67 GB
llama2_7b_chat_uncensored.Q8_0.gguf 6.67 GB e3525608 download
Q6_K 1 file 5.15 GB
llama2_7b_chat_uncensored.Q6_K.gguf 5.15 GB 0cfc0643 download
Q5 2 files 9.05 GB
llama2_7b_chat_uncensored.Q5_1.gguf 4.72 GB 52d808ac download
llama2_7b_chat_uncensored.Q5_0.gguf 4.33 GB b7b7c0c1 download
Q5_K 2 files 8.79 GB
llama2_7b_chat_uncensored.Q5_K_M.gguf 4.45 GB 79c30a5a download
llama2_7b_chat_uncensored.Q5_K_S.gguf 4.33 GB 1cabaf45 download
Q4 2 files 7.51 GB
llama2_7b_chat_uncensored.Q4_1.gguf 3.95 GB cb9f22b8 download
llama2_7b_chat_uncensored.Q4_0.gguf 3.56 GB 5d3a2bb8 download
Q4_K 2 files 7.39 GB
llama2_7b_chat_uncensored.Q4_K_M.gguf 3.80 GB 14a798ea download
llama2_7b_chat_uncensored.Q4_K_S.gguf 3.59 GB 3d9fabd4 download
Q3_K 3 files 9.17 GB
llama2_7b_chat_uncensored.Q3_K_L.gguf 3.35 GB 46578c61 download
llama2_7b_chat_uncensored.Q3_K_M.gguf 3.07 GB 4ff79df1 download
llama2_7b_chat_uncensored.Q3_K_S.gguf 2.75 GB 1ab763d5 download
Q2_K 1 file 2.36 GB
llama2_7b_chat_uncensored.Q2_K.gguf 2.36 GB a9e8beff download
Auxiliary files 2 files 4.99 KB
README.md 2.51 KB 5c0369d2 download
.gitattributes 2.48 KB d1d7680d download

README current version from Hugging Face


license: other
datasets:

  • georgesung/wizard_vicuna_70k_unfiltered

QuantFactory Banner

QuantFactory/llama2_7b_chat_uncensored-GGUF

This is quantized version of georgesung/llama2_7b_chat_uncensored created using llama.cpp

Original Model Card

Overview

Fine-tuned Llama-2 7B with an uncensored/unfiltered Wizard-Vicuna conversation dataset (originally from ehartford/wizard_vicuna_70k_unfiltered).
Used QLoRA for fine-tuning. Trained for one epoch on a 24GB GPU (NVIDIA A10G) instance, took ~19 hours to train.

The version here is the fp16 HuggingFace model.

GGML & GPTQ versions

Thanks to TheBloke, he has created the GGML and GPTQ versions:

Running in Ollama

https://ollama.com/library/llama2-uncensored

Prompt style

The model was trained with the following prompt style:

### HUMAN:
Hello

### RESPONSE:
Hi, how are you?

### HUMAN:
I'm fine.

### RESPONSE:
How can I help you?
...

Training code

Code used to train the model is available here.

To reproduce the results:

git clone https://github.com/georgesung/llm_qlora
cd llm_qlora
pip install -r requirements.txt
python train.py configs/llama2_7b_chat_uncensored.yaml

Fine-tuning guide

https://georgesung.github.io/ai/qlora-ift/

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 43.39
ARC (25-shot) 53.58
HellaSwag (10-shot) 78.66
MMLU (5-shot) 44.49
TruthfulQA (0-shot) 41.34
Winogrande (5-shot) 74.11
GSM8K (5-shot) 5.84
DROP (3-shot) 5.69

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-11-07Upload README.md with huggingface_hub6a1b6642.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration