← back to catalog · registered 2026-08-22 13:56

QuantFactory/Luna-AI-Llama2-Uncensored-GGUF

QuantFactory Llama GGUF 2K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/QuantFactory%2FLuna-AI-Llama2-Uncensored-GGUF"
Response includes
  • classification m8
  • files 16
  • hub_downloads_all_time 4,353
  • author_summary 48 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 2 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • author=quantfactory (M8 quantization producer)
  • is_gguf=1
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
4K
119 last 30d - cooling
Likes
2
Model age
2.1y ago
created 2024-09-02
Downloads over time
Now4.4K→from190↑2,213%
01.6K3.2K4.8K190 on Aug 28, 20244.4K on Oct 11Aug '24Dec '24Apr '25Aug '25Dec '25AprAug
Aug 28, 2024 → Oct 11 · 150 snapshots · spans 774 days

Metadata

Quantizations
Q2_K Q3_K Q4 Q4_K Q5 Q5_K Q6_K Q8_0
Tags
gguf license:cc-by-sa-4.0 endpoints_compatible region:us

Related

Total size
56.1 GB
Files
16
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2024-09-02 04:01

Files by quantization

Q8_0 1 file 6.67 GB
Luna-AI-Llama2-Uncensored.Q8_0.gguf 6.67 GB 0ae85885 download
Q6_K 1 file 5.15 GB
Luna-AI-Llama2-Uncensored.Q6_K.gguf 5.15 GB 99dfa4ea download
Q5 2 files 9.05 GB
Luna-AI-Llama2-Uncensored.Q5_1.gguf 4.72 GB 828e2457 download
Luna-AI-Llama2-Uncensored.Q5_0.gguf 4.33 GB 5b8476a0 download
Q5_K 2 files 8.79 GB
Luna-AI-Llama2-Uncensored.Q5_K_M.gguf 4.45 GB 9f5dc77e download
Luna-AI-Llama2-Uncensored.Q5_K_S.gguf 4.33 GB 86495e8f download
Q4 2 files 7.51 GB
Luna-AI-Llama2-Uncensored.Q4_1.gguf 3.95 GB 409c6d36 download
Luna-AI-Llama2-Uncensored.Q4_0.gguf 3.56 GB f277ad97 download
Q4_K 2 files 7.39 GB
Luna-AI-Llama2-Uncensored.Q4_K_M.gguf 3.80 GB cab6b14b download
Luna-AI-Llama2-Uncensored.Q4_K_S.gguf 3.59 GB 0b92bb3d download
Q3_K 3 files 9.17 GB
Luna-AI-Llama2-Uncensored.Q3_K_L.gguf 3.35 GB baa96708 download
Luna-AI-Llama2-Uncensored.Q3_K_M.gguf 3.07 GB 472755a2 download
Luna-AI-Llama2-Uncensored.Q3_K_S.gguf 2.75 GB 19bea421 download
Q2_K 1 file 2.36 GB
Luna-AI-Llama2-Uncensored.Q2_K.gguf 2.36 GB 49e7649b download
Auxiliary files 2 files 4.28 KB
.gitattributes 2.48 KB 1e874b3c download
README.md 1.80 KB 4865d598 download

README current version from Hugging Face


license: cc-by-sa-4.0


QuantFactory/Luna-AI-Llama2-Uncensored-GGUF

This is quantized version of Tap-M/Luna-AI-Llama2-Uncensored created using llama.cpp

Original Model Card

Model Description

“Luna AI Llama2 Uncensored” is a Llama2 based Chat model
fine-tuned on over 40,000 long form chat discussions
This model was fine-tuned by Tap, the creator of Luna AI.

Model Training

The fine-tuning process was performed on an 8x a100 80GB machine.
The model was trained on synthetic outputs which include multiple rounds of chats between Human & AI.

4bit GPTQ Version provided by @TheBloke - for GPU inference

GGML Version provided by @TheBloke - For CPU inference

Prompt Format

The model follows the Vicuna 1.1/ OpenChat format:

USER: I have difficulties in making friends, and I really need someone to talk to. Would you be my friend?

ASSISTANT: Of course! Friends are always here for each other. What do you like to do?

Benchmark Results

Task Version Metric Value Stderr
arc_challenge 0 acc_norm 0.5512 0.0146
hellaswag 0
mmlu 1 acc_norm 0.46521 0.036
truthfulqa_mc 1 mc2 0.4716 0.0155
Average - - 0.5114 0.0150

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-09-02Upload README.md with huggingface_hubc7657701.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration