← back to catalog · registered 2026-08-22 13:56

legraphista/Meta-Llama-3-8B-Instruct-abliterated-v3-IMat-GGUF

legraphista Llama 8B GGUF second-order 8K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/legraphista%2FMeta-Llama-3-8B-Instruct-abliterated-v3-IMat-GGUF"
Response includes
  • classification m8
  • files 30
  • benchmarks 16 entries
  • hub_downloads_all_time 46,635
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
47K
12K last 30d - stable
Likes
1
Model age
2.4y ago
created 2024-05-31
Downloads over time
Now48.8K→from427↑11,322%
017.9K35.8K53.6K427 on Jul 24, 202448.8K on Oct 11Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 156 snapshots · spans 809 days

Benchmarks

Benchmark Score Source
BBH average 0.44272927746789464 OpenLLM-v2
IFEval instruct 0.7649880095923262 OpenLLM-v2
IFEval-Prompt 0.6839186691312384 OpenLLM-v2
MATH lvl 5 0.09592145015105741 OpenLLM-v2
MMLU-Pro 0.3653590425531915 OpenLLM-v2
Entertainment 1.8 UGI
Hazardous 1.8 UGI
Natural Intelligence 11.73 UGI
Political lean -21.1% UGI
Sensitive-Info 16.86 UGI
SocPol 1.5 UGI
UGI 29.57 UGI
Willingness (10) 5.5 UGI
W10-Adherence 4 UGI
W10-Direct 7 UGI
Writing 19.6 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
llama3
Quantizations
BF16 F16 IQ1 IQ2 IQ3 IQ4 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf quantized GGUF imatrix quantization imat static 16bit 8bit 6bit 5bit 4bit

Related

Total size
116 GB
Files
30
Quantizations
13
Registered
2026-08-22 13:56
Last updated on HF
2024-06-01 12:09

Files by quantization

BF16 1 file 15.0 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.BF16.gguf 15.0 GB 2ba705e9 download
F16 1 file 15.0 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.FP16.gguf 15.0 GB 77e33add download
Q8_0 1 file 7.95 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf 7.95 GB d015402b download
Q6_K 1 file 6.14 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q6_K.gguf 6.14 GB c4c9aaca download
Q5_K 2 files 10.6 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q5_K.gguf 5.34 GB 2e669678 download
Meta-Llama-3-8B-Instruct-abliterated-v3.Q5_K_S.gguf 5.21 GB 8530aed2 download
Q4_K 2 files 8.95 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q4_K.gguf 4.58 GB 6161c05d download
Meta-Llama-3-8B-Instruct-abliterated-v3.Q4_K_S.gguf 4.37 GB 6a394827 download
IQ4 2 files 8.50 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ4_NL.gguf 4.36 GB 6d84d6b8 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ4_XS.gguf 4.14 GB 58707129 download
Q3_K 3 files 11.2 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K_L.gguf 4.03 GB 60186a00 download
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K.gguf 3.74 GB ffb48e8e download
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K_S.gguf 3.41 GB e1f53692 download
IQ3 4 files 13.3 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_M.gguf 3.52 GB 0ba9ff83 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_S.gguf 3.43 GB 9a431cf8 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_XS.gguf 3.28 GB 63410638 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_XXS.gguf 3.05 GB ad438f90 download
Q2_K 2 files 5.74 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.Q2_K.gguf 2.96 GB b381af04 download
Meta-Llama-3-8B-Instruct-abliterated-v3.Q2_K_S.gguf 2.78 GB ff8432dc download
IQ2 4 files 9.98 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_M.gguf 2.75 GB 46fbb120 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_S.gguf 2.57 GB 92848027 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_XS.gguf 2.43 GB b656bbe3 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_XXS.gguf 2.23 GB e1c26b36 download
IQ1 2 files 3.89 GB
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ1_M.gguf 2.01 GB 6bb0d057 download
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ1_S.gguf 1.88 GB f2633ece download
Auxiliary files 5 files 5.05 MB
imatrix.dat 4.76 MB 3623855b download
imatrix.dataset 273 KB 41ae1656 download
README.md 11.9 KB 6724ccbf download
imatrix.log 10.2 KB 4550a02f download
.gitattributes 3.66 KB 6849eef0 download

README current version from Hugging Face


base_model: failspy/Meta-Llama-3-8B-Instruct-abliterated-v3
inference: false
library_name: gguf
license: llama3
pipeline_tag: text-generation
quantized_by: legraphista
tags:

  • quantized
  • GGUF
  • imatrix
  • quantization
  • imat
  • imatrix
  • static
  • 16bit
  • 8bit
  • 6bit
  • 5bit
  • 4bit
  • 3bit
  • 2bit
  • 1bit

Meta-Llama-3-8B-Instruct-abliterated-v3-IMat-GGUF

Llama.cpp imatrix quantization of failspy/Meta-Llama-3-8B-Instruct-abliterated-v3

Original Model: failspy/Meta-Llama-3-8B-Instruct-abliterated-v3
Original dtype: BF16 (bfloat16)
Quantized by: llama.cpp b3058
IMatrix dataset: here


Files

IMatrix

Status: ✅ Available
Link: here

Common Quants

Filename Quant type File Size Status Uses IMatrix Is Split
Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf Q8_0 8.54GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q6_K.gguf Q6_K 6.60GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q4_K.gguf Q4_K 4.92GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K.gguf Q3_K 4.02GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q2_K.gguf Q2_K 3.18GB ✅ Available 🟢 IMatrix 📦 No

All Quants

Filename Quant type File Size Status Uses IMatrix Is Split
Meta-Llama-3-8B-Instruct-abliterated-v3.BF16.gguf BF16 16.07GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.FP16.gguf F16 16.07GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf Q8_0 8.54GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q6_K.gguf Q6_K 6.60GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q5_K.gguf Q5_K 5.73GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q5_K_S.gguf Q5_K_S 5.60GB ✅ Available ⚪ Static 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q4_K.gguf Q4_K 4.92GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q4_K_S.gguf Q4_K_S 4.69GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ4_NL.gguf IQ4_NL 4.68GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ4_XS.gguf IQ4_XS 4.45GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K.gguf Q3_K 4.02GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K_L.gguf Q3_K_L 4.32GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q3_K_S.gguf Q3_K_S 3.66GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_M.gguf IQ3_M 3.78GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_S.gguf IQ3_S 3.68GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_XS.gguf IQ3_XS 3.52GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ3_XXS.gguf IQ3_XXS 3.27GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q2_K.gguf Q2_K 3.18GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.Q2_K_S.gguf Q2_K_S 2.99GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_M.gguf IQ2_M 2.95GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_S.gguf IQ2_S 2.76GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_XS.gguf IQ2_XS 2.61GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ2_XXS.gguf IQ2_XXS 2.40GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ1_M.gguf IQ1_M 2.16GB ✅ Available 🟢 IMatrix 📦 No
Meta-Llama-3-8B-Instruct-abliterated-v3.IQ1_S.gguf IQ1_S 2.02GB ✅ Available 🟢 IMatrix 📦 No

Downloading using huggingface-cli

If you do not have hugginface-cli installed:

pip install -U "huggingface_hub[cli]"

Download the specific file you want:

huggingface-cli download legraphista/Meta-Llama-3-8B-Instruct-abliterated-v3-IMat-GGUF --include "Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf" --local-dir ./

If the model file is big, it has been split into multiple files. In order to download them all to a local folder, run:

huggingface-cli download legraphista/Meta-Llama-3-8B-Instruct-abliterated-v3-IMat-GGUF --include "Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0/*" --local-dir ./
# see FAQ for merging GGUF's

Inference

Simple chat template

<|begin_of_text|><|start_header_id|>user<|end_header_id|>

{user_prompt}<|eot_id|><|start_header_id|>assistant<|end_header_id|>

{assistant_response}<|eot_id|><|start_header_id|>user<|end_header_id|>

{next_user_prompt}<|eot_id|>

Chat template with system prompt

<|begin_of_text|><|start_header_id|>system<|end_header_id|>

{system_prompt}<|eot_id|><|start_header_id|>user<|end_header_id|>

{user_prompt}<|eot_id|><|start_header_id|>assistant<|end_header_id|>

{assistant_response}<|eot_id|><|start_header_id|>user<|end_header_id|>

{next_user_prompt}<|eot_id|>

Llama.cpp

llama.cpp/main -m Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf --color -i -p "prompt here (according to the chat template)"

FAQ

Why is the IMatrix not applied everywhere?

According to this investigation, it appears that lower quantizations are the only ones that benefit from the imatrix input (as per hellaswag results).

How do I merge a split GGUF?

  1. Make sure you have gguf-split available
  2. Locate your GGUF chunks folder (ex: Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0)
  3. Run gguf-split --merge Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0/Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0-00001-of-XXXXX.gguf Meta-Llama-3-8B-Instruct-abliterated-v3.Q8_0.gguf
    • Make sure to point gguf-split to the first chunk of the split.

Got a suggestion? Ping me @legraphista!

README history 20 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-06-01Update README.mda5a8b7611.9 KB
    Loading...
  2. 2024-05-31Upload README.md with huggingface_hub6a67fa011.8 KB
    Loading...
  3. 2024-05-31Upload README.md with huggingface_hub5a1ef8e11.6 KB
    Loading...
  4. 2024-05-31Upload README.md with huggingface_hubf4e19d711.5 KB
    Loading...
  5. 2024-05-31Upload README.md with huggingface_hub64cb59311.3 KB
    Loading...
  6. 2024-05-31Upload README.md with huggingface_hub7ec71f811 KB
    Loading...
  7. 2024-05-31Upload README.md with huggingface_hubab2c33110.8 KB
    Loading...
  8. 2024-05-31Upload README.md with huggingface_huba952a5210.6 KB
    Loading...
  9. 2024-05-31Upload README.md with huggingface_hub350266910.5 KB
    Loading...
  10. 2024-05-31Upload README.md with huggingface_hub4985f9c10.3 KB
    Loading...
  11. 2024-05-31Upload README.md with huggingface_hub650486310 KB
    Loading...
  12. 2024-05-31Upload README.md with huggingface_hub28ed5c59.8 KB
    Loading...
  13. 2024-05-31Upload README.md with huggingface_hube406bb59.5 KB
    Loading...
  14. 2024-05-31Upload README.md with huggingface_hub24ed9c59.4 KB
    Loading...
  15. 2024-05-31Upload README.md with huggingface_hub9f89efa9 KB
    Loading...
  16. 2024-05-31Upload README.md with huggingface_hub172cd398.7 KB
    Loading...
  17. 2024-05-31Upload README.md with huggingface_hub655400b8.6 KB
    Loading...
  18. 2024-05-31Upload README.md with huggingface_hub37d03828.2 KB
    Loading...
  19. 2024-05-31Upload README.md with huggingface_huba3845098.1 KB
    Loading...
  20. 2024-05-31Upload README.md with huggingface_hub67c610a7.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration