← back to catalog · registered 2026-08-22 13:56

Zoyd/TheSkullery_Aura-Llama-Abliterated-2_2bpw_exl2

Zoyd Llama second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Zoyd%2FTheSkullery_Aura-Llama-Abliterated-2_2bpw_exl2"
Response includes
  • classification m5
  • files 9
  • benchmarks 5 entries
  • hub_downloads_all_time 52
  • author_summary 77 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M5
Primary method

Mergekit merge

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • merge tag / mergekit / dare-ties in tags or name
  • no unusual architecture pattern (regular merge)
  • abliterated marker present
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
52
13 last 30d - stable
Likes
0
Model age
2.4y ago
created 2024-05-27
Downloads over time
Now57→from2↑2,750%
036721082 on Jul 24, 202457 on Oct 1198 on Sep 17, 2025Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 155 snapshots · spans 809 days

Benchmarks

Benchmark Score Source
BBH average 0.4065378129459074 OpenLLM-v2
IFEval instruct 0.6438848920863309 OpenLLM-v2
IFEval-Prompt 0.5378927911275416 OpenLLM-v2
MATH lvl 5 0.03172205438066465 OpenLLM-v2
MMLU-Pro 0.2741855053191489 OpenLLM-v2

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
transformers safetensors llama text-generation merge mergekit conversational base_model:failspy/Llama-3-8B-Instruct-abliterated base_model:quantized:failspy/Llama-3-8B-Instruct-abliterated license:apache-2.0 model-index text-generation-inference

Related

Total size
3.85 GB
Files
9
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2024-05-27 03:28

Files by quantization

Auxiliary files 9 files 3.86 GB
output.safetensors 3.85 GB 17039e04 download
tokenizer.json 8.66 MB b32575ff download
tokenizer_config.json 49.8 KB 1bfd1146 download
model.safetensors.index.json 30.5 KB 1585dae3 download
README.md 9.03 KB 02cdbeb1 download
.gitattributes 1.48 KB a6344aac download
config.json 1.03 KB ae026abd download
mergekit_config.yml 399 B b8b60f43 download
special_tokens_map.json 296 B 02ee80b6 download

README current version from Hugging Face


license: apache-2.0
tags:


Exllamav2 quant (exl2 / 2.2 bpw) made with ExLlamaV2 v0.0.21

Other EXL2 quants:

Quant Model Size lm_head
2.2
3944 MB
6
2.5
4258 MB
6
3.0
4829 MB
6
3.5
5403 MB
6
3.75
5688 MB
6
4.0
5975 MB
6
4.25
6260 MB
6
5.0
7115 MB
6
6.0
8369 MB
8
6.5
8934 MB
8
8.0
10593 MB
8
Aura-llama-3 Data Card

Aura-llama-3-Abliterated

Aura-llama-Abliterated Image

Now that the cute anime girl has your attention.

UPDATE: Model is now using the abliterated version of meta llama 3 8b

Aura-llama is using the methodology presented by SOLAR for scaling LLMs called depth up-scaling (DUS), which encompasses architectural modifications with continued pretraining. Using the solar paper as a base, I integrated Llama-3 weights into the upscaled layers, and In the future plan to continue training the model.

Aura-llama is a merge of the following models to create a base model to work from:

Abliterated Merged Evals (Has Not Been Finetuned):

Aura-llama-Abliterated

  • Avg: ?
  • ARC: ?
  • HellaSwag: ?
  • MMLU: ?
  • T-QA: ?
  • Winogrande: ?
  • GSM8K: ?

Non Abliterated Merged Evals (Has Not Been Finetuned):

Aura-llama-Original

  • Avg: 63.13
  • ARC: 58.02
  • HellaSwag: 77.82
  • MMLU: 65.61
  • T-QA: 51.94
  • Winogrande: 73.40
  • GSM8K: 52.01

🧩 Configuration


dtype: bfloat16
merge_method: passthrough
slices:
- sources:
  - layer_range: [0, 12]
    model: failspy/Llama-3-8B-Instruct-abliterated
- sources:
  - layer_range: [8, 20]
    model: failspy/Llama-3-8B-Instruct-abliterated
- sources:
  - layer_range: [16, 28]
    model: failspy/Llama-3-8B-Instruct-abliterated
- sources:
  - layer_range: [24, 32]
    model: failspy/Llama-3-8B-Instruct-abliterated
        

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 53.46
AI2 Reasoning Challenge (25-Shot) 49.23
HellaSwag (10-Shot) 72.27
MMLU (5-Shot) 55.71
TruthfulQA (0-shot) 46.63
Winogrande (5-shot) 69.30
GSM8k (5-shot) 27.60

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-05-27Upload folder using huggingface_hub46c2bd19 KB
    Loading...
  2. 2024-05-27Upload folder using huggingface_hubbbd389c7.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration