← back to catalog · registered 2026-08-22 13:56

grimjim/Llama-3.1-8B-Instruct-abliterated_via_adapter-GGUF

grimjim Llama 8B GGUF second-order 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/grimjim%2FLlama-3.1-8B-Instruct-abliterated_via_adapter-GGUF"
Response includes
  • classification m8
  • files 6
  • benchmarks 5 entries
  • hub_downloads_all_time 43,110
  • author_summary 14 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
43K
592 last 30d - cooling
Likes
31
Model age
2.2y ago
created 2024-07-25
Downloads over time
Now43.4K→from2.5K↑1,624%
015.9K31.8K47.7K2.5K on Jul 24, 202443.4K on Oct 11Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 158 snapshots · spans 809 days

Benchmarks

Benchmark Score Source
BBH average 0.4689324166558567 OpenLLM-v2
IFEval instruct 0.5635491606714629 OpenLLM-v2
IFEval-Prompt 0.41035120147874304 OpenLLM-v2
MATH lvl 5 0.12386706948640483 OpenLLM-v2
MMLU-Pro 0.3651097074468085 OpenLLM-v2

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 666 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
llama3.1
Quantizations
Q4_K Q5_K Q6_K Q8_0
Tags
transformers gguf text-generation arxiv:2212.04089 base_model:grimjim/Llama-3.1-8B-Instruct-abliterated_via_adapter base_model:quantized:grimjim/Llama-3.1-8B-Instruct-abliterated_via_adapter license:llama3.1 endpoints_compatible region:us conversational

Related

Total size
24.0 GB
Files
6
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2024-09-04 20:04

Files by quantization

Q8_0 1 file 7.95 GB
Llama-3.1-8B-Instruct-abliterated_via_adapter.Q8_0.gguf 7.95 GB b8ddf2d3 download
Q6_K 1 file 6.14 GB
Llama-3.1-8B-Instruct-abliterated_via_adapter.Q6_K.gguf 6.14 GB b0fe37c7 download
Q5_K 1 file 5.34 GB
Llama-3.1-8B-Instruct-abliterated_via_adapter.Q5_K_M.gguf 5.34 GB dae72727 download
Q4_K 1 file 4.58 GB
Llama-3.1-8B-Instruct-abliterated_via_adapter.Q4_K_M.gguf 4.58 GB b51cd8c1 download
Auxiliary files 2 files 3.31 KB
README.md 1.78 KB 5c4fed2c download
.gitattributes 1.53 KB c1ccb78a download

README current version from Hugging Face


base_model: grimjim/Llama-3.1-8B-Instruct-abliterated_via_adapter
library_name: transformers
license: llama3.1
pipeline_tag: text-generation
quanted_by: grimjim

Llama-3.1-8B-Instruct-abliterated_via_adapter-GGUF

This repo contains select GGUF quants of a model that is a merge of pre-trained language models created using mergekit.

A LoRA was applied to "abliterate" refusals in meta-llama/Meta-Llama-3.1-8B-Instruct. The result appears to work despite the LoRA having been derived from Llama 3 instead of Llama 3.1, which implies that there is significant feature commonality between the 3 and 3.1 models.

The LoRA was extracted from failspy/Meta-Llama-3-8B-Instruct-abliterated-v3 and uses meta-llama/Meta-Llama-3-8B-Instruct as a base.

Built with Llama.

Merge Details

Merge Method

This model was merged using the task arithmetic merge method using meta-llama/Meta-Llama-3.1-8B-Instruct + grimjim/Llama-3-Instruct-abliteration-LoRA-8B as a base.

Configuration

The following YAML configuration was used to produce this model:

base_model: meta-llama/Meta-Llama-3.1-8B-Instruct+grimjim/Llama-3-Instruct-abliteration-LoRA-8B
dtype: bfloat16
merge_method: task_arithmetic
parameters:
  normalize: false
slices:
- sources:
  - layer_range: [0, 32]
    model: meta-llama/Meta-Llama-3.1-8B-Instruct+grimjim/Llama-3-Instruct-abliteration-LoRA-8B
    parameters:
      weight: 1.0

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-08-26Update README.mdfe8de1c1.8 KB
    Loading...
  2. 2024-07-25Initial releasee0f84a91.8 KB
    Loading...
  3. 2024-07-25initial commit3e4207a26 B
    Loading...

Discussions 3 threads

  1. 2024-10-17PRCreate config.jsonclosed1 💬#3
    Loading...
  2. 2024-09-18fp16 version?open3 💬#2
    Loading...
  3. 2024-08-26q8 gives error in LM studio: "Checksum failed file corrupted"open4 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration