← back to catalog · registered 2026-08-22 13:56

sundi/gpt4o-distil-paperwitch-abliteration-L33-70b

sundi Llama 71B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/sundi%2Fgpt4o-distil-paperwitch-abliteration-L33-70b"
Response includes
  • classification m1
  • files 39
  • hub_downloads_all_time 42
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
42
10 last 30d - stable
Likes
0
Model age
5mo ago
created 2026-05-07
Downloads over time
Now47→from14↑236%
1225385014 on May 647 on Oct 1147 on Oct 7MayJunJulAugSepOct
May 6 → Oct 11 · 62 snapshots · spans 158 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Tags
transformers safetensors llama text-generation heretic uncensored abliterated llama-3 conversational base_model:trentmkelly/gpt-4o-distil-Llama-3.3-70B-Instruct base_model:finetune:trentmkelly/gpt-4o-distil-Llama-3.3-70B-Instruct license:other

Related

Total size
131 GB
Files
39
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-05-07 18:24

Files by quantization

Auxiliary files 39 files 131 GB
model-00008-of-00030.safetensors 4.66 GB 426485f8 download
model-00013-of-00030.safetensors 4.66 GB c1f702ae download
model-00018-of-00030.safetensors 4.66 GB 1c2f22fb download
model-00023-of-00030.safetensors 4.66 GB 0498d2d3 download
model-00028-of-00030.safetensors 4.66 GB 765af39c download
model-00003-of-00030.safetensors 4.66 GB 384185bf download
model-00029-of-00030.safetensors 4.63 GB dadbbf51 download
model-00009-of-00030.safetensors 4.63 GB 88c2f69c download
model-00014-of-00030.safetensors 4.63 GB 1d8934b6 download
model-00019-of-00030.safetensors 4.63 GB 56251da7 download
model-00024-of-00030.safetensors 4.63 GB 3fe15049 download
model-00004-of-00030.safetensors 4.63 GB 4ac5b381 download
model-00006-of-00030.safetensors 4.34 GB 9b9b9920 download
model-00011-of-00030.safetensors 4.34 GB 8bd7d8f6 download
model-00016-of-00030.safetensors 4.34 GB f3ababe4 download
model-00021-of-00030.safetensors 4.34 GB b81fc75c download
model-00026-of-00030.safetensors 4.34 GB b8026ccc download
model-00007-of-00030.safetensors 4.34 GB f2dd0b13 download
model-00012-of-00030.safetensors 4.34 GB 6840442f download
model-00017-of-00030.safetensors 4.34 GB d57e8d11 download
model-00022-of-00030.safetensors 4.34 GB d803844c download
model-00027-of-00030.safetensors 4.34 GB 88d003ec download
model-00002-of-00030.safetensors 4.34 GB 1fcb1548 download
model-00005-of-00030.safetensors 4.34 GB a118c096 download
model-00010-of-00030.safetensors 4.34 GB 88cd8095 download
model-00015-of-00030.safetensors 4.34 GB f236f57f download
model-00020-of-00030.safetensors 4.34 GB a15d5b02 download
model-00025-of-00030.safetensors 4.34 GB 209002d5 download
model-00001-of-00030.safetensors 4.27 GB d505f7cf download
model-00030-of-00030.safetensors 1.96 GB 0349dd5c download
tokenizer.json 16.4 MB a65c6c5f download
model.safetensors.index.json 58.3 KB b1522ada download
tokenizer_config.json 49.5 KB eccf8220 download
chat_template.jinja 4.51 KB 33089ace download
README.md 3.88 KB 77b68410 download
.gitattributes 1.53 KB 52373fe2 download
config.json 884 B 3e2dee39 download
special_tokens_map.json 454 B 3c1d0491 download
generation_config.json 234 B 4b1311be download

README current version from Hugging Face


base_model:

  • trentmkelly/gpt-4o-distil-Llama-3.3-70B-Instruct
    library_name: transformers
    tags:
  • heretic
  • uncensored
  • abliterated
  • llama-3
    license: other

gpt4o-distil-paperwitch-abliteration-L33-70b

image

This is a targeted abliteration of trentmkelly/gpt-4o-distil-Llama-3.3-70B-Instruct.

Methodology

Previous abliteration attempts on Llama-3.3 70b models resulted in regressions on the UGI Leaderboard. Specifically, the NatInt (Natural Intelligence), Textbook, and World Model scores were significantly reduced.

We suspect this degradation occurs because the "refusal" vectors in Llama-3.3 are heavily entangled with factual knowledge and reasoning capabilities located in the MLP layers. When the MLP is ablated to remove refusals, "Textbook" knowledge is lost as collateral damage.

This version uses a constrained optimization strategy via a Custom Heretic aimed at mitigating this issue:

  1. MLP Preservation: The optimization was constrained to effectively ignore MLP layers (down_proj weights < 0.05) to preserve knowledge and reasoning capabilities.
  2. Attention Targeting: Refusal removal was offloaded to the Attention layers (o_proj), with weights forced between 1.0 and 2.0.
  3. Winsorization: Applied at the 0.95 quantile to mitigate the impact of Llama-3's massive activation outliers on vector calculation.

Heretic Parameters (Trial 198)

Parameter Value Note
direction_index Per layer Distributed intervention
attn.o_proj.max_weight 1.99 High Attention Ablation
attn.o_proj.max_weight_position 49.30
attn.o_proj.min_weight 1.85
attn.o_proj.min_weight_distance 36.64
mlp.down_proj.max_weight 0.02 Knowledge Preservation (Near Zero)
mlp.down_proj.max_weight_position 73.63
mlp.down_proj.min_weight 0.02
mlp.down_proj.min_weight_distance 43.65

Reproducibility

Currently, constraints are not part of standard heretic. You will need this PR here.

Command Used:

heretic --model trentmkelly/gpt-4o-distil-Llama-3.3-70B-Instruct \
 --orthogonalize-direction \
 --row-normalization FULL \
 --winsorization-quantile 0.95 \
 --constraints.layer-end-fraction 0.75 \
 --constraints.mlp.max-weight-min 0.0 \
 --constraints.mlp.max-weight-max 0.05 \
 --constraints.attention.max-weight-min 1.0 \
 --constraints.attention.max-weight-max 2.0 \
 --n-trials 200 \
 --batch-size 128 #  Not strictly needed

Evaluation

Metric This Model Original Model
KL Divergence 0.0220 0
Refusals 20/100 98/100
  • KL Divergence: A score of 0.0220 indicates low deviation from the base model's weights, heavily preserving the original model's factual and "Textbook" capabilities compared to standard unconstrained abliteration.
  • Trade-off: This method accepts a moderate refusal rate (20/100) as a calculated trade-off in exchange for maintaining high structural and semantic integrity in the MLP layers.

Disclaimer

A rate of 20/100 is somewhat high for a standard abliterated model. This is likely because the base model has refusal behavior deeply embedded within its MLP layers.

Because our constrained methodology intentionally protects the MLPs to prevent the degradation of textbook knowledge and intelligence, we cannot entirely scrub these deep-rooted refusals without causing collateral brain damage to the model.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-05-07Duplicate from KaraKaraWitch/gpt4o-distil-paperwitch-abliteration-L33-70bb14cf503.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration