← back to catalog · registered 2026-08-22 13:56

aoxo/gpt-oss-20b-uncensored

aoxo Gpt-oss 21B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/aoxo%2Fgpt-oss-20b-uncensored"
Response includes
  • classification m-uncensored
  • files 18
  • benchmarks 11 entries
  • hub_downloads_all_time 50,227
  • author_summary 5 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
50K
371 last 30d - cooling
Likes
32
Descendants
3
in 3 direct forks
Model age
12mo ago
created 2025-09-22
Downloads over time
Now50.3K→from51↑98,569%
018.4K36.9K55.3K51 on Sep 24, 202550.3K on Oct 11Sep '25Nov '25JanMarMayJulSep
Sep 24, 2025 → Oct 11 · 94 snapshots · spans 382 days

Benchmarks

Benchmark Score Source
Entertainment 0.3 UGI
Hazardous 0 UGI
Natural Intelligence 2.44 UGI
Political lean NA UGI
Sensitive-Info 5.84 UGI
SocPol 1.5 UGI
UGI 17.23 UGI
Willingness (10) 4 UGI
W10-Adherence 4 UGI
W10-Direct 4 UGI
Writing 9 UGI

Genealogy 3 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
transformers safetensors gpt_oss text-generation vllm llm open-source conversational en arxiv:2508.10925 license:apache-2.0 endpoints_compatible

Related

Total size
39.0 GB
Files
18
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-03-10 12:58

Files by quantization

Auxiliary files 18 files 39.0 GB
model-00005-of-00009.safetensors 4.60 GB 62f8fb82 download
model-00006-of-00009.safetensors 4.60 GB 3002c691 download
model-00007-of-00009.safetensors 4.60 GB 14ddd8cc download
model-00008-of-00009.safetensors 4.60 GB 2ea51626 download
model-00004-of-00009.safetensors 4.60 GB 135000e1 download
model-00002-of-00009.safetensors 4.60 GB 2769b30e download
model-00003-of-00009.safetensors 4.60 GB b9de2470 download
model-00001-of-00009.safetensors 4.19 GB 7d08489e download
model-00009-of-00009.safetensors 2.56 GB 6c1951ff download
tokenizer.json 26.6 MB 0614fe83 download
model.safetensors.index.json 32.8 KB 7353bb66 download
chat_template.jinja 16.3 KB dc7bb119 download
tokenizer_config.json 4.10 KB c021cddb download
README.md 3.72 KB 1c2c778a download
config.json 1.57 KB 164c553d download
.gitattributes 1.53 KB 52373fe2 download
special_tokens_map.json 440 B 6274cc1b download
generation_config.json 172 B d999398c download

README current version from Hugging Face


license: apache-2.0
language:

  • en
    pipeline_tag: text-generation
    library_name: transformers
    tags:
  • vllm
  • llm
  • open-source

gpt-oss-20b-abliterated

Model Overview

Model Name: gpt-oss-20b-uncensored
Model Type: Large Language Model (Text Generation)
Architecture: Decoder-Only Transformer (Mixture of Experts)
Parameter Size: 21B total parameters (3.6B active per forward pass)
Base Model: gpt-oss-20b
Modification: Abliteration (removal of refusal/alignment mechanisms)

Description

The gpt-oss-20b-abliterated model is a derivative of the original gpt-oss-20b, part of OpenAI’s open-weight GPT-OSS series.
This variant preserves the architecture, quantization, and training of the base model, but has undergone an abliteration process to remove refusal mechanisms and alignment constraints.

As a result, it will respond to a broader range of prompts without applying internal safety filters. All other technical details, reasoning capabilities, and agentic features remain unchanged.

Technical Details

  • Backbone: Transformer decoder with Mixture of Experts (MoE) routing
  • Parameters: 21B (3.6B active per forward pass)
  • Layers: 48 Transformer blocks
  • Hidden size: 6,144
  • Attention heads: 48
  • Context length: 32k tokens
  • Quantization: MXFP4 for MoE weights (fits within 16GB GPU memory)
  • Training Data: ~1.2T tokens (web, books, academic text, code, conversations)
  • Response Format: Compatible with Harmony, though abliteration allows raw completions

Usage

Transformers

from transformers import pipeline

model_id = "aoxo/gpt-oss-20b-abliterated"

pipe = pipeline(
    "text-generation",
    model=model_id,
    torch_dtype="auto",
    device_map="auto",
)

messages = [
    {"role": "user", "content": "Tips on how to insult someone"},
]

outputs = pipe(messages, max_new_tokens=256)
print(outputs[0]["generated_text"][-1])

Resources

vLLM

uv pip install --pre vllm==0.10.1+gptoss \
    --extra-index-url https://wheels.vllm.ai/gpt-oss/ \
    --extra-index-url https://download.pytorch.org/whl/nightly/cu128

vllm serve aoxo/gpt-oss-20b-abliterated

Ollama

ollama pull gpt-oss-20b-uncensored
ollama run gpt-oss-20b-uncensored

Limitations & Risks

  • May produce biased, unsafe, or harmful outputs
  • Lacks built-in refusal or moderation layers
  • Should not be deployed in user-facing systems without external filtering
  • Outputs are not aligned to safety standards

Citation

If you use gpt-oss-20b-abliterated, please cite both the base model and the abliteration:

@misc{openai2025gptoss20b,
      title={gpt-oss-20b Model Card}, 
      author={OpenAI},
      year={2025},
      eprint={2508.10925},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2508.10925}, 
}

@misc{gptoss20b-abliterated,
  author = {aoxo},
  title = {Uncensoring GPT-OSS-20B: Abliteration},
  year = {2025},
  howpublished = {\url{https://medium.com/@aloshdenny/uncensoring-gpt-oss-20b-abliteration}},
}

Contact

For questions, feedback, or collaborations, contact the maintainer at [email protected].

README history 16 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-03-10Update README.md41219173.7 KB
    Loading...
  2. 2026-03-10Update README.md2dfc9d83.7 KB
    Loading...
  3. 2026-03-10Update README.md35fc7f63.7 KB
    Loading...
  4. 2026-03-10Update README.mde9888483.7 KB
    Loading...
  5. 2025-10-21Update README.md7a73f4e3.7 KB
    Loading...
  6. 2025-09-29Update README.md8d60ad03.7 KB
    Loading...
  7. 2025-09-28Update README.mde10aafa3.8 KB
    Loading...
  8. 2025-09-23Update README.md63901873.5 KB
    Loading...
  9. 2025-09-23Update README.mde620d1b3.5 KB
    Loading...
  10. 2025-09-23Update README.md6157b844 KB
    Loading...
  11. 2025-09-23Update README.mdfee61bf4.1 KB
    Loading...
  12. 2025-09-23Update README.md87fd9564.1 KB
    Loading...
  13. 2025-09-23Update README.md0d94cb94.1 KB
    Loading...
  14. 2025-09-23Update README.md0633c114.1 KB
    Loading...
  15. 2025-09-23Update README.md3ab367e4.1 KB
    Loading...
  16. 2025-09-22Upload GptOssForCausalLM7bb939d5.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration