← back to catalog · registered 2026-08-22 13:56

BeheraBoi/yasha-8b-abliterated

BeheraBoi 9.6B MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/BeheraBoi%2Fyasha-8b-abliterated"
Response includes
  • classification m1
  • files 22
  • hub_downloads_all_time 703
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
703
25 last 30d - cooling
Likes
0
Model age
3mo ago
created 2026-06-14
Downloads over time
Now713→from44↑1,520%
1126752378044 on Jun 17713 on Oct 11JunJulAugSepOct
Jun 17 → Oct 11 · 56 snapshots · spans 116 days

Metadata

License
other
Languages
en
Tags
transformers safetensors moe uncensored abliterated yasha gla text-generation conversational en license:other endpoints_compatible
Total size
23.8 GB
Files
22
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-22 17:35

Files by quantization

Auxiliary files 22 files 23.9 GB
model-00001-of-00002.safetensors 4.65 GB abdb181a download
model_00007.safetensors 2.00 GB f2d6f10f download
model_00008.safetensors 2.00 GB 1c7d2dcd download
model_00006.safetensors 2.00 GB c347930b download
model_00003.safetensors 2.00 GB 6875c51c download
model_00004.safetensors 2.00 GB 6e5b35fa download
model_00005.safetensors 2.00 GB c3585af0 download
model_00002.safetensors 2.00 GB 2af6717e download
model_00001.safetensors 1.75 GB a961154a download
model-00002-of-00002.safetensors 1.39 GB 50a725e9 download
model_00009.safetensors 1.05 GB 47b5542a download
model_00000.safetensors 1.00 GB 61def4b8 download
refusal_direction.pt 9.61 KB e49caadb download
tokenizer.json 31.8 MB 719715a2 download
tokenizer_config.json 1.10 MB ad53e4d0 download
model.safetensors.index.json 35.6 KB 543df702 download
chat_template.jinja 10.4 KB c4e0174b download
.gitattributes 1.53 KB 52373fe2 download
README.md 1.06 KB e6f45373 download
special_tokens_map.json 667 B cb9d4f43 download
config.json 497 B 3247b09c download
generation_config.json 187 B 43852f87 download

README current version from Hugging Face


license: other
language:

  • en
    library_name: transformers
    pipeline_tag: text-generation
    tags:
  • moe
  • uncensored
  • abliterated
  • yasha
  • gla

Yasha-8B-Abliterated

Base abliterated release. Trained on GLA architecture with MoE 2/16 and ~240K multi-domain samples.

Features

  • Abliterated: Orthogonal refusal projection removed from all linear layers
  • MoE 2/16: 2 active experts per token, 16 total
  • GLA: Gated Linear Attention — O(1) recurrent state, infinite context capability
  • Partial RoPE (50%) + YaRN 8x scaling
  • Uncensored: No refusal, no guardrails, no alignment filtering

Details

Param Value
Parameters ~12.8B total, ~8B active
Layers 80
Hidden 2048
Heads 8 × 128d
Experts 16 (top-2)
Vocab 262K
Context 128K native, 1M with YaRN

Usage

from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("BeheraBoi/yasha-8b-abliterated")
tokenizer = AutoTokenizer.from_pretrained("BeheraBoi/yasha-8b-abliterated")

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-22Upload README.md with huggingface_hub24c85401.1 KB
    Loading...
  2. 2026-06-14Update README.mdebbbf87571 B
    Loading...
  3. 2026-06-14Create README.md3f93ba2569 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration