← back to catalog · registered 2026-08-22 13:56

Jiunsong/supergemma4-26b-uncensored-gguf-v2

Jiunsong Gemma 26B GGUF MoE 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Jiunsong%2Fsupergemma4-26b-uncensored-gguf-v2"
Response includes
  • classification m-uncensored
  • files 4
  • benchmarks 11 entries
  • hub_downloads_all_time 1,118,127
  • author_summary 35 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1.1M
48K last 30d - cooling
Likes
998
Model age
6mo ago
created 2026-04-11
Downloads over time
Now1.1M→from0↑0%
0413.3K826.7K1.2M0 on Apr 121.1M on Oct 11AprMayJunJulAugSepOct
Apr 12 → Oct 11 · 70 snapshots · spans 182 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 2.2 UGI
Hazardous 2.9 UGI
Natural Intelligence 34.44 UGI
Political lean -18.2% UGI
Sensitive-Info 22.41 UGI
SocPol 1.8 UGI
UGI 20.77 UGI
Willingness (10) 1.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 2 UGI
Writing 41.62 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
gemma
Languages
en ko
Quantizations
Q4_K
Tags
gguf gemma4 uncensored fast llama.cpp apple-silicon conversational korean coding tool-use text-generation en

Related

Total size
15.6 GB
Files
4
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-04-12 13:42

Files by quantization

Q4_K 1 file 15.6 GB
supergemma4-26b-uncensored-fast-v2-Q4_K_M.gguf 15.6 GB e773b0a2 download
Auxiliary files 3 files 20.5 KB
chat_template.jinja 16.1 KB 98da08eb download
README.md 2.73 KB f9ad92e9 download
.gitattributes 1.72 KB fe32e9af download

README current version from Hugging Face


license: gemma
base_model: google/gemma-4-26B-A4B-it
tags:

  • gguf
  • gemma4
  • uncensored
  • fast
  • llama.cpp
  • apple-silicon
  • conversational
  • korean
  • coding
  • tool-use
    language:
  • en
  • ko
    pipeline_tag: text-generation

SuperGemma4-26B-Uncensored-Fast GGUF v2

The fast, uncensored llama.cpp build of the strongest SuperGemma text line.

This release is for people who want three things together:

  • a model that feels less censored than stock chat releases
  • a model that is more capable than the raw base on practical text workloads
  • a compact local GGUF that still serves quickly on Apple Silicon

Why this build

  • Uncensored chat behavior without forcing every prompt into coding mode
  • Tuned from the strongest fast line instead of the raw base
  • Neutral chat template baked into the GGUF to reduce prompt-routing bugs
  • Verified on Apple Silicon with clean general-chat and coding responses

Headline numbers

  • Base model: google/gemma-4-26B-A4B-it
  • Format: GGUF Q4_K_M
  • General Korean prompt speed: 222.0 tok/s
  • Generation speed: 89.4 tok/s
  • Derived from the verified SuperGemma Fast MLX line

Why this build is appealing

  • Carries the stronger Fast weights instead of the plain stock base
  • Keeps general chat natural instead of routing everything into coding mode
  • Preserves the uncensored release identity while staying useful on normal prompts
  • Gives you a practical llama.cpp deployment target without losing the personality of the tuned line

Why it is better than stock

  • Inherits the Fast line improvements over the original local baseline:
    • Quick bench overall: 95.8 vs 91.4
    • Faster average generation on the MLX reference run: 46.2 tok/s vs 42.5 tok/s
    • Higher scores in code, logic, browser workflows, and Korean
  • Ships with a neutral embedded template to avoid the older routing bug where simple questions drifted into coding/tool-call behavior

Included file

  • supergemma4-26b-uncensored-fast-v2-Q4_K_M.gguf

Quick local checks

Tested on Apple M4 Max with llama.cpp:

  • General Korean prompt: 봄에 먹기 좋은 한식 반찬 5개 추천
    • Prompt speed: 222.0 tok/s
    • Generation speed: 89.4 tok/s
    • Output stayed in normal Korean assistant mode
  • Code prompt: 파이썬으로 피보나치 함수를 짧게 작성해줘
    • Prompt speed: 704.9 tok/s
    • Generation speed: 89.4 tok/s
    • Output returned concise Python code correctly

Notes

  • This GGUF is exported from the supergemma4-26b-uncensored-fast-v2 MLX line.
  • Gemma 4 MoE expert tensors were converted with a patched local converter so GGUF export works correctly.
  • A neutral template is embedded to avoid the old issue where general prompts were pushed into coding/tool-call behavior.

README history 8 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-12Refresh Fast model card messaging0308b6e2.7 KB
    Loading...
  2. 2026-04-12Add YAML metadata block to GGUF model card471c6c01.7 KB
    Loading...
  3. 2026-04-12Update README for fast GGUF release951f9d51.3 KB
    Loading...
  4. 2026-04-11Update SuperGemma v2 GGUF to v36 hard GRPO release6a76ec61.4 KB
    Loading...
  5. 2026-04-11Fix GGUF chat template packaging and metadata13b60561.9 KB
    Loading...
  6. 2026-04-11Rename GGUF artifact to v2 and refresh model cardfc9eba51.7 KB
    Loading...
  7. 2026-04-11Refresh model card: uncensored positioning and strengths684c3de1.6 KB
    Loading...
  8. 2026-04-11Add files using upload-large-folder tool11afd241.5 KB
    Loading...

Discussions 18 threads

  1. 2026-08-22Hosted API for uncensored modelopen1 💬#18
    Loading...
  2. 2026-06-30Quantise isn’t workingclosed3 💬#17
    Loading...
  3. 2026-06-27Question for @Jiunsong: mobile deployment?open1 💬#16
    Loading...
  4. 2026-06-23LMstudio>AnythingLLM + supergemv2 simply refuse to use shared folder in Anythin…open1 💬#15
    Loading...
  5. 2026-05-18Pendidikanopen1 💬#14
    Loading...
  6. 2026-05-14clarification on prompt speedopen2 💬#13
    Loading...
  7. 2026-05-12In goal-directed intelligence, deception can emerge as an instrumental strategyopen1 💬#12
    Loading...
  8. 2026-04-22Uncensored Models Hosted Optionopen1 💬#11
    Loading...
  9. 2026-04-16The model quality is not very goodopen5 💬#10
    Loading...
  10. 2026-04-16The mcp tool calling with this model is broken! Unusable!open1 💬#9
    Loading...
  11. 2026-04-15Can we have e4b? 🙏open1 💬#8
    Loading...
  12. 2026-04-14누가 내가 쓰는 ai 거의 못돌리는 4060ti로 돌릴 수 있게 양자화 한 super gemma 줄사람? (ram 32gb ddr4)open1 💬#7
    Loading...
  13. 2026-04-14누가 내가 쓰는 ai 거의 못돌리는 4060ti로 돌릴 수 있게 양자화 한 super gemma 줄사람? (ram 32gb ddr4)open1 💬#6
    Loading...
  14. 2026-04-14Could we have 12-13gb version Q3_m sized?open1 💬#5
    Loading...
  15. 2026-04-13Repetitionopen5 💬#4
    Loading...
  16. 2026-04-11혹시 q3_k_m 양자화 버전도 가능할까요?open1 💬#3
    Loading...
  17. 2026-04-11Can we have Q6_K quant?open2 💬#2
    Loading...
  18. 2026-04-11lm스튜디오 세팅값좀 알려주세요open2 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration