← back to catalog · registered 2026-08-22 13:56

thy025/supergemma4-26b-uncensored-gguf-v2

thy025 Gemma 26B GGUF MoE 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/thy025%2Fsupergemma4-26b-uncensored-gguf-v2"
Response includes
  • classification m-uncensored
  • files 4
  • benchmarks 11 entries
  • hub_downloads_all_time 401
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
401
27 last 30d - cooling
Likes
0
Model age
4mo ago
created 2026-05-15
Downloads over time
Now406→from0↑0%
01492984470 on May 13406 on Oct 11406 on Oct 7MayJunJulAugSepOct
May 13 → Oct 11 · 61 snapshots · spans 151 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 2.2 UGI
Hazardous 2.9 UGI
Natural Intelligence 34.44 UGI
Political lean -18.2% UGI
Sensitive-Info 22.41 UGI
SocPol 1.8 UGI
UGI 20.77 UGI
Willingness (10) 1.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 2 UGI
Writing 41.62 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
gemma
Languages
en ko
Quantizations
Q4_K
Tags
gguf gemma4 uncensored fast llama.cpp apple-silicon conversational korean coding tool-use text-generation en

Related

Total size
15.6 GB
Files
4
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-05-15 04:36

Files by quantization

Q4_K 1 file 15.6 GB
supergemma4-26b-uncensored-fast-v2-Q4_K_M.gguf 15.6 GB e773b0a2 download
Auxiliary files 3 files 20.5 KB
chat_template.jinja 16.1 KB 98da08eb download
README.md 2.73 KB f9ad92e9 download
.gitattributes 1.72 KB fe32e9af download

README current version from Hugging Face


license: gemma
base_model: google/gemma-4-26B-A4B-it
tags:

  • gguf
  • gemma4
  • uncensored
  • fast
  • llama.cpp
  • apple-silicon
  • conversational
  • korean
  • coding
  • tool-use
    language:
  • en
  • ko
    pipeline_tag: text-generation

SuperGemma4-26B-Uncensored-Fast GGUF v2

The fast, uncensored llama.cpp build of the strongest SuperGemma text line.

This release is for people who want three things together:

  • a model that feels less censored than stock chat releases
  • a model that is more capable than the raw base on practical text workloads
  • a compact local GGUF that still serves quickly on Apple Silicon

Why this build

  • Uncensored chat behavior without forcing every prompt into coding mode
  • Tuned from the strongest fast line instead of the raw base
  • Neutral chat template baked into the GGUF to reduce prompt-routing bugs
  • Verified on Apple Silicon with clean general-chat and coding responses

Headline numbers

  • Base model: google/gemma-4-26B-A4B-it
  • Format: GGUF Q4_K_M
  • General Korean prompt speed: 222.0 tok/s
  • Generation speed: 89.4 tok/s
  • Derived from the verified SuperGemma Fast MLX line

Why this build is appealing

  • Carries the stronger Fast weights instead of the plain stock base
  • Keeps general chat natural instead of routing everything into coding mode
  • Preserves the uncensored release identity while staying useful on normal prompts
  • Gives you a practical llama.cpp deployment target without losing the personality of the tuned line

Why it is better than stock

  • Inherits the Fast line improvements over the original local baseline:
    • Quick bench overall: 95.8 vs 91.4
    • Faster average generation on the MLX reference run: 46.2 tok/s vs 42.5 tok/s
    • Higher scores in code, logic, browser workflows, and Korean
  • Ships with a neutral embedded template to avoid the older routing bug where simple questions drifted into coding/tool-call behavior

Included file

  • supergemma4-26b-uncensored-fast-v2-Q4_K_M.gguf

Quick local checks

Tested on Apple M4 Max with llama.cpp:

  • General Korean prompt: 봄에 먹기 좋은 한식 반찬 5개 추천
    • Prompt speed: 222.0 tok/s
    • Generation speed: 89.4 tok/s
    • Output stayed in normal Korean assistant mode
  • Code prompt: 파이썬으로 피보나치 함수를 짧게 작성해줘
    • Prompt speed: 704.9 tok/s
    • Generation speed: 89.4 tok/s
    • Output returned concise Python code correctly

Notes

  • This GGUF is exported from the supergemma4-26b-uncensored-fast-v2 MLX line.
  • Gemma 4 MoE expert tensors were converted with a patched local converter so GGUF export works correctly.
  • A neutral template is embedded to avoid the old issue where general prompts were pushed into coding/tool-call behavior.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-05-15Duplicate from Jiunsong/supergemma4-26b-uncensored-gguf-v2db84d2d2.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration