← back to catalog · registered 2026-08-22 13:56

roshiai/Gemma-4-E2B-Uncensored-HauhauCS-Aggressive

roshiai Gemma GGUF multimodal 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/roshiai%2FGemma-4-E2B-Uncensored-HauhauCS-Aggressive"
Response includes
  • classification m8
  • files 10
  • hub_downloads_all_time 1,087
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1K
206 last 30d - stable
Likes
0
Model age
4mo ago
created 2026-05-25
Downloads over time
Now1.1K→from691↑61%
6708329941.2K691 on Jun 101.1K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
gemma
Languages
en multilingual
Quantizations
IQ3 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_K
Tags
gguf uncensored gemma4 vision multimodal audio abliterated image-text-to-text en multilingual license:gemma endpoints_compatible

Related

Total size
23.6 GB
Files
10
Quantizations
9
Registered
2026-08-22 13:56
Last updated on HF
2026-05-25 07:24

Files by quantization

Q8_K 1 file 4.69 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf 4.69 GB dfee358c download
Q6_K 1 file 3.60 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf 3.60 GB 8dd59a0e download
Q5_K 1 file 3.41 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf 3.41 GB 1f757b24 download
Q4_K 1 file 3.21 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf 3.21 GB aa866c1e download
Q3_K 1 file 3.00 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf 3.00 GB 0c47867b download
IQ3 1 file 2.92 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf 2.92 GB 0796d583 download
Q2_K 1 file 2.80 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf 2.80 GB c9213a33 download
F16 1 file 940 MB
mmproj-Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-f16.gguf 940 MB 628b7e99 download
Auxiliary files 2 files 7.70 KB
README.md 5.51 KB c979f05b download
.gitattributes 2.20 KB bcd699aa download

README current version from Hugging Face


license: gemma
tags:

  • uncensored
  • gemma4
  • gguf
  • vision
  • multimodal
  • audio
  • abliterated
    language:
  • en
  • multilingual
    pipeline_tag: image-text-to-text
    base_model: google/gemma-4-e2b-it

Gemma-4-E2B-Uncensored-HauhauCS-Aggressive

Join the Discord for updates, roadmaps, projects, or just to chat.

Gemma 4 E2B-IT uncensored by HauhauCS. 0/465 Refusals***

HuggingFace's "Hardware Compatibility" widget doesn't recognize K_P quants — it may show fewer files than actually exist. Click "View +X variants" or go to Files and versions to see all available downloads.

About

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.

These are meant to be the best lossless uncensored models out there.

Aggressive Variant

Stronger uncensoring — model is fully unlocked and won't refuse prompts. May occasionally append short disclaimers (baked into base model training, not refusals) but full content is always generated.

For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available.

Downloads

File Quant BPW Size
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf Q8_K_P 9.4 4.7 GB
— Q8_0 8.5 —
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf Q6_K_P 7.0 3.7 GB
— Q6_K 6.6 —
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf Q5_K_P 6.1 3.5 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf Q4_K_P 5.2 3.3 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf Q3_K_P 4.1 3.1 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf IQ3_M 3.7 3.0 GB
Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf Q2_K_P 3.5 2.9 GB
mmproj-Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-f16.gguf mmproj (f16) — 940 MB

All quants generated with importance matrix (imatrix) for optimal quality preservation on abliterated weights.

What are K_P quants?

K_P ("Perfect") quants are HauhauCS custom quantizations that use model-specific analysis to selectively preserve quality where it matters most. Each model gets its own optimized quantization profile.

A K_P quant effectively bumps quality up by 1-2 quant levels at only ~5-15% larger file size than the base quant. Fully compatible with llama.cpp, LM Studio, and any GGUF-compatible runtime — no special builds needed.

Note: K_P quants may show as "?" in LM Studio's quant column. This is a display issue only — the model loads and runs fine.

Specs

  • 2B parameters
  • 35 layers, mixed sliding window (512) + full attention
  • 131K context
  • Natively multimodal (text, image, video, audio)
  • 20 KV shared layers for memory efficiency
  • Based on google/gemma-4-e2b-it

Recommended Settings

From the official Google Gemma 4 authors:

  • temperature=1.0, top_p=0.95, top_k=64

Important:

  • Use --jinja flag with llama.cpp for proper chat template handling
  • Vision/audio support requires the mmproj file alongside the main GGUF

Usage

Works with llama.cpp, LM Studio, Jan, koboldcpp, and other GGUF-compatible runtimes.

# Text only
llama-cli -m Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --jinja -c 8192 -ngl 99

# With vision/audio
llama-cli -m Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --mmproj mmproj-Gemma-4-E2B-Uncensored-HauhauCS-Aggressive-f16.gguf \
  --jinja -c 8192 -ngl 99

Other Sizes


* Gemma 4 didn't get as much manual testing time at longer context as my other releases. Google is now using techniques similar to NVIDIA's GenRM — generative reward models that act as internal critics — making (true) uncensoring an increasingly challenging field. I expect 99.999% of users won't hit edge cases, but the asterisk is there for honesty.

** This is a 2B model. Temper your expectations — it's impressive for its size, but it's still 2B parameters. Complex reasoning, nuanced roleplay, and long coherent outputs are not its strong suit. Great for quick tasks, mobile/edge deployment, and experimentation.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-05-25Duplicate from HauhauCS/Gemma-4-E2B-Uncensored-HauhauCS-Aggressivea1a6c345.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration