← back to catalog · registered 2026-08-22 13:56

RemySkye/Gemma-4-E2B-it-qat-abliterated-UD-Q4_K_XL-GGUF

RemySkye Gemma GGUF second-order 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/RemySkye%2FGemma-4-E2B-it-qat-abliterated-UD-Q4_K_XL-GGUF"
Response includes
  • classification m8
  • files 6
  • hub_downloads_all_time 1,213
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1K
611 last 30d - active
Likes
0
Model age
2mo ago
created 2026-08-07
Downloads over time
Now1.4K→from217↑541%
1586091.1K1.5K217 on Aug 191.4K on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
Q4_K
Tags
gguf gemma4 qat q4_0 abliterated uncensored llama.cpp base_model:huihui-ai/Huihui-gemma-4-E2B-it-qat-q4_0-unquantized-abliterated base_model:quantized:huihui-ai/Huihui-gemma-4-E2B-it-qat-q4_0-unquantized-abliterated license:apache-2.0 endpoints_compatible region:us

Related

Total size
2.46 GB
Files
6
Quantizations
3
Registered
2026-08-22 13:56
Last updated on HF
2026-08-15 15:57

Files by quantization

Q4_K 1 file 2.46 GB
Huihui-gemma-4-E2B-it-qat-UD-Q4_K_XL.gguf 2.46 GB 3dbe4f61 download
BF16 1 file 941 MB
mmproj-BF16.gguf 941 MB 38b33846 download
Auxiliary files 4 files 2.69 MB
imatrix_unsloth.gguf_file 2.69 MB 1828d3cc download
.gitattributes 1.75 KB 70738c3b download
README.md 1.32 KB 0f3dcd08 download
quantization_manifest.json 1.29 KB 683c1d69 download

README current version from Hugging Face


license: apache-2.0
base_model:

  • huihui-ai/Huihui-gemma-4-E2B-it-qat-q4_0-unquantized-abliterated
    tags:
  • gguf
  • gemma4
  • qat
  • q4_0
  • abliterated
  • uncensored
  • llama.cpp

Huihui Gemma 4 E2B QAT — UD-Q4_K_XL-like

Source:

huihui-ai/Huihui-gemma-4-E2B-it-qat-q4_0-unquantized-abliterated

Files

  • Huihui-gemma-4-E2B-it-qat-UD-Q4_K_XL-like.gguf — main model
  • mmproj-BF16.gguf — BF16 multimodal projector
  • imatrix_unsloth.gguf_file — exact Unsloth importance matrix used
  • quantization_manifest.json — hashes and provenance

Conversion

  • llama.cpp commit: fc6545d322a9ea6643f77439b31b66f403ba2cad
  • Main quant type: Q4_0
  • --pure: enabled
  • imatrix: unsloth/gemma-4-E2B-it-GGUF/imatrix_unsloth.gguf_file
  • mmproj: unsloth/gemma-4-E2B-it-qat-GGUF/mmproj-BF16.gguf

Important distinction

This is a publicly reproducible Unsloth-like QAT conversion, not the
official Unsloth UD-Q4_K_XL algorithm.

Unsloth's official Gemma 4 QAT GGUFs use additional scale/lattice recovery
logic when mapping the BF16 QAT lattice into llama.cpp's Q4_0 representation.
That exact recovery procedure is not exposed as a standard public
llama-quantize option.

llama.cpp example

llama-server \
  -m Huihui-gemma-4-E2B-it-qat-UD-Q4_K_XL-like.gguf \
  --mmproj mmproj-BF16.gguf \
  --jinja \
  -ngl 999

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-07Add files using upload-large-folder tool9ab0cb61.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration