license: apache-2.0
tags:
- gemma
- gemma-4
- heretic
- uncensored
- mxfp4
- gguf
- vision
- multimodal
base_model: llmfan46/gemma-4-E4B-it-ultra-uncensored-heretic
Gemma 4 E4B It Ultra Uncensored Heretic - MXFP4 GGUF
MXFP4 quantization of llmfan46/gemma-4-E4B-it-ultra-uncensored-heretic, an abliterated Gemma 4 E4B model with vision support.
About the Model
- Gemma 4 E4B — abliterated/uncensored variant for unrestricted use
- Vision support via separate mmproj vision projector
- Architecture: Gemma 4 with MoE (4B active parameters)
Quantization
Quantized from the BF16 safetensors using llama.cpp (build 537).
MXFP4 (Microscaling FP4) uses block-wise quantization with shared exponents.
Files
| File | Size | Description |
|---|---|---|
gemma-4-E4B-it-ultra-uncensored-heretic-mxfp4.gguf |
~4.7 GB | MXFP4 quantized model |
mmproj-gemma-4-E4B-it-ultra-uncensored-heretic-f16.gguf |
~0.97 GB | Vision projector (BF16) |
Usage
llama-server \
-m gemma-4-E4B-it-ultra-uncensored-heretic-mxfp4.gguf \
--mmproj mmproj-gemma-4-E4B-it-ultra-uncensored-heretic-f16.gguf \
-ngl 99 \
--host 0.0.0.0 \
--port 8080
License
Apache 2.0