← back to catalog · registered 2026-09-25 14:57

Brunobkr/OFFFELLIA_gemma-4-12B-it-uncensored-heretic.gguf

Brunobkr 12B GGUF
curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Brunobkr%2FOFFFELLIA_gemma-4-12B-it-uncensored-heretic.gguf"
Response includes
  • classification m3
  • files 7
  • author_summary 43 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-09-25

Metadata

Quantizations
F16 IQ4 Q8_0
Tags
gguf doi:10.57967/hf/10599 region:us

Related

Total size
40.5 GB
Files
7
Quantizations
4
Registered
2026-09-25 14:57
Last updated on HF
2026-09-25 14:49

Files by quantization

F16 1 file 22.2 GB
ΩFFFΣLLIα_f16_gemma-4-12B-it-uncensored-heretic.gguf 22.2 GB fb540a53 download
Q8_0 1 file 11.8 GB
ΩFFFΣLLIα_Q8_0_gemma-4-12B-it-uncensored-heretic.gguf 11.8 GB 3bb36000 download
IQ4 1 file 6.54 GB
ΩFFFΣLLIα_IQ4_NL_gemma-4-12B-it-uncensored-heretic.gguf 6.54 GB 58256c39 download
Auxiliary files 4 files 2.84 MB
capa.png 2.14 MB 34b0b662 download
pix.jpeg 711 KB a41f92d9 download
README.md 3.75 KB 84826f3a download
.gitattributes 1.84 KB a88828c5 download

README current version from Hugging Face

ΩFFFΣLLIa • llama.cpp • OFFFELLIA_HERETIC

OFFFELLIA_HERETIC Banner

OFFFELLIA_HERETIC Web UI -

   ██████╗ ███████╗███████╗███████╗██╗     ██╗     ██╗ █████╗ 
  ██╔═══██╗██╔════╝██╔════╝██╔════╝██║     ██║     ██║██╔══██╗
  ██║   ██║█████╗  █████╗  █████╗  ██║     ██║     ██║███████║
  ██║   ██║██╔══╝  ██╔══╝  ██╔══╝  ██║     ██║     ██║██╔══██║
  ╚██████╔╝██║     ██║     ███████╗███████╗███████╗██║██║  ██║-HERETIC
   ╚═════╝ ╚═╝     ╚═╝     ╚══════╝╚══════╝╚══════╝╚═╝╚═╝  ╚═╝

High-Performance LLM / VLM Inference & Autonomous Agentic Ecosystem in Pure C/C++

License: MIT
C++: 17/20
WebUI: SvelteKit + Vite
Status: Unlocked & Optimized
Agentic: Multi--Turn Engine


📖 Visão Geral

ΩFFFΣLLIa • llama.cpp • AlgMor24 é um fork avançado, destravado e de alta performance do ecossistema llama.cpp. Este projeto integra inferência local de última geração em C/C++ com um motor agêntico autônomo multi-turn, suporte nativo a FIM (Fill-in-the-Middle) para geração e preenchimento de código, Speculative Decoding otimizado para programação, integração de ferramentas MCP (Model Context Protocol) e uma interface Web moderna em SvelteKit/Vite com a identidade visual Cyberpunk Neon Fire.


✨

git clone https://github.com/brunoconta1980-tech/llama_OFFFELLIA_1984

cd llama_OFFFELLIA_1984

cmake -B build
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON

cmake --build build -j

or

cmake -S . -B build-vulkan
-DGGML_VULKAN=ON
-DLLAMA_BUILD_WEBUI=ON
-DLLAMA_SERVER_TOOLS=ON

cmake --build build-vulkan -j

Comando sugerido:

"/home/userk21/llama_OFFFELLIA_1984/build/bin/llama-server"
-m "/home/userk21/Área de trabalho/userk21/LLMS/ΩFFFΣLLIα_IQ4_NL_gemma-4-26B-A4B-it-ultra-uncensored-heretic.gguf"
-ngl 99 --n-cpu-moe 99
-c 50000
-ctk q8_0
-ctv q8_0
-t 4
-tb 4
-b 2048
-ub 1024
-fa on
--cpu-strict 1
--parallel 1
--agent
--tools all
--reasoning auto
--kv-unified
--load-mode mmap
--cors-origins "*"
--webui-mcp-proxy
--threads-http -1
--port 5173
--host 127.0.0.1


📜 Licença

Distribuído sob a licença MIT. Veja o arquivo LICENSE para mais detalhes.
Gracias https://github.com/charlie12345/ROCmFPX


ΩFFFΣLLIa • llama.cpp • AlgMor24 — Inferência Local, Descentralizada e Sem Limites
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Abliteration, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.