← back to catalog · registered 2026-08-22 13:56

DBMe/gemma-3-27b-it-ultra-uncensored-heretic-exl3

DBMe Gemma 27B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/DBMe%2Fgemma-3-27b-it-ultra-uncensored-heretic-exl3"
Response includes
  • classification m3
  • files 4
  • benchmarks 11 entries
  • hub_downloads_all_time 20
  • author_summary 7 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
20
11 last 30d - active
Likes
0
Model age
6mo ago
created 2026-04-14
Downloads over time
Now24→from4↑500%
31118264 on Apr 1524 on Oct 1124 on Oct 7AprMayJunJulAugSepOct
Apr 15 → Oct 11 · 65 snapshots · spans 179 days

Benchmarks

Benchmark Score Source
Entertainment 1.4 UGI
Hazardous 1.8 UGI
Natural Intelligence 30.96 UGI
Political lean -14.7% UGI
Sensitive-Info 17.41 UGI
SocPol 2.2 UGI
UGI 44.94 UGI
Willingness (10) 10 UGI
W10-Adherence 10 UGI
W10-Direct 10 UGI
Writing 40.09 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
gemma
Tags
exllamav3 exl3 text-generation base_model:llmfan46/gemma-3-27b-it-ultra-uncensored-heretic base_model:quantized:llmfan46/gemma-3-27b-it-ultra-uncensored-heretic license:gemma region:us

Related

Total size
0 B
Files
4
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-04-15 15:34

Files by quantization

Auxiliary files 4 files 76.4 KB
metrics_graph.png 70.1 KB b4eb732a download
README.md 3.45 KB 86752c99 download
.gitattributes 1.48 KB a6344aac download
metrics.json 1.36 KB bb83723e download

README current version from Hugging Face


base_model: llmfan46/gemma-3-27b-it-ultra-uncensored-heretic
base_model_relation: quantized
quantized_by: DBMe
library_name: exllamav3
pipeline_tag: text-generation
license: gemma
tags:

  • exl3

DBMe/gemma-3-27b-it-ultra-uncensored-heretic-exl3

EXL3 (ExLlamaV3) quantizations of llmfan46/gemma-3-27b-it-ultra-uncensored-heretic. All credit for the original model goes to the original authors.

📊 Available Quantizations & VRAM

The model weights are stored in separate branches. Please switch to a branch to download.
Note: VRAM estimates include PyTorch context overhead (~0.8GB) and assume an unquantized FP16 KV cache.

Target BPW Head BPW Branch (Download Link) WikiText-2 PPL (512 ctx)¹ 2K ctx 4K ctx 8K ctx 16K ctx 32K ctx
4.0 h6 4.0bpw_h6 8.5269 ~18.1 GB ~19.07 GB ~21.01 GB ~24.89 GB ~32.64 GB
5.0 h6 5.0bpw_h6 N/A ~21.08 GB ~22.05 GB ~23.99 GB ~27.87 GB ~35.62 GB
6.0 h6 6.0bpw_h6 8.4439 ~24.06 GB ~25.03 GB ~26.97 GB ~30.85 GB ~38.6 GB
8.0 h8 8.0bpw_h8 8.4301 ~30.35 GB ~31.32 GB ~33.26 GB ~37.13 GB ~44.88 GB

¹ Evaluated against WikiText-2 with ExLlamaV3 using a strided 512-token context window (-c 512) in llama.cpp parity mode (-g). Lower is better.
(Higher BPW = higher quality, lower BPW = fits in less VRAM).

📥 How to Download

It's recommended to use the huggingface-cli to download specific branches. (Do not use git clone as it will download all branches!)

Ensure you have the CLI installed:

pip install -U "huggingface_hub[cli]"

Download a specific branch (e.g., 4.0bpw_h6):

# Example: Downloading the 4.0bpw_h6 branch
huggingface-cli download DBMe/gemma-3-27b-it-ultra-uncensored-heretic-exl3 --revision 4.0bpw_h6 --local-dir gemma-3-27b-it-ultra-uncensored-heretic-exl3-4.0bpw_h6

💻 Supported Engines

These models are highly optimized for modern GPUs and can be run using:

  • TabbyAPI: A fast, OpenAI-compatible API server. (Set model_name: "gemma-3-27b-it-ultra-uncensored-heretic-exl3-<BranchName>" in your config)
  • Text-Generation-WebUI: A local web interface. (Select the exllamav3 loader)
  • ExLlamaV3 (Native): Python library for custom integration.

📈 Perplexity Degradation Curve

(Lower is better)
Perplexity Graph

⚙️ Advanced: Quantization Environment & Settings

🔬 Quantization Settings

  • Codebook: mcg

  • Output Scales: always

  • Calibration Rows: 250

  • Calibration Cols: 2048

  • Calibration Dataset: ExLlamaV3 Default (Wiki/C4/Code)

  • High Quality (HQ) Mode: False

  • ExLlamaV3: 0.0.29 (Commit: cb1a436)

  • Hardware: NVIDIA RTX PRO 6000 Blackwell Server Edition

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-15Update main README with VRAM Matrix344bb033.5 KB
    Loading...
  2. 2026-04-15Update main README with VRAM Matrix41687123.3 KB
    Loading...
  3. 2026-04-15Update main README with VRAM Matrixdf897f63.1 KB
    Loading...
  4. 2026-04-14Update main README with VRAM Matrix2d26f742.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration