← back to catalog · registered 2026-08-22 13:56

ailexleon/gemma-4-26B-A4B-it-qat-uncensored-heretic-mlx-lm-4Bit

ailexleon Gemma 25B MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/ailexleon%2Fgemma-4-26B-A4B-it-qat-uncensored-heretic-mlx-lm-4Bit"
Response includes
  • classification m3
  • files 11
  • hub_downloads_all_time 3,302
  • author_summary 15 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
3K
497 last 30d - stable
Likes
1
Model age
3mo ago
created 2026-07-05
Downloads over time
Now3.5K→from959↑268%
8311.8K2.8K3.8K959 on Jul 153.5K on Oct 11JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors gemma4 heretic uncensored decensored abliterated ara text-generation conversational en base_model:llmfan46/gemma-4-26B-A4B-it-qat-q4_0-unquantized-uncensored-heretic

Related

Total size
13.2 GB
Files
11
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-18 08:24

Files by quantization

Auxiliary files 11 files 13.3 GB
model-00002-of-00003.safetensors 4.99 GB 3aa49775 download
model-00001-of-00003.safetensors 4.95 GB 0c5faeaa download
model-00003-of-00003.safetensors 3.28 GB 2f4c28ce download
tokenizer.json 30.7 MB cc8d3a0c download
model.safetensors.index.json 136 KB 773ce92c download
chat_template.jinja 18.2 KB 4741bf6e download
config.json 10.2 KB 4e6c75b9 download
tokenizer_config.json 2.78 KB 3e9c7cde download
.gitattributes 1.53 KB 52373fe2 download
README.md 1.24 KB bcce0ec1 download
generation_config.json 218 B 44dfdff4 download

README current version from Hugging Face


base_model: llmfan46/gemma-4-26B-A4B-it-qat-q4_0-unquantized-uncensored-heretic
license: apache-2.0
license_link: https://ai.google.dev/gemma/docs/gemma_4_license
language: en
tags:

  • mlx
  • heretic
  • uncensored
  • decensored
  • abliterated
  • ara
    library_name: mlx
    pipeline_tag: text-generation

ailexleon/gemma-4-26B-A4B-it-qat-uncensored-heretic-mlx-lm-4Bit

Converted to MLX format from llmfan46/gemma-4-26B-A4B-it-qat-q4_0-unquantized-uncensored-heretic using mlx-lm version 0.31.3.

If you need vision capabilities, use ailexleon/gemma-4-26B-A4B-it-qat-uncensored-heretic-mlx-vlm-4Bit.

Use with mlx

pip install mlx-lm
from mlx_lm import load, generate

model, tokenizer = load("ailexleon/gemma-4-26B-A4B-it-qat-uncensored-heretic-mlx-lm-4Bit")

prompt = "hello"

if tokenizer.chat_template is not None:
    messages = [{"role": "user", "content": prompt}]
    prompt = tokenizer.apply_chat_template(
        messages, add_generation_prompt=True, return_dict=False,
    )

response = generate(model, tokenizer, prompt=prompt, verbose=True)

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-05Update README.md93decf31.2 KB
    Loading...
  2. 2026-07-05Update README.md75cea791.1 KB
    Loading...
  3. 2026-07-05Add files using upload-large-folder tool0067185633 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration