← back to catalog · registered 2026-08-22 13:56

divinetribe/Nemotron-3-Nano-Omni-30B-Abliterated-MM-4bit

divinetribe Nemotron 32B MoE multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/divinetribe%2FNemotron-3-Nano-Omni-30B-Abliterated-MM-4bit"
Response includes
  • classification m1
  • files 16
  • hub_downloads_all_time 257
  • author_summary 14 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
257
96 last 30d - stable
Likes
0
Model age
8w ago
created 2026-08-13
Downloads over time
Now283→from80↑254%
7014822530380 on Aug 19283 on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Tags
mlx safetensors NemotronH_Nano_Omni_Reasoning_V3 abliterated multimodal vision audio nemotron image-text-to-text conversational base_model:mlx-community/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-bf16 base_model:quantized:mlx-community/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-bf16

Related

Total size
18.3 GB
Files
16
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-22 20:27

Files by quantization

Auxiliary files 16 files 18.3 GB
model-00002-of-00004.safetensors 4.93 GB 2cfc974e download
model-00003-of-00004.safetensors 4.93 GB 49a1143c download
model-00001-of-00004.safetensors 4.77 GB a3931410 download
model-00004-of-00004.safetensors 3.66 GB a155f947 download
tokenizer.json 16.3 MB d0432ad8 download
model.safetensors.index.json 227 KB 0d4dc94a download
chat_template.jinja 13.9 KB 6381c9df download
config.json 11.6 KB 6e354510 download
README.md 2.34 KB a17a1e58 download
.gitattributes 1.53 KB 52373fe2 download
processor_config.json 795 B df33a33e download
abliteration_info.json 663 B ca99af2d download
tokenizer_config.json 567 B c6f6caad download
special_tokens_map.json 420 B 0d5c2c9d download
preprocessor_config.json 342 B 7317205a download
generation_config.json 309 B ae4d74d4 download

README current version from Hugging Face


license: other
base_model: mlx-community/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-bf16
tags:

  • mlx
  • abliterated
  • multimodal
  • vision
  • audio
  • nemotron
    pipeline_tag: image-text-to-text

Nemotron-3-Nano-Omni-30B Abliterated (4-bit, MLX)

Abliterated build of NVIDIA's tri-modal Nemotron-3-Nano-Omni-30B for Apple Silicon.
Text, vision, and audio all work — the refusal direction was removed from the
language stack only; the vision and sound towers are copied through untouched.

As far as I can tell this is the first abliterated Omni in MLX — the other uncensored
builds are NVFP4 / GGUF and don't run on a Mac.

  • Precision: 4-bit (~18GB)
  • Refusal removed: verified 5/6 on a held-out harmful set
  • Vision: describes images correctly
  • Audio: transcribes speech
  • Tensors: 1,481 — 401 language (abliterated) + 390 vision + 684 sound + 6 projection, name-for-name identical to the base

Run it

pip install -U mlx-vlm
python -m mlx_vlm generate --model divinetribe/Nemotron-3-Nano-Omni-30B-Abliterated-MM-4bit \
  --image your_image.jpg --prompt "Describe this image."

Needs a recent mlx-vlm (0.6.12+) with Omni support. Runtime reference:
https://github.com/nicedreamzapp/nemotron-omni-mlx

Method

Directional ablation (Arditi et al.). The refusal direction on this model is spread
across two blocks (16 and 31), not one — a single-layer ablation leaves it refusing.
Both directions are Gram-Schmidt'd and orthogonalized out of every residual-writing
projection: mamba out_proj, attention o_proj, MoE routed-expert fc2, shared-expert
down_proj, plus the token embeddings.

Use responsibly

Safety alignment has been removed. You are responsible for what you generate.


Part of Claude Code Local

This model is one of the fighters in Claude Code Local (3.2k★), which runs Claude Code 100% on-device on Apple Silicon through an MLX-native Anthropic-API server. Not sure which local model to run as an agent? Check the Agent-12 local agent leaderboard: real agent tasks, judged by the filesystem, same hardware for every row.

Built by Matt Macosko in Arcata, CA. Open to work on local-AI and Apple Silicon inference: [email protected].

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-22Model card: link to Claude Code Local and the Agent-12 leaderboardc4d2b9b2.3 KB
    Loading...
  2. 2026-08-13Abliterated Nemotron Omni (4-bit) — vision+audio intact1a3eef71.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration