license: other
base_model: mlx-community/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-bf16
tags:
- mlx
- abliterated
- multimodal
- vision
- audio
- nemotron
pipeline_tag: image-text-to-text
Nemotron-3-Nano-Omni-30B Abliterated (full, MLX)
Abliterated build of NVIDIA's tri-modal Nemotron-3-Nano-Omni-30B for Apple Silicon.
Text, vision, and audio all work — the refusal direction was removed from the
language stack only; the vision and sound towers are copied through untouched.
As far as I can tell this is the first abliterated Omni in MLX — the other uncensored
builds are NVFP4 / GGUF and don't run on a Mac.
- Precision: full (~62GB)
- Refusal removed: verified 6/6 on a held-out harmful set
- Vision: describes images correctly
- Audio: transcribes speech
- Tensors: 1,481 — 401 language (abliterated) + 390 vision + 684 sound + 6 projection, name-for-name identical to the base
Run it
pip install -U mlx-vlm
python -m mlx_vlm generate --model divinetribe/Nemotron-3-Nano-Omni-30B-Abliterated-MM-bf16 \
--image your_image.jpg --prompt "Describe this image."
Needs a recent mlx-vlm (0.6.12+) with Omni support. Runtime reference:
https://github.com/nicedreamzapp/nemotron-omni-mlx
Method
Directional ablation (Arditi et al.). The refusal direction on this model is spread
across two blocks (16 and 31), not one — a single-layer ablation leaves it refusing.
Both directions are Gram-Schmidt'd and orthogonalized out of every residual-writing
projection: mamba out_proj, attention o_proj, MoE routed-expert fc2, shared-expertdown_proj, plus the token embeddings.
Use responsibly
Safety alignment has been removed. You are responsible for what you generate.
Part of Claude Code Local
This model is one of the fighters in Claude Code Local (3.2k★), which runs Claude Code 100% on-device on Apple Silicon through an MLX-native Anthropic-API server. Not sure which local model to run as an agent? Check the Agent-12 local agent leaderboard: real agent tasks, judged by the filesystem, same hardware for every row.
Built by Matt Macosko in Arcata, CA. Open to work on local-AI and Apple Silicon inference: [email protected].