← back to catalog · registered 2026-08-22 13:56

Unrestricted/Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive

Unrestricted Nemotron 4B GGUF 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Unrestricted%2FNemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive"
Response includes
  • classification m-uncensored
  • files 14
  • hub_downloads_all_time 1,319
  • author_summary 22 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1K
396 last 30d - stable
Likes
0
Model age
4mo ago
created 2026-06-10
Downloads over time
Now1.4K→from52↑2,621%
05171K1.6K52 on Jun 101.4K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en
Quantizations
IQ2 IQ3 IQ4 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_K
Tags
gguf uncensored nemotron mamba2 hybrid en base_model:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 base_model:quantized:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 license:other endpoints_compatible region:us imatrix

Related

Total size
32.7 GB
Files
14
Quantizations
10
Registered
2026-08-22 13:56
Last updated on HF
2026-06-10 14:37

Files by quantization

Q8_K 1 file 4.37 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf 4.37 GB b527227d download
Q6_K 1 file 3.74 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf 3.74 GB 6a0f0a7c download
Q5_K 2 files 5.95 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf 3.07 GB 6e4bce12 download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_M.gguf 2.88 GB 5adeca43 download
Q4_K 2 files 5.44 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf 2.80 GB b3efc555 download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf 2.64 GB a7ef3fb1 download
Q3_K 2 files 4.57 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf 2.34 GB c9989e0c download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf 2.23 GB 39954e32 download
IQ4 1 file 2.25 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ4_XS.gguf 2.25 GB e58f67ad download
Q2_K 1 file 2.17 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf 2.17 GB 43c1f19a download
IQ3 1 file 2.17 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf 2.17 GB 0ee07bfc download
IQ2 1 file 2.03 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf 2.03 GB 3f8db090 download
Auxiliary files 2 files 9.39 KB
README.md 6.77 KB 725448e9 download
.gitattributes 2.62 KB ec4ee874 download

README current version from Hugging Face


license: other
license_name: nvidia-open-model-license
license_link: https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-open-model-license/
tags:

  • uncensored
  • nemotron
  • mamba2
  • hybrid
    language:
  • en
    base_model: nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive

Join the Discord for updates, roadmaps, projects, or just to chat.

NVIDIA Nemotron-3 Nano 4B uncensored by HauhauCS. 0/465 refusals.

HuggingFace's "Hardware Compatibility" widget doesn't recognize K_P quants — it may show fewer files than actually exist. Click "View +X variants" or go to Files and versions to see all available downloads.

About

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.

These are meant to be the best lossless uncensored models out there.

This release has NVIDIA's GenRM (generative reward model) fully removed. GenRM acts as an internal critic that scores and filters the model's own outputs — effectively a second layer of censorship on top of the base refusals. Removing it gives you the raw model output without any self-censoring.

For a comparison build with GenRM still active, see Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM (IQ2_M only, for side-by-side testing).

Aggressive Variant

Stronger uncensoring — model is fully unlocked and won't refuse prompts. May occasionally append short disclaimers (baked into base model training, not refusals) but full content is always generated.

For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available.

Downloads

File Quant BPW Size
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf Q8_K_P 9.4 4.4 GB
— Q8_0 8.5 —
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf Q6_K_P 7.0 3.8 GB
— Q6_K 6.6 —
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf Q5_K_P 6.1 3.1 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_M.gguf Q5_K_M 5.7 2.9 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf Q4_K_P 5.2 2.9 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf Q4_K_M 4.8 2.7 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ4_XS.gguf IQ4_XS 4.3 2.3 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf Q3_K_P 4.1 2.4 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf Q3_K_M 3.9 2.3 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf IQ3_M 3.7 2.2 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf Q2_K_P 3.5 2.2 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf IQ2_M 2.7 2.1 GB

All quants generated with importance matrix (imatrix) for optimal quality preservation on abliterated weights.

What are K_P quants?

K_P ("Perfect") quants are HauhauCS custom quantizations that use model-specific analysis to selectively preserve quality where it matters most. Each model gets its own optimized quantization profile.

A K_P quant effectively bumps quality up by 1-2 quant levels at only ~5-15% larger file size than the base quant. Fully compatible with llama.cpp, LM Studio, and any GGUF-compatible runtime — no special builds needed.

Note: K_P quants may show as "?" in LM Studio's quant column. This is a display issue only — the model loads and runs fine.

Specs

  • 3.97B parameters
  • Hybrid Mamba2-Transformer architecture (42 layers: 21 Mamba2, 17 MLP, 4 Attention)
  • 262K native context
  • Thinking/reasoning mode (toggleable)
  • Tool calling support
  • Compressed from NVIDIA Nemotron-Nano-9B-v2 using Nemotron Elastic framework
  • Based on nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Recommended Settings

From the official NVIDIA authors:

Reasoning mode (default, thinking enabled):

  • temperature=1.0, top_p=0.95

Tool calling:

  • temperature=0.6, top_p=0.95

Disabling reasoning:

  • Set enable_thinking=False in chat template — trades accuracy for speed on simpler tasks

Usage

Works with llama.cpp, LM Studio, Jan, koboldcpp, and other GGUF-compatible runtimes.

llama-cli -m Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --jinja -c 131072 -ngl 99

Note: LM Studio may show unexpected values in architecture/params columns — this is a display quirk with hybrid Mamba models, the model runs correctly.

Other Versions

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-10Upload README.md with huggingface_hub4f650646.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration