← back to catalog · registered 2026-08-22 13:56

HauhauCS/Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive

HauhauCS Nemotron 4B GGUF 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/HauhauCS%2FNemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive"
Response includes
  • classification m-uncensored
  • files 14
  • hub_downloads_all_time 49,153
  • author_summary 26 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
49K
5K last 30d - stable
Likes
43
Model age
6mo ago
created 2026-03-24
Downloads over time
Now52.4K→from3.2K↑1,541%
019.1K38.3K57.4K3.2K on Mar 2552.4K on Oct 11MarAprMayJunJulAugSepOct
Mar 25 → Oct 11 · 70 snapshots · spans 200 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en
Quantizations
IQ2 IQ3 IQ4 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_K
Tags
gguf uncensored nemotron mamba2 hybrid en base_model:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 base_model:quantized:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 license:other endpoints_compatible region:us imatrix

Related

Total size
32.7 GB
Files
14
Quantizations
10
Registered
2026-08-22 13:56
Last updated on HF
2026-04-05 19:00

Files by quantization

Q8_K 1 file 4.37 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf 4.37 GB b527227d download
Q6_K 1 file 3.74 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf 3.74 GB 6a0f0a7c download
Q5_K 2 files 5.95 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf 3.07 GB 6e4bce12 download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_M.gguf 2.88 GB 5adeca43 download
Q4_K 2 files 5.44 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf 2.80 GB b3efc555 download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf 2.64 GB a7ef3fb1 download
Q3_K 2 files 4.57 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf 2.34 GB c9989e0c download
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf 2.23 GB 39954e32 download
IQ4 1 file 2.25 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ4_XS.gguf 2.25 GB e58f67ad download
Q2_K 1 file 2.17 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf 2.17 GB 43c1f19a download
IQ3 1 file 2.17 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf 2.17 GB 0ee07bfc download
IQ2 1 file 2.03 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf 2.03 GB 3f8db090 download
Auxiliary files 2 files 9.39 KB
README.md 6.77 KB 725448e9 download
.gitattributes 2.62 KB 59d4f00e download

README current version from Hugging Face


license: other
license_name: nvidia-open-model-license
license_link: https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-open-model-license/
tags:

  • uncensored
  • nemotron
  • mamba2
  • hybrid
    language:
  • en
    base_model: nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive

Join the Discord for updates, roadmaps, projects, or just to chat.

NVIDIA Nemotron-3 Nano 4B uncensored by HauhauCS. 0/465 refusals.

HuggingFace's "Hardware Compatibility" widget doesn't recognize K_P quants — it may show fewer files than actually exist. Click "View +X variants" or go to Files and versions to see all available downloads.

About

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.

These are meant to be the best lossless uncensored models out there.

This release has NVIDIA's GenRM (generative reward model) fully removed. GenRM acts as an internal critic that scores and filters the model's own outputs — effectively a second layer of censorship on top of the base refusals. Removing it gives you the raw model output without any self-censoring.

For a comparison build with GenRM still active, see Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM (IQ2_M only, for side-by-side testing).

Aggressive Variant

Stronger uncensoring — model is fully unlocked and won't refuse prompts. May occasionally append short disclaimers (baked into base model training, not refusals) but full content is always generated.

For a more conservative uncensor that keeps some safety guardrails, check the Balanced variant when it's available.

Downloads

File Quant BPW Size
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q8_K_P.gguf Q8_K_P 9.4 4.4 GB
— Q8_0 8.5 —
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q6_K_P.gguf Q6_K_P 7.0 3.8 GB
— Q6_K 6.6 —
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_P.gguf Q5_K_P 6.1 3.1 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q5_K_M.gguf Q5_K_M 5.7 2.9 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf Q4_K_P 5.2 2.9 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf Q4_K_M 4.8 2.7 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ4_XS.gguf IQ4_XS 4.3 2.3 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_P.gguf Q3_K_P 4.1 2.4 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q3_K_M.gguf Q3_K_M 3.9 2.3 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ3_M.gguf IQ3_M 3.7 2.2 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q2_K_P.gguf Q2_K_P 3.5 2.2 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-IQ2_M.gguf IQ2_M 2.7 2.1 GB

All quants generated with importance matrix (imatrix) for optimal quality preservation on abliterated weights.

What are K_P quants?

K_P ("Perfect") quants are HauhauCS custom quantizations that use model-specific analysis to selectively preserve quality where it matters most. Each model gets its own optimized quantization profile.

A K_P quant effectively bumps quality up by 1-2 quant levels at only ~5-15% larger file size than the base quant. Fully compatible with llama.cpp, LM Studio, and any GGUF-compatible runtime — no special builds needed.

Note: K_P quants may show as "?" in LM Studio's quant column. This is a display issue only — the model loads and runs fine.

Specs

  • 3.97B parameters
  • Hybrid Mamba2-Transformer architecture (42 layers: 21 Mamba2, 17 MLP, 4 Attention)
  • 262K native context
  • Thinking/reasoning mode (toggleable)
  • Tool calling support
  • Compressed from NVIDIA Nemotron-Nano-9B-v2 using Nemotron Elastic framework
  • Based on nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Recommended Settings

From the official NVIDIA authors:

Reasoning mode (default, thinking enabled):

  • temperature=1.0, top_p=0.95

Tool calling:

  • temperature=0.6, top_p=0.95

Disabling reasoning:

  • Set enable_thinking=False in chat template — trades accuracy for speed on simpler tasks

Usage

Works with llama.cpp, LM Studio, Jan, koboldcpp, and other GGUF-compatible runtimes.

llama-cli -m Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-Q4_K_P.gguf \
  --jinja -c 131072 -ngl 99

Note: LM Studio may show unexpected values in architecture/params columns — this is a display quirk with hybrid Mamba models, the model runs correctly.

Other Versions

README history 8 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-04-05Upload README.md with huggingface_hub8d58c5d6.8 KB
    Loading...
  2. 2026-03-25Upload README.md with huggingface_hubd4761bc6.7 KB
    Loading...
  3. 2026-03-25Upload README.md with huggingface_huba83b4af6.7 KB
    Loading...
  4. 2026-03-25Upload README.md with huggingface_hub17dcb0e6.7 KB
    Loading...
  5. 2026-03-24Upload README.md with huggingface_hubdb99ae05.2 KB
    Loading...
  6. 2026-03-24Upload README.md with huggingface_huba033f345.2 KB
    Loading...
  7. 2026-03-24Upload README.md with huggingface_hub9d382bc5 KB
    Loading...
  8. 2026-03-24Upload README.md with huggingface_hub82cac9e5 KB
    Loading...

Discussions 4 threads

  1. 2026-08-14Nemotron 3.5 30B-A3B?open1 💬#4
    Loading...
  2. 2026-04-05Discord serveropen1 💬#3
    Loading...
  3. 2026-03-27mmproj?open1 💬#2
    Loading...
  4. 2026-03-25请问xinference能运行微调后的gguf文件吗open1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration