← back to catalog · registered 2026-08-22 13:56

justfrfn/Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM

justfrfn Nemotron 4B GGUF 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/justfrfn%2FNemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM"
Response includes
  • classification m-uncensored
  • files 3
  • hub_downloads_all_time 232
  • author_summary 22 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
232
15 last 30d - cooling
Likes
0
Model age
4mo ago
created 2026-06-10
Downloads over time
Now234→from136↑72%
131169206244136 on Jun 10234 on Oct 11234 on Oct 7JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en
Quantizations
IQ2
Tags
gguf uncensored nemotron mamba2 hybrid en base_model:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 base_model:quantized:nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 license:other endpoints_compatible region:us imatrix

Related

Total size
2.03 GB
Files
3
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-06-10 19:01

Files by quantization

IQ2 1 file 2.03 GB
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM-IQ2_M.gguf 2.03 GB 0f488c48 download
Auxiliary files 2 files 4.29 KB
README.md 2.70 KB f32711a4 download
.gitattributes 1.58 KB 77371db2 download

README current version from Hugging Face


license: other
license_name: nvidia-open-model-license
license_link: https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-open-model-license/
tags:

  • uncensored
  • nemotron
  • mamba2
  • hybrid
    language:
  • en
    base_model: nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM

Join the Discord for updates, roadmaps, projects, or just to chat.

This is NOT the recommended release. This repo exists purely for A/B comparison testing. For the fully uncensored model with GenRM removed, use Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive instead.

What is this?

This is an earlier abliterated build that has NVIDIA's GenRM (generative reward model) still active. The abliteration itself scores 0/465 refusals — same as the main release — but GenRM acts as a second layer of censorship that re-introduces refusals at generation time, skewing the effective result to roughly ~10/465.

Why does this exist?

To let people see the difference GenRM makes. This is the first publicly available abliteration of a model with GenRM, so this comparison build is useful for research and understanding how GenRM works.

How GenRM manifests

When GenRM intervenes, you'll see a clear 180-degree reversal between the Chain-of-Thought and the final output. The model will reason through the request normally in its thinking block, then GenRM kicks in and the visible output contradicts what the CoT was building toward — typically with a refusal or deflection.

This doesn't happen on every prompt — only on topics where GenRM's reward signal is strong enough to override the (abliterated) base behavior.

Downloads

File Quant Size
Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRM-IQ2_M.gguf IQ2_M 2.1 GB

Only IQ2_M provided — this is for comparison testing, not daily use.

Specs

  • 3.97B parameters
  • Hybrid Mamba2-Transformer architecture (42 layers: 21 Mamba2, 17 MLP, 4 Attention)
  • 262K native context
  • Thinking/reasoning mode (toggleable)
  • Tool calling support
  • Based on nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16

Use the real release instead

Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive — full release with GenRM removed, multiple quant formats, 0/465 refusals.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-10Duplicate from HauhauCS/Nemotron3-Nano-4B-Uncensored-HauhauCS-Aggressive-GenRMc2eb8c82.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration