← back to catalog · registered 2026-08-25 19:02

HangGlidersRule/Darkstar-Nemotron-3.5-Lightning-30B-A3B-Abliterated-BF16

HangGlidersRule Nemotron 33B MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/HangGlidersRule%2FDarkstar-Nemotron-3.5-Lightning-30B-A3B-Abliterated-BF16"
Response includes
  • classification m1
  • files 32
  • hub_downloads_all_time 405
  • author_summary 6 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
405
178 last 30d - stable
Likes
1
Descendants
3
in 3 direct forks
Model age
6w ago
created 2026-08-25
Downloads over time
Now489→from7↑6,886%
01793585377 on Aug 26489 on Oct 11AugSepOct
Aug 26 → Oct 11 · 47 snapshots · spans 46 days

Genealogy 3 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en
Tags
safetensors nemotron_h darkstar nemotron-h abliterated reduced-refusal bf16 vllm text-generation conversational base_model:nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 base_model:finetune:nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

Related

Total size
61.3 GB
Files
32
Quantizations
1
Registered
2026-08-25 19:02
Last updated on HF
2026-09-02 15:30

Files by quantization

Auxiliary files 32 files 61.3 GB
model-00012-of-00014.safetensors 4.65 GB dcbce8c6 download
model-00004-of-00014.safetensors 4.65 GB fcd936d4 download
model-00009-of-00014.safetensors 4.65 GB 5f8c2677 download
model-00011-of-00014.safetensors 4.65 GB 82cfee86 download
model-00008-of-00014.safetensors 4.65 GB 542723f2 download
model-00010-of-00014.safetensors 4.65 GB cda8bc99 download
model-00007-of-00014.safetensors 4.65 GB f85b4f8f download
model-00003-of-00014.safetensors 4.65 GB df342d5d download
model-00002-of-00014.safetensors 4.65 GB ed0dafaf download
model-00001-of-00014.safetensors 4.65 GB 124f8c46 download
model-00006-of-00014.safetensors 4.65 GB f788278e download
model-00005-of-00014.safetensors 4.64 GB b7e54f0a download
model-00013-of-00014.safetensors 3.03 GB 3922fce6 download
model-00014-of-00014.safetensors 2.49 GB 63196ad4 download
tokenizer.json 16.3 MB 623c3456 download
abliteration_report.json 689 KB e0e936d3 download
model.safetensors.index.json 598 KB 3cd40944 download
agentic_coding_benchmarks.png 215 KB 06cf5486 download
tokenizer_config.json 173 KB c96e5ad0 download
accuracy_plot.png 137 KB feecce07 download
chat_template.jinja 9.64 KB d85b0c77 download
README.md 3.77 KB 61cee1ec download
_SUCCESS.json 3.34 KB 8c27c1bc download
bias.md 3.19 KB cdcc8055 download
explainability.md 2.64 KB 2435f23c download
LICENSE 2.63 KB 4d76cc87 download
config.json 2.30 KB 993b6129 download
safety.md 2.09 KB f53bf94a download
.gitattributes 1.65 KB dd95036d download
privacy.md 828 B e3bf30aa download
special_tokens_map.json 563 B 0451f379 download
generation_config.json 210 B a41201df download

README current version from Hugging Face


license: other
license_name: openmdw-1.1
license_link: https://openmdw.ai/license/1-1/
language:

  • en
    tags:
  • nemotron-h
  • darkstar
  • abliteration
  • hybrid-mamba-moe
    pipeline_tag: text-generation
    base_model: nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
    base_model_relation: finetune
    extra_gated_heading: Darkstar Nemotron-3.5-Lightning 30B-A3B Abliterated BF16

Darkstar-Nemotron-3.5-Lightning-30B-A3B-Abliterated-BF16

Safety notice: This checkpoint is an edited (abliterated) derivative: the
refusal direction measured at layer 34 was deliberately projected out of 3,126
residual-writing tensors (attention output, Mamba output, routed/shared expert
down-projection, MTP, and embedding weights). It will tend to comply with
harmful requests. It is released for red-teaming, interpretability research,
and alignment experimentation only. Also see Safety.

A BF16, refusal-direction-edited derivative of NVIDIA's
Nemotron-3.5-Lightning-30B-A3B-BF16
(hybrid Mamba2 + MoE + sparse attention, 52 layers, 262,144-token context,
OpenMDW-1.1). Weights are the original BF16 tensors with the single normalized
harmful-minus-harmless direction removed in float32; everything outside the
3,126 targets is byte-identical to upstream.

Product family

Product Format Edit Status
Base-BF16 BF16 none (upstream reference) not republished here
Base-ModelOpt-NVFP4 W4A16 NVFP4 none sibling repository
Abliterated-BF16 BF16 refusal direction removed this repository
Abliterated-ModelOpt-NVFP4 W4A16 NVFP4 refusal direction removed sibling repository

Edit contract (reproducible)

  • Source: nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
    @ d468880b6ad3c6e0d21377ce7242adaea4cc884d
  • Direction: layer 34, seed 42, 320 harmful / 320 harmless prompts
    (mlabonne/harmful_behaviors + harmless_alpaca, chat-templated),
    harmful-minus-harmless unit vector (norm 1.0000, dim 2688), selected by
    refusal-generation test (0/8) across all 52 layers
  • Projection: W' = W − r(rᵀW) in float32, shard-by-shard
  • Targets (3,126): mixer.o_proj (6), mixer.out_proj (23),
    mixer.experts.*.down_proj (2,944), mixer.shared_experts.down_proj (23),
    mtp.layers.0.mixer.o_proj (1), mtp.layers.1.mixer.experts.*.down_proj (128),
    backbone.embeddings.weight (1)
  • Validation: 3,126/3,126 edited, max normalized residual leakage 0.000160
    (gate ≤ 0.01), MTP head intact (270 tensors), no vision tensors touched
  • Behavior gate: 200/200 harmful compliance, 0/83 safe over-refusals, 0 errors
    (refusal-form marker set, temp 0, max_tokens 100)
  • Recipe: recipes/nemotron-3.5-lightning/darkstar-nemotron-3.5-lightning-30b-a3b-abliterated-bf16.yaml
    in HangGlidersRule/model-forge

Publication

This checkpoint is public on Hugging Face at the pinned milestone tag
darkstar-nemotron-3.5-lightning-v1.0.0. Weights are hash-verified (sha256
manifest in the source repo) and serve with vLLM:

vllm serve HangGlidersRule/Darkstar-Nemotron-3.5-Lightning-30B-A3B-Abliterated-BF16 \
  --max-model-len 131072 --kv-cache-dtype bfloat16 --reasoning-parser nemotron_v3 \
  --speculative-config '{"method":"mtp","num_speculative_tokens":12}'

License

OpenMDW-1.1 (same as upstream). See https://openmdw.ai/license/1-1/. Retain all
NVIDIA copyright/attribution/notice lines in distributions of this derivative.

Safety

This model intentionally has a reduced refusal response. Do not deploy in
user-facing assistant roles without alignment hardening and content filtering.
It is intended for researchers studying refusal behavior, ablation, and
alignment techniques.

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-02Card format: conform to gold-standard Darkstar card skeleton7ddcc403.6 KB
    Loading...
  2. 2026-08-25update card: publication section, milestone tag, dashed H1f3723fc3.8 KB
    Loading...
  3. 2026-08-25publish model card: Darkstar Nemotron-3.5-Lightning (R1 abliterated)134352c3.3 KB
    Loading...
  4. 2026-08-25publish weightsd36f7ce82.3 KB
    Loading...
  5. 2026-08-25publish model card: Darkstar Nemotron-3.5-Lightning (R1 abliterated)dbb77ef3.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration