← back to catalog · registered 2026-08-22 13:56

Firworks/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated-nvfp4

Firworks Llama 7.4B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Firworks%2FL3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated-nvfp4"
Response includes
  • classification m3
  • files 13
  • hub_downloads_all_time 166
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
166
27 last 30d - stable
Likes
1
Model age
6mo ago
created 2026-03-29

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now181→from63↑187%
5710214819363 on Apr 1181 on Oct 11AprMayJunJulAugSepOct
Apr 1 → Oct 11 · 67 snapshots · spans 193 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
safetensors llama nvfp4 fp4 quantized dataset:xensive/roleplaydataset100k base_model:DavidAU/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated base_model:quantized:DavidAU/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated 8-bit compressed-tensors region:us

Related

Total size
9.73 GB
Files
13
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-03-29 05:13

Files by quantization

Auxiliary files 13 files 9.74 GB
model-00001-of-00003.safetensors 4.66 GB 8db69d6d download
model-00002-of-00003.safetensors 4.09 GB be97bec6 download
model-00003-of-00003.safetensors 1002 MB 9e8754d0 download
tokenizer.json 16.4 MB 8fc5ed64 download
model.safetensors.index.json 181 KB 1011179d download
tokenizer_config.json 49.5 KB 7bc54330 download
config.json 1.96 KB e4445cb8 download
.gitattributes 1.53 KB 52373fe2 download
README.md 1.52 KB a87b749f download
special_tokens_map.json 439 B 344c8261 download
chat_template.jinja 389 B 39bd0c9f download
recipe.yaml 303 B 1e407bf1 download
generation_config.json 142 B c4199cca download

README current version from Hugging Face


datasets:

  • xensive/roleplaydataset100k
    base_model:
  • DavidAU/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated
    tags:
    • nvfp4
    • fp4
    • quantized

L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated-nvfp4

Format: NVFP4 — weights & activations quantized to FP4 with dual scaling.
Base model: DavidAU/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated
How it was made: One-shot calibration with LLM Compressor (NVFP4 recipe), long-seq calibration (256 samples of 4096 length) with xensive/roleplaydataset100k.

Notes: Keep lm_head in high precision; calibrate on long, domain-relevant sequences.

Check the original model card for information about this model.

Running the model with VLLM in Docker

sudo docker run --runtime nvidia --gpus all -p 8000:8000 --ipc=host vllm/vllm-openai:nightly --model Firworks/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated-nvfp4 --dtype auto --max-model-len 8192

Running the model on the DGX Spark with VLLM in Docker

sudo docker run --gpus all --network host --ipc=host   nvcr.io/nvidia/vllm:26.02-py3   vllm serve Firworks/L3-Darkest-Planet-16B-HERETIC-Uncensored-Abliterated-nvfp4   --dtype auto   --max-model-len 8192

This was tested on a DGX Spark (GB10 Grace Blackwell Superchip, 128GB unified memory).

If there are other models you're interested in seeing quantized to NVFP4 for use on the DGX Spark, or other modern Blackwell (or newer) cards let me know. I'm trying to make more NVFP4 models available to allow more people to try them out.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-03-29Update README.md5c56f4e1.5 KB
    Loading...
  2. 2026-03-29Add NVFP4 quantized checkpoint2c5d3ef1.3 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration