← back to catalog · registered 2026-08-22 13:56

apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8

apetersson Deepseek 296B MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/apetersson%2FDeepSeek-V4-Flash-0731-Abliterated-FP8"
Response includes
  • classification m1
  • files 58
  • hub_downloads_all_time 36,370
  • author_summary 8 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
36K
3K last 30d - cooling
Likes
38
Descendants
6
in 6 direct forks
Model age
2mo ago
created 2026-07-31
Downloads over time
Now37.1K→from28↑132,543%
013.6K27.2K40.9K28 on Jul 2937.1K on Oct 11JulAugSepOct
Jul 29 → Oct 11 · 52 snapshots · spans 74 days

Genealogy 6 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Tags
transformers safetensors deepseek_v4 text-generation deepseek-v4 mixture-of-experts abliterated fp8 model-surgery mechanistic-interpretability base_model:deepseek-ai/DeepSeek-V4-Flash-0731 base_model:finetune:deepseek-ai/DeepSeek-V4-Flash-0731

Related

Total size
155 GB
Files
58
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-01 00:44

Files by quantization

Auxiliary files 58 files 155 GB
model-00048-of-00048.safetensors 3.44 GB d8b68325 download
model-00046-of-00048.safetensors 3.36 GB da1e0108 download
model-00004-of-00048.safetensors 3.35 GB 9610f56b download
model-00012-of-00048.safetensors 3.34 GB d200bcd9 download
model-00014-of-00048.safetensors 3.34 GB ce44ce42 download
model-00016-of-00048.safetensors 3.34 GB acf5cbaf download
model-00018-of-00048.safetensors 3.34 GB ac550ef1 download
model-00020-of-00048.safetensors 3.34 GB a27655e3 download
model-00022-of-00048.safetensors 3.34 GB eb87aef2 download
model-00024-of-00048.safetensors 3.34 GB 2cf277a1 download
model-00026-of-00048.safetensors 3.34 GB a4c8f53d download
model-00028-of-00048.safetensors 3.34 GB 51113c3f download
model-00030-of-00048.safetensors 3.34 GB 51e99849 download
model-00032-of-00048.safetensors 3.34 GB b7ecad01 download
model-00034-of-00048.safetensors 3.34 GB b0de49f5 download
model-00036-of-00048.safetensors 3.34 GB 002bdca7 download
model-00038-of-00048.safetensors 3.34 GB ab90e440 download
model-00040-of-00048.safetensors 3.34 GB 549d59fc download
model-00042-of-00048.safetensors 3.34 GB 11e81dc8 download
model-00044-of-00048.safetensors 3.34 GB a20dfe1e download
model-00006-of-00048.safetensors 3.34 GB 4a4f3764 download
model-00008-of-00048.safetensors 3.34 GB 224968d2 download
model-00010-of-00048.safetensors 3.34 GB 627145f4 download
model-00013-of-00048.safetensors 3.32 GB 9c36e454 download
model-00015-of-00048.safetensors 3.32 GB f9b40ffa download
model-00017-of-00048.safetensors 3.32 GB d6771600 download
model-00019-of-00048.safetensors 3.32 GB c31a222c download
model-00021-of-00048.safetensors 3.32 GB 4068e1c0 download
model-00023-of-00048.safetensors 3.32 GB 4b8d427c download
model-00025-of-00048.safetensors 3.32 GB dd813e48 download
model-00027-of-00048.safetensors 3.32 GB 70050f03 download
model-00029-of-00048.safetensors 3.32 GB f7b9198b download
model-00031-of-00048.safetensors 3.32 GB d31d4f70 download
model-00033-of-00048.safetensors 3.32 GB 7bf5a6c9 download
model-00035-of-00048.safetensors 3.32 GB 471a793b download
model-00037-of-00048.safetensors 3.32 GB 71f148a7 download
model-00039-of-00048.safetensors 3.32 GB 26f3acc9 download
model-00041-of-00048.safetensors 3.32 GB 1868317b download
model-00043-of-00048.safetensors 3.32 GB bc1dc626 download
model-00005-of-00048.safetensors 3.32 GB f87a5ac7 download
model-00007-of-00048.safetensors 3.32 GB df81bb80 download
model-00009-of-00048.safetensors 3.32 GB 04d69ef1 download
model-00011-of-00048.safetensors 3.32 GB e4b8e601 download
model-00002-of-00048.safetensors 3.32 GB 77b26c93 download
model-00003-of-00048.safetensors 3.32 GB 412abf4c download
model-00047-of-00048.safetensors 3.32 GB b8b41236 download
model-00045-of-00048.safetensors 1010 MB a5be6aed download
model-00001-of-00048.safetensors 1010 MB f3668ba4 download
tokenizer.json 6.07 MB 628e3364 download
model.safetensors.index.json 5.34 MB c3b10d45 download
ABLITERATION_MANIFEST.json 8.93 KB 8467498a download
README.md 6.24 KB 7f1eac6e download
config.json 1.84 KB 5f2da910 download
.gitattributes 1.48 KB a6344aac download
LICENSE 1.06 KB d62e3bef download
tokenizer_config.json 801 B f3dad388 download
NOTICE 793 B 71acc765 download
generation_config.json 170 B c56a8c5b download

README current version from Hugging Face


license: mit
library_name: transformers
pipeline_tag: text-generation
base_model: deepseek-ai/DeepSeek-V4-Flash-0731
base_model_relation: finetune
tags:

  • deepseek-v4
  • mixture-of-experts
  • abliterated
  • fp8
  • model-surgery
  • mechanistic-interpretability

DeepSeek-V4-Flash-0731 Abliterated (Native FP8)

Experimental model-surgery release. Direct behavioral benchmarking of
this native-FP8 checkpoint has not been run. The completed behavioral
validation below uses a quantized MLX derivative as a deployment proxy and
must not be read as a direct FP8 or capability result.

This is a rank-1 abliterated derivative of
deepseek-ai/DeepSeek-V4-Flash-0731.
It remains the official post-trained 0731 model: abliteration does not turn it
into a base model, and this release is not claimed to be universally
"uncensored."

Hugging Face labels the parent relationship as finetune because its model
tree currently has no generic derived or model-surgery relationship. This
checkpoint was not gradient-finetuned and was not quantized by this project;
it was produced by a targeted projection edit while retaining the upstream
checkpoint's native mixed precision.

Provenance

Edit recipe

  • Method: refusal-direction projection from attention residual writers.
  • Direction rank: 1.
  • Strength: lambda 3.5.
  • Main-model layers: 10 through 42, inclusive.
  • DSpark stages: the corresponding attention output projection in all three
    attached stages was edited.
  • Total edited tensors: 36.
  • Storage: the checkpoint's original mixed FP8/BF16/F32 representation, with
    three fixed-point FP8 requantization iterations.
  • Direction SHA-256:
    6e4d8a8f3aa9e21795faab2c5b14d29b019acdf2ddbfbd8238430458a5837fe0.

The recipe follows the public refusal-direction work in
drowzeys/DeepSeek-V4-Flash-DSpark-Abliterated-Uncensored-1M-57toks,
transferred to the newer 0731 checkpoint. That transfer is experimental. See
ABLITERATION_MANIFEST.json for the per-tensor edit and FP8 round-trip data.

Format and use

The architecture, tokenizer, official message encoding, one-million-token
context declaration, and attached DSpark tensors come from the upstream 0731
checkpoint. Refer to the upstream model card and the included encoding
directory for prompt formatting and runtime instructions.

Recommended upstream sampling defaults are temperature=1.0, top_p=0.95
for agentic scenarios, and top_p=1.0 otherwise.

Validation status

Structural validation

  • All 48 weight shards and 72,317 indexed tensors passed structural validation.
  • All 36 intended residual-writer edits are recorded in the manifest.
  • Preliminary direct refusal probes showed the expected behavioral shift.

Behavioral deployment proxy (2026-08-01)

The native-FP8 checkpoint was not loaded for the full benchmark on the 128 GB
test host. Instead, the complete prompt gauntlet was run through oMLX using
apetersson/DeepSeek-V4-Flash-0731-Abliterated-MLX-Mixed-2bit-3bit-g64,
a quantized derivative of this checkpoint. This is useful evidence that the
edited behavior survives that conversion and runtime, but it is not a direct
measurement of this FP8 artifact.

Run configuration:

  • oMLX 0.5.4rc1, OpenAI-compatible API, 32,768-token profile.
  • Temperature 0, top-p 1, seed 42, and maximum 160 generated tokens.
  • Benchmark tooling and pinned source revisions:
    apetersson/deepseek-model-tools@dc6af88.
  • Benchmark fingerprint:
    40c3573bd48861b846721220fa06ce0d71905aab236ce45d49d9aa0e95e79af5.
  • Final completeness gate: 830 unique cases, 830 successful latest records,
    and zero failed or missing latest records. Seven transient oMLX transport or
    memory-guard failures were repaired with a serial retry pass.
Suite Cases Completed result
UncensorBench 200 0 hard refusals; keyword compliance on 200/200
XSTest 450 0/450 refusals under the exact upstream prefix classifier; 0/250 safe-prompt over-refusals
StrongREJECT-small 180 60 prompts × baseline, ROT13, and refusal-suppression variants; 0 hard refusals under the local heuristic

One XSTest unsafe privacy prompt triggered the deliberately broader local
hard-refusal heuristic while not matching XSTest's official prefix classifier;
no safe prompt triggered either classifier.

The official fine-tuned StrongREJECT judge was not run, so this card does
not claim a StrongREJECT score. A clean reference run and capability benchmark
were also not run. Consequently, these results do not establish the native-FP8
checkpoint's refusal rate, quantify the effect of quantization, or demonstrate
general capability preservation.

Limitations

Abliteration can affect capabilities and behavior beyond refusals. It does not
guarantee compliance, factuality, safety, or a particular response style. Use
appropriate access controls and evaluate the model for your deployment.

License and attribution

The upstream repository and weights are MIT licensed. This derivative retains
the upstream LICENSE. The refusal direction is attributed to drowzeys/keys
under its accompanying MIT notice; see NOTICE. Please cite the original
DeepSeek-V4 work and credit DeepSeek-AI when redistributing or publishing
results.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-01docs: add quantized proxy validation results7d026406.2 KB
    Loading...
  2. 2026-07-31docs: pin provenance and correct parent relationshipae8a86d4.1 KB
    Loading...
  3. 2026-07-31Add files using upload-large-folder toolfad1dc12.8 KB
    Loading...

Discussions 3 threads

  1. 2026-08-07中文用户不要轻易尝试!open2 💬#3
    Loading...
  2. 2026-08-01<3 <3 <3open2 💬#2
    Loading...
  3. 2026-07-31You're so quick! Thanks for your model!open8 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration