← back to catalog · registered 2026-08-22 13:56

mumitrol/DeepSeek-V4-Flash-0731-Abliterated-NVFP4-vision

mumitrol Deepseek 277B multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/mumitrol%2FDeepSeek-V4-Flash-0731-Abliterated-NVFP4-vision"
Response includes
  • classification m1
  • files 64
  • hub_downloads_all_time 311
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
311
81 last 30d - stable
Likes
1
Model age
2mo ago
created 2026-08-06
Downloads over time
Now343→from28↑1,125%
1213325437528 on Aug 5343 on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Tags
safetensors deepseek_v4 vision-language deepseek-v4 moonvit nvfp4 fp8 abliterated image-text-to-text custom_code base_model:apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8 base_model:quantized:apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8

Related

Total size
164 GB
Files
64
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-06 19:25

Files by quantization

Auxiliary files 64 files 164 GB
model-00004-of-00048.safetensors 3.54 GB 5c1f3826 download
model-00012-of-00048.safetensors 3.53 GB 5746eb5e download
model-00014-of-00048.safetensors 3.53 GB 8a2aa1fc download
model-00016-of-00048.safetensors 3.53 GB e91093ab download
model-00018-of-00048.safetensors 3.53 GB f30b2538 download
model-00020-of-00048.safetensors 3.53 GB 98d1e424 download
model-00022-of-00048.safetensors 3.53 GB 28cfcb1d download
model-00024-of-00048.safetensors 3.53 GB ef360399 download
model-00026-of-00048.safetensors 3.53 GB 22b08641 download
model-00028-of-00048.safetensors 3.53 GB e3959f0d download
model-00030-of-00048.safetensors 3.53 GB 3ad29441 download
model-00032-of-00048.safetensors 3.53 GB 2d98626a download
model-00034-of-00048.safetensors 3.53 GB 85f30fe4 download
model-00036-of-00048.safetensors 3.53 GB c9aebc69 download
model-00038-of-00048.safetensors 3.53 GB 19883fa7 download
model-00040-of-00048.safetensors 3.53 GB a90130bc download
model-00042-of-00048.safetensors 3.53 GB 54b17db0 download
model-00044-of-00048.safetensors 3.53 GB f37e3353 download
model-00006-of-00048.safetensors 3.53 GB f77e49e0 download
model-00008-of-00048.safetensors 3.53 GB c11248a2 download
model-00010-of-00048.safetensors 3.53 GB 54ec8d6b download
model-00013-of-00048.safetensors 3.51 GB a04ea990 download
model-00015-of-00048.safetensors 3.51 GB 74f971c6 download
model-00017-of-00048.safetensors 3.51 GB 5d18f548 download
model-00019-of-00048.safetensors 3.51 GB c7e297ed download
model-00021-of-00048.safetensors 3.51 GB bfa5d8f7 download
model-00023-of-00048.safetensors 3.51 GB b01e2dba download
model-00025-of-00048.safetensors 3.51 GB 9e817714 download
model-00027-of-00048.safetensors 3.51 GB 633e625a download
model-00029-of-00048.safetensors 3.51 GB 0a93785c download
model-00031-of-00048.safetensors 3.51 GB a10e16e7 download
model-00033-of-00048.safetensors 3.51 GB 7e88ebaa download
model-00035-of-00048.safetensors 3.51 GB 11587de7 download
model-00037-of-00048.safetensors 3.51 GB 85326fcd download
model-00039-of-00048.safetensors 3.51 GB 50c21b09 download
model-00041-of-00048.safetensors 3.51 GB c99ad369 download
model-00043-of-00048.safetensors 3.51 GB 5ea0d4d1 download
model-00005-of-00048.safetensors 3.51 GB 4792d730 download
model-00007-of-00048.safetensors 3.51 GB e24d2e3f download
model-00009-of-00048.safetensors 3.51 GB 28ef22e7 download
model-00011-of-00048.safetensors 3.51 GB 140cd350 download
model-00002-of-00048.safetensors 3.51 GB 8a1d404e download
model-00003-of-00048.safetensors 3.51 GB 44d50465 download
model-00048-of-00048.safetensors 3.44 GB d8b68325 download
model-00046-of-00048.safetensors 3.36 GB da1e0108 download
model-00047-of-00048.safetensors 3.32 GB b8b41236 download
model-00045-of-00048.safetensors 1010 MB a5be6aed download
model-00001-of-00048.safetensors 1010 MB f3668ba4 download
vision_tower.safetensors 795 MB 1382c41f download
mm_projector.safetensors 76.5 MB 7024d9d5 download
model.safetensors.index.json 10.7 MB 71b3d998 download
tokenizer.json 6.07 MB 628e3364 download
modeling_deepseek_v4_vision.py 12.7 KB 7618d3cf download
ABLITERATION_MANIFEST.json 8.93 KB 8467498a download
config.json 7.95 KB 2c09eb62 download
SHA256SUMS 5.46 KB e66f0e9b download
hf_quant_config.json 4.44 KB 0c730748 download
README.md 2.55 KB 02b674cd download
.gitattributes 1.55 KB 495a7175 download
LICENSE 1.06 KB d62e3bef download
tokenizer_config.json 801 B f3dad388 download
NOTICE 793 B 71acc765 download
palette64.json 541 B 30e8db28 download
generation_config.json 170 B c56a8c5b download

README current version from Hugging Face


base_model:

  • deepseek-ai/DeepSeek-V4-Flash-0731
  • apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8
    pipeline_tag: image-text-to-text
    license: mit
    tags:
  • vision-language
  • deepseek-v4
  • moonvit
  • nvfp4
  • fp8
  • abliterated

DeepSeek-V4-Flash-0731 Abliterated + Vision (NVFP4)

Abliterated DeepSeek-V4-Flash-0731 (MOE, 284B total / 13B activated) with a
frozen MoonViT-3d vision tower (from Kimi-K2.6) and WebBrain's trained
PatchMerger projector — packaged for vLLM.

What this repo is

This is a fork of sakamakismile/DeepSeek-V4-Flash-0731-Abliterated-NVFP4
(the abliterated NVFP4 conversion) with the vision overlay from
webbrain-one/DeepSeek-V4-Flash-0731-Vision-NVFP4 spliced on:

Component Provenance Status
Text checkpoint sakamakismile ablit 0731 NVFP4 48 shards
Vision tower vision_tower.safetensors MoonViT frozen
Projector mm_projector.safetensors PatchMerger frozen
Routing palette64.json 64-id palette see below

How to serve with vLLM

The config carries auto_map so vLLM's Transformers modeling backend loads it
with remote code — no vLLM fork or custom image tag required:

docker run --rm --gpus all \
  -p 8000:8000 \
  vllm/vllm-openai:cu129-nightly \
  python -m vllm.entrypoints.openai.api_server \
    --model mumitrol/DeepSeek-V4-Flash-0731-Abliterated-NVFP4-vision \
    --model-impl transformers \
    --trust-remote-code \
    --tensor-parallel-size 4 \
    --dtype bfloat16 \
    --limit-mm-per-prompt image=1
  • Requires a Blackwell node (NVFP4); 1x B300 (288 GB) is enough for a
    single stream / moderate context, 2x for maximum context and throughput.

Status / validation (IMPORTANT)

This is a first-pass community port, not an official NVIDIA or DeepSeek
release. Validate on GPU before production use:

  • Tower attention matches MMEncoderAttention (head_dim^-0.5 scale, fused
    wqkv, no RoPE) — F.scaled_dot_product_attention with exact scale.
  • Positional-embedding interpolation mode matches MoonViT (bicubic).
  • Image-to-token seating is right (>= 512-token cap, 2x2 merge layout).
  • Hash-MoE routing for image tokens uses palette64.json (deterministic).
  • Abliteration + vision edge effects on refusing behavior are checked.

See modeling_deepseek_v4_vision.py for # -> implementation notes.

Licenses

Original DeepSeek weights MIT; Kimi-K2.6-derived tower under its own notice;
WebBrain projector per their repo; abliteration per apetersson/drowzeys notice.
Downstream users must satisfy all applicable licenses.

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-06Fix markdownlint table style20cee7f2.6 KB
    Loading...
  2. 2026-08-06Fix README table + line lengtha7dda3c2.7 KB
    Loading...
  3. 2026-08-06Add MoonViT vision overlay + custom modeling (auto_map)dae96132.7 KB
    Loading...
  4. 2026-08-06Duplicate from sakamakismile/DeepSeek-V4-Flash-0731-Abliterated-NVFP40f80ec918.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration