← back to catalog · registered 2026-08-22 13:56

cbert33/DeepSeek-V4-Flash-0731-abliterated-vision

cbert33 Deepseek 296B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/cbert33%2FDeepSeek-V4-Flash-0731-abliterated-vision"
Response includes
  • classification m1
  • files 64
  • hub_downloads_all_time 282
  • author_summary 8 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
282
Likes
1
Model age
7w ago
created 2026-08-21
Downloads over time
Now743→from30↑2,377%
027154381430 on Aug 19743 on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
transformers safetensors deepseek_v4 text-generation Safetensors deepseek-v4 multimodal vision-language abliterated uncensored vllm dgx-spark

Related

Total size
157 GB
Files
64
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-26 00:49

Files by quantization

Auxiliary files 64 files 157 GB
model-00048-of-00048.safetensors 3.44 GB cc43742b download
model-00046-of-00048.safetensors 3.36 GB 5db924ca download
model-00004-of-00048.safetensors 3.35 GB 9610f56b download
model-00012-of-00048.safetensors 3.34 GB 64ed4e5f download
model-00014-of-00048.safetensors 3.34 GB 45db2f54 download
model-00016-of-00048.safetensors 3.34 GB e0530b70 download
model-00018-of-00048.safetensors 3.34 GB e393fea9 download
model-00020-of-00048.safetensors 3.34 GB 9f556769 download
model-00022-of-00048.safetensors 3.34 GB decd67a4 download
model-00024-of-00048.safetensors 3.34 GB fc27aeb4 download
model-00026-of-00048.safetensors 3.34 GB 657b8931 download
model-00028-of-00048.safetensors 3.34 GB b2fd5cbb download
model-00030-of-00048.safetensors 3.34 GB 9ed3c317 download
model-00032-of-00048.safetensors 3.34 GB 16365384 download
model-00034-of-00048.safetensors 3.34 GB 0f949451 download
model-00036-of-00048.safetensors 3.34 GB 7e676142 download
model-00038-of-00048.safetensors 3.34 GB 137fa617 download
model-00040-of-00048.safetensors 3.34 GB 8bc93d8a download
model-00042-of-00048.safetensors 3.34 GB 4d19bf36 download
model-00044-of-00048.safetensors 3.34 GB 422d3889 download
model-00006-of-00048.safetensors 3.34 GB 4a4f3764 download
model-00008-of-00048.safetensors 3.34 GB 224968d2 download
model-00010-of-00048.safetensors 3.34 GB 627145f4 download
model-00013-of-00048.safetensors 3.32 GB 8dfe199d download
model-00015-of-00048.safetensors 3.32 GB 5810381a download
model-00017-of-00048.safetensors 3.32 GB ed111302 download
model-00019-of-00048.safetensors 3.32 GB a74ca4d3 download
model-00021-of-00048.safetensors 3.32 GB 1671cce7 download
model-00023-of-00048.safetensors 3.32 GB c61a3e17 download
model-00025-of-00048.safetensors 3.32 GB a66b6b8d download
model-00027-of-00048.safetensors 3.32 GB fb01f21a download
model-00029-of-00048.safetensors 3.32 GB 9ec2fdf9 download
model-00031-of-00048.safetensors 3.32 GB d5078c3f download
model-00033-of-00048.safetensors 3.32 GB f2cffd43 download
model-00035-of-00048.safetensors 3.32 GB 9cb6a316 download
model-00037-of-00048.safetensors 3.32 GB a59d662f download
model-00039-of-00048.safetensors 3.32 GB a29af1aa download
model-00041-of-00048.safetensors 3.32 GB fd312e7f download
model-00043-of-00048.safetensors 3.32 GB b7103842 download
model-00005-of-00048.safetensors 3.32 GB f87a5ac7 download
model-00007-of-00048.safetensors 3.32 GB df81bb80 download
model-00009-of-00048.safetensors 3.32 GB 04d69ef1 download
model-00011-of-00048.safetensors 3.32 GB e4b8e601 download
model-00002-of-00048.safetensors 3.32 GB 77b26c93 download
model-00003-of-00048.safetensors 3.32 GB 412abf4c download
model-00047-of-00048.safetensors 3.32 GB 62816173 download
model-overlay-00001-of-00001.safetensors 1.44 GB 20d25599 download
model-00045-of-00048.safetensors 1010 MB a5be6aed download
model-00001-of-00048.safetensors 1010 MB f3668ba4 download
tokenizer.json 6.07 MB 628e3364 download
model.safetensors.index.json 5.34 MB dffe4726 download
LICENSE 11.1 KB 261eeb9e download
README.md 9.78 KB adc0e4db download
MANIFEST.sha256 9.53 KB 9ba668fd download
LANGUAGE_WEIGHTS.sha256 4.84 KB f6522134 download
VALIDATION.md 3.50 KB d936e45f download
ASSEMBLY.md 3.00 KB 2a9e6fab download
config.json 2.55 KB 32460706 download
THIRD_PARTY_NOTICES.md 2.11 KB 3cc380e5 download
SOURCE_PINS.json 2.06 KB 5582b471 download
.gitattributes 1.48 KB a6344aac download
tokenizer_config.json 801 B f3dad388 download
generation_config.json 170 B c56a8c5b download
DOWNLOAD-COMPLETE.txt 144 B 63ad50dd download

README current version from Hugging Face


license: apache-2.0
base_model:

  • cebeuq/DeepSeek-V4-Flash-0731-abliterated
  • FlyCockpit/DeepSeek-V4-Flash-0731-vision
    pipeline_tag: image-text-to-text
    library_name: transformers
    tags:
  • deepseek-v4
  • multimodal
  • vision-language
  • screenshots
  • abliterated
  • uncensored
  • vllm
  • dgx-spark

DeepSeek-V4-Flash-0731 Abliterated Vision

We were happy with Cebeuq's DeepSeek-V4-Flash-0731 abliterated checkpoint, but we wanted vision too. We did not want to spend roughly another 8 GB on the Kimi-derived vision towers used by the other DeepSeek vision grafts we considered, so we built this smaller composition ourselves from the Cebeuq language model and FlyCockpit's DeepEncoderV2 tower and projector.

This model requires our companion dgx-spark-vllm-deepseek-v4-vision runtime repository to use its vision features. Stock Transformers and stock vLLM will not load the custom multimodal architecture correctly; use the pinned Anemll-based runtime and deployment instructions from that repository.

Uncensored model: the language checkpoint has undergone abliteration to reduce refusal behavior. Treat outputs as untrusted, apply application-level safeguards, and do not assume the model will decline harmful requests.

User responsibility: this model is provided without warranty. The maintainers are not responsible for what others generate, publish, deploy, or otherwise do with this abliterated model. Users must operate it responsibly, apply appropriate safeguards, comply with applicable law, and respect third-party rights.

What “merged” means here

This is a composition, not weight averaging or parameter blending:

  1. The complete language-model repository is copied from cebeuq/DeepSeek-V4-Flash-0731-abliterated at revision 21bd923c2574d9edcd7b914885024ce72fd5c076.[2]
  2. The DeepEncoderV2 tower and trained projector are copied from FlyCockpit/DeepSeek-V4-Flash-0731-vision at revision d8efc7dfaceee965164d95952e2498b60fee323c.[3]
  3. config.json is updated to declare DeepseekV4VisionForCausalLM and record the image-token and preprocessing contract.
  4. The original model.safetensors.index.json is required to remain byte-for-byte unchanged during assembly. No language tensor is rewritten by the vision composition step.
  5. At inference, each image is encoded by DeepEncoderV2, projected from 896 dimensions into the language model’s 4096-dimensional embedding space, and spliced into the prompt at image token ID 129279. A learned separator follows the view embeddings.

The underlying abliterated checkpoint itself is derived from DeepSeek’s official DeepSeek-V4-Flash-0731 release.[1][2]

Pinned components

Component Source Revision / digest
Language checkpoint cebeuq/DeepSeek-V4-Flash-0731-abliterated revision 21bd923c2574d9edcd7b914885024ce72fd5c076; weight-index SHA-256 a93ace48e04f89e17222353191d30a5ba5744cbb8176a7a4832221656ce2d545; 50-entry weight-manifest SHA-256 6ddeea70678d7f11a6ea72514cbb799b7070617e9fa3d62f990722e27ba371f4
Declared official base deepseek-ai/DeepSeek-V4-Flash-0731 the Cebeuq card does not declare the exact official-base revision used
Vision tower + projector FlyCockpit/DeepSeek-V4-Flash-0731-vision d8efc7dfaceee965164d95952e2498b60fee323c
Tower file vision/tower/deepencoder_v2_tower.safetensors SHA-256 9dcf6803d4c6b63acc4008bc2409e599a2ab6e3886e241f1727f61550c300df5
Projector file vision/adapter/merged-004800-5af0c5.pt SHA-256 6d0235333941210666bf347abb95e334943ef3f230dac65b83551186925468ec
Runtime base Locally supplied digest-pinned Anemll DGX Spark image registry digest sha256:a83948492cf13df455170fb42885f5ef4db54fefe0feff0f841ecbff464ac9d8
Runtime source Anemll GitHub repository ID 1301198905 47503f8e38dadd4dededca798150db2619594fce[4]
Vision plugin lineage FlyCockpit/DeepSeek-V4-Vision-2x-DGX-Sparks 7cb20472e0f007a0626bd22ed9f5e22a8825c7e1[5]

The machine-readable copy of these pins is in SOURCE_PINS.json. MANIFEST.sha256 covers every other file in the assembled repository; the manifest does not list itself.

Vision contract

  • Architecture: DeepEncoderV2 tower + MLP projector + learned view separator
  • Input image size: 1024 × 1024
  • Image token: <|image|> (129279)
  • Tokens per view: 256
  • Tiling: up to four local crops from a 2×2 grid, with one global view first and row-major crop order
  • Image token counts: 257, 769, or 1281, depending on aspect-aware tiling
  • Main language KV cache format in the qualified runtime: nvfp4_ds_mla

Qualified runtime

The release was exercised with:

  • two NVIDIA GB10 systems in tensor parallel (TP=2);
  • Anemll DSpark vLLM image 0.1.1 plus the included vision plugin;
  • served model name DeepSeek-V4-Flash-0731-Vision;
  • max_model_len=800000, max_num_seqs=2, and gpu_memory_utilization=0.88;
  • measured startup cache admission of 1,791,777 tokens, sufficient for two full 800K requests with 191,777 aggregate tokens of admission headroom.

The measured number is specific to this exact hardware/runtime/profile. Re-profile after changing the GPU, runtime, graph mode, batching limits, vision implementation, or cache format.

Serving

Use the required dgx-spark-vllm-deepseek-v4-vision runtime repository and point it at this full model directory. The essential launch settings are:

--served-model-name DeepSeek-V4-Flash-0731-Vision
--tensor-parallel-size 2
--kv-cache-dtype nvfp4_ds_mla
--block-size 256
--max-model-len 800000
--max-num-seqs 2
--gpu-memory-utilization 0.88

Example OpenAI-compatible request:

curl http://HEAD:8304/v1/chat/completions \
  -H 'Content-Type: application/json' \
  -d '{
    "model": "DeepSeek-V4-Flash-0731-Vision",
    "chat_template_kwargs": {"thinking": false},
    "messages": [{"role": "user", "content": [
      {"type": "text", "text": "Read the text in this image."},
      {"type": "image_url", "image_url": {"url": "data:image/png;base64,..."}}
    ]}],
    "max_tokens": 128
  }'

Validation performed

The qualified two-node deployment passed:

  • API health and model-list checks;
  • a deterministic text response;
  • two independent image/OCR prompts (ORBIT-7391 and NOVA-2846);
  • rank health checks with no OOM;
  • verification that the native DSpark/EAGLE3 path and nvfp4_ds_mla cache remained active.

This is a bounded deployment smoke test, not a comprehensive vision benchmark or safety evaluation. The redacted artifact identities, profile, observed outputs, and limitations are recorded in VALIDATION.md.

Limitations

  • The vision adapter is primarily described and tested for screenshots and UI understanding; it is not established as a production computer-use or coordinate-grounding model.[3][5]
  • Only one image per prompt was qualified.
  • Stock Transformers and stock vLLM are unsupported for this assembled layout.
  • The language component is abliterated/uncensored and may generate unsafe, misleading, or policy-violating content.[2]
  • OCR smoke tests do not establish broad visual reasoning quality.
  • The 800K × 2 result is an admission/cache measurement, not proof that every possible pair of 800K prompts will complete under every workload.

Licensing and redistribution

This assembled repository uses Apache-2.0 as its top-level license. The incorporated MIT notices are retained under LICENSES/, and third-party components retain their applicable upstream terms:

  • DeepSeek’s base model and the abliterated language checkpoint identify MIT terms.[1][2]
  • The FlyCockpit model repository is tagged Apache-2.0 and its card identifies the adapter packaging as Apache-2.0.[3]
  • The same card says the DeepEncoderV2 tower remains subject to upstream terms.[3]
  • The runtime overlay is derived from MIT-licensed FlyCockpit packaging and runs over Anemll’s runtime; vLLM-derived code in Anemll remains Apache-2.0.[4][5]

This top-level license choice does not remove or replace retained third-party notices. See THIRD_PARTY_NOTICES.md.

Reproducibility

See ASSEMBLY.md for the exact composition algorithm and verification gates. The release assembler:

  • rejects symlinked or overlapping input trees;
  • validates all 48 language shards, the overlay shard, and the language-model weight index against the pinned 50-entry LANGUAGE_WEIGHTS.sha256 manifest;
  • requires every weight-index reference to be present in that manifest;
  • validates pinned sizes and SHA-256 hashes for the tower, projector, and runtime wheel;
  • copies into a staging directory and atomically renames it;
  • rejects any change to the language-model weight index;
  • emits a complete SHA-256 manifest.

Citation

If you use this package, cite the original DeepSeek V4 release and credit both component repositories:

@misc{deepseekai2026deepseekv4,
  title  = {DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence},
  author = {DeepSeek-AI},
  year   = {2026}
}

Also cite the exact pinned source URLs below so the composition can be reconstructed.

Sources

[1] DeepSeek-V4-Flash-0731
[2] DeepSeek V4 Flash 0731 Abliterated
[3] DeepSeek V4 Flash 0731 Vision Assets
[4] Anemll DGX Spark vLLM source
[5] FlyCockpit DeepSeek V4 Vision Runtime

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-26Update README.md257f99410.5 KB
    Loading...
  2. 2026-08-26Update README.mdcac82ab10.5 KB
    Loading...
  3. 2026-08-25Update README.md70f37e09.8 KB
    Loading...
  4. 2026-08-21Public DGX Spark model releasec6e784a9.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration