← back to catalog · registered 2026-08-25 13:02

cbert33/DeepSeek-V4-Flash-0731-abliterated-vision-v2

cbert33 Deepseek 296B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/cbert33%2FDeepSeek-V4-Flash-0731-abliterated-vision-v2"
Response includes
  • classification m1
  • files 66
  • hub_downloads_all_time 837
  • author_summary 8 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
837
155 last 30d - stable
Likes
5
Model age
6w ago
created 2026-08-24
Downloads over time
Now890→from24↑3,608%
032665197724 on Aug 26890 on Oct 11AugSepOct
Aug 26 → Oct 11 · 47 snapshots · spans 46 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en zh
Tags
transformers safetensors deepseek_v4 text-generation Safetensors deepseek-v4 multimodal vision-language dspark dgx-spark anchored-tensors abliterated

Related

Total size
155 GB
Files
66
Quantizations
1
Registered
2026-08-25 13:02
Last updated on HF
2026-08-27 12:42

Files by quantization

Auxiliary files 66 files 155 GB
model-00048-of-00048.safetensors 3.44 GB cc43742b download
model-00046-of-00048.safetensors 3.36 GB 5db924ca download
model-00004-of-00048.safetensors 3.35 GB 9610f56b download
model-00012-of-00048.safetensors 3.34 GB 2030aa53 download
model-00014-of-00048.safetensors 3.34 GB 02d67665 download
model-00016-of-00048.safetensors 3.34 GB 5b7ce160 download
model-00018-of-00048.safetensors 3.34 GB d82fe222 download
model-00020-of-00048.safetensors 3.34 GB 52a5f5c4 download
model-00022-of-00048.safetensors 3.34 GB 62c0dab2 download
model-00024-of-00048.safetensors 3.34 GB bbcb99c3 download
model-00026-of-00048.safetensors 3.34 GB 60420b66 download
model-00028-of-00048.safetensors 3.34 GB 23842032 download
model-00030-of-00048.safetensors 3.34 GB 27844710 download
model-00032-of-00048.safetensors 3.34 GB 4396e5d6 download
model-00034-of-00048.safetensors 3.34 GB 1f946526 download
model-00036-of-00048.safetensors 3.34 GB a7bd1df3 download
model-00038-of-00048.safetensors 3.34 GB 137fa617 download
model-00040-of-00048.safetensors 3.34 GB 8bc93d8a download
model-00042-of-00048.safetensors 3.34 GB 4d19bf36 download
model-00044-of-00048.safetensors 3.34 GB 422d3889 download
model-00006-of-00048.safetensors 3.34 GB 4a4f3764 download
model-00008-of-00048.safetensors 3.34 GB 224968d2 download
model-00010-of-00048.safetensors 3.34 GB 627145f4 download
model-00013-of-00048.safetensors 3.32 GB 93f62c61 download
model-00015-of-00048.safetensors 3.32 GB e19c1dba download
model-00017-of-00048.safetensors 3.32 GB 7a4d3cfb download
model-00019-of-00048.safetensors 3.32 GB f07be062 download
model-00021-of-00048.safetensors 3.32 GB c25dbdd9 download
model-00023-of-00048.safetensors 3.32 GB 1429d6ef download
model-00025-of-00048.safetensors 3.32 GB 3184168c download
model-00027-of-00048.safetensors 3.32 GB bc60af4f download
model-00029-of-00048.safetensors 3.32 GB 59b2d435 download
model-00031-of-00048.safetensors 3.32 GB b6b017e5 download
model-00033-of-00048.safetensors 3.32 GB 71e97cf3 download
model-00035-of-00048.safetensors 3.32 GB b82dc25d download
model-00037-of-00048.safetensors 3.32 GB 79d9a398 download
model-00039-of-00048.safetensors 3.32 GB a29af1aa download
model-00041-of-00048.safetensors 3.32 GB fd312e7f download
model-00043-of-00048.safetensors 3.32 GB b7103842 download
model-00005-of-00048.safetensors 3.32 GB f87a5ac7 download
model-00007-of-00048.safetensors 3.32 GB df81bb80 download
model-00009-of-00048.safetensors 3.32 GB 04d69ef1 download
model-00011-of-00048.safetensors 3.32 GB e4b8e601 download
model-00002-of-00048.safetensors 3.32 GB 77b26c93 download
model-00003-of-00048.safetensors 3.32 GB 412abf4c download
model-00047-of-00048.safetensors 3.32 GB 62816173 download
model-00045-of-00048.safetensors 1010 MB a5be6aed download
model-00001-of-00048.safetensors 1010 MB f3668ba4 download
tokenizer.json 6.07 MB 628e3364 download
model.safetensors.index.json 5.34 MB c3b10d45 download
MANIFEST.sha256 8.46 KB f07f5b79 download
DOWNLOAD_MANIFEST.json 8.37 KB 54d6d8f8 download
README.md 6.67 KB 6bfb7353 download
LANGUAGE_WEIGHTS.sha256 4.73 KB c8a9b61c download
ABLIT_META.json 4.49 KB dafcf01a download
DSPARK_ANCHOR_MANIFEST.json 2.96 KB f4242b9b download
config.json 2.55 KB 32460706 download
VALIDATION.md 2.39 KB 8e9305fe download
SOURCE_PINS.json 2.30 KB 7b86e8a1 download
ASSEMBLY.md 2.20 KB c2ebe309 download
.gitattributes 1.48 KB a6344aac download
THIRD_PARTY_NOTICES.md 1.17 KB 133ead56 download
LICENSE 1.06 KB d62e3bef download
tokenizer_config.json 801 B f3dad388 download
RESPONSIBLE_USE.md 681 B 7dd30cb1 download
generation_config.json 170 B c56a8c5b download

README current version from Hugging Face


license: other
license_name: deepseek
license_link: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
base_model:

  • drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-Anchored-Tensors
    base_model_relation: finetune
    pipeline_tag: image-text-to-text
    library_name: transformers
    tags:
  • Safetensors
  • deepseek-v4
  • multimodal
  • vision-language
  • dspark
  • DGX Spark
  • anchored-tensors
  • abliterated
  • uncensored
  • fp8
  • nvfp4
  • vllm
    language:
  • en
  • zh

After creating the our original fintune (https://huggingface.co/cbert33/DeepSeek-V4-Flash-0731-abliterated-vision), Drwozeys came out with a more narrowly abliterated version of the Deepseek model. We used that as the base and grafted on a vision tower that was under 1gb in size. Note that standard vLLM, as well as Eugr and Anemll don't fully support this setup, so we have a custom one that does.

DeepSeek-V4-Flash-0731 Abliterated Vision 2.0

This repository composes the language checkpoint from drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-Anchored-Tensors with the DeepEncoderV2 tower and trained projector from FlyCockpit/DeepSeek-V4-Flash-0731-vision.

The operation is a component composition, not weight averaging or language-weight blending. The Drowzeys language-model shards remain byte-identical to their pinned source. The custom runtime loads the vision tower and projector from explicit paths and inserts projected image embeddings at the image token.

Uncensored model: safety refusals have been deliberately reduced in the language checkpoint. Treat outputs as untrusted, apply application-level safeguards, and do not assume the model will decline harmful requests. This model was created for research purposes only.

User responsibility: this model is provided without warranty. We are not responsible for content created by the model. Users are responsible for safe operation, legal compliance, and respecting third-party rights and licenses.

Serving

This model is intended to pair with the dgx-spark-vllm-deepseek-v4-vision 2.0 fork for vision serving with vLLM on two NVIDIA DGX Spark systems. The fork provides the integration required for the DeepEncoderV2 tower, projector, image-token handling, and distributed vision path used by this model.

At the time of release, upstream vLLM, Eugr, and Anemll do not provide the complete vision-serving path required by this model. Use the companion fork and point it at the complete model directory to enable image-conditioned requests.

Example request shape:

{
  "model": "DeepSeek-V4-Flash-0731-Vision",
  "chat_template_kwargs": {"thinking": false},
  "messages": [{
    "role": "user",
    "content": [
      {"type": "text", "text": "Read the text in this image."},
      {"type": "image_url", "image_url": {"url": "data:image/png;base64,..."}}
    ]
  }],
  "max_tokens": 128
}

Language checkpoint

The 2.0 language component is Drowzeys' DSpark-compatible anchored-tensor checkpoint:

  • Abliteration edits are limited to layers 10–35.
  • The edited tensors are attn.wo_b weights with λ 3.5 and one refusal direction.
  • Layers 36–42 are restored hash-identical to the official stock checkpoint.
  • DSpark target layers 40–42 remain stock.
  • MTP/draft tensors were not edited (edit_mtp: false).
    These properties are recorded in ABLIT_META.json and DSPARK_ANCHOR_MANIFEST.json and pinned in SOURCE_PINS.json.

Composition

  1. Copy the complete pinned Drowzeys language checkpoint at revision a1e69379517383be9cd78c67defb04e77ad6aa68.
  2. Preserve all 48 language shards and model.safetensors.index.json byte-for-byte.
  3. Copy the DeepEncoderV2 tower and projector from FlyCockpit at revision d8efc7dfaceee965164d95952e2498b60fee323c.
  4. Set the architecture to DeepseekV4VisionForCausalLM and add the explicit vision contract to config.json.
  5. Add the pinned custom runtime artifact and model metadata.
  6. Verify the complete output against MANIFEST.sha256.

Pinned identities

Component Revision / digest
Drowzeys language checkpoint a1e69379517383be9cd78c67defb04e77ad6aa68
Language weight-index SHA-256 98efab455cf08dfbbbaaba6f570e1bf10bf927d2b4c3c453a59c2f6f0e3be92b
Language weight-manifest SHA-256 467bbbf3167b80dedb30100427540dc3c01e021bc0b2cf7e0933b52cc6e02569
FlyCockpit vision revision d8efc7dfaceee965164d95952e2498b60fee323c
Vision tower SHA-256 9dcf6803d4c6b63acc4008bc2409e599a2ab6e3886e241f1727f61550c300df5
Projector SHA-256 6d0235333941210666bf347abb95e334943ef3f230dac65b83551186925468ec
Vision plugin lineage 7cb20472e0f007a0626bd22ed9f5e22a8825c7e1

Machine-readable pins are in SOURCE_PINS.json. MANIFEST.sha256 covers every packaged file except itself.

Vision contract

  • Architecture: DeepEncoderV2 tower, MLP projector, and learned view separator
  • Input image size: 1024 × 1024
  • Image token: <|image|> (129279)
  • Tokens per view: 256
  • Tiling: one global view plus up to four local crops from a 2×2 grid
  • Image-token counts: 257, 769, or 1281, depending on aspect-aware tiling
  • Qualified cache format: nvfp4_ds_mla

Validation and limits

The assembled Drowzeys + FlyCockpit artifact has been qualified through the custom runtime with text and image-conditioned requests. This is bounded deployment qualification, not a comprehensive visual, long-context, or safety benchmark. See VALIDATION.md.

  • One image per prompt was qualified.
  • OCR smoke tests do not establish broad visual-reasoning quality.
  • The language checkpoint is abliterated and may generate unsafe, false, or policy-violating content.
  • Maximum-context admission does not establish semantic quality at maximum length.

Reproducibility and notices

See ASSEMBLY.md for the composition and verification algorithm, SOURCE_PINS.json for immutable inputs, and THIRD_PARTY_NOTICES.md for attribution and licensing information.

Sources

  1. Official DeepSeek-V4-Flash-0731
  2. Drowzeys anchored-tensor language checkpoint
  3. FlyCockpit vision assets
  4. Vision plugin lineage

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-27Update README.md71cdaf07 KB
    Loading...
  2. 2026-08-25Update README.md5d34f976.7 KB
    Loading...
  3. 2026-08-25Update README.md1b5fe476.7 KB
    Loading...
  4. 2026-08-24Update README.md78cea816.2 KB
    Loading...
  5. 2026-08-24initial uploadce7566d6.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration