← back to catalog · registered 2026-08-22 13:56

moeshawky/DeepSeek-V4-Flash-Q4-mxfp4-0731-abliterated

moeshawky Deepseek 296B GGUF MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/moeshawky%2FDeepSeek-V4-Flash-Q4-mxfp4-0731-abliterated"
Response includes
  • classification m8
  • files 57
  • hub_downloads_all_time 306
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
306
75 last 30d - stable
Likes
1
Model age
2mo ago
created 2026-08-09
Downloads over time
Now333→from105↑217%
94181268356105 on Aug 5333 on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Tags
transformers safetensors deepseek_v4 text-generation deepseek deepseek-v4 moe mixed-expert fp4 fp8 quantization abliterated

Related

Total size
155 GB
Files
57
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-09 15:45

Files by quantization

Auxiliary files 57 files 155 GB
model-00048-of-00048.safetensors 3.44 GB 7a7c1b17 download
model-00046-of-00048.safetensors 3.36 GB 018d5de5 download
model-00004-of-00048.safetensors 3.35 GB 1fc49d5e download
model-00012-of-00048.safetensors 3.34 GB c75c12b6 download
model-00014-of-00048.safetensors 3.34 GB 6747058b download
model-00016-of-00048.safetensors 3.34 GB 556df99b download
model-00018-of-00048.safetensors 3.34 GB 20cf91d9 download
model-00020-of-00048.safetensors 3.34 GB b3340a6d download
model-00022-of-00048.safetensors 3.34 GB 5933803e download
model-00024-of-00048.safetensors 3.34 GB 8504a09b download
model-00026-of-00048.safetensors 3.34 GB cef434e7 download
model-00028-of-00048.safetensors 3.34 GB c8f63815 download
model-00030-of-00048.safetensors 3.34 GB ee07b51a download
model-00032-of-00048.safetensors 3.34 GB 04ec952d download
model-00034-of-00048.safetensors 3.34 GB ae8ba15e download
model-00036-of-00048.safetensors 3.34 GB 83fbe662 download
model-00038-of-00048.safetensors 3.34 GB afebf43a download
model-00040-of-00048.safetensors 3.34 GB 9a407641 download
model-00042-of-00048.safetensors 3.34 GB aac8decd download
model-00044-of-00048.safetensors 3.34 GB 661745fe download
model-00006-of-00048.safetensors 3.34 GB e39a9282 download
model-00008-of-00048.safetensors 3.34 GB 17503f33 download
model-00010-of-00048.safetensors 3.34 GB 5eb3f64f download
model-00013-of-00048.safetensors 3.32 GB 148cae75 download
model-00015-of-00048.safetensors 3.32 GB 7acdcce4 download
model-00017-of-00048.safetensors 3.32 GB 1ff417ec download
model-00019-of-00048.safetensors 3.32 GB f5fb6dcd download
model-00021-of-00048.safetensors 3.32 GB 6638e50f download
model-00023-of-00048.safetensors 3.32 GB b468a372 download
model-00025-of-00048.safetensors 3.32 GB a4128714 download
model-00027-of-00048.safetensors 3.32 GB baa9e6f8 download
model-00029-of-00048.safetensors 3.32 GB ad45159e download
model-00031-of-00048.safetensors 3.32 GB 94c0e8cc download
model-00033-of-00048.safetensors 3.32 GB 999feaed download
model-00035-of-00048.safetensors 3.32 GB 0a018b42 download
model-00037-of-00048.safetensors 3.32 GB c3b8f3a0 download
model-00039-of-00048.safetensors 3.32 GB 05b44e76 download
model-00041-of-00048.safetensors 3.32 GB 081d0981 download
model-00043-of-00048.safetensors 3.32 GB fcac7892 download
model-00005-of-00048.safetensors 3.32 GB e7564539 download
model-00007-of-00048.safetensors 3.32 GB 5340c507 download
model-00009-of-00048.safetensors 3.32 GB 2cd489a2 download
model-00011-of-00048.safetensors 3.32 GB 47f1045e download
model-00002-of-00048.safetensors 3.32 GB 395b26e9 download
model-00003-of-00048.safetensors 3.32 GB f2e8b14c download
model-00047-of-00048.safetensors 3.32 GB 61322f20 download
model-00045-of-00048.safetensors 1010 MB ed3f86e9 download
model-00001-of-00048.safetensors 1010 MB 7a599111 download
tokenizer.json 6.07 MB 628e3364 download
model.safetensors.index.json 5.21 MB aa52c7b9 download
config.json 1.84 KB 5f2da910 download
README.md 1.73 KB 155ee318 download
.gitattributes 1.48 KB a6344aac download
LICENSE 1.06 KB d62e3bef download
conversion_manifest.json 1.05 KB 36be5a3d download
tokenizer_config.json 801 B f3dad388 download
generation_config.json 170 B c56a8c5b download

README current version from Hugging Face


license: mit
library_name: transformers
pipeline_tag: text-generation
tags:

  • deepseek
  • deepseek-v4
  • moe
  • mixed-expert
  • fp4
  • fp8
  • quantization
  • safetensors
  • abliterated
    base_model: huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
    inference:
    parameters:
    do_sample: true

DeepSeek-V4-Flash Q4-MXFP4 — huihui-ai GGUF converted to safetensors + DSpark head

This is a conversion of the DeepSeek-V4-Flash Q4-MXFP4 GGUF from
hf://huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF/DeepSeek-V4-Flash-Q4-mxfp4-0731.gguf
to native safetensors format, with the DSpark 3-stage speculative drafter head
included on top.

  • Source: huihui-ai/Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
    → DeepSeek-V4-Flash-Q4-mxfp4-0731.gguf (Q4-MXFP4 quant, abliterated),
    converted to safetensors shards.
  • Main model: 43 layers in model-00001..00045-of-00048.safetensors
    (routed experts as native packed FP4, other quantized matrices as 128x128
    E4M3/E8M0 block FP8).
  • DSpark drafter: model-00046..00048-of-00048.safetensors
    (mtp.0 -> 46, mtp.1 -> 47, mtp.2 -> 48), mirroring the official repo's
    placement. Routed experts retain source MXFP4 values losslessly; other
    matrices use block FP8. The drafter weights are an additional inclusion,
    sourced from a separate DeepSeek-V4-Flash-DSpark draft GGUF (see
    conversion_manifest.json); they are not part of the huihui-ai GGUF.
  • config.json is the official DSpark config verbatim, laid out to load
    exactly like deepseek-ai/DeepSeek-V4-Flash-DSpark.
  • inference/config.json carries the official loader's n_mtp_layers: 3.

This is a format conversion of already-quantized weights, not a recovery of the
original FP8 checkpoint. See conversion_manifest.json for provenance.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-09Upload README.md with huggingface_hub5fa96161.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration