← back to catalog · registered 2026-08-22 13:56

mradermacher/DeepSeek-V4-Flash-0731-Abliterated-FP8-i1-GGUF

mradermacher Deepseek GGUF MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/mradermacher%2FDeepSeek-V4-Flash-0731-Abliterated-FP8-i1-GGUF"
Response includes
  • classification m8
  • files 57
  • hub_downloads_all_time 980
  • author_summary 3324 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 3 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=mradermacher (M8 quantization producer, never originator)
  • is_gguf=1
  • base_model='apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8' (base has 'abliterated' marker, assume M1 default)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
980
126 last 30d - stable
Likes
5
Model age
2mo ago
created 2026-08-02
Downloads over time
Now1K→from619↑63%
6007498981K619 on Aug 51K on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 126 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
mit
Languages
en
Quantizations
FP8
Tags
transformers gguf deepseek-v4 mixture-of-experts abliterated fp8 model-surgery mechanistic-interpretability en base_model:apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8 base_model:quantized:apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8 license:mit

Related

Total size
449 MB
Files
57
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-09-19 12:47

Files by quantization

FP8 1 file 449 MB
DeepSeek-V4-Flash-0731-Abliterated-FP8.imatrix.gguf 449 MB e1066d91 download
Auxiliary files 56 files 2.01 TB
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q6_K.gguf.part1of5 44.0 GB 6739df09 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q6_K.gguf.part2of5 44.0 GB be143d25 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q6_K.gguf.part3of5 44.0 GB 3b670cf9 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q6_K.gguf.part4of5 44.0 GB 11b87651 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_M.gguf.part1of3 43.0 GB ea301bd6 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_M.gguf.part2of3 43.0 GB 75b992bc download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_1.gguf.part1of4 42.0 GB d8105939 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_1.gguf.part2of4 42.0 GB 10505d15 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_1.gguf.part3of4 42.0 GB 4896754d download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q6_K.gguf.part5of5 41.3 GB 4de3963a download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_M.gguf.part1of4 41.0 GB 7c192efa download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_M.gguf.part2of4 41.0 GB 53a0a72f download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_M.gguf.part3of4 41.0 GB 47ab662f download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_M.gguf.part3of3 40.0 GB f4b115df download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_1.gguf.part4of4 39.7 GB e73b55d5 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_M.gguf.part1of3 39.0 GB 5de771fb download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_M.gguf.part2of3 39.0 GB 19314e4c download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_S.gguf.part1of3 39.0 GB 1d85849c download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_S.gguf.part2of3 39.0 GB f3a953b4 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_S.gguf.part1of3 39.0 GB 33059436 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_S.gguf.part2of3 39.0 GB 11aad6e0 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_0.gguf.part1of4 38.0 GB 13b2c407 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_0.gguf.part2of4 38.0 GB 719f200f download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_0.gguf.part3of4 38.0 GB ea15c67a download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_S.gguf.part1of4 38.0 GB 91bd9782 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_S.gguf.part2of4 38.0 GB e84b7ebd download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_S.gguf.part3of4 38.0 GB fd948ad1 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_M.gguf.part1of5 38.0 GB af2cb4f0 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_M.gguf.part2of5 38.0 GB 6453dce4 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_M.gguf.part3of5 38.0 GB 401870eb download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_M.gguf.part4of5 38.0 GB 33a125fb download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_M.gguf.part3of3 37.4 GB bce9331d download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_M.gguf.part4of4 37.0 GB a4328c12 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_S.gguf.part1of5 37.0 GB 5558ad8f download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_S.gguf.part2of5 37.0 GB c0366aed download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_S.gguf.part3of5 37.0 GB 400b7781 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_S.gguf.part4of5 37.0 GB 3837d906 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_K_S.gguf.part4of4 36.4 GB 8282ed39 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ3_S.gguf.part3of3 36.1 GB 7ccd0fba download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_S.gguf.part3of3 36.1 GB a8562ff4 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ4_XS.gguf.part1of4 36.0 GB 7250f735 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ4_XS.gguf.part2of4 36.0 GB b107f314 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ4_XS.gguf.part3of4 36.0 GB 337f0136 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_M.gguf.part5of5 35.8 GB de0dfa35 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q4_0.gguf.part4of4 35.8 GB e3d557c6 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_L.gguf.part1of4 35.0 GB e8b6f89c download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_L.gguf.part2of4 35.0 GB 9acaef14 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_L.gguf.part3of4 35.0 GB 1b1b7d48 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q5_K_S.gguf.part5of5 34.2 GB 0de82260 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-IQ4_XS.gguf.part4of4 32.9 GB 179382ea download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q2_K.gguf.part1of3 32.0 GB e791b708 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q2_K.gguf.part2of3 32.0 GB cc1458e1 download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q2_K.gguf.part3of3 32.0 GB 7071493b download
DeepSeek-V4-Flash-0731-Abliterated-FP8.i1-Q3_K_L.gguf.part4of4 31.3 GB fa1d7c0c download
README.md 11.9 KB 8d38e5fb download
.gitattributes 6.75 KB 511d32bf download

README current version from Hugging Face


base_model: apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8
language:

  • en
    library_name: transformers
    license: mit
    mradermacher:
    readme_rev: 1
    quantized_by: mradermacher
    tags:
  • deepseek-v4
  • mixture-of-experts
  • abliterated
  • fp8
  • model-surgery
  • mechanistic-interpretability

About

weighted/imatrix quants of https://huggingface.co/apetersson/DeepSeek-V4-Flash-0731-Abliterated-FP8

For a convenient overview and download list, visit our model page for this model.

static quants are available at https://huggingface.co/mradermacher/DeepSeek-V4-Flash-0731-Abliterated-FP8-GGUF

Usage

If you are unsure how to use GGUF files, refer to one of TheBloke's
READMEs
for
more details, including on how to concatenate multi-part files.

Provided Quants

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

Link Type Size/GB Notes
GGUF imatrix 0.6 imatrix file (for creating your own quants)
PART 1 PART 2 PART 3 i1-Q2_K 103.2 IQ3_XXS probably better
PART 1 PART 2 PART 3 i1-IQ3_S 122.6 beats Q3_K*
PART 1 PART 2 PART 3 i1-Q3_K_S 122.6 IQ3_XS probably better
PART 1 PART 2 PART 3 i1-IQ3_M 124.0
PART 1 PART 2 PART 3 i1-Q3_K_M 135.4 IQ3_S probably better
PART 1 PART 2 PART 3 PART 4 i1-Q3_K_L 146.5 IQ3_M probably better
PART 1 PART 2 PART 3 PART 4 i1-IQ4_XS 151.4
PART 1 PART 2 PART 3 PART 4 i1-Q4_0 160.9 fast, low quality
PART 1 PART 2 PART 3 PART 4 i1-Q4_K_S 161.6 optimal size/speed/quality
PART 1 PART 2 PART 3 PART 4 i1-Q4_K_M 171.9 fast, recommended
PART 1 PART 2 PART 3 PART 4 i1-Q4_1 178.0
P1 P2 P3 P4 P5 i1-Q5_K_S 195.7
P1 P2 P3 P4 P5 i1-Q5_K_M 201.7
P1 P2 P3 P4 P5 i1-Q6_K 233.4 practically like static Q6_K

Here is a handy graph by ikawrakow comparing some lower-quality quant
types (lower is better):

image.png

And here are Artefact2's thoughts on the matter:
https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9

FAQ / Model Request

See https://huggingface.co/mradermacher/model_requests for some answers to
questions you might have and/or if you want some other model quantized.

Thanks

I thank my company, nethype GmbH, for letting
me use its servers and providing upgrades to my workstation to enable
this work in my free time. Additional thanks to @nicoboss for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.

README history 11 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-19auto-patch README.md7b608f612.1 KB
    Loading...
  2. 2026-08-03auto-patch README.mdeb5500c11.9 KB
    Loading...
  3. 2026-08-03auto-patch README.md63e430e8.3 KB
    Loading...
  4. 2026-08-03auto-patch README.md006b2517.6 KB
    Loading...
  5. 2026-08-03auto-patch README.mdc6a7c016.4 KB
    Loading...
  6. 2026-08-02auto-patch README.md2c1d88b5.6 KB
    Loading...
  7. 2026-08-02auto-patch README.md850ef314.3 KB
    Loading...
  8. 2026-08-02auto-patch README.mdae7f77a3.6 KB
    Loading...
  9. 2026-08-02auto-patch README.md2dd4d493.1 KB
    Loading...
  10. 2026-08-02auto-patch README.mdd8b41c72.6 KB
    Loading...
  11. 2026-08-02uploaded from rich18c96805495 B
    Loading...

Discussions 1 thread

  1. 2026-08-07Is MTP supported?open1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration