← back to catalog · registered 2026-08-22 13:56

DavidAU/Qwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic

DavidAU Qwen 40B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/DavidAU%2FQwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic"
Response includes
  • classification m3
  • files 32
  • hub_downloads_all_time 1,279
  • author_summary 213 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
1K
395 last 30d - stable
Likes
69
Descendants
9
in 6 direct forks
Model age
2mo ago
created 2026-07-21
Downloads over time
Now1.4K→from20↑6,780%
05041K1.5K20 on Aug 51.4K on Oct 11AugSepOct
Aug 5 → Oct 11 · 51 snapshots · spans 67 days

Genealogy 6 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh
Tags
transformers safetensors qwen3_5 image-text-to-text MTP fine tune heretic uncensored abliterated multi-stage tuned. all use cases thinking

Related

Total size
74.4 GB
Files
32
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-09-16 23:52

Files by quantization

Auxiliary files 32 files 74.5 GB
model-00003-of-00017.safetensors 4.62 GB 4459ae1f download
model-00005-of-00017.safetensors 4.62 GB 85db46ab download
model-00007-of-00017.safetensors 4.62 GB 836c2237 download
model-00009-of-00017.safetensors 4.62 GB 38e56ad0 download
model-00011-of-00017.safetensors 4.62 GB 5f1c9ba2 download
model-00013-of-00017.safetensors 4.62 GB bd64e102 download
model-00015-of-00017.safetensors 4.62 GB 9cc55bc5 download
model-00014-of-00017.safetensors 4.60 GB c00c77fc download
model-00016-of-00017.safetensors 4.60 GB d8ea96d0 download
model-00008-of-00017.safetensors 4.59 GB 4256669a download
model-00010-of-00017.safetensors 4.59 GB a93d34f7 download
model-00004-of-00017.safetensors 4.59 GB 7d4e1456 download
model-00012-of-00017.safetensors 4.58 GB 333ecaa9 download
model-00006-of-00017.safetensors 4.57 GB f5270f8e download
model-00002-of-00017.safetensors 4.51 GB 10a0dc43 download
model-00017-of-00017.safetensors 2.96 GB 32be2eaf download
model-00001-of-00017.safetensors 2.37 GB 32b777cc download
model-mtp-restored.safetensors 100 MB 98348118 download
tokenizer.json 19.1 MB 06b95093 download
ff711-gone-60.gif 17.2 MB 75d0119d download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 152 KB 8842e857 download
tokenizer_config.json 18.9 KB 975f5521 download
chat_template-instruct.jinja 11.6 KB 3d56789c download
chat_template.jinja 11.5 KB 82faea87 download
README.md 9.12 KB afc4dae3 download
config.json 5.42 KB 7860d13e download
.gitattributes 1.59 KB fd80d73d download
processor_config.json 1.16 KB 33818c7f download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B d04042de download

README current version from Hugging Face


language:

  • en
  • zh
    license: apache-2.0
    tags:
  • MTP
  • fine tune
  • heretic
  • uncensored
  • abliterated
  • multi-stage tuned.
  • all use cases
  • thinking
  • reasoning
  • qwen3.6
  • coder
  • creative
  • writing
  • fiction
  • roleplaying
  • bfloat16
  • all use cases
    pipeline_tag: image-text-to-text
    base_model:
  • DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP
    library_name: transformers

For NEO Imatrix GGUFS - reg and MTP, as well as full model details/card please go here:

https://huggingface.co/DavidAU/Qwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF


Model #2 of this release series is here (10 stages of "construction", this is a FURTHER modified version of "Qwen3.6-40B-Grand-Intelligence"):

https://huggingface.co/DavidAU/Qwen3.6-40B-Fable-Fusion-6-Core-Deckard-Eleanor-Heretic-Uncensored

For NEO Imatrix GGUFS - reg and MTP, as well as full model details/card please go here:

https://huggingface.co/DavidAU/Qwen3.6-40B-Fable-Fusion-6-Core-Deckard-Eleanor-Heretic-Uncensored-NM-DAU-NEO-MAX-MTP-GGUF

"40B Eleanor" takes "40B-Grand-Intelligence" to the next level in terms of reduction of thinking tokens / quality and detail of output generation.


Qwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic

40B : "There is a BIGGER storm coming... and it will take no prisoners."

1290 Tensors, 96 layers : 50% larger than 27B Qwen 3.6 it is based on.

We have selected 2 models for release, these 2 compliment each other in terms of benchs
and operations as well as thinking block/style and general overall performance.

  1. Qwen3.6-40B-Fable-Fusion-6-Core-Deckard-Eleanor-Heretic-Uncensored ( Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune W "The Deckard 40B" )
  2. Qwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic ( Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune )

These models contain MULTIPLE 711 and 717 Qwen 3.6 27B Fable Fusion models, with the "Deckard" version
containing all the unqiue "DNA" of the Deckard 40B model (previously released).

The "benches" far and above exceed Qwen 3.6 27B benches, and previously released 40B models by us.

Although the final benches are not as strong as 27B FF711, they are in the "ballpark" and human testing/results
show these models are stronger in some areas, which do not show up in the "general bench marking" results.

NOTE: #1 will have a separate repo :

https://huggingface.co/DavidAU/Qwen3.6-40B-Fable-Fusion-6-Core-Deckard-Eleanor-Heretic-Uncensored


! Approaching release status ; there are 3 models being considered.

One is listed here (Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune W "The Deckard 40B"), with 13k generation from testing (testing/benching still ongoing):

https://huggingface.co/DavidAU/Qwen3.6-40B-007-The-Deckard-FF

ABOUT:

40B version(s) based on the wildly powerful "711" (1500+ likes, 1.6 million+ downloads) which operates in/near "OpenAI, Claude and Gemini" closed source intelligence levels:

https://huggingface.co/DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF

(Benchmarks of the "711" at the above repo that surpass the Qwen 3.6 27B in all respects)

Additional models used / being considered:

  • The Deckard/Claude Opus 40B, 685+ likes, over 1 million downloads.
  • 717 Unreleased "711" Fable fusion lab version.
  • Other FF711 type / 40B Trained models from the lab.

TESTING:

  • Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune W "The Deckard 40B" ("Fable Fusion 711" Pipeline stage2 ) IN TESTING/BENCHING (two versions).
  • Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune ("Fable Fusion 711" Pipeline) IN TESTING/BENCHING.
  • Finishing Alpha stage in 48 hrs then moving to BETA(s) (secondary stage post expand tuning).
  • Qwen3.6-40B-Grand-Intelligence-Six-711-717-rawb (717 is an unreleased ARC-C 717 model.) now in Alpha Training.
  • Qwen3.6-40B-Grand-Intelligence-Six-711-717-raw + rawb ; in testing.
  • Grand Intelligence, Mix 5 Raw in testing.
  • Prototype 1, light tune for testing of TWO 700+ ARC-C (OpenAi,Claude, and Gemini level intelligence) Qwen 3.6B super tunes joined "at the hip" to make a 40B monster.
  • Prototype 4, Model "Fable-Fusion-711" joined "at the hip" with Deckard 27B (Qwen 3.6) to make a 40B monster in pre-tune testing.
  • Testing / benching and post expansion tuning in progress.

IMPORTANT:

  • THIS IS A WORK in PROGRESS, and will highlight some parts of this process (40B version of "711") as we proceed.
  • NAME of this repo will CHANGE as the project proceeds.
  • RUNNING Benchmarks [subject to change] below.

PROJECT NOTES:

  • If you want to join the waitlist, you will be notified by email when the FINAL, fully tested and optimized version releases.
  • During expansion is normal to lose some benchmarks levels, which are then restored during the post expansion tuning.
  • It will (likely) take a number of tuning rounds/steps/stages to bring the 40B up to "700" club status. This can take several days due to training and eval time. We are only using local hardware.

RUNNING NOTES [reverse order]:

  • BETA of Six-711-717-rawb and BETA Six-711-717-rawb [level 2] W The Deckard in testing/benching.
  • Six-711-717-rawb -> Alpha Training, (complete), entering eval/benching and maybe multi-stage pipeline pending results.
  • Six-711-717-raw / rawb in testing...
  • Alpha 4 ("Qwen3.6-40B-Grand-Intelligence-Four-raw") in training/testing.
  • Alpha 2 and 3, 3b in testing // additional "non trained" expanded also in testing.
  • Selecting model(s) for Beta staging in progress.
  • Prelim benchmarks for "Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha" (tuned expansion) posted, moving on to next stage(s). Other Alphas are pending too.
  • Benchmark below for "Qwen3.6-40B-Grand-Intelligence-One" (the root model for "Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha") BEFORE post expansion (27B to 40B) training.

BENCHMARKS by Nightmedia

           arc/c arc/e boolq hswag obkqa piqa  wino

[#1] "Qwen3.6-40B-Fable-Fusion-6-Core-Deckard-Eleanor-Heretic-Uncensored"
Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune
W "The Deckard Claude-Opus 40B" ("Fable Fusion 711" Pipeline stage 2 )
mxfp8      0.687,0.857,0.908,0.825,0.500,0.818,0.771

REMARKS:
Off the scale detail. Stable. Intelligent, never seen before
generation quality/detail. Depth of thought too.
VASTLY reduced thinking tokens/thinking block ; auto-variable.
(1/10 to 1/2 the number of thinking tokens VS "norm" Qwen)


[#2] "Qwen3.6-40B-Grand-Intelligence-Fable-Fusion-Uncensored-Heretic"
(Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune)
[FF711 Pipeline]
mxfp8      0.698,0.862,0.904

REMARKS: Strong generation, Strong Intelligence, High detail. Stable.


Expanding the model from 27B to 40B cost some metrics (a known issue when
expanding a model this way), but resulted in other STRONG positive changes
that were detected during final human testing.

---

Qwen3.6-40B-Grand-Intelligence-Six-711-717-rawb-Alpha-Tune
[complex tune, mid, Heretic 2.0]
mxfp8      0.692,0.859,0.903,...
REMARKS: Strong, close to perfect at this stage.

Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha4
[complex tune, mid]
mxfp8      0.675,0.856,0.906,...
REMARKS: 711/Deckard mix. Stable. But metrics drop too far at this stage.

Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha3b
[mid-light repair, deeper tune only]
mxfp8      0.695,0.864,0.902,0.819,0.494,0.814,0.772
REMARKS: Slightly lower, however 100% stable. SOTA IQ/power at 40B.

Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha3
[mid-light repair tune only]
mxfp8      0.701,0.862,0.903,...
REMARKS: Excellent, but unstable with some prompts.

Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha2
[light repair tune only]
mxfp8      0.690,0.864,0.908,...

Qwen3.6-40B-Grand-Intelligence-One-IQ-tune-alpha
[light repair tune only]
mxfp8      0.689,0.859,0.903,...

------------------------------------------------------------
RAW EXPANDED MODEL(s) - pre stage to be tuned/adjusted.
------------------------------------------------------------

Qwen3.6-40B-Grand-Intelligence-Six-711-717-rawb
mxfp8      0.677,0.858,0.893

Qwen3.6-40B-Grand-Intelligence-Six-711-717-raw
[NOT TRAINED YET, expansion only]
mxfp8      0.673,0.856,0.894

Qwen3.6-40B-Grand-Intelligence-Four-raw
[NOT TRAINED YET, expansion only]
mxfp8      0.679,0.853,0.905

Qwen3.6-40B-Grand-Intelligence-One 
[NOT TRAINED YET, expansion only]
mxfp8      0.675,0.860,0.900,0.795,0.478,0.801,0.752

------------------------------------------------------------
ORG MODELS FROM QWEN, no tuning, non heretic.
------------------------------------------------------------

Qwen3.6-27B-Instruct: [base, non heretic]
mxfp8      0.647,0.803,0.910,0.773,0.450,0.806,0.742

Qwen3.6-35B-A3B-Instruct [base, non heretic]
mxfp8      0.581,0.757,0.892,0.751,0.428,0.803,0.688

Qwen3.5-27B-Instruct: [base, non heretic]
mxfp8      0.557,0.711,0.868,0.533,0.452,0.706,0.695

NOTES:

  • Models are tested in "Instruct" mode because this generally works better with the testing harness.
  • Testing via "thinking" mode also shows the metrics (and changes) but not the true extent.
  • In actual fact when the model IS in thinking mode, it will exceed INSTRUCT benchmark scores in most cases.

README history 20 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-16Update README.mda9175ce10.7 KB
    Loading...
  2. 2026-08-13Update README.md8442b919.1 KB
    Loading...
  3. 2026-08-10Update README.md28567429.1 KB
    Loading...
  4. 2026-08-10Update README.mdbcc6b459.2 KB
    Loading...
  5. 2026-08-09Update README.md6afa3c99.3 KB
    Loading...
  6. 2026-08-09Update README.md9883fbf9.2 KB
    Loading...
  7. 2026-08-09Update README.mdcb5431c8.5 KB
    Loading...
  8. 2026-08-09Update README.md5c7972a8.5 KB
    Loading...
  9. 2026-08-09Update README.mdd09d11a8.4 KB
    Loading...
  10. 2026-08-09Update README.md469738c8.2 KB
    Loading...
  11. 2026-08-09Update README.mddf9acf38.2 KB
    Loading...
  12. 2026-08-09Update README.mdb6f157c8.2 KB
    Loading...
  13. 2026-08-08Update README.md99623cc8.1 KB
    Loading...
  14. 2026-08-08Update README.mda6e765d8 KB
    Loading...
  15. 2026-08-08Update README.md9a286817.9 KB
    Loading...
  16. 2026-08-08Update README.md0319fdd7.4 KB
    Loading...
  17. 2026-08-08Update README.md6e34e576.7 KB
    Loading...
  18. 2026-08-07Update README.md469cbbd6.6 KB
    Loading...
  19. 2026-08-07Update README.md75c269c6.4 KB
    Loading...
  20. 2026-08-07Update README.mda20c41a6.4 KB
    Loading...

Discussions 10 threads

  1. 2026-09-23New 'ToMoE' approach, converting dense model to MoE model -- what do you think …open1 💬#10
    Loading...
  2. 2026-08-13This thing is bizarreopen4 💬#9
    Loading...
  3. 2026-08-12Feedback compared to the 27b model from a young adult fantasy novel perspectiveopen2 💬#8
    Loading...
  4. 2026-08-12Muse Glimmer as a base possible?open2 💬#7
    Loading...
  5. 2026-08-07Process transferrabilityopen2 💬#6
    Loading...
  6. 2026-08-07Six Seven Elevenopen1 💬#5
    Loading...
  7. 2026-08-06I can offer compute if that helps?open8 💬#4
    Loading...
  8. 2026-07-30NVFP4 Safetensor Quant?open2 💬#3
    Loading...
  9. 2026-07-25when u opening up the repo?open2 💬#2
    Loading...
  10. 2026-07-23Thank you for your great job.open3 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration