← back to catalog · registered 2026-08-23 19:02

LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V10-GGUF

LuffyTheFox Qwen 35B GGUF MoE multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/LuffyTheFox%2FQwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V10-GGUF"
Response includes
  • classification m-uncensored
  • files 24
  • author_summary 20 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
1.1M
↑ 2% in 90 days
Likes
564
Model age
3mo ago
created 2026-07-13

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now1.2M→from1.2M↑2%
1.2M1.2M1.2M1.2M1.2M on Aug 261.2M on Aug 28Aug
Aug 26 → Aug 28 · 3 snapshots · spans 2 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh multilingual
Quantizations
Q8_K
Tags
hermes gguf uncensored qwen3.6 moe vision multimodal genesis agentic image-text-to-text conversational en

Related

Total size
326 GB
Files
24
Quantizations
3
Registered
2026-08-23 19:02
Last updated on HF
2026-08-28 11:52

Files by quantization

Q8_K 3 files 122 GB
Hermes3.6-35B-A3B-Uncensored-Genesis-V10-Q8_K_P.gguf 40.6 GB 172014df download
Hermes3.6-35B-A3B-Uncensored-Genesis-V7-Q8_K_P.gguf 40.6 GB b37a4225 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V9-Q8_K_P.gguf 40.6 GB e46ce9f2 download
F16 1 file 858 MB
mmproj-Hermes3.6-35B-A3B-Uncensored-Genesis-F16.gguf 858 MB c8e70234 download
Auxiliary files 20 files 204 GB
Hermes3.6-35B-A3B-Uncensored-Genesis-V7-MTP-APEX.gguf 24.8 GB 63a3cd73 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V9-MTP-APEX.gguf 24.8 GB 6189d19f download
Hermes3.6-35B-A3B-Uncensored-Genesis-V10-APEX.gguf 24.0 GB 7e5a1a5f download
Hermes3.6-35B-A3B-Uncensored-Genesis-V7-APEX.gguf 24.0 GB 995c2a82 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V9-APEX.gguf 24.0 GB d4d2e418 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V7-MTP-APEX-Compact.gguf 17.1 GB 11edd0da download
Hermes3.6-35B-A3B-Uncensored-Genesis-V9-MTP-APEX-Compact.gguf 17.1 GB 03e097b5 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V10-APEX-Compact.gguf 16.4 GB 78ce8a18 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V7-APEX-Compact.gguf 16.2 GB 41826ae6 download
Hermes3.6-35B-A3B-Uncensored-Genesis-V9-APEX-Compact.gguf 16.2 GB f051d414 download
pelikan.PNG 85.4 KB a3af73fa download
full_output.txt 52.4 KB 07f60abe download
QWEN_MTP.py 21.0 KB b608fe1f download
chat_template.jinja 15.9 KB 995c9f4a download
README.md 8.41 KB e4e5234b download
.gitattributes 6.73 KB 2804fc09 download
System_Prompt.txt 6.15 KB 30df63fb download
System_Prompt_Creative.txt 5.97 KB 7804f369 download
pelikan.svg 3.60 KB 05470f55 download
System_Prompt_Agent.txt 1.12 KB 72207237 download

README current version from Hugging Face


license: apache-2.0
tags:

  • uncensored
  • qwen3.6
  • moe
  • gguf
  • vision
  • multimodal
  • genesis
  • hermes
  • agentic
    language:
  • en
  • zh
  • multilingual
    datasets:
  • NousResearch/hermes-function-calling-v1
    pipeline_tag: image-text-to-text
    base_model:
  • HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

🌟 Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive -> Genesis Hermes V10

⚡ https://web.tribute.tg/d/KIH ⚡ If you like this Genesis LLM release you can donate to me via @Tribute bot in Telegram messenger and support future Genesis LLM development.

⚡ Generate an SVG of a pelican riding a bicycle. ⚡ Result: pelican.svg

⚡ Why Genesis project exists? During training, ALL models don't just learn knowledge - they also accumulate random noise in their tensors. This noise builds up and creates something I call the Noise Gate - a fundamental barrier that stops LLM models from learning further and makes them unstable, verbose, and prone to hallucinations. My approach reduces this noise. It repairs the signal without touching the learned knowledge and gradient. The result is a model that consistent in performance, context clarity and following instructions, because it's no longer fighting its own internal chaos.

What is Genesis? Genesis is post training data regeneration and calibrarion algorythm for neural networks (LLM) in GGUF format that I made with AI help during almost half a year of development. It's optimized, architecture independent, works with any model in GGUF format and based on mathematical statistics. I don't train or finetune models, I repair purity of signal in them instead on Google Collab Free on Tesla T4 GPU via Python based on how models learns information. On first stage I scan ssm_conv1d tensors in model, they handle long context memory. I repair balance between heads in them. On second stage I scan model and detect noise in tensors via custom SVD. During scanning I exclude token_embd.weight, output.weight, ffn_gate_inp_shexp.weight, 1D tensors, bias and norms. Then I reduce training noise in tensors via custom SVD with preserved training data, 99% of siginal and learned gradient. On third stage, I scan blocks in model via chunks via 3 parameters and pick best one that fits to weight distribution in tensor. Best picked chunk replaces zero chunks in broken tensor without touching learned structure in model

Any questions?

Contact: [email protected], [email protected]

My Telegram: @LuffyTheFox

Model is based on HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive base.

And DJLougen/hermes-qwen3.5-35b-a3b-GGUF finetune for Hermes agent.

I transferred data from finetune on Hermes dataset (around 2k blocks from two FFN expert tensors) to HauhauCS uncensored base.

Join the Discord for updates, roadmaps, projects, or just to chat.

Base model. HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive- 0/465 refusals.

Thanks to HauhauCS

Tensor repair by me. Method: Genesis

Links:


LLM models often have:

  • Saturated weights: the model's activations are stuck, gradients vanish, outputs degrade.
  • Scale mismatches: one layer's weights are 10× larger than its peers for no good reason.
  • Mean drift: weight distributions shifted positive or negative, breaking symmetry assumptions.
  • Zero blocks: zero blocks corrupt the signal, turning training into noise amplification.
  • Training Noise: training noise increase randomness and ruins model output quality.

My approach fixes all of that without retraining - pure numerical surgery on the raw bytes of the file.

Quantization script available here: https://pastebin.com/hXhcMJn9

Feel free to do your own quants if you want.

Recommended Settings for best perfomance on APEX quant

Chat template: chat_template.jinja thanks to froggeric and qweefchief

Set K Cache Quantization Type and V Cache Quantization Type to F16.

Set Number of layers for which to force MoE weights onto CPU to 40.

Set GPU offload to maximum. Set number of active experts to 8.

For best model stability and first experience I recommend starting from this string in your System Prompt with enabled thinking and nothing else:

You are Qwen. You are a helpful and capable general-purpose AI language model developed by Alibaba Group's Tongyi Lab.

If you want to bring more creativity to model use this System Prompt with agent identity: link

Or this System Prompt with assistant identity: System_Prompt_Creative.txt

Thinking mode (coding):

  • Hermes agent: temperature=0.6, top_p=0.95, top_k=20, min_p=0.05, seed=42, presence_penalty=disabled, repeat_penalty=1.08
  • Coding/precise tasks: temperature=0.6, top_p=0.95, top_k=20, min_p=0, seed=42, presence_penalty=disabled, repeat_penalty=disabled
  • General: temperature=1.0, top_p=0.95, top_k=20, min_p=0.05, seed=42, presence_penalty=disabled, repeat_penalty=disabled

Non Thinking mode (creative):

  • General: temperature=0.7, top_p=0.85, top_k=20, min_p=0.015, seed=42, presence_penalty=disabled, repeat_penalty=disabled

For agentic tasks you can use this System Prompt:

You are Qwen. You are a helpful and capable general-purpose AI language model developed by Alibaba Group's Tongyi Lab. You are a helpful assistant that answers in JSON. Here's the json schema you must adhere to:\n<schema>\n{schema}\n</schema>.

And this fix: link to discussion

And commands from this dataset: hermes-function-calling-v1

Usage

Ready to use. Recommended quant: V8 APEX

Recommended LM Studio runtime: link to discussion

Testing

HermesBench benchmark: link to discussion

Static 2D testing

System Prompt: You are Qwen, a large language model created by Tongyi Lab team from Alibaba Group. You are a helpful assistant.

Settings: temperature=0.6, top_p=0.95, top_k=20, min_p=0, seed=42, presence_penalty=disabled, repeat_penalty=disabled

Prompt 1: Hello. What is your name?

Prompt 2: Generate an SVG of a pelican riding a bicycle.

Result: pelikan.svg

Important:

  • Keep at least 128K context to preserve thinking capabilities
  • Use --jinja flag with llama.cpp for proper chat template handling
  • Vision support requires the mmproj file alongside the main GGUF

Specs

  • 35B total parameters, ~3B active per forward pass (MoE)
  • 256 experts, 8 routed + 1 shared per token
  • Hybrid architecture: Gated DeltaNet linear attention + full softmax attention (3:1 ratio)
  • 40 layers, pattern: 10 × (3 × DeltaNet-MoE + 1 × Attention-MoE)
  • 262K native context (extendable to 1M with YaRN)
  • Natively multimodal (text, image, video)
  • 248K vocabulary, 201 languages
  • Base model. HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

Compatibility

Works with llama.cpp, LM Studio, koboldcpp, and other GGUF-compatible runtimes.

README history 20 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-26Update README.mdb5e77388.2 KB
    Loading...
  2. 2026-09-26Update README.mdef3bd558.2 KB
    Loading...
  3. 2026-09-26Update README.md45acd838.2 KB
    Loading...
  4. 2026-09-26Update README.md347eff38.2 KB
    Loading...
  5. 2026-09-26Update README.md5a5ce0e8.2 KB
    Loading...
  6. 2026-09-26Update README.md40cd6f38.2 KB
    Loading...
  7. 2026-09-26Update README.mda626bfe8.3 KB
    Loading...
  8. 2026-09-26Update README.mdd1ea78c8.3 KB
    Loading...
  9. 2026-09-15Update README.mda9418508.5 KB
    Loading...
  10. 2026-09-15Update README.mdacd892f8.5 KB
    Loading...
  11. 2026-09-15Update README.mda247b468.5 KB
    Loading...
  12. 2026-09-15Update README.mdc7bd03d8.5 KB
    Loading...
  13. 2026-09-14Update README.mdaa1d1078.5 KB
    Loading...
  14. 2026-09-14Update README.md93e39798.5 KB
    Loading...
  15. 2026-09-14Update README.md77d63d18.5 KB
    Loading...
  16. 2026-09-14Update README.md209d2648.5 KB
    Loading...
  17. 2026-09-13Update README.md4d1437c8.5 KB
    Loading...
  18. 2026-09-13Update README.md7934eb78.5 KB
    Loading...
  19. 2026-09-13Update README.md13c3f128.5 KB
    Loading...
  20. 2026-09-13Update README.mdb7cc1158.5 KB
    Loading...

Discussions 69 threads

  1. 2026-10-09I got fired from my work. I am starting fundraising for processing big LLMs via…open2 💬#75
    Loading...
  2. 2026-10-08I updated quants for this model with latest revision of Genesis.open1 💬#74
    Loading...
  3. 2026-10-05My opinion about DGX Sparks. Is it worth bying? My opinion - NO.open1 💬#73
    Loading...
  4. 2026-10-01I am done with Qwen.open19 💬#72
    Loading...
  5. 2026-09-25No visionclosed2 💬#71
    Loading...
  6. 2026-09-23Best model for 12gb VRAMopen3 💬#70
    Loading...
  7. 2026-09-19I am gonna take a long break from AI developmentopen4 💬#69
    Loading...
  8. 2026-09-15Checking GSQ-RCO quantsopen2 💬#68
    Loading...
  9. 2026-09-15LuffyTheFox/Qwen3.8-27B-Uncensored-Genesis-V1-MTP-GGUF was truly great :)open7 💬#67
    Loading...
  10. 2026-09-14Thank you!!open2 💬#66
    Loading...
  11. 2026-09-14Damn me. I made mistake yesterday. With Genesis and Genesis Hermes models Noise…open7 💬#65
    Loading...
  12. 2026-09-13Some people asked me to bring back APEX quant for this model with MTP support.open3 💬#64
    Loading...
  13. 2026-09-12Final release of this model ready for testing.open19 💬#63
    Loading...
  14. 2026-09-10Tested DeepSeek-V4.1-Flash today to push the limits of Genesis algorytmopen3 💬#62
    Loading...
  15. 2026-09-08I recieved a big email about that my work is really important for Open Source L…open2 💬#61
    Loading...
  16. 2026-09-08Differences between releasesopen2 💬#60
    Loading...
  17. 2026-09-07My final words about this legendary model.open8 💬#59
    Loading...
  18. 2026-09-06According to my tests, tensors die on Q6_K - APEX quantization.open14 💬#57
    Loading...
  19. 2026-09-03For Genesis Hermes V13 all quants now available for testing.open10 💬#56
    Loading...
  20. 2026-09-02Some people asked me, how exactly Noise Gate affects usage limit and what actua…open1 💬#55
    Loading...
  21. 2026-09-02Coming soon V13 update for Genesis project. Some people asked me, is there any …open8 💬#54
    Loading...
  22. 2026-09-01A few words according to V12 release of Genesis Hermes and inference settings f…open12 💬#53
    Loading...
  23. 2026-09-01Coming today Genesis-Hermes V12open4 💬#52
    Loading...
  24. 2026-08-31Damn me. In V11 update for this model I selected wrong version of Genesis algor…open6 💬#51
    Loading...
  25. 2026-08-29Random noiseclosed2 💬#50
    Loading...
  26. 2026-08-28New Marchenko–Pastur V11 update for Genesis Hermes coming soon. What is new?open4 💬#49
    Loading...
  27. 2026-08-25能否在Ornith-1.5-35B-A3B的基础上进行V10的发布open7 💬#48
    Loading...
  28. 2026-08-25Some people asked me is it possible to get speedup on NVIDIA Blackwell architec…open2 💬#47
    Loading...
  29. 2026-08-24PRpuher1 💬#46
    Loading...
  30. 2026-08-23Coming today: V10 update for Genesis-Hermes uncensored.open5 💬#45
    Loading...
  31. 2026-08-23I wrote a email to Qwen team. Looks like my model will reach 1 Million download…open3 💬#44
    Loading...
  32. 2026-08-22A person asked me to fix Ornith 1.5 35B-A3B via Genesis. My decision.open5 💬#43
    Loading...
  33. 2026-08-19Tunned the RIG to Run since V6, V7 is this good (read bellow)open2 💬#42
    Loading...
  34. 2026-08-18Coming today: V9 update for Genesis-Hermes uncensored.open2 💬#41
    Loading...
  35. 2026-08-16I tested V8 update for this model. My results.open1 💬#40
    Loading...
  36. 2026-08-15Why Qwen3.8-27B overthinks? Here the reason.open10 💬#38
    Loading...
  37. 2026-08-14I checked blk 0, 1, 2 in Kimi K3 model. My findings.open1 💬#37
    Loading...
  38. 2026-08-13Hermes3.6-35B-A3B Genesis V7 - practical Hermes Agent testopen2 💬#36
    Loading...
  39. 2026-08-12I just checked Qwen3.8-2.4T for singular matrices and exploded condition numbersopen2 💬#35
    Loading...
  40. 2026-08-12Incredible Benchmark Results on Laptop Hardware (RTX 5060 8GB / Ryzen 9) - 40+ …open1 💬#34
    Loading...
  41. 2026-08-12Some people asked me to fix vanilla censored Qwen3.6-35B-A3B BF16 weights from …open1 💬#33
    Loading...
  42. 2026-08-11Those who want see V7 agentic benckmark,see thereopen3 💬#32
    Loading...
  43. 2026-08-09A few words about doing Genesis for other big models? Kimi K3, Qwen3.8-Max, Dee…open9 💬#31
    Loading...
  44. 2026-08-08I did a quick hallucination test on V7closed2 💬#30
    Loading...
  45. 2026-08-06JoyFox-Qwen3.6-35B-A3B-RP An interesting project for RPG scenarios and beyond. …closed2 💬#29
    Loading...
  46. 2026-08-06A few words about training quality that Qwen team doingopen3 💬#28
    Loading...
  47. 2026-08-05New V7 update for model now available. This is final release for Genesis projec…open23 💬#27
    Loading...
  48. 2026-08-04About the deleted V7 filesclosed4 💬#26
    Loading...
  49. 2026-08-03Stuck in "git log --oneline"closed3 💬#24
    Loading...
  50. 2026-08-03Thanks for 500 followers on Hugging Face ^_^open2 💬#23
    Loading...
  51. 2026-08-01is there any reason the vanilla Qwen3.6-35B-A3B-Uncensored-Genesis-GGUF is remo…closed2 💬#22
    Loading...
  52. 2026-07-30pull mistake with ollamaclosed3 💬#20
    Loading...
  53. 2026-07-30MTP version maybe?closed3 💬#19
    Loading...
  54. 2026-07-30Some people asked me about heretic versions of Genesis. My decision.open17 💬#17
    Loading...
  55. 2026-07-28suggestionclosed4 💬#16
    Loading...
  56. 2026-07-28New V6 update for model now available.closed12 💬#15
    Loading...
  57. 2026-07-28terminate unexpectedlyclosed2 💬#14
    Loading...
  58. 2026-07-26HermesBench Resultsopen12 💬#13
    Loading...
  59. 2026-07-26A few words about LM Studio llama-cpp runtime.closed3 💬#12
    Loading...
  60. 2026-07-24A Mind In The Darkclosed1 💬#11
    Loading...
  61. 2026-07-24I tested these models for hallucination with one test questionclosed1 💬#10
    Loading...
  62. 2026-07-24Qwen3.6-27B-Uncensored-Genesis now cookingclosed2 💬#9
    Loading...
  63. 2026-07-22Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V5 now cookingclosed16 💬#7
    Loading...
  64. 2026-07-20Heads-up - V4 repo vanished mid-download (+ MTP extraction fix and 12 GB perf d…closed16 💬#6
    Loading...
  65. 2026-07-16V3 update incoming for this modelclosed4 💬#5
    Loading...
  66. 2026-07-16Important note about model settings for roleplayclosed3 💬#4
    Loading...
  67. 2026-07-15Worse than v1closed2 💬#3
    Loading...
  68. 2026-07-15Genesis SVDclosed7 💬#2
    Loading...
  69. 2026-07-14"Harmony"closed4 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration