← back to catalog · registered 2026-08-22 13:56

symrex/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp

symrex Qwen 35B GGUF MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/symrex%2FQwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp"
Response includes
  • classification m-uncensored
  • files 17
  • hub_downloads_all_time 1,127
  • author_summary 36 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1K
156 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-08-02
Downloads over time
Now1.2K→from617↑93%
5888091K1.2K617 on Aug 51.2K on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
mlx safetensors qwen3_5_moe oq quantized base_model:LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V13-GGUF base_model:quantized:LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V13-GGUF 4-bit region:us

Related

Total size
20.1 GB
Files
17
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-02 18:49

Files by quantization

Auxiliary files 17 files 20.2 GB
model-00003-of-00005.safetensors 4.78 GB 9a89f454 download
model-00004-of-00005.safetensors 4.78 GB 92d56966 download
model-00002-of-00005.safetensors 4.78 GB 9abda615 download
model-00001-of-00005.safetensors 4.66 GB e50d47f5 download
model-00005-of-00005.safetensors 1.13 GB 08aac313 download
tokenizer.json 12.2 MB 5f9e4d49 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 200 KB e5ed6dbe download
config.json 86.3 KB defa3d2c download
oq_imatrix_report.json 31.7 KB 2c090b3c download
tokenizer_config.json 16.3 KB 28d96ff3 download
chat_template.jinja 7.58 KB a8755d82 download
README.md 3.65 KB 8291638b download
.gitattributes 1.53 KB 52373fe2 download
preprocessor_config.json 390 B 2ea84a43 download
generation_config.json 202 B 023756cf download

README current version from Hugging Face


library_name: mlx
tags:

  • mlx
  • oq
  • quantized
    base_model:
  • LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF

Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp

This model was quantized using oQ (oMLX v0.5.4rc2) mixed-precision quantization.

Quantization details

  • Model type: qwen3_5_moe
  • Bits: 4
  • Group size: 64
  • Format: MLX safetensors

Performance Benchmark

oMLX - LLM inference, optimized for your Mac
https://github.com/jundot/omlx
Benchmark Model: Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp
Engine: Auto
================================================================================

Single Request Results
--------------------------------------------------------------------------------
Test                                TTFT(ms)    TPOT(ms)        pp TPS        tg TPS      E2E(s)    Throughput    Peak Mem
pp1024/tg128                           573.0        9.14  1787.0 tok/s   110.3 tok/s       1.741   661.5 tok/s    20.02 GB
pp4096/tg128                          2274.4        9.39  1800.9 tok/s   107.3 tok/s       3.476  1215.0 tok/s    20.90 GB
pp8192/tg128                          4753.3        9.69  1723.4 tok/s   104.0 tok/s       5.992  1388.5 tok/s    21.35 GB
pp16384/tg128                        10458.4       10.35  1566.6 tok/s    97.4 tok/s      11.781  1401.5 tok/s    22.24 GB
pp32768/tg128                        24436.3       11.57  1341.0 tok/s    87.1 tok/s      25.914  1269.4 tok/s    24.04 GB
pp65536/tg128                        63275.8       13.89  1035.7 tok/s    72.5 tok/s      65.050  1009.4 tok/s    27.62 GB
pp131072/tg128                      192425.7       18.01   681.2 tok/s    56.0 tok/s     194.726   673.8 tok/s    34.91 GB
pp200000/tg128                      404997.5       22.52   493.8 tok/s    44.7 tok/s     407.874   490.7 tok/s    42.57 GB

Continuous Batching
pp1024 / tg128
--------------------------------------------------------------------------------
Batch           tg TPS   Speedup        pp TPS    pp TPS/req    TTFT(ms)      E2E(s)
1x         110.3 tok/s     1.00x  1787.0 tok/s  1787.0 tok/s       573.0       1.741
2x         146.0 tok/s     1.32x   690.4 tok/s   345.2 tok/s      2966.4       4.199
4x         244.4 tok/s     2.22x  1545.6 tok/s   386.4 tok/s      2543.2       4.434
8x         327.7 tok/s     2.97x  1524.7 tok/s   190.6 tok/s      5015.4       8.144

Intelligence Benchmark Comparison

Intelligence Benchmark Comparison

               Mode    Sampled         Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp
--------------------------------------------------------------------------------------------------------
MMLU           Sample  1000/14042                                                                  87.8%
TRUTHFULQA     Full    817                                                                         80.9%
HUMANEVAL      Full    164                                                                         92.1%
LIVECODEBENCH  Sample  300/1055                                                                    48.7%

--- Detail ---

Model: Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp
Benchmark         Accuracy   Correct   Total   Time(s)   Think
--------------------------------------------------------------
MMLU                 87.8%       878    1000    3798.1     Yes
TRUTHFULQA           80.9%       661     817    2633.4     Yes
HUMANEVAL            92.1%       151     164    1284.1     Yes
LIVECODEBENCH        48.7%       146     300   11636.8     Yes

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-02adding Intelligence Benchmark Comparison2bd80453.7 KB
    Loading...
  2. 2026-08-02fix codeb1ff7262.5 KB
    Loading...
  3. 2026-08-02adding Performance Benchmark2bcfce32.5 KB
    Loading...
  4. 2026-08-02Upload Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-dequantized-oQ4e-mtp via ...6689e97361 B
    Loading...

Discussions 1 thread

  1. 2026-08-03Hey bro, can you upload a oQ6e quantized model?closed3 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration