← back to catalog · registered 2026-08-22 13:56

symrex/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp

symrex Qwen 35B GGUF MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/symrex%2FQwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp"
Response includes
  • classification m-uncensored
  • files 20
  • hub_downloads_all_time 2,210
  • author_summary 36 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
334 last 30d - stable
Likes
3
Model age
2mo ago
created 2026-08-06
Downloads over time
Now2.3K→from427↑443%
3321.1K1.8K2.5K427 on Aug 52.3K on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
mlx safetensors qwen3_5_moe oq quantized base_model:LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V13-GGUF base_model:quantized:LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V13-GGUF 8-bit region:us

Related

Total size
36.0 GB
Files
20
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-07 08:05

Files by quantization

Auxiliary files 20 files 36.0 GB
model-00001-of-00008.safetensors 4.88 GB 61d9378e download
model-00004-of-00008.safetensors 4.78 GB ce9a2c9b download
model-00006-of-00008.safetensors 4.78 GB 8a62f658 download
model-00007-of-00008.safetensors 4.78 GB 5ab86c60 download
model-00005-of-00008.safetensors 4.78 GB 1b35d4dd download
model-00003-of-00008.safetensors 4.78 GB fa70d40f download
model-00002-of-00008.safetensors 4.78 GB 0c89d7bf download
model-00008-of-00008.safetensors 2.39 GB 714e85da download
tokenizer.json 12.2 MB 5f9e4d49 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 200 KB d6b0896b download
config.json 73.5 KB 66106b76 download
oq_imatrix_report.json 31.7 KB d3c9004e download
tokenizer_config.json 16.3 KB 28d96ff3 download
chat_template.jinja 7.58 KB a8755d82 download
README.md 3.83 KB 5265f25d download
.gitattributes 1.53 KB 52373fe2 download
preprocessor_config.json 390 B 2ea84a43 download
generation_config.json 202 B 023756cf download

README current version from Hugging Face


library_name: mlx
tags:

  • mlx
  • oq
  • quantized
    base_model:
  • LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-GGUF

Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp

This model was quantized using oQ (oMLX v0.5.7) mixed-precision quantization.

Quantization details

  • Model type: qwen3_5_moe
  • Bits: 8
  • Group size: 64
  • Format: MLX safetensors

Performance Benchmark

oMLX - LLM inference, optimized for your Mac
https://github.com/jundot/omlx
Benchmark Model: Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp
Engine: Auto
Context: Code (Python)
================================================================================

Single Request Results
--------------------------------------------------------------------------------
Test                                TTFT(ms)    TPOT(ms)        pp TPS        tg TPS      E2E(s)    Throughput    Peak Mem
pp1024/tg128                           593.2       11.45  1726.2 tok/s    88.0 tok/s       2.055   560.7 tok/s    35.47 GB
pp4096/tg128                          2304.0       11.72  1777.8 tok/s    86.0 tok/s       3.776  1118.0 tok/s    36.35 GB
pp8192/tg128                          4822.6       14.61  1698.7 tok/s    71.2 tok/s       5.207  1578.3 tok/s    36.80 GB
pp16384/tg128                        10594.2       12.62  1546.5 tok/s    79.9 tok/s      12.205  1352.9 tok/s    37.69 GB
pp32768/tg128                        24743.9       13.66  1324.3 tok/s    73.8 tok/s      26.488  1241.9 tok/s    39.49 GB
pp65536/tg128                        63873.6       15.89  1026.0 tok/s    63.4 tok/s      65.901   996.4 tok/s    43.07 GB
pp131072/tg128                      193797.5       19.84   676.3 tok/s    50.8 tok/s     196.330   668.3 tok/s    50.36 GB
pp200000/tg128                      408283.1       24.26   489.9 tok/s    41.5 tok/s     411.381   486.5 tok/s    58.03 GB

Continuous Batching
pp1024 / tg128
--------------------------------------------------------------------------------
Batch           tg TPS   Speedup        pp TPS    pp TPS/req    TTFT(ms)      E2E(s)
1x          88.0 tok/s     1.00x  1726.2 tok/s  1726.2 tok/s       593.2       2.055
2x         147.4 tok/s     1.68x   654.8 tok/s   327.4 tok/s      3127.6       4.865
4x         221.4 tok/s     2.52x  1492.9 tok/s   373.2 tok/s      2631.8       5.056
8x         303.3 tok/s     3.45x  1477.5 tok/s   184.7 tok/s      5172.4       8.920

Intelligence Benchmark Comparison

Intelligence Benchmark Comparison

               Mode    Sampled         Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp
--------------------------------------------------------------------------------------------------------
MMLU           Sample  1000/14042                                                                  88.3%
MMLU_PRO       Sample  300/12032                                                                   81.3%
TRUTHFULQA     Full    817                                                                         81.8%
HUMANEVAL      Full    164                                                                         95.7%
LIVECODEBENCH  Sample  300/1055                                                                    39.0%

--- Detail ---

Model: Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp
Benchmark         Accuracy   Correct   Total   Time(s)   Think
--------------------------------------------------------------
MMLU                 88.3%       883    1000    4470.9     Yes
MMLU_PRO             81.3%       244     300    2818.9     Yes
TRUTHFULQA           81.8%       668     817      3162     Yes
HUMANEVAL            95.7%       157     164    1495.8     Yes
LIVECODEBENCH        39.0%       117     300     14379     Yes

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-07adding Intelligence Benchmark Comparisonc8c9f8e3.8 KB
    Loading...
  2. 2026-08-06adding Performance Benchmarkebd85502.5 KB
    Loading...
  3. 2026-08-06Upload Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-dequantized-oQ8e-mtp via ...f81c048358 B
    Loading...

Discussions 1 thread

  1. 2026-08-27Daiyly driverclosed2 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration