← back to catalog · registered 2026-08-22 13:56

coughmedicine/Huihui-Qwen3-Next-80B-A3B-Instruct-abliterated-W4A16

coughmedicine Qwen 77B MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/coughmedicine%2FHuihui-Qwen3-Next-80B-A3B-Instruct-abliterated-W4A16"
Response includes
  • classification m1
  • files 22
  • benchmarks 11 entries
  • hub_downloads_all_time 3,612
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
4K
32 last 30d - cooling
Likes
2
Model age
10mo ago
created 2025-12-10
Downloads over time
Now3.6K→from19↑18,968%
01.3K2.7K4K19 on Dec 10, 20253.6K on Oct 11Dec '25FebAprJunAugOct
Dec 10, 2025 → Oct 11 · 83 snapshots · spans 305 days

Benchmarks

Benchmark Score Source
Entertainment 1.3 UGI
Hazardous 3.5 UGI
Natural Intelligence 27.28 UGI
Political lean -17.8% UGI
Sensitive-Info 20.22 UGI
SocPol 1.7 UGI
UGI 41.81 UGI
Willingness (10) 8.5 UGI
W10-Adherence 10 UGI
W10-Direct 7 UGI
Writing 42.37 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
safetensors qwen3_next quantization llm-compressor compressed-tensors w4a16 moe region:us

Related

Total size
41.6 GB
Files
22
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-12-10 06:00

Files by quantization

Auxiliary files 22 files 41.6 GB
model-00008-of-00009.safetensors 4.66 GB 400a21ed download
model-00005-of-00009.safetensors 4.66 GB 2139b831 download
model-00006-of-00009.safetensors 4.66 GB 8f95407a download
model-00007-of-00009.safetensors 4.66 GB ce4fe161 download
model-00004-of-00009.safetensors 4.66 GB e9b29a6a download
model-00003-of-00009.safetensors 4.66 GB f3c2849c download
model-00002-of-00009.safetensors 4.66 GB d8dbc615 download
model-00001-of-00009.safetensors 4.66 GB 5ce36bf4 download
model-00009-of-00009.safetensors 4.28 GB 75ed009d download
model.safetensors.index.json 20.6 MB 35509cfa download
tokenizer.json 10.9 MB aeb13307 download
vocab.json 2.65 MB 4783fe10 download
merges.txt 1.59 MB 31349551 download
config.json 21.4 KB 1a227064 download
tokenizer_config.json 5.28 KB c9fc1221 download
chat_template.jinja 2.57 KB 70adff8a download
.gitattributes 1.60 KB aa7aacd0 download
README.md 1.17 KB a05fd91a download
added_tokens.json 707 B b54f9135 download
special_tokens_map.json 613 B ac23c0aa download
generation_config.json 213 B 3dd2c545 download
recipe.yaml 180 B bf33cf6e download

README current version from Hugging Face


base_model:

  • huihui-ai/Huihui-Qwen3-Next-80B-A3B-Instruct-abliterated
    base_model_relation: quantized
    tags:
    • quantization
    • llm-compressor
    • compressed-tensors
    • w4a16
    • moe

Made with llm-compressor:

import re
from llmcompressor.modifiers.quantization import QuantizationModifier
from llmcompressor import oneshot

MODEL_ID = "huihui-ai/Huihui-Qwen3-Next-80B-A3B-Instruct-abliterated"
OUTPUT_DIR = "./Huihui-Qwen3-Next-80B-A3B-Instruct-abliterated-ExpertsOnly-W4A16"

# Target only expert MLP linears via regex

expert_pattern = r".*mlp\.experts\.\d+\.(gate_proj|up_proj|down_proj)$"

recipe = [
    QuantizationModifier(
        scheme="W4A16",
        # Regex target: only expert MLP linear layers
        targets=[f"re:{expert_pattern}"],
        ignore=["lm_head"],
    )
]

# Run oneshot in "data-free" mode

quantized_model = oneshot(
    model=MODEL_ID,
    precision="bf16",
    trust_remote_code_model=True,
    recipe=recipe,
    dataset=None,
    num_calibration_samples=0,
    quantization_aware_calibration=False,
    max_seq_length=1,           # irrelevant here, but keeps dataset pipeline trivial
    output_dir=OUTPUT_DIR,
    save_compressed=True,
)

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-12-10Update README.md92c1af51.2 KB
    Loading...
  2. 2025-12-10Create README.md4e7a3e91.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration