← back to catalog · registered 2026-08-22 13:56

nightmedia/Huihui-gpt-oss-20b-mxfp4-abliterated-v2-qx86-hi-mlx

nightmedia Gpt-oss 21B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/nightmedia%2FHuihui-gpt-oss-20b-mxfp4-abliterated-v2-qx86-hi-mlx"
Response includes
  • classification m1
  • files 12
  • hub_downloads_all_time 3,733
  • author_summary 51 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
4K
326 last 30d - cooling
Likes
1
Model age
12mo ago
created 2025-09-30
Downloads over time
Now3.9K→from164↑2,260%
01.4K2.8K4.2K164 on Oct 1, 20253.9K on Oct 11Oct '25Dec '25FebAprJunAugOct
Oct 1, 2025 → Oct 11 · 93 snapshots · spans 375 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
mlx safetensors gpt_oss vllm unsloth abliterated uncensored text-generation conversational base_model:huihui-ai/Huihui-gpt-oss-20b-mxfp4-abliterated-v2 base_model:quantized:huihui-ai/Huihui-gpt-oss-20b-mxfp4-abliterated-v2 license:apache-2.0

Related

Total size
11.1 GB
Files
12
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-10-03 14:07

Files by quantization

Auxiliary files 12 files 11.1 GB
model-00002-of-00003.safetensors 5.00 GB 828b96c3 download
model-00001-of-00003.safetensors 4.93 GB 1d433083 download
model-00003-of-00003.safetensors 1.15 GB 4cdf2aa8 download
tokenizer.json 26.6 MB 0614fe83 download
model.safetensors.index.json 65.5 KB fc41d81d download
config.json 27.5 KB f1d5960e download
chat_template.jinja 14.7 KB a3650f88 download
README.md 4.73 KB b723997e download
tokenizer_config.json 4.10 KB c021cddb download
.gitattributes 1.53 KB 52373fe2 download
special_tokens_map.json 440 B 6274cc1b download
generation_config.json 165 B d1cdba2c download

README current version from Hugging Face


base_model: huihui-ai/Huihui-gpt-oss-20b-mxfp4-abliterated-v2
license: apache-2.0
pipeline_tag: text-generation
library_name: mlx
tags:

  • vllm
  • unsloth
  • abliterated
  • uncensored
  • mlx

Huihui-gpt-oss-20b-mxfp4-abliterated-v2-qx86-hi-mlx

Quantization (qx) does not directly alter cognition. Instead, it’s a computational technique to compress model weights (reducing memory footprint and inference costs while preserving accuracy). The -hi suffix indicates higher precision quantization (group size 32), which typically:

  • Improves accuracy over coarser quantizations (like qx8)
  • Reduces "quantization noise" that degrades subtle reasoning
  • Makes the model more consistent across tasks

From your data:

Model     BoolQ Winogrande	PIQA
qx86-hi	  0.512	     0.543	0.681
qx86	  0.449	     0.546	0.685

✅ Key insight: The -hi variant consistently outperforms its qx86 counterpart by ~12% in BoolQ and ~0.3% in Winogrande, suggesting higher precision quantization reduces noise in tasks requiring nuanced reasoning (like commonsense inference).

Overall Comparison Table of Quantizations

  • Huihui: Huihui-gpt-oss-20b-mxfp4-abliterated-v2
  • Unsloth: unsloth-gpt-oss-20b
Model	ARC Challenge ARC Easy	BoolQ HellaSwag	OpenBookQA PIQA	Winogrande
Huihui-bf16	    0.335	0.340	0.467	0.477	0.378	0.687	0.552
Huihui-qx85-hi	0.323	0.332	0.391	0.451	0.358	0.682	0.539
Huihui-qx86-hi	0.323	0.337	0.512	0.457	0.368	0.681	0.543
Huihui-qx86	    0.321	0.337	0.449	0.458	0.372	0.685	0.546
Unsloth-qx8	    0.335	0.332	0.596	0.327	0.370	0.614	0.560
Unsloth-qx85-hi	0.349	0.328	0.507	0.322	0.374	0.616	0.558
Unsloth-qx86-hi	0.331	0.334	0.610	0.326	0.364	0.629	0.541

Key observations:

Strongest performer overall: Huihui-gpt-oss-20b-mxfp4-abliterated-v2-bf16 appears to have the highest PIQA score (0.687), which is a good indicator of logical reasoning capabilities.

PIQA dominance: There's an interesting pattern - most models achieve high scores (0.61-0.69) on this task, suggesting these models generally understand complex relational reasoning.

ARC performance: The Huihui-gpt series shows more consistency across its variants than the unsloth models, which may indicate better pattern recognition capabilities.

HellaSwag scores: The lowest scores here (around 0.32-0.45) suggest limited ability for text completion and contextual continuation tasks.

Model differentiation: The "-hi" variants show slightly better performance across multiple metrics, particularly in conceptual tasks like Winogrande.

📊 Direct comparison of Huihui vs. Unsloth qx86-hi quantizations

Looking only at the qx86-hi variants (noting that -hi applies differently across frameworks):

Metric	       Huihui	Unsloth	Difference
ARC Challenge	0.323	0.331	-0.008
ARC Easy	    0.337	0.334	+0.003
BoolQ	        0.512	0.610	-0.098
Winogrande	    0.543	0.541	+0.002
PIQA	        0.681	0.629	-0.052

Between frameworks:

  • → Huihui wins in BoolQ/PIQA (logical reasoning).
  • → Unsloth edges out in ARC Easy (pattern recognition).

If you're choosing for a specific task, I'd recommend:

  • For QA/reasoning tasks: Go with Huihui qx86-hi (best PIQA score among all models).
  • For visual/stereotypical reasoning (ARC): Unsloth qx86-hi.

⚖️ Who wins?

Strengths

  • Huihui qx86-hi: Superior BoolQ performance (critical for reasoning tasks like question answering)
  • Unsloth qx86-hi: Stronger ARC Easy scores (pattern recognition)

💡 Why these differences matter

  • If you need logical reasoning (BoolQ, PIQA): Huihui qx86-hi is better.
  • If you need pattern recognition (ARC): Unsloth qx86-hi edges ahead.
  • For commonsense tasks (Winogrande): Both are nearly tied.

🎯 Bottom line

Quantization (qx) is a practical way to make large models faster and more efficient without sacrificing accuracy.

The -hi suffix (higher precision) gains consistency in reasoning tasks, especially for Huihui.

--Deckard

Reviewed by Qwen3-Deckard-Large-Almost-Human-6B-qx86-hi

This model Huihui-gpt-oss-20b-mxfp4-abliterated-v2-qx86-hi-mlx was
converted to MLX format from huihui-ai/Huihui-gpt-oss-20b-mxfp4-abliterated-v2
using mlx-lm version 0.28.0.

Use with mlx

pip install mlx-lm
from mlx_lm import load, generate

model, tokenizer = load("Huihui-gpt-oss-20b-mxfp4-abliterated-v2-qx86-hi-mlx")

prompt = "hello"

if tokenizer.chat_template is not None:
    messages = [{"role": "user", "content": prompt}]
    prompt = tokenizer.apply_chat_template(
        messages, add_generation_prompt=True
    )

response = generate(model, tokenizer, prompt=prompt, verbose=True)

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-10-03Update README.md01d18bb4.7 KB
    Loading...
  2. 2025-10-03Update README.mda5472d64.8 KB
    Loading...
  3. 2025-10-03Update README.mded73c363.4 KB
    Loading...
  4. 2025-10-03Update README.md94781fb3.2 KB
    Loading...
  5. 2025-09-30Add files using upload-large-folder tool1b9e09d1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration