← back to catalog · registered 2026-08-22 13:56

huihui-ai/Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF

huihui-ai Qwen 9B GGUF 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/huihui-ai%2FHuihui-Ornith-1.0-9B-abliterated-MTP-GGUF"
Response includes
  • classification m8
  • files 7
  • hub_downloads_all_time 54,122
  • author_summary 185 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of layer-wise ablation inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=huihui-ai + is_gguf=1
  • M3 (huihui abliteration) wrapped in M8 (GGUF quantization)
Refusal direction extracted via
Extraction technique

huihui-ai layer-band extraction

Confidence
HIGH
Why we say so
producer=huihui-ai (documented layer-band methodology in model cards)
Downloads · lifetime
54K
6K last 30d - stable
Likes
29
Model age
3mo ago
created 2026-07-01
Downloads over time
Now54.9K→from4.5K↑1,133%
1.9K21.3K40.6K60K4.5K on Jul 554.9K on Oct 11JulAugSepOct
Jul 5 → Oct 11 · 54 snapshots · spans 98 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Quantizations
BF16 IQ3 Q4_K Q8_0
Tags
transformers gguf llama.cpp speculative-decoding mtp multi-token-prediction qwen3.5 abliterated uncensored GGUF huihui text-generation

Related

Total size
40.5 GB
Files
7
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2026-07-05 08:22

Files by quantization

BF16 1 file 17.1 GB
ornith-9b-mtp-kl-BF16.gguf 17.1 GB 6e800973 download
Q8_0 2 files 11.4 GB
ornith-9b-mtp-kl-Q8_0.gguf 9.11 GB d8f694f5 download
mtp-ornith-9b-mtp-kl-Q8_0.gguf 2.26 GB e7d45965 download
Q4_K 1 file 6.41 GB
ornith-9b-mtp-kl-Q4_K_M.gguf 6.41 GB 79a11db2 download
IQ3 1 file 5.58 GB
ornith-9b-mtp-kl-IQ3_M.gguf 5.58 GB b76d7725 download
Auxiliary files 2 files 7.38 KB
README.md 5.58 KB 865ce1be download
.gitattributes 1.80 KB 54a88e4c download

README current version from Hugging Face


library_name: transformers
license: mit
pipeline_tag: text-generation
base_model:

  • deepreinforce-ai/Ornith-1.0-9B
    tags:
  • gguf
  • llama.cpp
  • speculative-decoding
  • mtp
  • multi-token-prediction
  • qwen3.5
  • abliterated
  • uncensored
  • GGUF
  • huihui

extra_gated_prompt: >-
Usage Warnings

“**Risk of Sensitive or Controversial Outputs**“: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.

“**Not Suitable for All Audiences**:“ Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.

“**Legal and Ethical Responsibilities**“: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.

“**Research and Experimental Use**“: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.

“**Monitoring and Review Recommendations**“: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.

“**No Default Safety Guarantees**“: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.

huihui-ai/Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF

This is an uncensored version of deepreinforce-ai/Ornith-1.0-9B created with abliteration (see remove-refusals-with-transformers to know more about it).
This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.

Note We will directly perform ablation on the GGUF files from protoLabsAI/Ornith-1.0-9B-MTP-GGUF. Some weights quantized in Q3_K, Q4_K, Q5_K and Q6_K will be converted to Q8_0. This time, the weights of the first 5 layers, so the final weights will increase, but the change will not be significant.

We performed a simple quantization comparison between Q8_0 and MXFP4 on the Q3 version of protoLabsAI/Ornith-1.0-9B-MTP-GGUF.

Clearly, Q8_0 has the largest weight size, but its PPL is smaller than the original.

While MXFP4 can reduce the model weight size, it increases the PPL.

name Size PPL weight
Ornith-1.0-9B-MTP-GGUF/ornith-9b-mtp-kl-IQ3_M.gguf 4.67GB 7.2065 +/- 0.04554 Q6_K(output), IQ3_S(ffn_down/ssm_out)
Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF-Q8/ornith-9b-mtp-kl-IQ3_M.gguf 5.99GB 7.1733 +/- 0.04588 Q8_0(output/ffn_down/ssm_out)
Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF/ornith-9b-mtp-kl-IQ3_M.gguf 4.84GB 7.2943 +/- 0.04648 Q6_K(output), Q8_0(ffn_down/ssm_out)
Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF-MXFP4/ornith-9b-mtp-kl-IQ3_M.gguf 4.42GB 7.3632 +/- 0.04687 MXFP4(output/ffn_down/ssm_out)

Download

Use the latest llama.cpp,

huggingface-cli download huihui-ai/Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF --local-dir ./huihui-ai/Huihui-Ornith-1.0-9B-abliterated-MTP-GGUF --token xxx

Usage Warnings

  • Risk of Sensitive or Controversial Outputs: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.

  • Not Suitable for All Audiences: Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.

  • Legal and Ethical Responsibilities: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.

  • Research and Experimental Use: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.

  • Monitoring and Review Recommendations: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.

  • No Default Safety Guarantees: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.

Donation

If you like it, please click 'like' and follow us for more updates.
You can follow x.com/support_huihui to get the latest model information from huihui.ai.

Your donation helps us continue our further development and improvement, a cup of coffee can do it.
  • bitcoin(BTC):
  bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-05Update README.md8796bfc5.6 KB
    Loading...
  2. 2026-07-05Update README.mdb11833f5.7 KB
    Loading...
  3. 2026-07-01Update README.mdd56ef5e5.6 KB
    Loading...
  4. 2026-07-01Update README.md1d0abce5.6 KB
    Loading...
  5. 2026-07-01Update README.md4a40e4f5.6 KB
    Loading...
  6. 2026-07-01Create README.mdb0409984.4 KB
    Loading...

Discussions 1 thread

  1. 2026-07-0235b moe abliteratedopen1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration