← back to catalog · registered 2026-10-08 04:58

huihui-ai/Huihui-Kolibri-1-BF16-abliterated-GGUF

huihui-ai GGUF MoE
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/huihui-ai%2FHuihui-Kolibri-1-BF16-abliterated-GGUF"
Response includes
  • classification unknown
  • files 10
  • author_summary 183 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
208
Likes
12
Model age
1d ago
created 2026-10-06

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 208 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
apache-2.0
Languages
de en
Quantizations
BF16 Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
vllm gguf abliterated uncensored huihui reasoning moe text-generation de en base_model:Aleph-Alpha/Kolibri-1-BF16 base_model:quantized:Aleph-Alpha/Kolibri-1-BF16

Related

Total size
440 GB
Files
10
Quantizations
8
Registered
2026-10-08 04:58
Last updated on HF
2026-10-08 04:45

Files by quantization

BF16 1 file 146 GB
Huihui-Kolibri-1-abliterated-bf16.gguf 146 GB aee267a9 download
Q8_0 1 file 77.4 GB
Huihui-Kolibri-1-abliterated-Q8_0.gguf 77.4 GB 4f046a8a download
Q6_K 1 file 59.8 GB
Huihui-Kolibri-1-abliterated-Q6_K.gguf 59.8 GB e9c282c8 download
Q5_K 1 file 51.8 GB
Huihui-Kolibri-1-abliterated-Q5_K.gguf 51.8 GB 4a363048 download
Q4_K 1 file 44.2 GB
Huihui-Kolibri-1-abliterated-Q4_K_M.gguf 44.2 GB f0facf28 download
Q3_K 1 file 34.9 GB
Huihui-Kolibri-1-abliterated-Q3_K.gguf 34.9 GB cd671bba download
Q2_K 1 file 26.7 GB
Huihui-Kolibri-1-abliterated-Q2_K.gguf 26.7 GB 4b22f23a download
Auxiliary files 3 files 34.3 KB
kolibri1-llama.cpp.patch 29.1 KB d07c2702 download
README.md 3.05 KB b2a4476d download
.gitattributes 2.18 KB d00bfbf7 download

README current version from Hugging Face


license: "apache-2.0"
pipeline_tag: text-generation
library_name: vllm
base_model: Aleph-Alpha/Kolibri-1-BF16
language:

  • de
  • en
    tags:
  • abliterated
  • uncensored
  • huihui
  • reasoning
  • moe

huihui-ai/Huihui-Kolibri-1-BF16-abliterated-GGUF

This is an uncensored version of Aleph-Alpha/Kolibri-1-BF16 created with abliteration.

This ablation was performed entirely using llama.cpp's GGUF model, without relying on Transformers,

For the specific ablation method, please refer to cvector-generator.

Latest update

We re-quantized the models using llama.cpp and have uploaded the Q2, Q3, Q4, Q5, Q6, Q8, and BF16 quantized models.

Note

GGUFs come from Hob-forge/Kolibri-1-GGUF.

The ablation did not modify all expert modules.

llama.cpp

Requires a llama.cpp patch

Stock llama.cpp does not support the kolibri1 architecture yet. Apply kolibri1-llama.cpp.patch (included in this repo) to llama.cpp at upstream commit 836d571, then build:

git clone https://github.com/ggml-org/llama.cpp && cd llama.cpp
git checkout 836d571
git am /path/to/kolibri1-llama.cpp.patch
cmake -B build -DCMAKE_BUILD_TYPE=Release   # add -DGGML_CUDA=ON etc. for your GPU
cmake --build build -j --target llama-server llama-cli

Running

llama-cli -m huihui-ai/Huihui-Kolibri-1-abliterated-GGUF/Huihui-Kolibri-1-abliterated-Q4_K_M.gguf -c 262144

Usage Warnings

  • Risk of Sensitive or Controversial Outputs: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.

  • Not Suitable for All Audiences: Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.

  • Legal and Ethical Responsibilities: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.

  • Research and Experimental Use: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.

  • Monitoring and Review Recommendations: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.

  • No Default Safety Guarantees: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.

Donation

Your donation helps us continue our further development and improvement, a cup of coffee can do it.
  • bitcoin:
  bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge
  • Support our work on Ko-fi!

README history 7 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-10-08Update README.mdd160d1d3.1 KB
    Loading...
  2. 2026-10-08Update README.md89256643 KB
    Loading...
  3. 2026-10-08Update README.mdec7dec93 KB
    Loading...
  4. 2026-10-07Update README.md49ec0162.9 KB
    Loading...
  5. 2026-10-07Update README.md8aa9daf2.9 KB
    Loading...
  6. 2026-10-06Update README.mdcff6db82.9 KB
    Loading...
  7. 2026-10-06Create README.mdc35a6a02.8 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration