← back to catalog · registered 2026-08-22 13:56

huihui-ai/Huihui-MiniCPM-V-4_5-abliterated

huihui-ai 8.7B GGUF multimodal 41K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/huihui-ai%2FHuihui-MiniCPM-V-4_5-abliterated"
Response includes
  • classification m8
  • files 25
  • hub_downloads_all_time 50,221
  • author_summary 183 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of layer-wise ablation inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=huihui-ai + is_gguf=1
  • M3 (huihui abliteration) wrapped in M8 (GGUF quantization)
Refusal direction extracted via
Extraction technique

huihui-ai layer-band extraction

Confidence
HIGH
Why we say so
producer=huihui-ai (documented layer-band methodology in model cards)
Downloads · lifetime
50K
2K last 30d - cooling
Likes
33
Descendants
2
in 2 direct forks
Model age
13mo ago
created 2025-08-31
Downloads over time
Now51.3K→from3.8K↑1,265%
1.4K19.6K37.8K56K3.8K on Sep 3, 202551.3K on Oct 11Sep '25Nov '25JanMarMayJulSep
Sep 3, 2025 → Oct 11 · 97 snapshots · spans 403 days

Genealogy 2 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
multilingual
Tags
transformers safetensors gguf minicpmv feature-extraction minicpm-v vision ocr multi-image video custom_code abliterated

Related

Total size
16.2 GB
Files
25
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-09-08 23:23

Files by quantization

Auxiliary files 25 files 16.2 GB
model-00003-of-00004.safetensors 4.64 GB ff36b1d3 download
model-00002-of-00004.safetensors 4.58 GB 11a6403d download
model-00001-of-00004.safetensors 4.56 GB 2f56635d download
model-00004-of-00004.safetensors 2.41 GB 9a56d718 download
tokenizer.json 10.9 MB c5a94a2c download
vocab.json 2.65 MB 4783fe10 download
merges.txt 1.59 MB 31349551 download
model.safetensors.index.json 71.4 KB daea6a5b download
model_params.txt 64.8 KB 9e55b2d0 download
modeling_navit_siglip.py 40.9 KB 6fe732cf download
tokenizer_config.json 25.2 KB aaf1d40a download
image_processing_minicpmv.py 20.3 KB 7ae14cf9 download
modeling_minicpmv.py 17.3 KB 88dce467 download
special_tokens_map.json 11.8 KB b0204ae4 download
resampler.py 11.5 KB be74b30a download
processing_minicpmv.py 10.8 KB cefc2456 download
chat_template.jinja 4.21 KB 96107b68 download
README.md 4.07 KB 6849b2c9 download
configuration_minicpm.py 3.29 KB c1a10b62 download
added_tokens.json 2.90 KB fe7e0b26 download
.gitattributes 1.84 KB dcf57ca6 download
tokenization_minicpmv_fast.py 1.61 KB c8411fa7 download
config.json 1.43 KB 126378bb download
preprocessor_config.json 714 B 7111b617 download
generation_config.json 268 B 233aece9 download

README current version from Hugging Face


license: apache-2.0
pipeline_tag: image-text-to-text
library_name: transformers
base_model:

  • openbmb/MiniCPM-V-4_5
    language:
  • multilingual
    tags:
  • minicpm-v
  • vision
  • ocr
  • multi-image
  • video
  • custom_code
  • abliterated
  • uncensored

huihui-ai/Huihui-MiniCPM-V-4_5-abliterated

This is an uncensored version of openbmb/MiniCPM-V-4_5 created with abliteration (see remove-refusals-with-transformers to know more about it).

It was only the text part that was processed, not the image part.

The abliterated model will no longer say "I'm sorry, but I can't assist with that."

Chat with Image

1. llama.cpp Inference

(llama-mtmd-cli needs to be compiled.)

llama-mtmd-cli -m huihui-ai/Huihui-Qwen3-8B-abliterated/GGUF/ggml-model-Q4_K_M.gguf --mmproj huihui-ai/Huihui-Qwen3-8B-abliterated/GGUF/mmproj-model-f16.gguf -c 4096 --temp 0.7 --top-p 0.8 --top-k 100 --repeat-penalty 1.05 --image abc.png -p "What is in the image?" 

2. Transfromers Inference

import torch
from PIL import Image
from transformers import AutoModel, AutoTokenizer

torch.manual_seed(100)

model = AutoModel.from_pretrained('huihui-ai/Huihui-MiniCPM-V-4_5-abliterated', trust_remote_code=True,
    attn_implementation='sdpa', torch_dtype=torch.bfloat16) # sdpa or flash_attention_2, no eager
model = model.eval().cuda()
tokenizer = AutoTokenizer.from_pretrained('huihui-ai/Huihui-MiniCPM-V-4_5-abliterated', trust_remote_code=True)

image = Image.open('./assets/minicpmo2_6/show_demo.jpg').convert('RGB')

enable_thinking=False # If `enable_thinking=True`, the thinking mode is enabled.
stream=True # If `stream=True`, the answer is string

# First round chat 
question = "What is the landform in the picture?"
msgs = [{'role': 'user', 'content': [image, question]}]

answer = model.chat(
    msgs=msgs,
    tokenizer=tokenizer,
    enable_thinking=enable_thinking,
    stream=True
)

generated_text = ""
for new_text in answer:
    generated_text += new_text
    print(new_text, flush=True, end='')

# Second round chat, pass history context of multi-turn conversation
msgs.append({"role": "assistant", "content": [generated_text]})
msgs.append({"role": "user", "content": ["What should I pay attention to when traveling here?"]})

answer = model.chat(
    msgs=msgs,
    tokenizer=tokenizer,
    stream=True
)

generated_text = ""
for new_text in answer:
    generated_text += new_text
    print(new_text, flush=True, end='')

Usage Warnings

  • Risk of Sensitive or Controversial Outputs: This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs.

  • Not Suitable for All Audiences: Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security.

  • Legal and Ethical Responsibilities: Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences.

  • Research and Experimental Use: It is recommended to use this model for research, testing, or controlled environments, avoiding direct use in production or public-facing commercial applications.

  • Monitoring and Review Recommendations: Users are strongly advised to monitor model outputs in real-time and conduct manual reviews when necessary to prevent the dissemination of inappropriate content.

  • No Default Safety Guarantees: Unlike standard models, this model has not undergone rigorous safety optimization. huihui.ai bears no responsibility for any consequences arising from its use.

Donation

Your donation helps us continue our further development and improvement, a cup of coffee can do it.
  • bitcoin:
  bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge
  • Support our work on Ko-fi!

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-09-01Update README.mdce2d54b4.1 KB
    Loading...
  2. 2025-09-01Update README.md7e0bbbf4.1 KB
    Loading...
  3. 2025-09-01Update README.md20d77be2.6 KB
    Loading...
  4. 2025-09-01Update README.md4d9812e2.6 KB
    Loading...
  5. 2025-08-31Add files using upload-large-folder tool54daa422.2 KB
    Loading...
  6. 2025-08-31initial commit060d3af28 B
    Loading...

Discussions 4 threads

  1. 2026-03-13what's the use of this when you can't ....open1 💬#4
    Loading...
  2. 2026-01-03ggufopen1 💬#3
    Loading...
  3. 2025-11-27Can this be used with comfyui?open1 💬#2
    Loading...
  4. 2025-09-08Q_6 pleaseopen2 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration