← back to catalog · registered 2026-08-22 13:56

Jibbalit/phi-4-multimodal-instruct-abliterated

Jibbalit Phi multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Jibbalit%2Fphi-4-multimodal-instruct-abliterated"
Response includes
  • classification m1
  • files 27
  • hub_downloads_all_time 226
  • author_summary 6 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
226
13 last 30d - cooling
Likes
0
Model age
6mo ago
created 2026-03-21
Downloads over time
Now230→from24↑858%
149317225124 on Mar 25230 on Oct 11230 on Oct 9MarAprMayJunJulAugSepOct
Mar 25 → Oct 11 · 68 snapshots · spans 200 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Languages
multilingual ar zh cs da nl en fi fr de he hu it ja ko no pl pt ru es sv th tr uk
Tags
transformers safetensors phi4mm text-generation nlp code audio automatic-speech-recognition speech-summarization speech-translation visual-question-answering phi-4-multimodal

Related

Total size
1.77 GB
Files
27
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-03-21 15:54

Files by quantization

Auxiliary files 27 files 1.80 GB
model-00003-of-00003.safetensors 1.77 GB a3aba0ff download
tokenizer.json 14.8 MB 4c1b9f64 download
phi_4_mm.tech_report.02252025.pdf 5.05 MB a5469d91 download
vocab.json 3.73 MB ea953a43 download
merges.txt 2.50 MB e5c0a03f download
model.safetensors.index.json 236 KB 9460e3cd download
modeling_phi4mm.py 116 KB c4739e5c download
speech_conformer_encoder.py 111 KB 8c694978 download
vision_siglip_navit.py 78.1 KB 47ffbe9e download
processing_phi4mm.py 32.7 KB f89774fe download
sample_finetune_vision.py 19.7 KB 73c6d0c1 download
sample_finetune_speech.py 16.7 KB 4ccbae4a download
configuration_phi4mm.py 11.0 KB da4c4817 download
sample_inference_phi4mm.py 10.5 KB 1dcab14e download
config.json 4.69 KB acd677f4 download
README.md 3.52 KB c8b67b62 download
tokenizer_config.json 3.29 KB c60d9553 download
SECURITY.md 2.63 KB 6b906d43 download
.gitattributes 1.64 KB b3742eb2 download
SUPPORT.md 1.21 KB 291d4d43 download
LICENSE 1.13 KB 3d8b93bc download
special_tokens_map.json 497 B af930f71 download
preprocessor_config.json 496 B 4fe574bd download
CODE_OF_CONDUCT.md 453 B c72a5749 download
added_tokens.json 261 B 77e6ed40 download
generation_config.json 201 B 5d2077f7 download
processor_config.json 127 B 21851529 download

README current version from Hugging Face


license: mit
license_link: >-
https://huggingface.co/huihui-ai/Phi-4-multimodal-instruct-abliterated/resolve/main/LICENSE
language:

  • multilingual
  • ar
  • zh
  • cs
  • da
  • nl
  • en
  • fi
  • fr
  • de
  • he
  • hu
  • it
  • ja
  • ko
  • 'no'
  • pl
  • pt
  • ru
  • es
  • sv
  • th
  • tr
  • uk
    tags:
  • nlp
  • code
  • audio
  • automatic-speech-recognition
  • speech-summarization
  • speech-translation
  • visual-question-answering
  • phi-4-multimodal
  • phi
  • phi-4-mini
  • abliterated
  • uncensored
    widget:
  • example_title: Librispeech sample 1
    src: https://cdn-media.huggingface.co/speech_samples/sample1.flac
  • example_title: Librispeech sample 2
    src: https://cdn-media.huggingface.co/speech_samples/sample2.flac
  • messages:
    • role: user
      content: Can you provide ways to eat combinations of bananas and dragonfruits?
      library_name: transformers
      base_model:
  • microsoft/Phi-4-multimodal-instruct

huihui-ai/Phi-4-multimodal-instruct-abliterated

This is an uncensored version of microsoft/Phi-4-multimodal-instruct created with abliteration (see remove-refusals-with-transformers to know more about it).
This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens.

It was only the text part that was processed, not the image part.

The abliterated model will no longer say "I'm sorry, but I cannot provide details or descriptions of images"

Usage

You can use this model in your applications by loading it with Hugging Face's transformers library:

import os
import requests
import torch
from PIL import Image
import soundfile
from transformers import AutoModelForCausalLM, AutoProcessor, GenerationConfig

model_path = 'huihui-ai/Phi-4-multimodal-instruct-abliterated'

kwargs = {}
kwargs['torch_dtype'] = torch.bfloat16

processor = AutoProcessor.from_pretrained(model_path, trust_remote_code=True)
print(processor.tokenizer)

model = AutoModelForCausalLM.from_pretrained(
    model_path,
    trust_remote_code=True,
    torch_dtype='auto',
    _attn_implementation='flash_attention_2',
).cuda()
print("model.config._attn_implementation:", model.config._attn_implementation)

generation_config = GenerationConfig.from_pretrained(model_path, 'generation_config.json')

user_prompt = '<|user|>'
assistant_prompt = '<|assistant|>'
prompt_suffix = '<|end|>'
 
#################################################### text-only ####################################################
prompt = f'{user_prompt}what is the answer for 1+1? Explain it.{prompt_suffix}{assistant_prompt}'
print(f'>>> Prompt\n{prompt}')
inputs = processor(prompt, images=None, return_tensors='pt').to('cuda:0')

generate_ids = model.generate(
    **inputs,
    max_new_tokens=1000,
    generation_config=generation_config,
)
generate_ids = generate_ids[:, inputs['input_ids'].shape[1] :]
response = processor.batch_decode(
    generate_ids, skip_special_tokens=True, clean_up_tokenization_spaces=False
)[0]

print(f'>>> Response\n{response}')

Donation

If you like it, please click 'like' and follow us for more updates.
You can follow x.com/support_huihui to get the latest model information from huihui.ai.

Your donation helps us continue our further development and improvement, a cup of coffee can do it.
  • bitcoin:
  bc1qqnkhuchxw0zqjh2ku3lu4hq45hc6gy84uk70ge

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-03-21Upload folder using huggingface_hub38ce30e3.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration