← back to catalog · registered 2026-08-22 13:56

prithivMLmods/Qwen3-VL-30B-A3B-Instruct-abliterated-v1

prithivMLmods Qwen 31B MoE multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/prithivMLmods%2FQwen3-VL-30B-A3B-Instruct-abliterated-v1"
Response includes
  • classification m1
  • files 17
  • hub_downloads_all_time 23,157
  • author_summary 98 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
23K
20 last 30d - cooling
Likes
3
Descendants
4
in 4 direct forks
Model age
12mo ago
created 2025-10-16
Downloads over time
Now23.2K→from246↑9,320%
08.5K17K25.5K246 on Nov 12, 202523.2K on Oct 1123.2K on Oct 10Nov '25JanMarMayJulSep
Nov 12, 2025 → Oct 11 · 87 snapshots · spans 333 days

Genealogy 4 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
transformers safetensors qwen3_vl_moe image-text-to-text text-generation-inference abliterated v1.0 conversational en base_model:Qwen/Qwen3-VL-30B-A3B-Instruct base_model:finetune:Qwen/Qwen3-VL-30B-A3B-Instruct license:apache-2.0

Related

Total size
57.9 GB
Files
17
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-01 10:57

Files by quantization

Auxiliary files 17 files 57.9 GB
model-00001-of-00003.safetensors 27.9 GB ******** download
model-00002-of-00003.safetensors 27.9 GB ******** download
model-00003-of-00003.safetensors 2.12 GB ******** download
tokenizer.json 10.9 MB ******** download
vocab.json 2.65 MB 4783fe10 download
merges.txt 1.59 MB 31349551 download
model.safetensors.index.json 78.9 KB 03880686 download
tokenizer_config.json 5.32 KB fec7f182 download
chat_template.jinja 5.17 KB 12438680 download
README.md 4.43 KB 623a95cd download
config.json 1.70 KB c98ad83b download
.gitattributes 1.53 KB 52373fe2 download
video_preprocessor_config.json 817 B e32b1d90 download
preprocessor_config.json 782 B 2fa65535 download
added_tokens.json 707 B b54f9135 download
special_tokens_map.json 613 B ac23c0aa download
generation_config.json 213 B bdb4e037 download

README current version from Hugging Face


license: apache-2.0
base_model:

  • Qwen/Qwen3-VL-30B-A3B-Instruct
    language:
  • en
    pipeline_tag: image-text-to-text
    library_name: transformers
    tags:
  • text-generation-inference
  • abliterated
  • v1.0

1

Qwen3-VL-30B-A3B-Instruct-abliterated

Qwen3-VL-30B-A3B-Instruct-abliterated is an abliterated (v1.0) variant of Qwen3-VL-30B-A3B-Instruct**, designed for Abliterated Reasoning and Captioning.
This model leverages the Qwen3-VL-MoE (Mixture of Experts) architecture to deliver deeply descriptive, context-rich, and reasoning-oriented multimodal outputs. It handles complex, sensitive, and nuanced visual content while maintaining balanced interpretive coherence and multilingual adaptability.

1

Key Highlights

  • Abliterated / Uncensored Captioning and Reasoning
    Fine-tuned to bypass standard content filters while preserving factual accuracy, descriptive depth, and logical reasoning.

  • High-Fidelity Reasoning and Visual Understanding
    Generates detailed captions and structured reasoning for diverse visual categories—artistic, technical, abstract, or low-context.

  • Mixture of Experts (MoE) Efficiency
    Built on Qwen3-VL-MoE, dynamically routing computation through specialized experts for enhanced precision and scalability.

  • Aspect-Ratio Robustness
    Performs consistently across wide, tall, square, panoramic, and irregular visual formats.

  • Variational Detail Control
    Supports both concise summaries and highly detailed reasoning narratives, depending on prompt configuration.

  • Multilingual Output Capability
    Defaults to English but adaptable for multilingual use through prompt engineering.


Base Model Signatures:

This model has been re-sharded and optimized for the latest Transformers version from the base model: https://huggingface.co/huihui-ai/Huihui-Qwen3-VL-30B-A3B-Instruct-abliterated.


Quick Start with Transformers

from transformers import Qwen3VLMoeForConditionalGeneration, AutoProcessor
from qwen_vl_utils import process_vision_info
import torch

model = Qwen3VLMoeForConditionalGeneration.from_pretrained(
    "prithivMLmods/Qwen3-VL-30B-A3B-Instruct-abliterated-v1",
    torch_dtype="auto",
    device_map="auto"
)

processor = AutoProcessor.from_pretrained("prithivMLmods/Qwen3-VL-30B-A3B-Instruct-abliterated-v1")

messages = [
    {
        "role": "user",
        "content": [
            {
                "type": "image",
                "image": "https://qianwen-res.oss-cn-beijing.aliyuncs.com/Qwen-VL/assets/demo.jpeg",
            },
            {"type": "text", "text": "Provide a detailed caption and reasoning for this image."},
        ],
    }
]

text = processor.apply_chat_template(
    messages, tokenize=False, add_generation_prompt=True
)
image_inputs, video_inputs = process_vision_info(messages)

inputs = processor(
    text=[text],
    images=image_inputs,
    videos=video_inputs,
    padding=True,
    return_tensors="pt",
).to("cuda")

generated_ids = model.generate(**inputs, max_new_tokens=128)

generated_ids_trimmed = [
    out_ids[len(in_ids):] for in_ids, out_ids in zip(inputs.input_ids, generated_ids)
]

output_text = processor.batch_decode(
    generated_ids_trimmed,
    skip_special_tokens=True,
    clean_up_tokenization_spaces=False
)

print(output_text)

Intended Use

This model is suited for:

  • Generating detailed, uncensored captions and reasoning for complex or creative visual datasets.
  • Research in multimodal reasoning, safety evaluation, and content moderation studies.
  • Enabling descriptive captioning and analytical reasoning for datasets excluded from mainstream models.
  • Creative applications such as narrative generation, artistic interpretation, and visual storytelling.
  • Advanced reasoning over diverse visual structures and aspect ratios.

Limitations

  • May produce explicit, sensitive, or offensive content depending on input and prompt.
  • Not recommended for deployment in production systems that require strict moderation or filtering.
  • Style, tone, and reasoning detail can vary based on prompt phrasing.
  • May show variable performance on synthetic, abstract, or highly stylized visual inputs.
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration