← back to catalog · registered 2026-08-22 13:56

linjian257/qwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257

linjian257 Minimax
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/linjian257%2Fqwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257"
Response includes
  • classification m-uncensored
  • files 3
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
50
Model age
2mo ago
created 2026-08-05
Downloads over time
Now0→from0↑0%
00110 on Aug 50 on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Tags
comfyui minimax-h3 qwen3-vl text-encoder safetensors int8 convrot quantized uncensored base_model:Comfy-Org/MiniMax-H3 base_model:finetune:Comfy-Org/MiniMax-H3 license:other
Total size
24.0 GB
Files
3
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-05 14:43

Files by quantization

Auxiliary files 3 files 24.0 GB
qwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257.safetensors 24.0 GB e385249a download
README.md 10.5 KB bab1d502 download
.gitattributes 1.48 KB a6344aac download

README current version from Hugging Face


license: other
license_name: personal-entertainment-use-only
base_model:

  • MiniMaxAI/MiniMax-H3
  • Comfy-Org/MiniMax-H3
    tags:
  • comfyui
  • minimax-h3
  • qwen3-vl
  • text-encoder
  • safetensors
  • int8
  • convrot
  • quantized
  • uncensored

qwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257

中文 | English

模型简介

这是一个供 MiniMax-H3 ComfyUI 工作流使用的 Qwen3-VL-32B INT8 TensorWise + ConvRot 社区修改版文本编码器。该版本由 linjian257 从修改后的 BF16 权重量化得到,重点是在保持 ComfyUI 工作流兼容性的同时明显降低权重文件体积。

本文件只是 MiniMax-H3 的文本编码器组件,并不是完整的视频生成模型,也不是可直接在 LM Studio、llama.cpp 或 vLLM 中作为聊天模型加载的完整 Qwen3-VL 模型。

关键信息

项目 内容
文件名 qwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257.safetensors
用途 MiniMax-H3 的 ComfyUI 文本编码器
量化格式 INT8 TensorWise + ConvRot
ConvRot group size 256
Scale FP32、per-channel、沿输出特征轴
文件大小 25,772,287,417 bytes(约 25.77 GB / 24.00 GiB)
总张量数量 1,838
量化权重张量 468
ConvRot 张量 359(另有 109 个量化张量未旋转)
保留 BF16 张量 434
文本编码器输出 第 50 层之后的未归一化隐藏状态
修改者 linjian257
上游仓库 MiniMaxAI/MiniMax-H3 / Comfy-Org/MiniMax-H3

量化信息

  • 算法:int8_tensorwise,版本 0.1.0
  • 存储类型:signed INT8 权重
  • 缩放方式:每输出通道一个 FP32 scale(per_channel / out_features)
  • Scale 计算:绝对值最大值(amax)
  • 舍入:nearest-even
  • 量化范围:model.* 与 visual.* 中选定的二维权重;bias 等不适合量化的张量保持原精度
  • ConvRot:启用,group size 为 256

该文件是从 linjian257 制作的 BF16 社区修改版权重重新量化得到的衍生版本,并非官方 qwen3vl_32b_minimax_h3_int8_convrot.safetensors 的改名或逐字节镜像。

“Uncensored”用于区分版本,不代表对所有提示词、语言、输入形式或工作流效果作出保证。

安装与使用

  1. 使用支持 MiniMax-H3 及原生 INT8 ConvRot 的较新版本 ComfyUI。如果出现未知量化格式、缺少 weight_scale 或 ConvRot 相关错误,请先更新 ComfyUI 及其依赖。

  2. 将文件放入:

    ComfyUI/models/text_encoders/
    
  3. 下载 MiniMax-H3 工作流所需的扩散模型和 VAE 文件。官方整理版见 Comfy-Org/MiniMax-H3。

  4. 打开官方 T2V、I2V 或 R2V 工作流。

  5. 在 CLIPLoader 节点中选择本文件,并保持模型类型为 minimax。

如果 ComfyUI 找不到该文件,请确认它位于 models/text_encoders 而不是 models/checkpoints 或 models/diffusion_models,然后刷新模型列表或重启 ComfyUI。

已知限制

  • 本仓库只提供文本编码器权重;运行 MiniMax-H3 还需要匹配的扩散模型、视频 VAE,以及在需要音频时使用的音频 VAE。
  • INT8 通常能减少权重存储和内存压力,但实际显存占用、速度与兼容性取决于 GPU、后端、Torch/ComfyUI 版本、卸载设置和完整工作流;本模型卡不承诺特定加速倍数。
  • ConvRot 需要运行时正确识别并应用对应旋转元数据。较旧或不兼容的加载器可能报错,也可能错误加载并产生异常输出。
  • INT8 量化和社区修改可能相对原始 BF16 版本带来细微质量变化,包括提示词遵循、稳定性、事实性或生成质量方面的差异。
  • 当前未随本模型卡提供系统化的 BF16/官方 INT8 对比、安全评测或端到端视频质量基准。
  • 该权重不是官方 Comfy-Org 文件的原样镜像;请勿使用官方文件哈希校验本衍生版本。

安全与责任

该版本可能生成冒犯性、危险、违法、误导或不适宜的内容。使用者必须自行评估输出、遵守适用法律和平台规则,并为生成、发布和传播的内容负责。请勿将其用于伤害未成年人、侵犯隐私、欺诈、仇恨、军事或其他违法有害用途。

使用许可

本权重采用个人娱乐使用许可 v1.0:仅允许个人娱乐、个人学习、非商业研究和非商业测试。未经 linjian257 书面许可,不得用于商业用途、收费服务、商业部署、转售、再分发或转授权。

致谢

  • MiniMaxAI — MiniMax-H3
  • Qwen — Qwen3-VL-32B
  • Comfy-Org — ComfyUI 适配、INT8 运行支持及权重整理
  • linjian257 — 本衍生版本的处理、量化与发布

English

Model description

This is a community-modified and quantized INT8 TensorWise + ConvRot Qwen3-VL-32B text encoder intended for MiniMax-H3 workflows in ComfyUI. It was produced by linjian257 from modified BF16 weights, with the goal of substantially reducing the weight-file size while retaining compatibility with the ComfyUI workflow.

This file is only the text-encoder component of MiniMax-H3. It is not a complete video-generation model and is not a standalone Qwen3-VL chat model that can be loaded directly in LM Studio, llama.cpp, or vLLM.

Key facts

Field Value
Filename qwen3vl_32b_minimax_h3_int8_convrot_uncensored-by-linjian257.safetensors
Intended role ComfyUI text encoder for MiniMax-H3
Quantization INT8 TensorWise + ConvRot
ConvRot group size 256
Scales FP32, per-channel, along the output-feature axis
File size 25,772,287,417 bytes (about 25.77 GB / 24.00 GiB)
Total tensor count 1,838
Quantized weight tensors 468
ConvRot tensors 359 (109 other quantized tensors are not rotated)
BF16 tensors retained 434
Encoder output Unnormalized hidden state after layer 50
Modified by linjian257
Upstream MiniMaxAI/MiniMax-H3 / Comfy-Org/MiniMax-H3

Quantization details

  • Algorithm: int8_tensorwise, version 0.1.0
  • Storage: signed INT8 weights
  • Scaling: one FP32 scale per output channel (per_channel / out_features)
  • Scale method: absolute maximum (amax)
  • Rounding: nearest-even
  • Selection: eligible 2D weights under model.* and visual.*; unsuitable tensors such as biases remain at their source precision
  • ConvRot: enabled with group size 256

This is a newly quantized community derivative of the modified BF16 weights produced by linjian257. It is not a rename or byte-identical mirror of the official qwen3vl_32b_minimax_h3_int8_convrot.safetensors file.

“Uncensored” distinguishes this release and does not guarantee identical behavior across all prompts, languages, input types, or workflows.

Installation and usage

  1. Use a recent ComfyUI build with MiniMax-H3 and native INT8 ConvRot support. If you see unknown quantization-format, missing weight_scale, or ConvRot-related errors, update ComfyUI and its dependencies first.

  2. Place the file in:

    ComfyUI/models/text_encoders/
    
  3. Download the matching MiniMax-H3 diffusion model and VAE files. The official ComfyUI repack is available at Comfy-Org/MiniMax-H3.

  4. Open the official T2V, I2V, or R2V workflow.

  5. Select this file in the CLIPLoader node and keep the model type set to minimax.

If the file is not listed, verify that it is under models/text_encoders, not models/checkpoints or models/diffusion_models, then refresh the model list or restart ComfyUI.

Limitations

  • This repository provides only the text encoder. MiniMax-H3 also requires a compatible diffusion model, video VAE, and an audio VAE when audio generation is used.
  • INT8 generally reduces weight storage and memory pressure, but actual VRAM use, speed, and compatibility depend on the GPU, backend, Torch/ComfyUI version, offloading settings, and complete workflow. No specific speedup is guaranteed.
  • ConvRot requires a loader that recognizes and applies the rotation metadata correctly. Older or incompatible loaders may fail or may produce invalid output.
  • INT8 quantization and community modifications can introduce small quality changes relative to the original BF16 version, including differences in prompt adherence, stability, factuality, or generation quality.
  • No systematic BF16/official-INT8 comparison, safety evaluation, or end-to-end video-quality benchmark is provided with this model card.
  • This is a modified derivative, not a byte-identical mirror of the official Comfy-Org file. Do not validate it against the official file hash.

Safety and responsibility

This derivative may produce offensive, dangerous, illegal, misleading, or otherwise inappropriate content. Users are responsible for reviewing outputs, complying with applicable law and platform policies, and controlling how generated material is used or distributed. Do not use it to harm minors, violate privacy, commit fraud, promote hate, support military purposes, or engage in other illegal or harmful activity.

License

These weights are released under the Personal Entertainment Use License v1.0. They may be used only for personal entertainment, personal learning, non-commercial research, and non-commercial testing. Commercial use, paid services, commercial deployment, resale, redistribution, and sublicensing require prior written permission from linjian257.

Credits

  • MiniMaxAI — MiniMax-H3
  • Qwen — Qwen3-VL-32B
  • Comfy-Org — ComfyUI integration, INT8 runtime support, and repackaged weights
  • linjian257 — modification, quantization, and release of this derivative

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-05Update README.md19a1c2010.5 KB
    Loading...
  2. 2026-08-05initial commita1027de28 B
    Loading...

Discussions 1 thread

  1. 2026-08-20give error using it!!open1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration