library_name: mlx-vlm
pipeline_tag: image-text-to-text
base_model: huihui-ai/Huihui-Qwen3.6-27B-abliterated
tags:
- mlx
- mlx-vlm
- qwen3.6
- multimodal
- vision-language
- abliterated
Huihui-Qwen3.6-27B-abliterated-mlx-bf16
Local MLX-VLM conversion of huihui-ai/Huihui-Qwen3.6-27B-abliterated.
Overview
- Format: MLX-VLM
- Precision: bf16
- Size: 51G
- q_mode: bf16
- q_bits: bf16
- Source model type: Qwen3_5ForConditionalGeneration
- Source pipeline: image-text-to-text
- Intended runtime: mlx-vlm, LM Studio
Validation
- text generation smoke test: passed
- image generation path smoke test: passed
Quantization
- Quantization: none (bf16 export)
Test Results (2026-04-23)
Test runner: mlx_vlm.generate
Text prompt:你好,请用一句自然中文回应。
Image prompt:请用一句中文描述这张图片。
Image asset:
64x64 pure red square
Observed results:
| Case | Prompt TPS | Generation TPS | Peak Memory |
|---|---|---|---|
| Text | 11.309 | 6.471 | 54.861 GB |
| Image | 26.533 | 6.165 | 55.004 GB |
Notes:
- Both text and image tests completed successfully.
- The current
mlx_vlmruntime still emits<think>content even with--processor-kwargs '{"enable_thinking": false}'.