library_name: transformers
license: apache-2.0
license_link: https://huggingface.co/Qwen/Qwen3.6-35B-A3B/blob/main/LICENSE
pipeline_tag: image-text-to-text
tags:
- heretic
- uncensored
- decensored
- abliterated
- mlx
- mlx-my-repo
base_model: llmfan46/Qwen3.6-35B-A3B-uncensored-heretic
ijwfly/Qwen3.6-35B-A3B-uncensored-heretic-mlx-8Bit
The Model ijwfly/Qwen3.6-35B-A3B-uncensored-heretic-mlx-8Bit was converted to MLX format from llmfan46/Qwen3.6-35B-A3B-uncensored-heretic using mlx-vlm version 0.4.4.
Use with mlx-openai-server
pip install mlx-openai-server
You have to choose trade-off between model types (model-type parameter):
lm: prefix caching works, but vision is not supportedmultimodal: vision works, but prefix caching is unavailable
mlx-openai-server launch \
--model-path ijwfly/Qwen3.6-35B-A3B-uncensored-heretic-mlx-8Bit \
--reasoning-parser qwen3_5 \
--model-type lm \
--tool-call-parser qwen3_coder \
--context-length 65535 \
--port 8432 \
--temperature 0.6 \
--top-p=0.95 \
--top-k=20 \
--min-p=0.0 \
--presence-penalty=0.0 \
--repetition-penalty=1.0