license: apache-2.0
base_model: wangzhang/Qwen3.6-35B-A3B-abliterated-v2
tags:
- gguf
- rocmfpx
- rocmfp4
- mtp
- text-generation
- conversational
pipeline_tag: text-generation
Qwen3.6-35B-A3B Abliterated (ROCmFP4)
This is an abliterated (uncensored) Qwen 3.6 35B MoE model converted to GGUF and quantized to ROCmFP4 (Q4_0_ROCMFP4_FAST_COHERENT) using the ROCmFPX toolchain.
Optimized for AMD Radeon GPUs (e.g., RX 9060 XT 16GB / RX 7900 XT 20GB) via Vulkan or ROCm backends with Multi-Token Prediction (MTP) speculative decoding enabled.
How to Run with llama-server
llama-server.exe -m qwen3.6-35b-v2-ROCMFP4.gguf --spec-type draft-mtp -c 32768 --port 8080