library_name: transformers
pipeline_tag: image-text-to-text
license: apache-2.0
base_model: orcarouter/Qwen3.8-27B-Uncensored
base_model_relation: quantized
tags:
- Qwen3.8
- nvfp4
- fp4
- w4a4
- compressed-tensors
- 1cat-vllm
- multimodal
- mtp
Qwen3.8-27B-Uncensored NVFP4 for 1Cat-vLLM
An all-transformer-linear NVFP4 W4A4 PTQ build oforcarouter/Qwen3.8-27B-Uncensored,
exported in the compressed-tensors nvfp4-pack-quantized layout used byQUASAR-QAT/Qwen3.8-27B-QUASAR-NVFP4
and supported by 1CatAI/1Cat-vLLM.
- 496/496 transformer linears: NVFP4 W4A4, group size 16
lm_head, token embeddings, vision tower, and MTP tensors: BF16- Calibration: 256 UltraChat samples, maximum sequence length 2,048
- LLM Compressor revision:
c8676edd2f6cddd0be2427f034825ca0d476fa72
This is post-training quantization, not QUASAR quantization-aware training.
It matches the runtime/checkpoint layout but does not claim QUASAR's benchmark
quality. See quantization-audit.json and calibration.json.
The source model has substantially reduced safety alignment. Use it only for
lawful, controlled purposes and add suitable safeguards before deployment.