library_name: transformers
license: apache-2.0
license_link: https://huggingface.co/Qwen/Qwen3.4-9B/blob/main/LICENSE
pipeline_tag: image-text-to-text
base_model:
- Qwen/Qwen3.5-4B-Base
tags: - fp32
- qwen
- uncensored
- caption
Qwen3.5-4B
USE FP32 version for quantization or conversion to BF16/FP16
[!Note]>
This model was trained in FP32 with Adam on Vision, Chain of Thought and Reasoning.
This model shows a massive increase in uncensored and unfiltered results using with THINK ON or with THINK OFF
In testing with think mode off 98% of results where uncensored. With THINK on the results are less consistent but far more unique.
Qwen3.5 Highlights
Qwen3.5 features the following enhancement:
Unified Vision-Language Foundation: Early fusion training on multimodal tokens achieves cross-generational parity with Qwen3 and outperforms Qwen3-VL models across reasoning, coding, agents, and visual understanding benchmarks.
Efficient Hybrid Architecture: Gated Delta Networks combined with sparse Mixture-of-Experts deliver high-throughput inference with minimal latency and cost overhead.
Scalable RL Generalization: Reinforcement learning scaled across million-agent environments with progressively complex task distributions for robust real-world adaptability.
Global Linguistic Coverage: Expanded support to 201 languages and dialects, enabling inclusive, worldwide deployment with nuanced cultural and regional understanding.
Next-Generation Training Infrastructure: Near-100% multimodal training efficiency compared to text-only training and asynchronous RL frameworks supporting massive-scale agent scaffolds and environment orchestration.