Qwen3.5 4B is a compact open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output.
Added Aug 16, 2026
Context Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.11/1M
Output Price (Auto)
$0.21/1M
Cache Read (Auto)
$0.053/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Qwen3.5 4B with similar models from the same provider or model family.
Qwen3.5 0.8B
qwen3.5-0.8bQwen3.5 0.8B is a lightweight open-weight multimodal model from Alibaba for fast reasoning, visual understanding, tool use, and JSON output.
Qwen3.5 2B
qwen3.5-2bQwen3.5 2B is a small open-weight multimodal model from Alibaba for efficient reasoning, coding, visual understanding, tool use, and JSON output.
Qwen3.8 27B
qwen3.8-27bQwen3.8 27B is an open-weight multimodal model from Alibaba for coding, visual understanding, tool use, and structured output. This variant keeps thinking disabled for faster direct responses.
Qwen3.8 27B Thinking
qwen3.8-27b:thinkingQwen3.8 27B is an open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output. This variant enables thinking by default.
Qwen3.8 Max
qwen3.8-maxQwen3.8 Max is Qwen's 2.4T-parameter flagship model for coding, knowledge work, full-stack development, data analysis, and long-running agent workflows in non-thinking mode. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.
Qwen3.8 Max Thinking
qwen3.8-max:thinkingQwen3.8 Max Thinking enables generation-time reasoning for deeper coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, tool calling, structured output, and a near-million-token context window.