
Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation. This makes it well-suited for multi-turn chat, instruction following, and complex agent workflows.
Modalities
Context
128K
Released
Apr 30, 2025
Knowledge Cutoff
Mar 2025
Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation.
Qwen3 4B has a 128,000 token context window.
Qwen3.8 27B, Qwen3.8 2.4T A95B, Qwen3.8 Max and 47 more are other text models from Qwen.
Qwen3 4B was released on April 30, 2025. Its knowledge cutoff is March 31, 2025.
Token volume and request traffic to this model over time.