Qwen: Qwen3 32B
qwen/qwen3-32b
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for tasks like math, coding, and logical inference, and a "non-thinking" mode for faster, general-purpose conversation. The model demonstrates strong performance in instruction-following, agent tool use, creative writing, and multilingual tasks across 100+ languages and dialects. It natively handles 32K token contexts and can extend to 131K tokens using YaRN-based scaling.
Model specifications
- Context
- 40,960 tokens
- Max output
- 32,768 tokens
- Input price
- $0.08 / 1M tokens
- Output price
- $0.24 / 1M tokens
- Released
- 2026-02-05
Capabilities
- Streaming
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
2 available providers · tier input average $0.13 / 1M tokens · tier output average $0.92 / 1M tokens
AtlasCloud
Tier: Standard · Region: US
Pricing
- Input
- $0.1 / 1M tokens
- Output