Qwen: Qwen3.8 Max
qwen/qwen3.8-max
2.4-trillion-parameter MoE flagship delivering a comprehensive leap in coding and professional work. Autonomously codes and delivers complete projects spanning 10+ days. Handles hundreds of specialized tasks across legal, financial, design, and other professional domains, producing production-grade results end-to-end in a single conversation. Native visual understanding runs through the full cycle of planning, execution, and verification, enabling deep semantic analysis of ultra-long documents and extended video content. In long-horizon tasks, plans autonomously, iterates through closed feedback loops, and continuously evolves.
Model specifications
- Input
- text, image, video
- Output
- text
- Context
- 1,000,000 tokens
- Max output
- 128,000 tokens
- Input price
- $2 / 1M tokens
- Output price
- $6 / 1M tokens
- Released
- 2026-08-03
Capabilities
- Streaming
- Function calling
- Vision
- JSON mode
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Flex service tier
1 available provider · discounts up to 50% · tier input average $0.825 / 1M tokens · tier output average $2.4755 / 1M tokens
Alibaba Cloud Int.(CN)
Tier: Flex · Region: CN
Pricing
- Input
- $0.825 / 1M tokens
- Output
- $2.4755 / 1M tokens
- Cache read
- $0.103 / 1M tokens
Discounts
- Input: 50% off — list