Qwen: Qwen3 Next 80B A3B Thinking
qwen/qwen3-next-80b-a3b-thinking
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic planning, and reports strong results across knowledge, reasoning, coding, alignment, and multilingual evaluations. Compared with prior Qwen3 variants, it emphasizes stability under long chains of thought and efficient scaling during inference, and it is tuned to follow complex instructions while reducing repetitive or off-task behavior. The model is suitable for agent frameworks and tool use (function calling), retrieval-heavy workflows, and standardized benchmarking where step-by-step solutions are required. It supports long, detailed completions and leverages throughput-oriented techniques (e.g., multi-token prediction) for faster generation. Note that it operates in thinking-only mode.
Model specifications
- Context
- 262,140 tokens
- Max output
- 32,768 tokens
- Input price
- $0.15 / 1M tokens
- Output price
- $1.5 / 1M tokens
- Released
- 2026-02-05
Capabilities
- Streaming
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
2 available providers · discounts up to 50% · tier input average $0.1125 / 1M tokens · tier output average $1.05 / 1M tokens
Alibaba Cloud Int.(SG)
Tier: Standard · Region: SG · Quantization: int8
Pricing
- Input
- $0.075 / 1M tokens
- Output
- $0.6 / 1M tokens
Discounts
- Input: 50% off — list $0.15 / 1M tokens, discounted $0.075 / 1M tokens, effective $0.075 / 1M tokens
- Output: 50% off — list $1.2 / 1M tokens, discounted $0.6 / 1M tokens, effective $0.6 / 1M tokens
Data privacy and compliance
- Region
- SG
- Zero data retention
- No
- Data retention
- 30-day retention
- Used for training
- No
- Data collection
- Moderated
- Yes
- GDPR compliant
- Yes
- HIPAA compliant
- No
- SOC 2 certified
- Yes
- BYOK supported
- Yes
Privacy policy · Terms · Official website · Documentation · Status · Support
AtlasCloud
Tier: Standard · Region: US
Pricing
- Input
- $0.15 / 1M tokens
- Output
- $1.5 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- No
- Data retention
- 7-day retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- No
- HIPAA compliant
- No
- SOC 2 certified
- No
- BYOK supported
- No
Privacy policy · Terms · Official website · Documentation · Support
Frequently asked questions
- What is Qwen: Qwen3 Next 80B A3B Thinking?
- Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic planning, and reports strong results across knowledge, reasoning, coding, alignment, and multilingual evaluations. Compared with prior Qwen3 variants, it emphasizes stability under long chains of thought and efficient scaling during inference, and it is tuned to follow complex instructions while reducing repetitive or off-task behavior. The model is suitable for agent frameworks and tool use (function calling), retrieval-heavy workflows, and standardized benchmarking where step-by-step solutions are required. It supports long, detailed completions and leverages throughput-oriented techniques (e.g., multi-token prediction) for faster generation. Note that it operates in thinking-only mode.
- How much does Qwen: Qwen3 Next 80B A3B Thinking cost?
- Input costs start at $0.15 / 1M tokens and output costs start at $1.5 / 1M tokens. Provider-level prices vary by service tier.
- Are provider discounts available for Qwen: Qwen3 Next 80B A3B Thinking?
- Yes. Current provider offers include discounts of up to 50% from list price. The provider table shows list, discounted, and effective prices.
- What is the context length of Qwen: Qwen3 Next 80B A3B Thinking?
- Qwen: Qwen3 Next 80B A3B Thinking supports a 262,140 token context window and up to 32,768 output tokens.
- What capabilities does Qwen: Qwen3 Next 80B A3B Thinking support?
- Qwen: Qwen3 Next 80B A3B Thinking supports Streaming, Playground.
- Which providers offer Qwen: Qwen3 Next 80B A3B Thinking?
- Qwen: Qwen3 Next 80B A3B Thinking is available from Alibaba Cloud Int.(SG), AtlasCloud.
- How do providers handle data privacy for Qwen: Qwen3 Next 80B A3B Thinking?
- 2 of 2 providers report that customer data is not used for training, and 0 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.