Qwen: Qwen3 Embedding 4B
qwen/qwen3-embedding-4b
The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capabilities, long-text understanding, and reasoning skills of its foundational model. The Qwen3 Embedding series represents significant advancements in multiple text embedding and ranking tasks, including text retrieval, code retrieval, text classification, text clustering, and bitext mining.
Model specifications
- Input
- text
- Output
- text
- Context
- 32,000 tokens
- Max output
- 32,000 tokens
- Input price
- $0.02 / 1M tokens
- Output price
- $0 / 1M tokens
- Released
- 2026-02-18
Capabilities
- Streaming
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
1 available provider · tier input average $0.02 / 1M tokens
DeepInfra
Tier: Standard · Region: US
Pricing
- Input
- $0.02 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- Yes
- Data retention
- Zero retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- Yes
- HIPAA compliant
- No
- SOC 2 certified
- Yes
- BYOK supported
- Yes
Privacy policy · Terms · Official website · Documentation · Status · Support
Frequently asked questions
- What is Qwen: Qwen3 Embedding 4B?
- The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series inherits the exceptional multilingual capabilities, long-text understanding, and reasoning skills of its foundational model. The Qwen3 Embedding series represents significant advancements in multiple text embedding and ranking tasks, including text retrieval, code retrieval, text classification, text clustering, and bitext mining.
- How much does Qwen: Qwen3 Embedding 4B cost?
- Input costs start at $0.02 / 1M tokens and output costs start at $0 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of Qwen: Qwen3 Embedding 4B?
- Qwen: Qwen3 Embedding 4B supports a 32,000 token context window and up to 32,000 output tokens.
- What capabilities does Qwen: Qwen3 Embedding 4B support?
- Qwen: Qwen3 Embedding 4B supports Streaming, Playground.
- Which providers offer Qwen: Qwen3 Embedding 4B?
- Qwen: Qwen3 Embedding 4B is available from DeepInfra.
- How do providers handle data privacy for Qwen: Qwen3 Embedding 4B?
- 1 of 1 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.