Deepseek: Deepseek R1 Distill Llama 70B
deepseek/deepseek-r1-distill-llama-70b
DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across multiple benchmarks, including: AIME 2024 pass@1: 70.0 MATH-500 pass@1: 94.5 CodeForces Rating: 1633 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.
Model specifications
- Input
- text
- Output
- text
- Context
- 131,072 tokens
- Max output
- 16,400 tokens
- Input price
- $0.7 / 1M tokens
- Output price
- $0.8 / 1M tokens
- Released
- 2026-02-18
Capabilities
- Streaming
- Function calling
- JSON mode
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
1 available provider · tier input average $0.7 / 1M tokens · tier output average $0.8 / 1M tokens
DeepInfra
Tier: Standard · Region: US
Pricing
- Input
- $0.7 / 1M tokens
- Output
- $0.8 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- Yes
- Data retention
- Zero retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- Yes
- HIPAA compliant
- No
- SOC 2 certified
- Yes
- BYOK supported
- Yes
Privacy policy · Terms · Official website · Documentation · Status · Support
Frequently asked questions
- What is Deepseek: Deepseek R1 Distill Llama 70B?
- DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across multiple benchmarks, including: AIME 2024 pass@1: 70.0 MATH-500 pass@1: 94.5 CodeForces Rating: 1633 The model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.
- How much does Deepseek: Deepseek R1 Distill Llama 70B cost?
- Input costs start at $0.7 / 1M tokens and output costs start at $0.8 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of Deepseek: Deepseek R1 Distill Llama 70B?
- Deepseek: Deepseek R1 Distill Llama 70B supports a 131,072 token context window and up to 16,400 output tokens.
- What capabilities does Deepseek: Deepseek R1 Distill Llama 70B support?
- Deepseek: Deepseek R1 Distill Llama 70B supports Streaming, Function calling, JSON mode, Playground.
- Which providers offer Deepseek: Deepseek R1 Distill Llama 70B?
- Deepseek: Deepseek R1 Distill Llama 70B is available from DeepInfra.
- How do providers handle data privacy for Deepseek: Deepseek R1 Distill Llama 70B?
- 1 of 1 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.