NVIDIA: Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with open-weights, datasets and recipes so developers can easily customize, optimize, and deploy the model on their infrastructure for maximum privacy and security. Note: For the free endpoint, all prompts and output are logged to improve the provider's model and its product and services. Please do not upload any personal, confidential, or otherwise sensitive information. This is a trial use only. Do not use for production or business-critical systems.
Model specifications
- Input
- text
- Output
- text
- Context
- 262,144 tokens
- Max output
- 262,144 tokens
- Input price
- $0.05 / 1M tokens
- Output price
- $0.2 / 1M tokens
- Released
- 2026-02-11
Capabilities
- Streaming
- Function calling
- JSON mode
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
1 available provider · tier input average $0.05 / 1M tokens · tier output average $0.2 / 1M tokens
DeepInfra
Tier: Standard · Region: US · Quantization: fp4
Pricing
- Input
- $0.05 / 1M tokens
- Output
- $0.2 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- Yes
- Data retention
- Zero retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- Yes
- HIPAA compliant
- No
- SOC 2 certified
- Yes
- BYOK supported
- Yes
Privacy policy · Terms · Official website · Documentation · Status · Support
Frequently asked questions
- What is NVIDIA: Nemotron 3 Nano 30B A3B?
- NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with open-weights, datasets and recipes so developers can easily customize, optimize, and deploy the model on their infrastructure for maximum privacy and security. Note: For the free endpoint, all prompts and output are logged to improve the provider's model and its product and services. Please do not upload any personal, confidential, or otherwise sensitive information. This is a trial use only. Do not use for production or business-critical systems.
- How much does NVIDIA: Nemotron 3 Nano 30B A3B cost?
- Input costs start at $0.05 / 1M tokens and output costs start at $0.2 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of NVIDIA: Nemotron 3 Nano 30B A3B?
- NVIDIA: Nemotron 3 Nano 30B A3B supports a 262,144 token context window and up to 262,144 output tokens.
- What capabilities does NVIDIA: Nemotron 3 Nano 30B A3B support?
- NVIDIA: Nemotron 3 Nano 30B A3B supports Streaming, Function calling, JSON mode, Playground.
- Which providers offer NVIDIA: Nemotron 3 Nano 30B A3B?
- NVIDIA: Nemotron 3 Nano 30B A3B is available from DeepInfra.
- How do providers handle data privacy for NVIDIA: Nemotron 3 Nano 30B A3B?
- 1 of 1 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.