NVIDIA: Nemotron 3 Nano 30B A3B

nvidia/nemotron-3-nano-30b-a3b

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with open-weights, datasets and recipes so developers can easily customize, optimize, and deploy the model on their infrastructure for maximum privacy and security. Note: For the free endpoint, all prompts and output are logged to improve the provider's model and its product and services. Please do not upload any personal, confidential, or otherwise sensitive information. This is a trial use only. Do not use for production or business-critical systems.

Model specifications

Input
text
Output
text
Context
262,144 tokens
Max output
262,144 tokens
Input price
$0.05 / 1M tokens
Output price
$0.2 / 1M tokens
Released
2026-02-11

Capabilities

  • Streaming
  • Function calling
  • JSON mode
  • Playground

Provider pricing, discounts and data privacy

Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.

Standard service tier

1 available provider · tier input average $0.05 / 1M tokens · tier output average $0.2 / 1M tokens

DeepInfra

Tier: Standard · Region: US · Quantization: fp4

Pricing
Input
$0.05 / 1M tokens
Output
$0.2 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
Yes
Data retention
Zero retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
Yes
HIPAA compliant
No
SOC 2 certified
Yes
BYOK supported
Yes

Frequently asked questions

What is NVIDIA: Nemotron 3 Nano 30B A3B?
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with open-weights, datasets and recipes so developers can easily customize, optimize, and deploy the model on their infrastructure for maximum privacy and security. Note: For the free endpoint, all prompts and output are logged to improve the provider's model and its product and services. Please do not upload any personal, confidential, or otherwise sensitive information. This is a trial use only. Do not use for production or business-critical systems.
How much does NVIDIA: Nemotron 3 Nano 30B A3B cost?
Input costs start at $0.05 / 1M tokens and output costs start at $0.2 / 1M tokens. Provider-level prices vary by service tier.
What is the context length of NVIDIA: Nemotron 3 Nano 30B A3B?
NVIDIA: Nemotron 3 Nano 30B A3B supports a 262,144 token context window and up to 262,144 output tokens.
What capabilities does NVIDIA: Nemotron 3 Nano 30B A3B support?
NVIDIA: Nemotron 3 Nano 30B A3B supports Streaming, Function calling, JSON mode, Playground.
Which providers offer NVIDIA: Nemotron 3 Nano 30B A3B?
NVIDIA: Nemotron 3 Nano 30B A3B is available from DeepInfra.
How do providers handle data privacy for NVIDIA: Nemotron 3 Nano 30B A3B?
1 of 1 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.

Browse all AI models