Nemotron 3 Nano 30B

NVIDIA's compact Nemotron 3 MoE model with 30B total and 3B active parameters, offering toggleable reasoning and long context at low cost.

nemotron-3-nano-30b
STABLEGet StartedView uptime
262,144 context
Released December 15, 2025
Starting at $0.06/M input tokens
Starting at $0.24/M output tokens
Streaming
Tools
Reasoning
No ratings yetSign in to rate

Select Provider

All Providers for Nemotron 3 Nano 30B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Nebius AI
Context: 262.1kQuant: fp8
Input
$0.06
/M tokens
Cache Read
/M tokens
Output
$0.24
/M tokens
Get Started

Frequently asked questions

What is Nemotron 3 Nano 30B?

NVIDIA's compact Nemotron 3 MoE model with 30B total and 3B active parameters, offering toggleable reasoning and long context at low cost. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.

How much does Nemotron 3 Nano 30B cost?

Pricing for Nemotron 3 Nano 30B on LLM Gateway starts at $0.06 per million input tokens and $0.24 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.

What is the context length of Nemotron 3 Nano 30B?

Nemotron 3 Nano 30B supports a context window of up to 262,144 tokens on its largest provider deployment.

Which providers serve Nemotron 3 Nano 30B?

Nemotron 3 Nano 30B is served by Nebius AI through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.

Does Nemotron 3 Nano 30B support tool calling and structured outputs?

Nemotron 3 Nano 30B supports tool (function) calling, but not structured JSON output mode.

When was Nemotron 3 Nano 30B released?

Nemotron 3 Nano 30B was released on December 15, 2025.