Skip to content

Nemotron 3 Nano 30B A3B

NVIDIA · Open weights · 262K context · released Dec 14, 2025 · checked 2h ago

Ways to buy 4

Way to buyIn / out per 1M
Crusoe
fp8 · cached $0.03
$0.05 / $0.2Cheapest
DeepInfra
fp4 · cached $0.025
$0.05 / $0.2Cheapest
Novita AI
fp4
$0.05 / $0.2Cheapest
Nebius
fp8
$0.06 / $0.24+20% vs cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $9.00 direct.

A dash: not on that gateway's own model list. No published token fee: Martian Gateway.

Related models

In / out per 1M
Nemotron 3.5 Lightning
NVIDIA · 1M
$0.065 / $0.18
Nemotron 3 Ultra
NVIDIA · 262K
$0.5 / $2.2
Nemotron 3.5 Content Safety
NVIDIA · 131K
$0.2 / $0.2
Nemotron 3 Super
NVIDIA · 262K
$0.085 / $0.4
Muse Spark 1.3 Contributor
Meta · 1M
$0.1 / $0.2
GLM 5.3 Flash
Z.ai · 1.31M
$0.045 / $0.14
Muse Spark 1.2 Contributor
Meta · 1M
$0.1 / $0.2
DeepSeek V4 Flash 0731
DeepSeek · 1.31M
$0.044 / $0.132

About Nemotron 3 Nano 30B A3B

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems.

Takes text in and returns text. First listed Sep 26, 2026. Up to 236K output tokens. Blended price $0.088 per 1M tokens at 3 input to 1 output. Model id nemotron-3-nano-30b-a3b.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Across fru.devCompany profile

Questions

How much does Nemotron 3 Nano 30B A3B cost?

$0.05 per million input tokens and $0.2 per million output tokens, with cached input at $0.03.

What is the context window of Nemotron 3 Nano 30B A3B?

262K tokens, with up to 236K tokens of output.

Where is Nemotron 3 Nano 30B A3B cheapest?

Of 4 routes tracked, Crusoe is cheapest at $0.05 input and $0.2 output per million tokens.

Price changes by email

Thursdays, only in weeks when an LLM API list price changed.

Double opt-in. Unsubscribe any time.