Nvidia

NVIDIA Nemotron 3 Nano 30B

nvidia
Chat

NVIDIA Nemotron 3 Nano 30B is a compact and efficient language model from NVIDIA, optimized for fast inference while maintaining strong performance across diverse tasks.

Modalities

In / out price

$0.075 / $0.30

per 1M

Context

128K

Max output

16K

Released

Jan 27, 2026

Updated

Sep 7, 2026

Providers

1 route

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVenice
Private$0.075$0.30128K16K

Privacy

2 / 4

Private routing

Private tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed. Learn more

Features

4
Streaming
Tool calling
JSON/schema
Web search

Uptime

NVIDIA Nemotron 3 Nano 30B — AnonRouter Models