← Back to Directory

NVIDIA: Nemotron 3 Nano 30B A3B

NVIDIA Developer Architecture Profile

Intelligence (ELO)1416Chatbot Arena Verified
Max Context Limit262,144Max Out: 235,929 tokens
Prompt Caching40% OFF$0.03 / 1M cached
Standard API / 1M$0.25Blended Standard Rate

Model Capabilities & Contract SLAs

  • Coding & Logic
  • Fictional
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

Granular Pricing Matrix

Input Tokens (Standard Prompt)$0.05 / 1M
⚡ Cached Input (Prompt Cache Read)40% OFF$0.03 / 1M
Output Tokens (Completion)$0.20 / 1M

Pricing data via OpenRouter. Sync: 10/2/2026

Evaluate Competitors