Skip to content
AI Primer

NVIDIA Nemotron 3 Nano Omni

Audio, Video, Image, Text → Text

A multimodal NVIDIA Nemotron 3 Nano release described as a 30B-parameter, 3B-active model for text, image, audio, and video understanding with text output and a 256K-token context window.

Pricing

Model profile · Current snapshot
Input / 1M
$0.09
Output / 1M
$0.36
Blended / 1M
$0.158
Output TPS
0
TTFT (s)
0

Model Intelligence

Context window
262,144 tokens
Arena ranking
10
Benchmarkable
Yes
Model level
release
Intelligence Index
10.3
Coding Index
13.8
GPQA
0.47
HLE
0.05
IFBench
0.63
LCR
0.4
TerminalBench Hard
0.08
TAU2
0.45

Recent stories

1 linked story
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.