Nvidia
NVIDIA: Nemotron Nano 12B 2 VL (free)
Nvidia
Input / M
Routed
Output / M
Context
Max Output
128K
Usage
Unranked
Quality
Unscored
Providers
1
Best Uptime 30m
71.63%
Latency (p50)
1059 ms
Throughput
19.0 tok/s
Released
October 28, 2025
Modalities
image, text, video
About
NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s.
Capabilities
Vision
Tools
Reasoning
Streaming
JSON Mode
Host offerings · 1
Per-host pricing, uptime, latency, throughput
| Host | Input/M | Output/M | Context | Uptime |
|---|---|---|---|---|
| NvidiaBest Nvidia | nvidia/nemotron-nano-12b-v2-vl:free | Free | Free | 128K | 71.6% |
Performance History
200 snapshots
Price History
Last 7 days · 0 changes
No creator-price changes in the last 7 days.
Free shows the last 7 days. Full history is a Pro feature.Get Pro
Activity Timeline
- No events yet.
Versus
Nearest peers by price, context, and capabilities. One click to compare.
No comparable models found.
Show this WM Score
Free live badge for your README, docs or site
Live badge — it always shows the current WM Score. Free to use anywhere, no key needed.
Markdown
[](https://whatmodel.app/models/68b391a5-6625-4364-b590-0f5803f8ffe7)HTML
<a href="https://whatmodel.app/models/68b391a5-6625-4364-b590-0f5803f8ffe7"><img src="https://whatmodel.app/api/public/embed/score/68b391a5-6625-4364-b590-0f5803f8ffe7" alt="NVIDIA: Nemotron Nano 12B 2 VL (free) WM Score" height="26" /></a>Image URL
https://whatmodel.app/api/public/embed/score/68b391a5-6625-4364-b590-0f5803f8ffe7First seen 43d ago · Last verified 9m ago
