Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Back to explorer
Nvidia

NVIDIA: Nemotron 3 Ultra (free)

Nvidia

Input / M
Routed
Output / M
Context
Max Output
66K
Usage
Unranked
Quality
Unscored
Providers
1
Best Uptime 30m
99.89%
Latency (p50)
1686 ms
Throughput
30.0 tok/s
Released
June 4, 2026
Modalities
text

About

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it.

Capabilities

Vision
Tools
Reasoning
Streaming
JSON Mode

Host offerings · 1

Per-host pricing, uptime, latency, throughput
HostInput/MOutput/MContextUptime
NvidiaBest
Nvidia | nvidia/nemotron-3-ultra-550b-a55b-20260604:free
FreeFree1.0M99.9%

Performance History

200 snapshots

Price History

Last 7 days · 0 changes
No creator-price changes in the last 7 days.
Free shows the last 7 days. Full history is a Pro feature.Get Pro

Activity Timeline

  • No events yet.

Versus

Nearest peers by price, context, and capabilities. One click to compare.
No comparable models found.

Show this WM Score

Free live badge for your README, docs or site
NVIDIA: Nemotron 3 Ultra (free) WM Score badge

Live badge — it always shows the current WM Score. Free to use anywhere, no key needed.

Markdown
[![NVIDIA: Nemotron 3 Ultra (free) WM Score](https://whatmodel.app/api/public/embed/score/0c3c870a-ea86-4c66-8d15-0215cf000747)](https://whatmodel.app/models/0c3c870a-ea86-4c66-8d15-0215cf000747)
HTML
<a href="https://whatmodel.app/models/0c3c870a-ea86-4c66-8d15-0215cf000747"><img src="https://whatmodel.app/api/public/embed/score/0c3c870a-ea86-4c66-8d15-0215cf000747" alt="NVIDIA: Nemotron 3 Ultra (free) WM Score" height="26" /></a>
Image URL
https://whatmodel.app/api/public/embed/score/0c3c870a-ea86-4c66-8d15-0215cf000747
First seen 43d ago · Last verified 9m ago