LLM Leaderboard
NVIDIA logo

Nemotron 3 Ultra

by NVIDIA Reasoning

RunFree Score
insufficient data
Blended price
$1.35
per 1M tokens
Context
512K
tokens
Max output
tokens
Output speed
tokens / sec
Latency (TTFT)
first token
Uptime
99.4%
last 24h
Providers
4
serving this model

Capability profile

Not enough categorised benchmark data to chart a profile yet.

Compare Nemotron 3 Ultra

See it side by side with any other model — benchmarks, price and specs.

Benchmark results

We don't have verified public benchmark results for Nemotron 3 Ultra yet. We only publish scores with a primary source — no estimates.

Every score links to its original source. See our methodology for how the RunFree Score is computed.

API pricing

Input
$0.60 / 1M
Output
$3.60 / 1M
Cached input
$0.20 / 1M
Blended (3:1)
$1.35 / 1M

Specifications

Provider
NVIDIA
Context window
512K
Max output
Reasoning model
Yes
Modality
text
Released
Jun 4, 2026

Related models

← Back to the full LLM Leaderboard