NVIDIA model benchmarks.

Compare NVIDIA models across answer quality, coding, complex tasks, speed, price, and context size.

NVIDIA

Quality

Highest overall scores

  1. 1NVIDIA: Nemotron 3 Ultra38.3
  2. 2NVIDIA: Nemotron 3.5 Lightning23.6
  3. 3NVIDIA: Nemotron 3 Nano 30B A3B14.5

Response speed

Fastest output in this sample

  1. 1NVIDIA: Nemotron 3 Nano 30B A3B135 t/s
  2. 2NVIDIA: Nemotron 3.5 Lightning132 t/s
  3. 3NVIDIA: Nemotron 3 Ultra121 t/s

Input price

Lowest price per 1M tokens

  1. 1NVIDIA: Nemotron 3 Nano 30B A3B$0.05/1M
  2. 2NVIDIA: Nemotron 3.5 Lightning$0.10/1M
  3. 3NVIDIA: Nemotron 3 Ultra$0.60/1M

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better

Quality benchmarks

Compare NVIDIA model capability.

Performance and price

Compare practical tradeoffs.

All NVIDIA models.

3 models in the current catalogue.

NVIDIA technical benchmarks.

Compare named evaluation records available for NVIDIA models. Missing results remain blank rather than being estimated.

BenchmarkAreaNemotron 3 UltraNemotron 3.5 LightningNemotron 3 Nano 30B A3BSource
Intelligence IndexoverallQuality38.323.614.5OpenRouter / artificial-analysis
Coding Indexcoding49.326.814.4OpenRouter / artificial-analysis
Agentic Indexagentic27.513.82OpenRouter / artificial-analysis
GPQAGPQA Diamond86.774.375.7OpenRouter model benchmarks
Humanity's Last ExamHLE28.410.611.4OpenRouter model benchmarks
IFBenchIFBench81.4—71.1OpenRouter model benchmarks
τ²-Bench Telecomτ²-Bench Telecom83.3—40.9OpenRouter model benchmarks
AA-LCRAA-LCR7155.337.3OpenRouter model benchmarks
GDPval-AAGDPval-AA33.216.20OpenRouter model benchmarks
CritPtCritPt3.100.9OpenRouter model benchmarks
SciCodeSciCode39.931.629.6OpenRouter model benchmarks
Terminal-Bench HardTerminal-Bench Hard36.4—13.6OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy22.614.417.3OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate70.362.416.7OpenRouter model benchmarks
MMLU-ProKnowledge and reasoning——78.3OpenEvals/leaderboard-data