Anthropic: Claude Opus 4.8.

See how Anthropic: Claude Opus 4.8 compares with other broadly useful AI models across quality, speed, price, context, and practical capabilities.

CompareVisit model

Quality

#6/76

57.3

Overall quality score

Speed

#37/76

70 t/s

Output tokens per second

Input price

#72/76

$5.00/1M

Per 1M tokens

Output price

#71/76

$25.00/1M

Per 1M tokens

Context

#12/76

1M

Content read at once

Comparison summary

Anthropic: Claude Opus 4.8 ranks #6 of 76 for overall quality and #37 for response speed in the current catalogue.

Its listed input price is $5.00/1M, output price is $25.00/1M, and it can work with up to 1M of context at once.

Its strongest areas in the current data are coding, math and logic, following instructions.

Practical specifications

Company
Anthropic
Reasoning
Yes
Open weights
No
Input
Text, Image
Output
Text
Context window
1M

Anthropic: Claude Opus 4.8 technical benchmarks.

Named evaluation records for technical comparison. Results retain their original benchmark names and sources.

BenchmarkAreaClaude Opus 4.8Source
Intelligence IndexoverallQuality57.3OpenRouter / artificial-analysis
Coding Indexcoding74.3OpenRouter / artificial-analysis
Agentic Indexagentic49.4OpenRouter / artificial-analysis
GPQAGPQA Diamond92OpenRouter model benchmarks
Humanity's Last ExamHLE48.7OpenRouter model benchmarks
IFBenchIFBench62.2OpenRouter model benchmarks
τ²-Bench Telecomτ²-Bench Telecom94.4OpenRouter model benchmarks
AA-LCRAA-LCR73OpenRouter model benchmarks
GDPval-AAGDPval-AA54.3OpenRouter model benchmarks
CritPtCritPt20.9OpenRouter model benchmarks
SciCodeSciCode53.5OpenRouter model benchmarks
Terminal-Bench HardTerminal-Bench Hard58.3OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy48.8OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate60.7OpenRouter model benchmarks

Quality benchmarks

See where Anthropic: Claude Opus 4.8 stands.

The darker bar marks Anthropic: Claude Opus 4.8; the lighter bars provide context from other leading models.

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better