OpenRating is an independent, open-source AI rating platform. Join the community

OOpenRating
All models

Anthropic · Released May 15, 2026

Claude Opus 5 (max)

Anthropic's flagship frontier model with industry-leading autonomous agentic capabilities, complex systems architecture reasoning, and deep analytical synthesis.

Proprietary

OpenRating Index

63.4

0–100 composite

Reasoning Index

89.4

0–100 composite

Cost / Task

$2.34

benchmark run

Output speed

59 tok/s

blended tok/s

Time to first token

1.1s

median latency

Input price

$4.8

per 1M tokens

Output price

$24

per 1M tokens

Context window

200K

tokens

Strengths

  • Top-ranked OpenRating Index for multi-step reasoning

  • Superior long-context comprehension with minimal hallucination

  • Native tool orchestration and agentic workflow execution

How we rate Claude Opus 5 (max)

OpenRating scores every model on the same axes. The OpenRating Index is a weighted composite of reasoning, knowledge, and agentic benchmarks, normalised to a 0–100 scale. Speed is blended output measured in tokens per second under standardised load, with time-to-first-token reported separately. Prices are list input and output prices per million tokens, and Task Cost measures end-to-end workload spending.

OpenRating Index

63.4

Reasoning Index

89.4

Token Pricing

Input per 1M tokens

$4.8

Output per 1M tokens

$24

Cost of 1M input + 1M output

$29

Task Cost Economics

Empirical cost breakdown for a standard OpenRating Index benchmark task ($2.34 total).

Reasoning (Chain of Thought)
$1.14
Answer generation
$0.18
Cache write overhead
$0.42
Cache hit queries
$0.55
Input payload
$0.042