Skip to content
Back to overall leaderboard

Claude Opus 5

Anthropic · 2026-07-24 released · 10/01 14:05 updated

Overall consensus index93.9Overall rank #8
Category scores4/ 4 categories
Recorded scores20evaluations
Context window—Token
API input / output · per million tokens¥33.52 / ¥167.61Cached ¥3.35List $5 / $25Official vendor price
CAPABILITY PROFILE

See where each model excels.

Each capability is scored separately; insufficient evidence is left blank.

Scores reflect ranking support within each reference group and cannot be added or compared as absolute capability levels.

UNDERSTANDING THE RANK

How stable is the overall rank?

Remove an evaluation or organization, adjust weights and error handling, then observe the rank.

Rank after eligibility checks6-8 High confidence

All completed comparisons retained eligibility. With the original candidate set it is #7-8. This is not a confidence interval and excludes unpublished scores.

BEHIND THE SCORE

Every score has a source.

Publicly reported scores for this model. Expand a row to see configuration and usage.

综合评测与体验7 items

编程3 items

推理5 items

知识2 items

专业办公2 items

视觉理解1 items

Some things remain unknown

Missing evaluations are not counted as zero. Ranks can change with new evidence; close scores should not be overinterpreted.

See the methodology →