Back to overall leaderboardCAPABILITY PROFILE UNDERSTANDING THE RANK BEHIND THE SCORE
Claude Opus 5
Anthropic · 2026-07-24 released · 10/01 14:05 updated
Overall consensus index93.9Overall rank #8
Category scores4/ 4 categories
Recorded scores20evaluations
Context window—Token
API input / output · per million tokens¥33.52 / ¥167.61Cached ¥3.35List $5 / $25Official vendor price
See where each model excels.
Each capability is scored separately; insufficient evidence is left blank.
Scores reflect ranking support within each reference group and cannot be added or compared as absolute capability levels.
How stable is the overall rank?
Remove an evaluation or organization, adjust weights and error handling, then observe the rank.
Rank after eligibility checks6-8 High confidence
All completed comparisons retained eligibility. With the original candidate set it is #7-8. This is not a confidence interval and excludes unpublished scores.
Every score has a source.
Publicly reported scores for this model. Expand a row to see configuration and usage.
综合评测与体验7 items
编程3 items
推理5 items
知识2 items
专业办公2 items
视觉理解1 items
Some things remain unknown
Missing evaluations are not counted as zero. Ranks can change with new evidence; close scores should not be overinterpreted.
See the methodology →