The AI tools index that doesn't waste your time.不浪费你时间的 AI 工具索引。
Head-to-head正面对比

Claude Opus 5 vs Claude Opus 4.8

Anthropic meets Anthropic: sourced benchmark scores side by side, dimension by dimension, with prices and specs. Every conclusion below is derived from the published dataset — no opinions, no sponsorship.

Anthropic 对 Anthropic:带来源的 benchmark 成绩逐项并排,附价格与规格。下面每条结论都从公开数据集推导——没有主观意见,没有厂商赞助。

The verdict, derived from data结论(全部由数据推导)

ReaCodKnoMatAgtPre

Claude Opus 5   Claude Opus 4.8 · missing dimensions count as 0 in the chart缺数据的维度在图上按 0 计

Spec by spec逐项规格

Claude Opus 5Claude Opus 4.8
Provider厂商AnthropicAnthropic
Released发布2026-07-242026-05-27
Intelligence Score智能评分8884.4
Benchmark coveragebenchmark 覆盖100%85%
Reasoning推理96.489.4
Coding代码85.990
Knowledge知识96.4
Math数学68.177.2
Agent智能体78.686.2
Preference偏好10043.8
API input $/1M输入价 $/百万$5.00$5.00
API output $/1M输出价 $/百万$25$25
Blended $/1M混合价 $/百万$15$15
Value (score per $)性价比(分数/美元)5.95.6
Context window上下文1M1M
Max output最大输出128K128K
Reasoning tiers推理档位low / medium / high / x-highlow / medium / high / x-high
Open weights开放权重NoNo
CN-direct国内直连NoNo
Free tier免费层NoNo

Benchmark by benchmark逐项 benchmark

Benchmark基准Claude Opus 5Claude Opus 4.8
Humanity's Last Exam54.4⚡max46⚡max
GPQA Diamond93.794.3
SWE-bench Verified9688.6
AIME 20259798.3
LiveCodeBench89.0387.82
FrontierMath31.2547.24
Terminal-Bench43.5⚡max74.6
τ²-bench82.7894.4
LMArena (Chatbot Arena)1699⚡max1474⚡high
WebArena6871.2

Claude Opus 5 dossier → 档案 → · Claude Opus 4.8 dossier → 档案 → · How we score评分方法