The AI tools index that doesn't waste your time.不浪费你时间的 AI 工具索引。
Head-to-head正面对比

Claude Opus 5 vs GPT-5.6 Sol

Anthropic meets OpenAI: sourced benchmark scores side by side, dimension by dimension, with prices and specs. Every conclusion below is derived from the published dataset — no opinions, no sponsorship.

Anthropic 对 OpenAI:带来源的 benchmark 成绩逐项并排,附价格与规格。下面每条结论都从公开数据集推导——没有主观意见,没有厂商赞助。

The verdict, derived from data结论(全部由数据推导)

ReaCodKnoMatAgtPre

Claude Opus 5   GPT-5.6 Sol · missing dimensions count as 0 in the chart缺数据的维度在图上按 0 计

Spec by spec逐项规格

Claude Opus 5GPT-5.6 Sol
Provider厂商AnthropicOpenAI
Released发布2026-07-242026-07-09
Intelligence Score智能评分8893.4
Benchmark coveragebenchmark 覆盖100%100%
Reasoning推理96.490.4
Coding代码85.996.3
Knowledge知识96.493.8
Math数学68.198.4
Agent智能体78.692.9
Preference偏好10081
API input $/1M输入价 $/百万$5.00$2.00
API output $/1M输出价 $/百万$25$10
Blended $/1M混合价 $/百万$15$6.00
Value (score per $)性价比(分数/美元)5.915.6
Context window上下文1M1.1M
Max output最大输出128K128K
Reasoning tiers推理档位low / medium / high / x-highnone / minimal / low / medium / high / xhigh
Open weights开放权重NoNo
CN-direct国内直连NoNo
Free tier免费层NoNo

Benchmark by benchmark逐项 benchmark

Benchmark基准Claude Opus 5GPT-5.6 Sol
Humanity's Last Exam54.4⚡max47.2⚡high
GPQA Diamond93.794.1⚡max
MMLU-Pro91.5989.1
SWE-bench Verified9696.2⚡max
AIME 20259797
LiveCodeBench89.0381
FrontierMath31.2589
Terminal-Bench43.5⚡max91.9
τ²-bench82.7885.1
LMArena (Chatbot Arena)1699⚡max1623⚡max
WebArena6892.2

Claude Opus 5 dossier → 档案 → · GPT-5.6 Sol dossier → 档案 → · How we score评分方法