The AI tools index that doesn't waste your time.不浪费你时间的 AI 工具索引。
Head-to-head正面对比

Grok 4.5 vs Gemini 3.1 Pro Preview

xAI meets Google: sourced benchmark scores side by side, dimension by dimension, with prices and specs. Every conclusion below is derived from the published dataset — no opinions, no sponsorship.

xAI 对 Google:带来源的 benchmark 成绩逐项并排,附价格与规格。下面每条结论都从公开数据集推导——没有主观意见,没有厂商赞助。

The verdict, derived from data结论(全部由数据推导)

ReaCodKnoMatAgtPre

Grok 4.5   Gemini 3.1 Pro Preview · missing dimensions count as 0 in the chart缺数据的维度在图上按 0 计

Spec by spec逐项规格

Grok 4.5Gemini 3.1 Pro Preview
Provider厂商xAIGoogle
Released发布2026-07-082026-02-19
Intelligence Score智能评分80.582.4
Benchmark coveragebenchmark 覆盖95%100%
Reasoning推理83.790.5
Coding代码88.685.6
Knowledge知识93.994.5
Math数学74.158.5
Agent智能体3885.7
Preference偏好42.546.8
API input $/1M输入价 $/百万$2.00$2.00
API output $/1M输出价 $/百万$6.00$12
Blended $/1M混合价 $/百万$4.00$7.00
Value (score per $)性价比(分数/美元)20.111.8
Context window上下文500K1M
Max output最大输出450K66K
Reasoning tiers推理档位low / medium / high (cannot disable)minimal / low / medium / high
Open weights开放权重NoNo
CN-direct国内直连NoNo
Free tier免费层NoNo

Benchmark by benchmark逐项 benchmark

Benchmark基准Grok 4.5Gemini 3.1 Pro Preview
Humanity's Last Exam40.346.44⚡high
GPQA Diamond92.995.5
MMLU-Pro89.289.8
SWE-bench Verified86.680.6
AIME 202591.7⚡high93.3
LiveCodeBench7991.7
FrontierMath4816.7
Terminal-Bench83.368.5
LMArena (Chatbot Arena)14691486
WebArena3569

Grok 4.5 dossier → 档案 → · Gemini 3.1 Pro Preview dossier → 档案 → · How we score评分方法