Model dossier模型档案
Grok 4.3
Grok 4.3 is a high-tier model from xAI (Intelligence Score 72.3/100). Strongest in agent (98.6) and reasoning (77.7); best fit: tool-use and multi-step agents. It trails in preference (35.8). Mid-priced: $1.25 in / $2.50 out per 1M tokens. Context window 1M tokens — long-document friendly.
Grok 4.3 是 xAI 的高水平梯队模型(智能评分 72.3/100)。 最强项是智能体(98.6分),其次是推理(77.7分);适合工具调用与多步 Agent。 短板是偏好(35.8分)。 定价属中档定价:每百万 token 输入 $1.25 / 输出 $2.50。 上下文 1M token,长文档友好。
72.3
Intelligence Score · benchmark coverage 智能评分 · benchmark 覆盖 80%
◈
Benchmark scoresBenchmark 成绩
9 entries 条| Benchmark基准 | Score成绩 | Index指数 | Dated日期 | Measured by测评方 | |
|---|---|---|---|---|---|
| Humanity's Last ExamReasoning | 35 | 61 | — | 3rd-party第三方 | source ↗×2 |
| GPQA DiamondReasoning | 90.1⚡high | 94 | — | 3rd-party第三方 | source ↗×4 |
| SWE-bench VerifiedCoding | 71.4 | 74 | — | 3rd-party第三方 | source ↗×2 |
| AIME 2025Math | 89 | 89 | — | 3rd-party第三方 | source ↗×3⚠±11 |
| LiveCodeBenchCoding | 84.49 | 90 | — | 3rd-party第三方 | source ↗×3⚠±22.19 |
| FrontierMathMath | 38 | 43 | — | 3rd-party第三方 | source ↗×2⚠±26 |
| Terminal-BenchCoding | 37.9 | 41 | — | 3rd-party第三方 | source ↗×2⚠±4.05 |
| τ²-benchAgent | 97.7 | 99 | — | 3rd-party第三方 | source ↗×3 |
| LMArena (Chatbot Arena)Preference | 1442 | 36 | 2026-08-12 | official官方榜 | source ↗×3⚠±95 |
≡
Specs & pricing规格与价格
| Provider厂商 | xAI |
|---|---|
| Released发布 | 2026-04-30 |
| Context window上下文 | 1M tokens |
| Max output最大输出 | 900K tokens |
| Modality模态 | text+image+file->text |
| Reasoning tiers推理档位 | ⚡low ⚡medium ⚡high default默认 high · low / medium / high (cannot disable) |
| API input price输入价格 | $1.25 / 1M tokens |
| API output price输出价格 | $2.50 / 1M tokens |
⛓
Tools using this model使用它的工具
How we score →评分方法 → · Catalog & pricing via the official OpenRouter API.模型目录与价格来自 OpenRouter 官方 API。