Model dossier模型档案
Step 3.7 Flash
Step 3.7 Flash is a high-tier model from StepFun (Intelligence Score 78.8/100). Strongest in agent (99.4) and math (97.3); best fit: tool-use and multi-step agents. It trails in preference (24). Budget-priced: $0.200 in / $1.15 out per 1M tokens. Open weights — self-hostable. Reachable via CN-direct endpoints.
Step 3.7 Flash 是 StepFun 的高水平梯队模型(智能评分 78.8/100)。 最强项是智能体(99.4分),其次是数学(97.3分);适合工具调用与多步 Agent。 短板是偏好(24分)。 定价属低价档:每百万 token 输入 $0.200 / 输出 $1.15。 开放权重,可自托管。 支持国内直连。
open weights开放权重 ⌂ CN direct国内直连
78.8
Intelligence Score · benchmark coverage 智能评分 · benchmark 覆盖 73%
◈
Benchmark scoresBenchmark 成绩
8 entries 条| Benchmark基准 | Score成绩 | Index指数 | Dated日期 | Measured by测评方 | |
|---|---|---|---|---|---|
| Humanity's Last ExamReasoning | 49.5 | 86 | — | 3rd-party第三方 | source ↗×2 |
| GPQA DiamondReasoning | 76.6 | 80 | — | 3rd-party第三方 | source ↗×3⚠±4.3 |
| ↳ ⚡high | 67.1 | — | — | vendor-reported厂商自报 | source ↗ |
| SWE-bench VerifiedCoding | 76.3 | 79 | — | 3rd-party第三方 | source ↗×2 |
| AIME 2025Math | 97.3 | 97 | — | 3rd-party第三方 | source ↗ |
| LiveCodeBenchCoding | 69 | 74 | — | 3rd-party第三方 | source ↗×3⚠±17.4 |
| Terminal-BenchCoding | 59.5 | 65 | — | 3rd-party第三方 | source ↗ |
| τ²-benchAgent | 98.5 | 99 | — | 3rd-party第三方 | source ↗×3 |
| LMArena (Chatbot Arena)Preference | 1395 | 24 | — | official官方榜 | source ↗×2 |
≡
Specs & pricing规格与价格
| Provider厂商 | StepFun |
|---|---|
| Released发布 | 2026-05-28 |
| Context window上下文 | 262K tokens |
| Max output最大输出 | 230K tokens |
| Modality模态 | text+image+video->text |
| Reasoning tiers推理档位 | not documented未见公开文档 |
| API input price输入价格 | $0.200 / 1M tokens |
| API output price输出价格 | $1.15 / 1M tokens |
⛓
Tools using this model使用它的工具
How we score →评分方法 → · Catalog & pricing via the official OpenRouter API.模型目录与价格来自 OpenRouter 官方 API。