The AI tools index that doesn't waste your time.不浪费你时间的 AI 工具索引。
Model dossier模型档案

Grok 4.20

Grok 4.20 is a high-tier model from xAI (Intelligence Score 85.1/100). Strongest in agent (98) and knowledge (93.2); best fit: tool-use and multi-step agents. Mid-priced: $1.25 in / $2.50 out per 1M tokens. Context window 2M tokens — long-document friendly.

Grok 4.20 是 xAI 的高水平梯队模型(智能评分 85.1/100)。 最强项是智能体(98分),其次是知识(93.2分);适合工具调用与多步 Agent。 定价属中档定价:每百万 token 输入 $1.25 / 输出 $2.50。 上下文 2M token,长文档友好。

85.1
Intelligence Score · benchmark coverage 智能评分 · benchmark 覆盖 95%
ReaCodKnoMatAgtPre

Benchmark scoresBenchmark 成绩

10 entries
Benchmark基准Score成绩Index指数Dated日期Measured by测评方
Humanity's Last ExamReasoning50.7883rd-party第三方source ↗×2⚠±18.5
GPQA DiamondReasoning91953rd-party第三方source ↗×4⚠±3.5
MMLU-ProKnowledge86.3933rd-party第三方source ↗×2⚠±8.7
SWE-bench VerifiedCoding76.7⚡high802026-04-113rd-party第三方source ↗×2
AIME 2025Math91.7923rd-party第三方source ↗×3
LiveCodeBenchCoding84.27903rd-party第三方source ↗
FrontierMathMath67.1713rd-party第三方source ↗
Terminal-BenchCoding47.1⚡high513rd-party第三方source ↗
τ²-benchAgent96.5⚡high982026-06-103rd-party第三方source ↗⚠±26.9
⚡off69.62026-06-103rd-party第三方source ↗
LMArena (Chatbot Arena)Preference1475632026-08-12official官方榜source ↗×3⚠±38

Specs & pricing规格与价格

Provider厂商xAI
Released发布2026-03-31
Context window上下文2M tokens
Max output最大输出1.8M tokens
Modality模态text+image+file->text
Reasoning tiers推理档位⚡low ⚡medium ⚡high default默认 high · low / medium / high (cannot disable)
API input price输入价格$1.25 / 1M tokens
API output price输出价格$2.50 / 1M tokens

Tools using this model使用它的工具

How we score →评分方法 → · Catalog & pricing via the official OpenRouter API.模型目录与价格来自 OpenRouter 官方 API。