Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

Grok 4.20 (4.20 0309 Reasoning)

xAI · Status: Active · Released 2026-02-17 · Accessibility: API access

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Xentir Snapshot

Best forScores in the top 11% of models measured on Epoch Capabilities Index (ECI) (152.6).
Watch out forNot yet measured on 4 of the 9 benchmarks Xentir tracks — see the breakdown below.

Benchmark breakdown

BenchmarkScoreSource
Epoch Capabilities Index (ECI)152.6Epoch AI Benchmarking Hub
SWE-bench VerifiedNot measuredEpoch AI Benchmarking Hub
GPQA Diamond0.893Epoch AI Benchmarking Hub
MATH Level 5Not measuredEpoch AI Benchmarking Hub
OTIS Mock AIME 2024-20250.922Epoch AI Benchmarking Hub
FrontierMathNot measuredEpoch AI Benchmarking Hub
FrontierMath Tier 4Not measuredEpoch AI Benchmarking Hub
SimpleQA Verified0.351Epoch AI Benchmarking Hub
Chess Puzzles0.240Epoch AI Benchmarking Hub
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality.

Pricing

No licensed per-token pricing exists for this model yet. Not shown as an estimate — see Data Sources.

History

History begins 2026-07-26 — one data point so far. A trend needs at least two.

Compare

Latest Xentir coverage

No Xentir coverage of this model yet. See the latest AI news.

Sources and methodology

Scores on this page come from Epoch AI Benchmarking Hub, licensed CC BY 4.0. See the full methodology and data sources pages for how rankings, variants and ties are computed, and the full AI Model Rankings table for how Grok 4.20 (4.20 0309 Reasoning) compares to every other measured model.