Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

GPT-5.1

OpenAI · Status: Active · Released 2025-11-13 · Accessibility: API access

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Xentir Snapshot

Best forScores in the top 15% of models measured on Epoch Capabilities Index (ECI) (149.7).
Watch out forNot yet measured on 1 of the 9 benchmarks Xentir tracks — see the breakdown below.

Benchmark breakdown

BenchmarkScoreSource
Epoch Capabilities Index (ECI)149.7Epoch AI Benchmarking Hub
SWE-bench Verified0.669 (high)Epoch AI Benchmarking Hub
GPQA Diamond0.876 (high)Epoch AI Benchmarking Hub
MATH Level 5Not measuredEpoch AI Benchmarking Hub
OTIS Mock AIME 2024-20250.886 (high)Epoch AI Benchmarking Hub
FrontierMath0.310 (high)Epoch AI Benchmarking Hub
FrontierMath Tier 40.125 (high)Epoch AI Benchmarking Hub
SimpleQA Verified0.489 (high)Epoch AI Benchmarking Hub
Chess Puzzles0.320 (high)Epoch AI Benchmarking Hub
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality.

Pricing

No licensed per-token pricing exists for this model yet. Not shown as an estimate — see Data Sources.

History

History begins 2026-07-26 — one data point so far. A trend needs at least two.

Compare

Latest Xentir coverage

Sources and methodology

Scores on this page come from Epoch AI Benchmarking Hub, licensed CC BY 4.0. See the full methodology and data sources pages for how rankings, variants and ties are computed, and the full AI Model Rankings table for how GPT-5.1 compares to every other measured model.