Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

o3-mini

OpenAI · Status: Active · Released 2025-01-31 · Accessibility: API access

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Xentir Snapshot

Best forScores in the top 9% of models measured on MATH Level 5 (0.965 (high)).
Watch out forNot yet measured on 2 of the 9 benchmarks Xentir tracks — see the breakdown below.

Benchmark breakdown

BenchmarkScoreSource
Epoch Capabilities Index (ECI)141.4Epoch AI Benchmarking Hub
SWE-bench VerifiedNot measuredEpoch AI Benchmarking Hub
GPQA Diamond0.770 (high)Epoch AI Benchmarking Hub
MATH Level 50.965 (high)Epoch AI Benchmarking Hub
OTIS Mock AIME 2024-20250.769 (high)Epoch AI Benchmarking Hub
FrontierMath0.124 (high)Epoch AI Benchmarking Hub
FrontierMath Tier 40.042 (high)Epoch AI Benchmarking Hub
SimpleQA VerifiedNot measuredEpoch AI Benchmarking Hub
Chess Puzzles0.170 (high)Epoch AI Benchmarking Hub
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality.

Pricing

No licensed per-token pricing exists for this model yet. Not shown as an estimate — see Data Sources.

History

History begins 2026-07-26 — one data point so far. A trend needs at least two.

Compare

Latest Xentir coverage

Sources and methodology

Scores on this page come from Epoch AI Benchmarking Hub, licensed CC BY 4.0. See the full methodology and data sources pages for how rankings, variants and ties are computed, and the full AI Model Rankings table for how o3-mini compares to every other measured model.