Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

Claude 3.5 Sonnet (Jun 2024)

Anthropic · Status: Active · Released 2024-06-20 · Accessibility: API access

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Xentir Snapshot

Watch out forNot yet measured on 3 of the 9 benchmarks Xentir tracks — see the breakdown below.

Benchmark breakdown

BenchmarkScoreSource
Epoch Capabilities Index (ECI)130.0Epoch AI Benchmarking Hub
SWE-bench VerifiedNot measuredEpoch AI Benchmarking Hub
GPQA Diamond0.540Epoch AI Benchmarking Hub
MATH Level 50.517Epoch AI Benchmarking Hub
OTIS Mock AIME 2024-20250.065Epoch AI Benchmarking Hub
FrontierMath0.010Epoch AI Benchmarking Hub
FrontierMath Tier 40.000Epoch AI Benchmarking Hub
SimpleQA VerifiedNot measuredEpoch AI Benchmarking Hub
Chess PuzzlesNot measuredEpoch AI Benchmarking Hub
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality.

Pricing

No licensed per-token pricing exists for this model yet. Not shown as an estimate — see Data Sources.

History

History begins 2026-07-26 — one data point so far. A trend needs at least two.

Compare

Latest Xentir coverage

Sources and methodology

Scores on this page come from Epoch AI Benchmarking Hub, licensed CC BY 4.0. See the full methodology and data sources pages for how rankings, variants and ties are computed, and the full AI Model Rankings table for how Claude 3.5 Sonnet (Jun 2024) compares to every other measured model.