Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

Claude Opus 4.6 vs GPT-5.6 Luna

Claude Opus 4.6 · Anthropic, released 2026-02-05
GPT-5.6 Luna · OpenAI, released 2026-07-09

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Head-to-head

Measured on 5 of 9 tracked benchmarks for both models — Claude Opus 4.6 leads on 2, GPT-5.6 Luna leads on 3.

BenchmarkClaude Opus 4.6 GPT-5.6 LunaLeads
Epoch Capabilities Index (ECI)155.4155.3Claude Opus 4.6
SWE-bench Verified0.772Not measuredNot measured on both
GPQA Diamond0.905 (32k)0.916 (max)GPT-5.6 Luna
MATH Level 5Not measuredNot measuredNot measured on both
OTIS Mock AIME 2024-20250.944 (64k)0.983 (max)GPT-5.6 Luna
FrontierMath0.407 (max)Not measuredNot measured on both
FrontierMath Tier 40.229 (max)Not measuredNot measured on both
SimpleQA Verified0.465 (32k)0.417 (max)Claude Opus 4.6
Chess Puzzles0.170 (32k)0.400 (max)GPT-5.6 Luna
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality. A model with no row here for a benchmark has not been measured on it — that is left blank, never estimated.

Full profiles

Claude Opus 4.6 → · GPT-5.6 Luna → · the full AI Model Rankings table