Claude Opus 4.5 · Anthropic, released 2025-11-24
Qwen 3.6 Max (Preview) · Alibaba, released 2026-04-20
“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources
Measured on 8 of 9 tracked benchmarks for both models — Qwen 3.6 Max (Preview) leads on 6, tied on 2.
| Benchmark | Claude Opus 4.5 | Qwen 3.6 Max (Preview) | Leads |
|---|---|---|---|
| Epoch Capabilities Index (ECI) | 149.6 | 149.7 | Qwen 3.6 Max (Preview) |
| SWE-bench Verified | 0.767 | 0.767 | Tied |
| GPQA Diamond | 0.860 (32k) | 0.891 | Qwen 3.6 Max (Preview) |
| MATH Level 5 | Not measured | Not measured | Not measured on both |
| OTIS Mock AIME 2024-2025 | 0.861 (32k) | 0.911 | Qwen 3.6 Max (Preview) |
| FrontierMath | 0.207 | 0.231 | Qwen 3.6 Max (Preview) |
| FrontierMath Tier 4 | 0.042 | 0.042 | Tied |
| SimpleQA Verified | 0.418 (32k) | 0.569 | Qwen 3.6 Max (Preview) |
| Chess Puzzles | 0.120 (32k) | 0.170 | Qwen 3.6 Max (Preview) |
Claude Opus 4.5 → · Qwen 3.6 Max (Preview) → · the full AI Model Rankings table