Gemini 2.0 Flash Thinking (Jan 2025) vs o1-preview
Gemini 2.0 Flash Thinking (Jan 2025) · Google DeepMind,Google, released 2025-01-21
o1-preview · OpenAI, released 2024-09-12
“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-09-09 · Methodology · Data Sources
Gemini 2.0 Flash Thinking (Jan 2025) leads on more of the 3 benchmarks where both are measured — Gemini 2.0 Flash Thinking (Jan 2025) on 3. The remaining 6 of 9 tracked benchmarks cannot be compared: 1 has a result for only one of the two models, 5 for neither. Nothing on this page is estimated to fill those in.
Head-to-head
Measured on 3 of 9 tracked benchmarks for both models — Gemini 2.0 Flash Thinking (Jan 2025) leads on 3.
| Benchmark | Gemini 2.0 Flash Thinking (Jan 2025) | o1-preview | Leads |
|---|---|---|---|
| Epoch Capabilities Index (ECI) | 135.4 | 134.8 | Gemini 2.0 Flash Thinking (Jan 2025) |
| SWE-bench Verified | Not measured | Not measured | Not measured on both |
| GPQA Diamond | 0.571 | 0.503 | Gemini 2.0 Flash Thinking (Jan 2025) |
| MATH Level 5 | Not measured | 0.816 | Comparison unavailable — Gemini 2.0 Flash Thinking (Jan 2025) has no published result |
| OTIS Mock AIME 2024-2025 | 0.578 | 0.311 | Gemini 2.0 Flash Thinking (Jan 2025) |
| FrontierMath | Not measured | Not measured | Not measured on both |
| FrontierMath Tier 4 | Not measured | Not measured | Not measured on both |
| SimpleQA Verified | Not measured | Not measured | Not measured on both |
| Chess Puzzles | Not measured | Not measured | Not measured on both |
What remains uncertain
- Gemini 2.0 Flash Thinking (Jan 2025) has no published result on MATH Level 5.
- 5 tracked benchmarks have no published result for either model.
A missing result means Xentir has not found one published by the source it cites, not that the model performs badly. See Data Sources.
Full profiles
Gemini 2.0 Flash Thinking (Jan 2025) → · o1-preview → · the full AI Model Rankings table