Xentir Media Source-backed AI & tech
LIVE INTELLIGENCE · part of the Xentir AI Index

DeepSeek v4 Pro vs Kimi K2.5 (instant)

DeepSeek v4 Pro OPEN WEIGHTS · DeepSeek, released 2026-04-24
Kimi K2.5 (instant) OPEN WEIGHTS · Moonshot, released 2026-01-27

“Epoch AI, 'AI Benchmarking Hub'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks' [online resource].” · epoch.ai/benchmarks · Licensed CC BY 4.0 · data last verified 2026-07-26 · Methodology · Data Sources

Head-to-head

Measured on 6 of 9 tracked benchmarks for both models — DeepSeek v4 Pro leads on 6.

BenchmarkDeepSeek v4 Pro Kimi K2.5 (instant)Leads
Epoch Capabilities Index (ECI)148.7148.3DeepSeek v4 Pro
SWE-bench Verified0.776 (max)0.738DeepSeek v4 Pro
GPQA Diamond0.896 (max)0.876DeepSeek v4 Pro
MATH Level 5Not measuredNot measuredNot measured on both
OTIS Mock AIME 2024-20250.967 (max)0.922DeepSeek v4 Pro
FrontierMathNot measured0.279Not measured on both
FrontierMath Tier 4Not measured0.042Not measured on both
SimpleQA Verified0.570 (max)0.339DeepSeek v4 Pro
Chess Puzzles0.200 (max)0.120DeepSeek v4 Pro
Benchmark results measure specific tasks and should not be treated as a universal measure of model quality. A model with no row here for a benchmark has not been measured on it — that is left blank, never estimated.

Full profiles

DeepSeek v4 Pro → · Kimi K2.5 (instant) → · the full AI Model Rankings table