DeepSeek V5 (Leaked Staging) vs GPT-6.1 Sol

A direct, empirical evaluation of DeepSeek V5 (Leaked Staging) (DeepSeek AI) against GPT-6.1 Sol (OpenAI).

DeepSeek V5 (Leaked Staging)

Developed by DeepSeek AI · 1.8 Trillion (Sparse MoE) parameters

Autonomous coding agents scoring 81.4% on SWE-bench Verified, zero-telemetry self-hosting, sub-cent 1M-token processing, and formal mathematical proofs.

GPT-6.1 Sol

Developed by OpenAI · Proprietary Frontier Workhorse Engine parameters

Near-Astra reasoning at 1/5th price ($1.25/$5.00), SWE-bench Verified 74.8%, OSWorld 2.0 desktop agent automation, and sub-80ms Ultrafast Mode.

Metric DeepSeek V5 (Leaked Staging) GPT-6.1 Sol
SWE-bench Verified 81.4% 74.8%
MMLU-Pro 87.5% 92.4%
MATH-500 96.8% 96.2%
Context Window 1,000,000 tokens 1,000,000 tokens
Input Pricing (per 1M) $0.12 $1.25
Output Pricing (per 1M) $0.48 $5.00