in the loop

DeepSeek-V4-Pro
#61 of 274 scored.

DeepSeekApr 24, 2026Open weights

Capabilities Index

Range 147.3–150.7

149.1

Context

Takes text

1.05M

Input

Per 1M tokens

$0.21

Output

Per 1M tokens

$0.42

Benchmarks

How it scores.

  • GPQA Diamond87.9%

    30th of 186

  • SWE-bench Verified77.6%

    6th of 32

  • FrontierMath45.3%

    49th of 81

  • SimpleQA Verified47.0%

    33rd of 77

  • Mock AIME96.7%

    28th of 176

Everything else Epoch has for it

  • DTBench84.5%
  • WeirdML48.9%
  • LMCA48.5%
  • Surface Evolver Bench40.0%
  • FrontierCode17.6%
  • ProofBench16.0%
  • Chess Puzzles15.8%
  • CL-bench Life13.5%
  • Mystery Game Puzzles8.6%
  • FrontierMath-Tier-4-v2-Private2.4%

On the index

Around it.

  1. #58Qwen 3.8 27BAlibaba · $0.42 in149.4
  2. #59Qwen 3.6 Max (Preview)Alibaba149.2
  3. #60Grok 4.3 BetaxAI149.2
  4. #61DeepSeek-V4-ProDeepSeek · $0.21 in149.1
  5. #62GPT-5.4 MiniOpenAI · $0.75 in148.8
  6. #63InklingThinking Machines · $1.00 in148.5
  7. #64Kimi K2.5Moonshot · $0.45 in148.0

Sources