in the loop

GPT-4.5
#134 of 274 scored.

OpenAIFeb 27, 2025Closed

Capabilities Index

Range 133.4–138.7

136.7

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • GPQA Diamond58.3%

    103rd of 186

  • ARC-AGI-20.8%

    68th of 79

  • Humanity's Last Exam0.7%

    34th of 40

  • Mock AIME37.7%

    117th of 176

Everything else Epoch has for it

  • MATH level 578.6%
  • Lech Mazur Writing75.6%
  • Fiction.LiveBench63.9%
  • Aider polyglot44.9%
  • WeirdML39.4%
  • SimpleBench21.4%
  • Cybench17.5%
  • VPCT17.5%
  • ARC-AGI10.3%

On the index

Around it.

  1. #131DeepSeek-R1-Distill-Qwen-32BDeepSeek137.4
  2. #132Qwen3-30B-A3B-Instruct (Jul 2025)Alibaba137.4
  3. #133GPT-4.1OpenAI · $2.00 in136.8
  4. #134GPT-4.5OpenAI136.7
  5. #135Qwen3-30B-A3BAlibaba · $0.12 in136.2
  6. #136Qwen3-8BAlibaba136.2
  7. #137DeepSeek-V3 (Mar 2025)DeepSeek135.9

Sources