in the loop

GPT-4.1 mini
#142 of 274 scored.

OpenAIApr 14, 2025Closed

Capabilities Index

Range 131.3–136.7

135.0

Context

Takes image, text, file

1.05M

Input

Per 1M tokens

$0.40

Output

Per 1M tokens

$1.60

Benchmarks

How it scores.

  • GPQA Diamond54.5%

    108th of 186

  • FrontierMath6.7%

    76th of 81

  • ARC-AGI-20.0%

    72nd of 79

  • SimpleQA Verified12.7%

    71st of 77

  • Mock AIME44.7%

    115th of 176

Everything else Epoch has for it

  • MATH level 587.3%
  • DTBench48.0%
  • Fiction.LiveBench44.4%
  • WeirdML37.6%
  • Aider polyglot32.4%
  • LMCA24.9%
  • CadEval16.0%
  • ARC-AGI3.5%
  • Chess Puzzles2.1%
  • Mystery Game Puzzles0.0%

On the index

Around it.

  1. #139DeepSeek-R1-Distill-Qwen-14BDeepSeek135.4
  2. #140Gemini 2.0 Flash Thinking (Jan 2025)Google135.4
  3. #141Gemini 2.0 ProGoogle135.1
  4. #142GPT-4.1 miniOpenAI · $0.40 in135.0
  5. #143o1-previewOpenAI134.8
  6. #144Gemini 2.0 Flash (Feb 2025)Google134.7
  7. #145Gemini 2.0 Flash (Dec 2024)Google134.7

Sources