in the loop

Qwen3-Max
#105 of 274 scored.

AlibabaSep 24, 2025Closed

Capabilities Index

Range 140.1–144.8

142.4

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • GPQA Diamond63.5%

    97th of 186

  • FrontierMath18.9%

    71st of 81

  • SimpleQA Verified48.8%

    28th of 77

  • Mock AIME73.3%

    85th of 176

Everything else Epoch has for it

  • MATH level 597.1%
  • DTBench70.2%
  • Fiction.LiveBench66.7%
  • LMCA33.3%
  • CL-bench14.5%
  • Chess Puzzles0.0%
  • Mystery Game Puzzles0.0%

On the index

Around it.

  1. #102Qwen3.5-35B-A3BAlibaba · $0.08 in142.5
  2. #103Gemini 2.5 Pro (May 2025)Google142.5
  3. #104Claude Haiku 4.5Anthropic · $1.00 in142.4
  4. #105Qwen3-MaxAlibaba142.4
  5. #106o1OpenAI · $15 in141.9
  6. #107Gemma 4 26B A4BGoogle · $0.068 in141.8
  7. #108Claude Sonnet 4Anthropic · $3.00 in141.7

Sources