in the loop

Qwen3.7-Max
#40 of 274 scored.

AlibabaMay 19, 2026Closed

Capabilities Index

Range 151.7–155.8

153.7

Context

Takes text

1M

Input

Per 1M tokens

$1.48

Output

Per 1M tokens

$4.42

Benchmarks

How it scores.

  • GPQA Diamond87.9%

    30th of 186

  • SWE-bench Verified77.3%

    7th of 32

  • FrontierMath64.6%

    31st of 81

  • SimpleQA Verified55.8%

    19th of 77

  • Mock AIME95.5%

    32nd of 176

Everything else Epoch has for it

  • DTBench87.1%
  • SimpleBench64.5%
  • LMCA51.8%
  • FrontierMath-Tier-4-v2-Private34.2%
  • ProofBench26.0%
  • Mystery Game Puzzles25.1%
  • Chess Puzzles14.8%
  • EBR-bench9.5%

On the index

Around it.

  1. #37Gemini 3.6 FlashGoogle · $0.75 in154.3
  2. #38Muse Spark 1.1Meta · $1.25 in154.2
  3. #39Grok 4.5xAI · $2.00 in153.9
  4. #40Qwen3.7-MaxAlibaba · $1.48 in153.7
  5. #41Grok 4.7xAI · $2.00 in153.5
  6. #42GPT-5.2OpenAI · $1.75 in153.4
  7. #43Gemini 3 ProGoogle152.9

Sources