in the loop

Grok 4.3 Beta
#60 of 274 scored.

xAIApr 17, 2026Closed

Capabilities Index

Range 147.6–150.7

149.2

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • GPQA Diamond85.1%

    45th of 186

  • FrontierMath42.8%

    52nd of 81

  • SimpleQA Verified33.2%

    56th of 77

  • Mock AIME93.3%

    41st of 176

Everything else Epoch has for it

  • DTBench84.5%
  • WeirdML49.9%
  • LMCA45.1%
  • Chess Puzzles21.1%
  • FrontierMath-Tier-4-v2-Private14.6%
  • ProofBench11.0%

On the index

Around it.

  1. #57GPT-5.1OpenAI · $1.25 in149.6
  2. #58Qwen 3.8 27BAlibaba · $0.42 in149.4
  3. #59Qwen 3.6 Max (Preview)Alibaba149.2
  4. #60Grok 4.3 BetaxAI149.2
  5. #61DeepSeek-V4-ProDeepSeek · $0.21 in149.1
  6. #62GPT-5.4 MiniOpenAI · $0.75 in148.8
  7. #63InklingThinking Machines · $1.00 in148.5

Sources