in the loop

GPT-3.5 Turbo (Jan 2024)
#216 of 274 scored.

OpenAIJan 25, 2024Closed

Capabilities Index

Range 107.4–118.9

115.7

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • GPQA Diamond2.9%

    177th of 186

  • FrontierMath0.0%

    81st of 81

  • Mock AIME2.1%

    162nd of 176

Everything else Epoch has for it

  • MMLU56.4%
  • DTBench14.2%
  • MATH level 511.6%
  • LMCA11.3%
  • WeirdML3.5%
  • Chess Puzzles0.0%
  • Mystery Game Puzzles0.0%

On the index

Around it.

  1. #213Llama 3-8BMeta116.5
  2. #214Qwen2.5-Coder-14BAlibaba116.4
  3. #215Gemma 3 4BGoogle · $0.05 in116.0
  4. #216GPT-3.5 Turbo (Jan 2024)OpenAI115.7
  5. #217PaLM 2-LGoogle115.0
  6. #218Llama 2-70BMeta113.8
  7. #219GPT-3.5 Turbo (Jun 2023)OpenAI113.4

Sources