in the loop

Llama 3.2 90B
#182 of 274 scored.

MetaSep 24, 2024Open weights

Capabilities Index

Range 119.4–127.3

125.5

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • GPQA Diamond21.4%

    154th of 186

  • Mock AIME2.5%

    158th of 176

Everything else Epoch has for it

  • MMLU73.7%
  • GeoBench52.0%
  • MATH level 539.4%
  • Balrog27.3%

On the index

Around it.

  1. #179GPT-4 Turbo (Nov 2023)OpenAI126.5
  2. #180Llama 3.1-70BMeta125.9
  3. #181GPT-4 (Mar 2023)OpenAI125.9
  4. #182Llama 3.2 90BMeta125.5
  5. #183Qwen2-72BAlibaba125.3
  6. #184DeepSeek-V2 (MoE-236B, May 2024)DeepSeek124.8
  7. #185Amazon Nova ProAmazon123.8

Sources