in the loop

PaLM 2-M
#231 of 274 scored.

GoogleMay 17, 2023

Capabilities Index

Range 98.6–116.8

108.3

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

No independent scores yet.

Epoch AI hasn't run PaLM 2-M yet. Its scores show up here the morning after they do.

Everything else Epoch has for it

  • TriviaQA81.7%
  • HellaSwag78.7%
  • Winogrande58.4%
  • ARC AI253.2%
  • OpenBookQA43.2%

On the index

Around it.

  1. #228Falcon 2 11BTII109.6
  2. #229Mistral 7B v0.3Mistral109.0
  3. #230DeepSeek-Coder-V2-Lite-BaseDeepSeek108.9
  4. #231PaLM 2-MGoogle108.3
  5. #232Phi-2Microsoft107.9
  6. #233Qwen2.5-Coder-3BAlibaba107.7
  7. #234Nemotron-4 15BNvidia107.7

Sources