in the loop

Phi-2
#232 of 274 scored.

MicrosoftDec 12, 2023Open weights

Capabilities Index

Range 88.6–113.8

107.9

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

No independent scores yet.

Epoch AI hasn't run Phi-2 yet. Its scores show up here the morning after they do.

Everything else Epoch has for it

  • ARC AI267.9%
  • OpenBookQA64.8%
  • BBH45.9%
  • TriviaQA45.2%
  • MMLU44.5%
  • HellaSwag38.1%
  • ANLI13.8%
  • Winogrande9.4%

On the index

Around it.

  1. #229Mistral 7B v0.3Mistral109.0
  2. #230DeepSeek-Coder-V2-Lite-BaseDeepSeek108.9
  3. #231PaLM 2-MGoogle108.3
  4. #232Phi-2Microsoft107.9
  5. #233Qwen2.5-Coder-3BAlibaba107.7
  6. #234Nemotron-4 15BNvidia107.7
  7. #235Yi-9B01.AI107.6

Sources