in the loop

Phi-1.5
#265 of 274 scored.

MicrosoftSep 11, 2023Open weights

Capabilities Index

Range 71.6–100.8

91.5

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

No independent scores yet.

Epoch AI hasn't run Phi-1.5 yet. Its scores show up here the morning after they do.

Everything else Epoch has for it

  • Winogrande46.8%
  • HellaSwag30.1%
  • ARC AI225.9%
  • MMLU16.8%
  • OpenBookQA16.3%

On the index

Around it.

  1. #262XGen-7BSalesforce93.4
  2. #263Qwen-1_8BAlibaba92.9
  3. #264open_llama_7bMeta91.7
  4. #265Phi-1.5Microsoft91.5
  5. #266Baichuan1-7BBaichuan90.5
  6. #267RedPajama-INCITE-7B-BaseTogether90.4
  7. #268Dolly 2.0-12bDatabricks89.7

Sources