in the loop

vicuna-13b-v1.1
#258 of 274 scored.

LMSYSApr 12, 2023

Capabilities Index

Range 78.9–102.6

94.7

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

No independent scores yet.

Epoch AI hasn't run vicuna-13b-v1.1 yet. Its scores show up here the morning after they do.

Everything else Epoch has for it

  • PIQA54.8%
  • HellaSwag43.7%
  • Winogrande41.6%
  • GSM8K28.1%
  • ARC AI224.3%
  • BBH24.1%
  • OpenBookQA10.7%

On the index

Around it.

  1. #255DeepSeek Coder 33BDeepSeek96.3
  2. #256Falcon-7BTII95.1
  3. #257CodeQwen1.5-7BAlibaba94.9
  4. #258vicuna-13b-v1.1LMSYS94.7
  5. #259MPT-7BDatabricks94.6
  6. #260Gemma 2BGoogle94.2
  7. #261StarCoder 2 7BHugging Face93.6

Sources