in the loop

Qwen2.5-Coder-14B
#214 of 274 scored.

AlibabaSep 18, 2024

Capabilities Index

Range 106.5–121.8

116.4

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

No independent scores yet.

Epoch AI hasn't run Qwen2.5-Coder-14B yet. Its scores show up here the morning after they do.

Everything else Epoch has for it

  • GSM8K88.7%
  • HellaSwag73.6%
  • MMLU66.9%
  • ARC AI254.7%
  • Winogrande53.6%

On the index

Around it.

  1. #211Gemini 1.0 ProGoogle117.0
  2. #212Llama 3.1-8BMeta116.6
  3. #213Llama 3-8BMeta116.5
  4. #214Qwen2.5-Coder-14BAlibaba116.4
  5. #215Gemma 3 4BGoogle · $0.05 in116.0
  6. #216GPT-3.5 Turbo (Jan 2024)OpenAI115.7
  7. #217PaLM 2-LGoogle115.0

Sources