in the loop

Kimi K2 (Jul 2025)
#116 of 274 scored.

MoonshotJul 12, 2025Open weights

Capabilities Index

Range 136.6–142.0

140.1

Context

Tokens

—

Input

Per 1M tokens

—

Output

Per 1M tokens

—

Benchmarks

How it scores.

  • Terminal-Bench27.8%

    27th of 35

Everything else Epoch has for it

  • Lech Mazur Writing85.6%
  • Fiction.LiveBench61.1%
  • Aider polyglot59.1%
  • WeirdML39.4%
  • SimpleBench11.6%
  • GSO-Bench4.9%

On the index

Around it.

  1. #113Gemini 2.5 Flash (Jun 2025)Google140.8
  2. #114Grok-3 minixAI140.3
  3. #115o3-miniOpenAI · $1.10 in140.3
  4. #116Kimi K2 (Jul 2025)Moonshot140.1
  5. #117Gemini 2.5 Flash (Apr 2025)Google140.0
  6. #118gpt-oss-120bOpenAI · $0.037 in139.9
  7. #119DeepSeek-V3.1DeepSeek · $0.25 in139.9

Sources