Grok 4.20 Multi Agent Beta 0309vsKimi K2.5 Thinking

Grok 4.20 Multi Agent Beta 0309 leads the overall Podium Score by 1.0 points (74.0 vs 73.0).Category wins: Grok 4.20 Multi Agent Beta 0309 0 — 0 Kimi K2.5 Thinking.Kimi K2.5 Thinking is the cheaper pick ($3/M vs $15/M per 1M output tokens). Grok 4.20 Multi Agent Beta 0309 is faster (50 t/s vs 50 t/s).

MetricGrok 4.20 Multi Agent Beta 0309Kimi K2.5 Thinking
Podium Score74.0 ▲73.0
Arena Elo
Intelligence Index
Coding
Math
Reasoning
Agentic
Knowledge
Multimodal
Long-context
Output speed50 t/s50 t/s
Output price$15/M$3/M ▲
Context window128K128K