LLM Arena

Compare models head-to-head across all benchmarks and metrics.

GPT-4o
OpenAI · Proprietary
  • Composite89.2
  • Arena ELO1287
  • MMLU88.7%
  • HumanEval90.2%
  • SWE-Bench38.4%
  • Speed110 t/s
  • TTFT320ms
  • Output Price$10/M
VS
Claude 3.5 Sonnet
Anthropic · Proprietary
  • Composite88.4
  • Arena ELO1272
  • MMLU88.7%
  • HumanEval92%
  • SWE-Bench49%
  • Speed82 t/s
  • TTFT380ms
  • Output Price$15/M
Full Comparison →