Summary:
Arena launches multi-modal AI leaderboards
Why It Matters:
The platform has expanded from simple text comparisons to ranking how AI models perform across code, vision, video, and image generation.
What Changed:
- Changed the site title to "Arena Leaderboard | Compare & Benchmark the Best Frontier AI Models."
- Added a comprehensive "Leaderboard Overview" page featuring top-10 rankings for eight new categories.
- Introduced specialized tabs for Text, Code, Vision, Text-to-Image, Image Edit, Search, Text-to-Video, and Image-to-Video.
- Updated the main ranking table to include 594 models with granular scores for "Hard Prompts," "Creative Writing," and "Instruction Following."
- Added links to external model providers like Google (Gemini 3), xAI (Grok 4.1), and Anthropic (Claude 4.5).