Arena.ai launches redesigned Leaderboard Overview
The new overview gathers live signals on newly added models, category performance, capability notes, and Arena news, and can show cross-arena scores for models such as GPT-6 Sol and Claude Opus 5.5.
The new overview gathers live signals on newly added models, category performance, capability notes, and Arena news, and can show cross-arena scores for models such as GPT-6 Sol and Claude Opus 5.5.
Arena.ai places Claude Opus 5.5 (Max) first in Code Arena WebDev at 1,818 points, 26 ahead of GPT-6 Astra (Max) and 126 above Opus 5 (Max). Claude says the model matches Claude Fable 5.1 on most tasks and costs 40% less to run than Opus 5.
Arena.ai scores SpaceXAI’s Grok 4.7 (xHigh) at 1,632 points, 16 points and six places above Grok 4.6 (High). Simulations and Consumer Product also move up.
The multilingual Controlled Voice Arena leaderboards compare text-to-speech models in 9 new languages. A language-selectable leaderboard, a public voting arena, and the benchmarking methodology are available.
No single open-weight text-to-speech model leads all nine languages in Artificial Analysis’s rankings. Higgs Audio V3 TTS, OpenAudio S1 Mini, Voxtral TTS, and Breeze TTS 2 each top the open-weight field in different languages, and each board includes only 2 to 6 open-weight models.
The Controlled Voice Arena now ranks text-to-speech models in Japanese, Mandarin Chinese, Hindi, Spanish, German, French, Portuguese, Vietnamese, and Arabic. Cartesia’s Sonic family leads eight of the nine boards, while Inworld AI’s Realtime TTS-2 ranks first in Mandarin.
Arena.ai gives Qwen-Image-2.1 1,228 points and first place among open models, 17th overall, four points behind Gemini-3-pro-image-preview and 11 behind GPT-Image-1.5-high-fidelity.