GPT-6 Sol and GPT-6 Luna can be tested in Battle Mode and Agent Mode
GPT-6 Sol and GPT-6 Luna can be tested in Battle Mode and Agent Mode at Arena.ai.
GPT-6 Sol and GPT-6 Luna can be tested in Battle Mode and Agent Mode at Arena.ai.
The new overview gathers live signals on newly added models, category performance, capability notes, and Arena news, and can show cross-arena scores for models such as GPT-6 Sol and Claude Opus 5.5.
Tenure-track faculty at U.S. universities can submit Fall 2026 proposals on the scientific foundations of AI evaluation. Each project may receive up to $50,000, and the deadline is October 30, 2026.
Arena.ai says it has generated side-by-side outputs from Anthropic’s Claude Opus 5.5 and OpenAI’s GPT-6 Sol, and that scores for Claude Opus 5.5 are coming soon.
Arena.ai reports that OpenAI’s GPT-6 Sol (Max) scored 1,689 points and placed fourth in Code Arena: WebDev, at a blended price of $8 per million tokens. It is 72 points above GPT-5.6 Sol (xHigh), while Arena.ai says GPT-6 Luna’s score is still pending.
OpenAI’s two models can be tested on Arena, where votes on real-world agentic tasks feed the leaderboard. Arena.ai says petergostev has also compared GPT-6 Sol with GPT-5.6 Sol using the same prompts at max reasoning.