Via latimes.com
Anthropic’s Claude Opus 5 tops Fullstack Code Arena leaderboard, signaling AI’s next frontier for developers
The new benchmark tests AI models on real-world full-stack development tasks, and Anthropic's latest model is pulling ahead of OpenAI's GPT-5.6 Sol by a meaningful margin.
Anthropic’s Claude Opus 5, running in its Max configuration, has claimed the top spot on the newly launched Fullstack Code Arena leaderboard with a score of 1,699 points. It’s one thing to ace a coding quiz. It’s another to build an entire web application from scratch, wire up a database, and deploy it. That’s exactly what this benchmark measures.
The Fullstack Code Arena, introduced in early August 2026 on Arena.ai, represents a significant upgrade in how the industry evaluates AI coding abilities. Instead of testing whether a model can spit out a React component or solve a LeetCode problem, it assesses end-to-end web development capabilities: multi-step reasoning, tool use, database integration, and API orchestration.
The leaderboard landscape
Opus 5 isn’t competing in a vacuum. OpenAI’s GPT-5.6 Sol scored approximately 1,638 points in the same evaluations, putting it roughly 61 points behind Anthropic’s flagship. Moonshot AI’s Kimi K3 has also emerged as a competitive presence on the leaderboard, though its exact score wasn’t specified in the latest rankings.
Claude Opus 5 also performs well beyond this single benchmark. The model ranks highly on the Artificial Analysis leaderboard, confirming that its strong showing isn’t just an artifact of one evaluation framework.
AI, tech, and the markets they move—in one daily briefing.
Daily. Free. Join 34,000+ readers across crypto, finance, and policy.
Pricing and positioning
Anthropic released Claude Opus 5 on July 24, 2026, with pricing set at $5 per million input tokens and $25 per million output tokens. That positions it competitively against its predecessor, Opus 4.8, while reportedly delivering near-Fable 5 intelligence at a significantly lower cost point.