Twitter/X

Code Arena launched the Fullstack Code Arena on 2026-07-28 to evaluate AI models…

Brief

Code Arena announced the Fullstack Code Arena (2026-07-28), expanding from frontend prototypes to full-stack web development evaluation. The environment lets models act as agents using structured tool calls to plan, execute, and iterate with real databases, API keys, and deployment; early leaderboard leaders are Kimi K3 (Max), GPT 5.6 Sol (xHigh), and Claude Fable 5.

Why it matters

Code Arena launched the Fullstack Code Arena on 2026-07-28 to evaluate AI models on full-stack web development tasks — multi-step reasoning, tool use, and end-to-end app generation — including databases, API keys, and fast deployments with models operating as agents via structured tool calls.

Key details

  • Current leaderboard positions reported: Kimi K3 (Max) #1, GPT 5.6 Sol (xHigh) #2, Claude Fable 5 #3; full rankings available at arena.ai/leaderboard/code/we…
Source evidence

Code Arena now measures fullstack capabilities!

View overall rankings across AI models on full-stack web development tasks: multi-step reasoning, tool use, and end-to-end app generation.

  • Kimi K3 (Max) takes #1
  • GPT 5.6 Sol (xHigh) at #2
  • Claude Fable 5 at #3
    See more scores at: arena.ai/leaderboard/code/we…

Arena.ai (@arena)

Code Arena just leveled up with fullstack capabilities 🚀

Introducing the new Fullstack Code Arena. We’re moving beyond frontend prototypes to fullstack development complete with databases, API keys, and fast deployments. Build, iterate, and ship real-world software — all in one place.

Models now act as agents in the Code Arena, using structured tool calls to plan, execute, and refine in real time with real world tasks.

Read more about it in the thread 🧵

Video

— https://nitter.net/arena/status/2072713730711023673#m