A coding-model evaluation arena where users compare models by having them produce code and vote on results; its coding and frontend comparisons support leaderboard evaluation.

Recent stories
2 linked stories
releaseSECONDARY2026-09-01
Alibaba releases Qwen3.8-Max-0902 with 1M-token context
Alibaba released Qwen3.8-Max-0902 through QwenCloud with a 1M-token context window and 2.4T parameters. The company prices input at $2 per million tokens and says the model leads Code Arena's WebDev leaderboard.
newsSECONDARY2026-07-18
Code Arena users report Kaleb shows Qwen-like token quirks
Multiple testers reported Kaleb and torenia-alpha appearing in Code Arena/LMArena, and several pointed to Qwen-like token quirks in Kaleb. Follow-up tests described strong 3D output and a late-2025 or early-2026 cutoff, but the model identity remains unconfirmed.