Skip to content

Frontend benchmarks

Frontend Code Arena

Snapshot 2026-07-16 · undefined tasks · Source · Elo/scores from Code Arena; costs are estimates.

Illustrative snapshot from arena.ai Code Arena | WebDev (Elo from blind human votes on frontend tasks, Jul 16 2026). X-axis is blended public API $/MTok — preference Elo has no $/task. Not live OpenCodex metering.

Costs on this board are estimates or API-price blends — score/$ ranking and Best value are disabled until every row has a measured cost-per-task from the source.

ModelEffortScoreCost / taskUse case
kimi-k31679 ±17$9Frontier, Workhorse
claude-fable-51631 ±13$30Planner, Frontier
gpt-5.6-solxhigh1618 ±13$17.5Frontier, Planner
glm-5.2max1587 ±10$2.9Workhorse, Cheap subagent
claude-opus-4.8thinking1562 ±9$15Frontier, Planner
grok-4.51558 ±13$4Workhorse, Frontier
claude-opus-4.7thinking1558 ±7$15Frontier
claude-opus-4.71555 ±7$15Frontier
claude-opus-4.6thinking1542 ±6$15Workhorse
claude-sonnet-5high1542 ±12$6Workhorse
claude-opus-4.61536 ±6$15Workhorse
claude-opus-4.81534 ±9$15Frontier
glm-5.11526 ±9$2.9Workhorse, Cheap subagent
claude-sonnet-4.61522 ±6$9Workhorse
kimi-k2.61515$2.13Cheap subagent, Workhorse
gpt-5.5xhigh1504 ±7$17.5Frontier, Workhorse
gpt-5.5high1482 ±7$17.5Workhorse
gpt-5.4high1457 ±17$8.75Workhorse