Skip to content

Intelligence benchmarks

AA Intelligence Index

Snapshot 2026-07-16 · undefined tasks · Source · Figures from Artificial Analysis public pages; not affiliated with AA.

Illustrative snapshot inspired by Artificial Analysis Intelligence Index cost-per-task (Answer / Reasoning / Cache write / Cache hit / Input). Totals match published headline figures where noted; segment splits are approximate for charting. Not live OpenCodex metering.

ModelEffortScoreCost / taskScore / $Use case
claude-fable-5max60$3.2518.5Planner, Frontier
gpt-5.6-solmax59$1.0456.7Frontier, Planner
gpt-5.6-solxhigh58$0.8865.9Frontier, Workhorse
claude-opus-4.8max57$1.7832.0Frontier, Planner
gpt-5.5xhigh56$0.9956.6Frontier, Workhorse
gpt-5.6-terramax55$0.55100.0Workhorse
gpt-5.6-lunamax51$0.21242.9Workhorse, Cheap subagent
gemini-3.1-propreview46$0.32143.8Frontier, Fast
deepseek-v4-proBest valuemax44$0.041100.0Cheap subagent, Workhorse
grok-4.3high48$0.28171.4Workhorse, Fast
glm-5.2max42$0.12350.0Workhorse, Cheap subagent
claude-sonnet-4.6max49$1.1542.6Workhorse