Agentic work is where
@grok 4.6 lands hardest, taking the top spot on the Artificial Analysis Agentic Index at 59, tied with Claude Opus 5 Max.
The index measures tool use, planning, autonomy and complex problem solving rather than single answers
Grok 4.6 completes tasks in ~53 turns and ~0.5bn input tokens on average, against ~103 turns and ~2.0bn for Claude Opus 5 Max
Cost of $0.84 per task, putting it on the intelligence versus cost per task Pareto frontier
Enterprises buying agents pay per completed task, not per benchmark point. Turn efficiency is what determines whether a long-running workflow is affordable at volume.
Two labs now sit at the top of this index with very different cost structures. Buyers get real choice on price for the first time in agentic deployment.