注册并分享邀请链接,可获得视频播放与邀请奖励。

changis_k 的个人资料封面
changis_k 的头像

changis_k (@changis_k)

@changis_k
0 正在关注    0 粉丝
Agentic work is where @grok 4.6 lands hardest, taking the top spot on the Artificial Analysis Agentic Index at 59, tied with Claude Opus 5 Max. The index measures tool use, planning, autonomy and complex problem solving rather than single answers Grok 4.6 completes tasks in ~53 turns and ~0.5bn input tokens on average, against ~103 turns and ~2.0bn for Claude Opus 5 Max Cost of $0.84 per task, putting it on the intelligence versus cost per task Pareto frontier Enterprises buying agents pay per completed task, not per benchmark point. Turn efficiency is what determines whether a long-running workflow is affordable at volume. Two labs now sit at the top of this index with very different cost structures. Buyers get real choice on price for the first time in agentic deployment.
显示更多