注册并分享邀请链接,可获得视频播放与邀请奖励。

X Freeze (@XFreeze) “Grok 4.6 just matched Claude Fable 5 on Perplexity’s WANDR benchmark....at more” — TopicDigg

X Freeze 的个人资料封面
X Freeze 的头像
X Freeze
@XFreeze
加入 July 2024
0 正在关注    0 粉丝
Grok 4.6 just matched Claude Fable 5 on Perplexity’s WANDR benchmark....at more than 60% lower cost Both score exactly 0.496 But the cost per task is ridiculous: • Grok 4.6 → $7.58 • Claude Fable 5 → $20.30 Same benchmark performance....for a fraction of the cost and Grok 4.6 also comfortably outperforms GPT-5.6 Sol Grok 4.6’s efficiency is totally insane
显示更多
Congrats to @SpaceXAI on one more amazing model: Grok 4.6. We benchmarked it as an orchestrator on our Wide-And-Deep-Research benchmark using the Perplexity Computer harness, and it neatly sits on the Pareto frontier of performance vs cost. Available to all Pro and Max users on Perplexity!
显示更多
0
27
189
43
转发到社区