注册并分享邀请链接,可获得视频播放与邀请奖励。

AlphaSense (@AlphaSenseInc) “Every few weeks, a new model tops a public leaderboard. But in finance and busin” — TopicDigg

AlphaSense 的个人资料封面
AlphaSense 的头像
AlphaSense
@AlphaSenseInc
加入 September 2013
0 正在关注    0 粉丝
Every few weeks, a new model tops a public leaderboard. But in finance and business research, the model isn't the bottleneck anymore. Context is. Continuous benchmarking at AlphaSense shows GPT-5.6 Sol leading the pack, Opus 5 underperforming its predecessor at 5x the cost, and Gemma 4-31B matching Sonnet 5 quality at 40x lower cost. The lesson: newer isn't automatically better, token price doesn't equal question cost, and per-task model routing beats any single-model strategy by 2.8x. Chris Ackerson and Daniel Campos break down what months of head-to-head model testing reveal about frontier AI, and the two engineering programs designed to widen the gap even further. Read the full article:
显示更多
0
11
333
16
转发到社区