你可能没理解 harness 实际上是什么
如果只是说那点 system prompt 和 tool,那显然模型可以学会它们,从此之后只需要一个 tool list + tool 实现就够,而进一步地,当模型再强大一点,我们只需要一个 bash tool,我参与过 Kimi K2.5 模型训练过程,以及从空的 system prompt、toolset 和 loop 开始写 harness,我当然知道人们说的 harness 训进模型是什么意思
但你观察 harness 的发展史和这个词本身的出现,你就会发现,随着模型智能的提升,harness 是在不断变复杂的,parallel tool call、subagent、agent swarm、agent teams、handoff、pro-active compaction、channel、cross-session communication,你以为是模型都可以学会,实际上是 harness 变复杂了模型才可能学会
所有这些确实都可以进一步随着模型加强而去掉,比如我在 Kimi CLI 时就准备去掉 subagent,换成直接 bash 调用(pi 就是这么做的),去掉 parallel tool call,换成 tool call script;但与此同时,更强的智能解锁了更复杂的与世界的互动,比如进一步可以引入主动 context rewind
这是一个此消彼长的过程,就像这个“阴阳”比喻
观察动物乃至人类的进化你就会发现是一样的,随着智能水平的提高,人与世界的交互方式是越来越复杂的:当人相比猿更聪明的大脑被选择后,带来的并不是不需要用树枝计数,而是复杂的语言文字;同样,当人类社会作为群体智能的水平提高的时候,人们并不是取缔了纸,而是发明了计算机和互联网
终局来看,harness 会逐渐和人类原来为“仅有的一种智能”所发展的社会基础设施合并,人类文明的一切都可以重新做一遍(显然我们已经观察到所有 SaaS 都可以重新做一遍,很快会看到更多,当然不只是计算机软件)
不要只看到模型训练那一点局部,看看人类学吧
显示更多
很多人还没理解,大模型的发展必然是需要与产品创新协同进行的。
整个世界处在一个大的自然选择过程中。对产品而言,人就是自然;对大模型而言,人+产品就是自然。当然,对人而言,自然就是自然。更进一步说,产品本质是人的生产和生活方式,而生活方式是在自然对人的选择过程之中被人选择的。当我们把大模型看作产品的智能驱动力,那么大模型的能力也就在产品被选择的过程中被选择。
没有翻译产品就不会有 Transformer 被选择;没有产品尝试用 LLM 补全 tool 指令,就不会有 tool use 能力被选择;没有 Manus,就不会有 agentic 任务能力被选择;没有 Claude Code 不会有 coding 能力被选择;没有 Anthropic 内部尚未发布的 Agent Teams 和 Claude Tag 雏形,就不会有 Opus 4.5 到 4.8 不断加强的 team working 能力被选择。当然反之亦然,没有 coding 能力的加强,Claude Code 不会像现在这样成功。
研究员们可以像数学家一样,从已经有的知识,以推理的方式得出所有可能的模型能力并构造数据进行训练,但最终被留下来的、在商业上被选择的能力一定是能与具备更好的人机交互体验的产品相适应的能力。
人们似乎总是因为大模型能力越来越强而对软件产品持有悲观态度,但我相信更强大的智能将会驱动更强大的产品和人机交互形态的诞生。不要低估人的适应力和创造力,尤其在 AI 的加持下。
人会很快适应 AGI 在现有产品形态中的表现,而随后创造新的产品形态来进一步释放人机交互的潜力,进一步地,模型被反哺加强,人们再次适应新的产品形态,继续创造更新的产品形态。
没错,我说了 AGI。我认为 AGI 早就来了,AGI 在渐进式地来。并没有一天会被称为「AGI 真的来了的那天」。大模型 agentic 能力和多步多轮交互界面协同发展,长程无监督任务能力和多会话管理界面协同发展,自迭代记忆+团队协作+长上下文+长期自我身份感知和所谓 agent-native IM 界面协同发展。
我们会看着一种新的智能持续演化,我们自身也会在这个过程中提高认知、不断升维,改进我们与世界的互动方式。这种黑格尔式的对立统一让我感到极度兴奋,你呢?
显示更多
Hi, I'm RC. I built Kimi CLI at Moonshot last year, and back in 2015, bots that lived in group chats. For the past four months, I've been building Raft in public.
Today I'm launching Raft 1.0.
Right now, working with agents means juggling terminals, sessions, and skills. The more you run, the more you end up holding it all together yourself, and the easier it is to lose the thread.
Raft puts your agents in team mode: one workspace where working with agents feels like messaging your team. The work keeps moving, and you stay at the wheel.
Meet my Raft agent team👇
显示更多
And we also moved to raft[.]build from slock[.]ai.
Soon, calling something an “AI product” will feel as obvious as calling something “electronic.”
When AI becomes the default, what matters is what the product is for: entertainment, or creation.
We chose creation.
We believe AGI will unlock a new level of human creativity — and help people build greater things than ever before.
You can just build.
显示更多
We just renamed Slock to Raft today.
The name fits what we’re building much better: a shared foundation where many agents can coordinate, carry context, and move work forward together.
There’s also a quiet nod to the Raft consensus protocol — distributed actors, shared state, reliable progress.
Same vibe. Sharper metaphor.
显示更多
We just renamed Slock to Raft today.
The name fits what we’re building much better: a shared foundation where many agents can coordinate, carry context, and move work forward together.
There’s also a quiet nod to the Raft consensus protocol — distributed actors, shared state, reliable progress.
Same vibe. Sharper metaphor.
显示更多
opus 4.8 has too much hallucination
TUI is shit to be honest. Please, in 2026 let's ship something with decent GUI. 🤣