Anthropic's Lance Martin just walked through how they build agents that run for hours with no human in the loop.
His talk on async, long-horizon agents in 5 timestamps:
3:19 – split the brain from the hands
6:20 – build agent vs verifier agent in a loop
8:18 – 20 iterations to clear an ML benchmark, unsupervised
13:19 – "dreaming" fixes memories the agent got wrong
16:08 – Claude Tag as an org-level harness, not a slackbot
The point: past the 1-hour mark, architecture (not just model size) is what makes async agents work.
Solid 25 min if you're building anything long-running.
显示更多