所以很多时候,在具体的生产场景下,训练一个强化学习的黑盒,往往没有直接让大模型写一段代码脚本来得成本低、效果好。
不是每一个场景都需要硬塞一个大模型进去。千万不要手里拿着大模型这个锤子,看啥都是钉子。
Very exciting to see the cool result! Same pattern, different physics: Codex-grown heuristics matching or beating DRL agents in fluid dynamics, while staying readable, maintainable, and transferable.
Heuristics were not dead. They were under-maintained.
显示更多