评论认为 OpenAI 的强化学习在数学超智能方向领先
一条推特观点:OpenAI 靠 RL 冲数学超智能,而不是堆编程 Agent,思路跟别人不太一样,可以看看.
OpenAI 用强化学习训练模型冲击数学领域的超智能,评论认为其在这条路线上目前领先其他机构。这条路线更接近几年前 AGI 的原始构想。文中还将它和押注软件工程 Agent 的路线做了对比。
OpenAI RL machine is fearsome they want math superintelligence and they're currently closer to it than anyone else This was the original idea of AGI, a couple years ago. Perhaps it's a more principled path to the goal than SWE agent-maxxing.