跳到正文
原文
The Decoder· Jonathan Kemper·· 2 小时前AI 评分62

Google 研究提出 RRSI 方法,防止自改进 AI 智能体记住测试任务

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google 研究人员提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),在让语言模型自动重写智能体 harness 的循环中加入约束,避免其死记测试任务。

来源:The Decoder · the-decoder.com