The Decoder· Jonathan Kemper·· 2 小时前AI 评分62
Google 研究提出 RRSI 方法,防止自改进 AI 智能体记住测试任务
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google 研究人员提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),在让语言模型自动重写智能体 harness 的循环中加入约束,避免其死记测试任务。
来源:The Decoder · the-decoder.com