跳到正文
原文
The Decoder· Jonathan Kemper·· 4 小时前AI 评分60

Google 研究人员提出 RRSI,抑制自我改进 AI 智能体记忆测试任务

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google 研究人员提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),通过收缩编辑预算、记录失败尝试并过滤基准特定技巧,减少自我改进 AI 智能体对测试任务的记忆。

来源:The Decoder · the-decoder.com