The Decoder· Jonathan Kemper·· 14 小时前AI 评分70
Google研究人员提出RRSI,减少自我改进型AI代理对测试任务的记忆
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google研究人员提出RRSI,通过逐步缩小编辑预算、记录失败尝试并由严格审查器筛选修改,减少AI代理对测试任务的记忆。RRSI在八个基准上使用冻结的Claude Opus 4.8,训练任务最高提升14.1分,五个未见基准最高提升4.7分,运行token比未正则化版本少约30%。研究代码已发布到GitHub,但结果仅覆盖冻结模型,不涉及模型权重变化。
来源:The Decoder · the-decoder.com