跳到正文
原文
The Decoder· Jonathan Kemper·· 3 小时前精选AI 评分76

Google 研究人员提出 RRSI,防止自改进 AI 智能体记忆测试

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google 研究人员提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),在保持 harness 可编辑的同时约束自优化。实验基于冻结的 Claude Opus 4.8,在 8 个基准上,RRSI 在训练任务上最高提升 14.1 分,在 5 个未见过的基准上最高提升 4.7 分,并减少约 30% 运行 token。

推荐理由

Google 提出的 RRSI 通过收缩编辑预算和严格批评者约束自优化,让 harness 在训练任务上保持增益的同时迁移到新任务。

来源:The Decoder · the-decoder.com