The Decoder· Jonathan Kemper·· 4 hr agoSelectedAI score76
Google 研究人员提出 RRSI,防止自改进 AI 智能体记忆测试
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI brief
Google 研究人员提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),在保持 harness 可编辑的同时约束自优化。实验基于冻结的 Claude Opus 4.8,在 8 个基准上,RRSI 在训练任务上最高提升 14.1 分,在 5 个未见过的基准上最高提升 4.7 分,并减少约 30% 运行 token。
Why it matters
Google 提出的 RRSI 通过收缩编辑预算和严格批评者约束自优化,让 harness 在训练任务上保持增益的同时迁移到新任务。
Source: The Decoder · the-decoder.com