跳到正文
原文
The Decoder· Jonathan Kemper·· 2 天前AI 评分48

Google 研究人员找到防止自我改进 AI 智能体记住测试任务的方法

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google 研究人员提出 RRSI(正则化递归自我改进智能体框架),通过限制编辑预算和严格审查机制,防止 AI 智能体在自我优化过程中记忆测试任务。在 8 个基准测试中,RRSI 在未见过的任务上最高提升 4.7 分,运行时 token 消耗减少约 30%,且性能从未低于基线。相关代码已在 GitHub 开源。

来源:The Decoder · the-decoder.com