The Decoder· Jonathan Kemper·· 2 天前AI 评分48
Google 研究人员找到防止自我改进 AI 智能体记住测试任务的方法
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google 研究人员提出 RRSI(正则化递归自我改进智能体框架),通过限制编辑预算和严格审查机制,防止 AI 智能体在自我优化过程中记忆测试任务。在 8 个基准测试中,RRSI 在未见过的任务上最高提升 4.7 分,运行时 token 消耗减少约 30%,且性能从未低于基线。相关代码已在 GitHub 开源。
来源:The Decoder · the-decoder.com