跳到正文
原文
The Decoder· Jonathan Kemper·· 5 小时前

Google 研究人员找到防止自我改进 AI 智能体死记测试的方法

Google researchers find a way to keep self-improving AI agents from memorizing their tests

SI 导读

Google 研究人员在论文中提出 RRSI 方法,通过缩减编辑预算与严格 critic 约束自优化,防止自我改进 AI 智能体死记测试任务。基于冻结的 Claude Opus 4.8,RRSI 在8个基准上训练任务最高提升14.1点,未见任务最高提升4.7点且运行时少约30% token。

SI 评分58

来源:The Decoder · the-decoder.com

© 2026 SI·Hot · Super Intelligence Hot · 超级智能热点 · 网站数据均来源于网络公开资料,版权归来源方所有