The Decoder· Jonathan Kemper·· 16 小时前AI 评分56
Google 研究者提出 RRSI 方法,防止自改进 AI 智能体记忆测试任务
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google 研究者提出了一种名为 RRSI 的方法,旨在防止自改进 AI 智能体在优化过程中记忆其测试任务。该方法通过在优化循环的两端施加约束来实现:限制每次提议的修改量,并使用一个严格的评判器来剔除那些硬编码任务名称或解决方案的修改。
来源:The Decoder · the-decoder.com