跳到正文
The Decoder· Jonathan Kemper·· 4 小时前AI 评分54

Google 研究人员找到防止自改进 AI 智能体记忆测试任务的方法

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google 研究人员提出了一种名为 RRSI 的方法,旨在防止自改进的 AI 智能体在优化其控制程序(harness)时记忆并过拟合有限的测试任务。该方法通过限制编辑预算和设置严格的审查规则来保持控制程序的通用性,在编码、办公和工程设计等八个基准测试中,RRSI 在未见任务上获得了最高 4.7 分的提升,同时运行时 token 使用量减少了约 30%。

来源:The Decoder · the-decoder.com