The Decoder· Jonathan Kemper·· 4 天前AI 评分74
Google 研究人员提出 RRSI 方法防止 AI 智能体在自我改进时记忆测试
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google 研究人员提出 RRSI(正则化递归自我改进智能体框架)方法,通过限制编辑预算和严格审查机制,防止 AI 智能体的自我改进过程过度拟合训练任务。在 8 个基准测试上,RRSI 在训练任务上提升高达 14.1 分,在 5 个未见过的任务上提升达 4.7 分,同时运行时 token 使用减少约 30%。
来源:The Decoder · the-decoder.com