The Decoder· Jonathan Kemper·2026-10-04 20:40· 1 天前AI 评分62Google 研究者提出 RRSI,防止自我改进智能体记住测试任务AI 导读Google 研究者提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),通过随时间收缩的编辑预算和严格评审机制,阻止智能体在自我优化中记住测试任务。正文阅读原文 当前提供 AI 导读,完整正文请阅读原文。来源:The Decoder · the-decoder.com#论文/研究#Agent#Google#数据/训练查看事件全部后续