跳到正文
原文
The Decoder· Jonathan Kemper·· 1 天前AI 评分62

Google 研究者提出 RRSI,防止自我改进智能体记住测试任务

AI 导读

Google 研究者提出 RRSI(Regularized Recursive Self-Improvement of Agent Harnesses),通过随时间收缩的编辑预算和严格评审机制,阻止智能体在自我优化中记住测试任务。

当前提供 AI 导读,完整正文请阅读原文。

来源:The Decoder · the-decoder.com