跳到正文
原文
The Decoder· Jonathan Kemper·· 2 小时前AI 评分59

Google研究者提出防止AI智能体记忆测试的正则化自改进方法

Google researchers find a way to keep self-improving AI agents from memorizing their tests

AI 导读

Google研究者提出正则化递归自我改进(RRSI)方法,用于优化AI智能体的评测框架(harness)。传统自改进方法会让智能体记忆训练任务,导致在未见任务上性能下降。

来源:The Decoder · the-decoder.com