The Decoder· Jonathan Kemper·· 2 小时前AI 评分59
Google研究者提出防止AI智能体记忆测试的正则化自改进方法
Google researchers find a way to keep self-improving AI agents from memorizing their tests
AI 导读
Google研究者提出正则化递归自我改进(RRSI)方法,用于优化AI智能体的评测框架(harness)。传统自改进方法会让智能体记忆训练任务,导致在未见任务上性能下降。
来源:The Decoder · the-decoder.com