跳到正文
The Decoder·· 1 天前

Google researchers find a way to keep self-improving AI agents from memorizing their tests

Google researchers find a way to keep self-improving AI agents from memorizing their tests

摘要

Self-improving AI agents tend to memorize their test tasks, so their gains shrink or disappear on new ones. RRSI, a new method from Google researchers, reins in this effect and lifts scores on unseen benchmarks by up to 4.7 points while using about 30 percent fewer tokens than an unregularized version. The article Google researchers find a way to keep self-improving AI agents from memorizing their tests appeared first on The Decoder .

应来源方要求,这里只提供摘要与原文入口。完整内容请阅读原文。

来源:The Decoder · the-decoder.com