Loading…

Self-improving AI agents tend to memorize their test tasks, so their gains shrink or disappear on new ones. RRSI, a new method from Google researchers, reins in this effect and lifts scores on unseen benchmarks by up to 4.7 points while using about 30 percent fewer tokens than…
To respect copyright, we link to the source rather than republishing the full text. Read the complete article on The Decoder.