Reading up on Google Cloud AI Research
1 deep · digging since sep 23
- RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
RRSI regularizes the search loop of agent harnesses, yielding +4.0 pts on evolved benchmarks, +3.4 pts on held‑out ones, and a 36% reduction in policy tokens.