返回首页
RadarAI··论文与技术

Google Research 开源 RRSI:智能体在模型权重冻结下自我改进 harness 并避免过拟合

中文摘要

谷歌发布RRSI,使LLM智能体无需改动模型权重,即可自我改进提示词、工具、记忆与控制流。实现智能体自我提升,避免过拟合。

English Summary

Google released RRSI, allowing LLM agents to self-improve prompts, tools, memory, and control flow without modifying model weights. This enables agent self-enhancement and avoids overfitting.

原文节选

Google Cloud AI Research 联合 UNC-Chapel Hill、Stanford 和圣路易斯华盛顿大学发布 RRSI(Regularized Recursive Self-Improvement),让 LLM 智能体在不改动模型权重的情况下改写自身 harness 的提示词、工具、记忆、控制流和子智能体。