Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement

  • 类型:arxiv
  • 标识:2609.13406
  • 链接:https://arxiv.org/abs/2609.13406
  • 主分类:agent
  • 形态:method
  • TLDR:When we speak of recursive self-improvement (RSI), are we speaking of a phenomenon, a mechanism, or a prospect? Towards autonomous and evolving intelligence, RSI is being claimed at many scales, while no single framework that formally describes these emerging instances exists. Its counterpart in the classical realm, iterative policy improvement, is characterized by generalized policy iteration (GPI), a framework of broad applicability with well-understood theoretical properties, but only where the update principle and the evaluation base lie outside the agent. In this paper, we propose General
  • 待LLM分类:否
  • 来源文件
  • /inbox/tom/_candidates/2026-09-17-agent-rag-longcontext-candidates.json