Building the Foundations of Self-Improving LLM Agents @allenai
Building the Foundations of Self-Improving LLM Agents  @allenai
Uploaded February 2025 | Updated September 2026, 3 days ago
Abstract:
Large language model (LLM) Agents are powerful tools for completing complex tasks but remain underexplored in their ability to self-improve through feedback, adaptation, and exploration. In this talk, I present key advances in these three directions. First, I show that by drawing an analogy between optimization and interactive learning, feedback emerges as a powerful driver of iterative improvement for LLM agents. In particular, I highlight how directional feedback enables stable and efficient performance across a wide range of optimization tasks. Second, I introduce a novel optimization framework, "Optimization with Trace Oracle (OPTO),” that leverages execution traces and rich feedback to optimize LLM agents with complex workflows, akin to how AutoDiff enables differentiable optimization. Finally, I investigate LLMs' exploration capabilities in uncertain decision-making scenarios, proposing algorithm-guided methods that enable
smaller models (Gemini-1.5 Flash) to outperform larger ones (Gemini-1.5 Pro) in exploratory efficiency. Together, these insights outline a foundation for LLM agents that can learn, adapt, and explore autonomously, paving the way for the next generation of interactive AI systems.

Bio:
anie.me/about
Building the Foundations of Self-Improving LLM AgentsMolmo 2 | Complex video question answeringTest-Time Adaptation: Next Steps for Robust Visual RecognitionLanguage Priors for Visual IntelligenceAi2 Live StreamReading and Writing Interfaces with LLMsLMQL Programming Large Language ModelsBiomedical AI for Precision HealthEntailer: Answering Questions with Faithful and Truthful Chains of ReasoningLearning for Never-before-seen BiomedicineOpen AI: considering the ethical upsides and downsides of Open AI developmentBuilding robotics systems in simulation and on real robots
Ai2 |

Building the Foundations of Self-Improving LLM Agents

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER