Reflective Optimization: Self-Improving AI Agents with Textual Feedback
Lakshya A. Agrawal developed reflective optimization to address the limitations of conventional AI training, which requires massive datasets like trillions of tokens or hundreds of thousands of rollouts.
Traditional methods like gradient descent typically demand massive datasets, often requiring trillions of tokens or hundreds of thousands of rollouts, which are resource-intensive.


