LLM Agents & Reinforcement Learning
RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
A detailed analysis of RLEF and how reinforcement learning teaches code LLMs to condition future generations on execution feedback rather than simply generating another independent answer.
Oct 06, 202619 pages · Paper Analysis