AI Dynamics

Global AI News Aggregator

About

Experiential Reinforcement Learning: Teaching LLMs Self-Reflection

Making LLMs truly learn from its experience "Experiential Reinforcement Learning (ERL)" ERL makes an agent attempt -> get sparse feedback -> write a self-reflection -> retry All by distilling the improved retry back into the base policy so the correction sticks without needing

→ View original post on X — @askalphaxiv