Can agents improve their own skills without generating any new rollouts? This paper introduces SkillRefiner, which learns entirely from historical agent traces. It turns past mistake patterns into targeted edits to the agent’s existing skill, by basically treat deployment
SkillRefiner: Agents Improve Skills from Historical Traces
By
–
