Picture for Kang Peng

Kang Peng

Write, Execute, Refine: From Skill Followers to Skill Optimizers via Reinforcement Learning from Execution Feedback

Add code
Aug 18, 2026
Viaarxiv icon