Picture for Haoyun Deng

Haoyun Deng

Understanding Knowledge Distillation in Post-Training: When It Helps and When It Fails

Add code
Jun 22, 2026
Viaarxiv icon

CM2: Reinforcement Learning with Checklist Rewards for Multi-Turn and Multi-Step Agentic Tool Use

Add code
Feb 12, 2026
Viaarxiv icon

Complex Logical Instruction Generation

Add code
Aug 12, 2025
Viaarxiv icon