Picture for Shusheng Xu

Shusheng Xu

When Personal Memory Has No Single Answer: Evaluating LLM Agents under Irreducible Conflict

Add code
Aug 14, 2026
Viaarxiv icon

Building Multi-Task Agentic LLMs via Two-Phase Distillation

Add code
Jun 29, 2026
Viaarxiv icon

AREAL-DTA: Dynamic Tree Attention for Efficient Reinforcement Learning of Large Language Models

Add code
Jan 31, 2026
Viaarxiv icon

From Self-Evolving Synthetic Data to Verifiable-Reward RL: Post-Training Multi-turn Interactive Tool-Using Agents

Add code
Jan 30, 2026
Viaarxiv icon

Beyond Ten Turns: Unlocking Long-Horizon Agentic Search with Large-Scale Asynchronous RL

Add code
Aug 13, 2025
Viaarxiv icon

How Far Are We from Optimal Reasoning Efficiency?

Add code
Jun 08, 2025
Viaarxiv icon

AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning

Add code
May 30, 2025
Viaarxiv icon

On Designing Effective RL Reward at Training Time for LLM Reasoning

Add code
Oct 19, 2024
Figure 1 for On Designing Effective RL Reward at Training Time for LLM Reasoning
Figure 2 for On Designing Effective RL Reward at Training Time for LLM Reasoning
Figure 3 for On Designing Effective RL Reward at Training Time for LLM Reasoning
Figure 4 for On Designing Effective RL Reward at Training Time for LLM Reasoning
Viaarxiv icon

Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study

Add code
Apr 16, 2024
Figure 1 for Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Figure 2 for Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Figure 3 for Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Figure 4 for Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
Viaarxiv icon

Language-Guided Generation of Physically Realistic Robot Motion and Control

Add code
Jun 18, 2023
Figure 1 for Language-Guided Generation of Physically Realistic Robot Motion and Control
Figure 2 for Language-Guided Generation of Physically Realistic Robot Motion and Control
Figure 3 for Language-Guided Generation of Physically Realistic Robot Motion and Control
Figure 4 for Language-Guided Generation of Physically Realistic Robot Motion and Control
Viaarxiv icon