Picture for Aoqi Hu

Aoqi Hu

EchoRec: Multi-Item Prediction-Empowered Generative Recommendation via Cycle-Consistent Preference Alignment

Add code
Aug 14, 2026
Viaarxiv icon

Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforcement Learning

Add code
Jul 24, 2026
Viaarxiv icon

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

Add code
Jun 30, 2026
Viaarxiv icon