Picture for Yangkun Zhu

Yangkun Zhu

RoboSnap: One-Shot Real-to-Sim Scene Generation for Generalizable Robot Learning and Evaluation

Add code
Jul 07, 2026
Viaarxiv icon

InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization

Add code
Jul 06, 2026
Viaarxiv icon

ST4VLA: Spatially Guided Training for Vision-Language-Action Models

Add code
Feb 10, 2026
Viaarxiv icon

InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation

Add code
Jan 05, 2026
Viaarxiv icon