Picture for Yujin Zhou

Yujin Zhou

SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent

Add code
Aug 07, 2026
Viaarxiv icon

S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation

Add code
Jul 17, 2026
Viaarxiv icon

Not Just the Destination, But the Journey: Reasoning Traces Causally Shape Generalization Behaviors

Add code
Mar 12, 2026
Viaarxiv icon

What, Whether and How? Unveiling Process Reward Models for Thinking with Images Reasoning

Add code
Feb 09, 2026
Viaarxiv icon

Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning

Add code
Jan 20, 2026
Viaarxiv icon

LRAS: Advanced Legal Reasoning with Agentic Search

Add code
Jan 12, 2026
Viaarxiv icon

AM$^3$Safety: Towards Data Efficient Alignment of Multi-modal Multi-turn Safety for MLLMs

Add code
Jan 08, 2026
Viaarxiv icon