Picture for Xiaoyan Cong

Xiaoyan Cong

Ms. Forcing: Efficient Streaming Video Generation with Multi-Scale Patchification and Attention

Add code
Jul 23, 2026
Viaarxiv icon

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

Add code
Jul 06, 2026
Viaarxiv icon

UMO: Unified In-Context Learning Unlocks Motion Foundation Model Priors

Add code
Mar 16, 2026
Viaarxiv icon

PackUV: Packed Gaussian UV Maps for 4D Volumetric Video

Add code
Feb 26, 2026
Viaarxiv icon

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation

Add code
Jan 30, 2026
Viaarxiv icon

VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization

Add code
Dec 18, 2025
Figure 1 for VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization
Figure 2 for VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization
Figure 3 for VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization
Figure 4 for VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization
Viaarxiv icon

GenHSI: Controllable Generation of Human-Scene Interaction Videos

Add code
Jun 24, 2025
Viaarxiv icon

Art3D: Training-Free 3D Generation from Flat-Colored Illustration

Add code
Apr 14, 2025
Viaarxiv icon

Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method

Add code
Nov 17, 2024
Figure 1 for Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
Figure 2 for Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
Figure 3 for Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
Figure 4 for Oscillation Inversion: Understand the structure of Large Flow Model through the Lens of Inversion Method
Viaarxiv icon

4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations

Add code
Jun 14, 2024
Figure 1 for 4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
Figure 2 for 4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
Figure 3 for 4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
Figure 4 for 4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
Viaarxiv icon