Picture for Siteng Huang

Siteng Huang

RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

Add code
Aug 10, 2026
Viaarxiv icon

RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Model

Add code
Jul 20, 2026
Viaarxiv icon

RynnWorld-4D: 4D Embodied World Models for Robotic Manipulation

Add code
Jul 07, 2026
Viaarxiv icon

RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation

Add code
Jul 07, 2026
Viaarxiv icon

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon

Add code
Jul 02, 2026
Viaarxiv icon

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

Add code
Jun 01, 2026
Viaarxiv icon

MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation

Add code
Mar 27, 2026
Viaarxiv icon

Articulat3D: Reconstructing Articulated Digital Twins From Monocular Videos with Geometric and Motion Constraints

Add code
Mar 12, 2026
Viaarxiv icon

RynnBrain: Open Embodied Foundation Models

Add code
Feb 13, 2026
Viaarxiv icon

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

Add code
Dec 10, 2025
Figure 1 for HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
Figure 2 for HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
Figure 3 for HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
Figure 4 for HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
Viaarxiv icon