Picture for Yuzhi Huang

Yuzhi Huang

TAU-Bench: From Anomaly Instance Tracking to Fine-Grained Video Anomaly Understanding

Add code
Aug 06, 2026
Viaarxiv icon

ChainVLA: Chaining Vision-Language-Action Queries through a Unified Execution State for Long-Horizon Manipulation

Add code
Aug 03, 2026
Viaarxiv icon

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs

Add code
Jul 14, 2026
Viaarxiv icon

RoboStream: Weaving Spatio-Temporal Reasoning with Memory in Vision-Language Models for Robotics

Add code
Mar 13, 2026
Viaarxiv icon

Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World

Add code
Mar 13, 2026
Viaarxiv icon

Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline

Add code
Jun 05, 2025
Figure 1 for Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline
Figure 2 for Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline
Figure 3 for Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline
Figure 4 for Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline
Viaarxiv icon

Underwater Object Detection in the Era of Artificial Intelligence: Current, Challenge, and Future

Add code
Oct 08, 2024
Figure 1 for Underwater Object Detection in the Era of Artificial Intelligence: Current, Challenge, and Future
Figure 2 for Underwater Object Detection in the Era of Artificial Intelligence: Current, Challenge, and Future
Figure 3 for Underwater Object Detection in the Era of Artificial Intelligence: Current, Challenge, and Future
Figure 4 for Underwater Object Detection in the Era of Artificial Intelligence: Current, Challenge, and Future
Viaarxiv icon