Picture for Zhiheng Wu

Zhiheng Wu

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding

Add code
Jul 30, 2026
Viaarxiv icon

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

Add code
Jun 08, 2026
Viaarxiv icon

MedHorizon: Towards Long-context Medical Video Understanding in the Wild

Add code
May 07, 2026
Viaarxiv icon

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

Add code
Apr 27, 2026
Viaarxiv icon

M$^{2}$GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

Add code
Apr 21, 2026
Viaarxiv icon

UC-OWOD: Unknown-Classified Open World Object Detection

Add code
Jul 23, 2022
Figure 1 for UC-OWOD: Unknown-Classified Open World Object Detection
Figure 2 for UC-OWOD: Unknown-Classified Open World Object Detection
Figure 3 for UC-OWOD: Unknown-Classified Open World Object Detection
Figure 4 for UC-OWOD: Unknown-Classified Open World Object Detection
Viaarxiv icon