Picture for Shiyu Dong

Shiyu Dong

MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding

Add code
Aug 18, 2026
Viaarxiv icon

Perception Encoder: The best visual embeddings are not at the output of the network

Add code
Apr 17, 2025
Figure 1 for Perception Encoder: The best visual embeddings are not at the output of the network
Figure 2 for Perception Encoder: The best visual embeddings are not at the output of the network
Figure 3 for Perception Encoder: The best visual embeddings are not at the output of the network
Figure 4 for Perception Encoder: The best visual embeddings are not at the output of the network
Viaarxiv icon

UniPlane: Unified Plane Detection and Reconstruction from Posed Monocular Videos

Add code
Jul 04, 2024
Viaarxiv icon