Picture for Zuyan Liu

Zuyan Liu

Hy-Embodied-VLM-1.0: Efficient Physical-World Agents

Add code
Jul 14, 2026
Viaarxiv icon

ViQ: Text-Aligned Visual Quantized Representations at Any Resolution

Add code
Jun 25, 2026
Viaarxiv icon

GEM: Generative Supervision Helps Embodied Intelligence

Add code
May 27, 2026
Viaarxiv icon

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Add code
Apr 08, 2026
Viaarxiv icon

PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning

Add code
Mar 27, 2026
Viaarxiv icon

Insight-V++: Towards Advanced Long-Chain Visual Reasoning with Multimodal Large Language Models

Add code
Mar 18, 2026
Viaarxiv icon

GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization

Add code
Nov 19, 2025
Viaarxiv icon

Vision Generalist Model: A Survey

Add code
Jun 11, 2025
Viaarxiv icon

SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs

Add code
Jun 05, 2025
Viaarxiv icon

Ola: Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment

Add code
Feb 06, 2025
Figure 1 for Ola: Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment
Figure 2 for Ola: Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment
Figure 3 for Ola: Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment
Figure 4 for Ola: Pushing the Frontiers of Omni-Modal Language Model with Progressive Modality Alignment
Viaarxiv icon