Picture for Zhan Chen

Zhan Chen

Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study

Add code
Jul 15, 2026
Viaarxiv icon

Ego-Dynamics-Augmented World Model for Autonomous Driving with Zero-Shot Cross-Chassis Adaptation

Add code
Jul 15, 2026
Viaarxiv icon

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs

Add code
Jun 01, 2026
Viaarxiv icon

MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models

Add code
Apr 09, 2026
Viaarxiv icon

AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving

Add code
Mar 16, 2026
Viaarxiv icon

RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention

Add code
Dec 25, 2025
Figure 1 for RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
Figure 2 for RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
Figure 3 for RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
Figure 4 for RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
Viaarxiv icon

Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs

Add code
Feb 20, 2025
Viaarxiv icon

UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images

Add code
Aug 26, 2024
Figure 1 for UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 2 for UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 3 for UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 4 for UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Viaarxiv icon

UNetMamba: Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images

Add code
Aug 21, 2024
Figure 1 for UNetMamba: Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 2 for UNetMamba: Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 3 for UNetMamba: Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Figure 4 for UNetMamba: Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
Viaarxiv icon

Secrets of RLHF in Large Language Models Part II: Reward Modeling

Add code
Jan 12, 2024
Figure 1 for Secrets of RLHF in Large Language Models Part II: Reward Modeling
Figure 2 for Secrets of RLHF in Large Language Models Part II: Reward Modeling
Figure 3 for Secrets of RLHF in Large Language Models Part II: Reward Modeling
Figure 4 for Secrets of RLHF in Large Language Models Part II: Reward Modeling
Viaarxiv icon