Picture for Haotian Xia

Haotian Xia

Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Question Answering

Add code
Aug 04, 2026
Viaarxiv icon

SportD: Can VLMs Physically Strategize?

Add code
Jul 16, 2026
Viaarxiv icon

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios?

Add code
Jun 29, 2026
Viaarxiv icon

StoryAlign: Evaluating and Training Reward Models for Story Generation

Add code
May 06, 2026
Viaarxiv icon

SportR: A Benchmark for Multimodal Large Language Model Reasoning in Sports

Add code
Nov 17, 2025
Viaarxiv icon

DeepSport: A Multimodal Large Language Model for Comprehensive Sports Video Reasoning via Agentic Reinforcement Learning

Add code
Nov 17, 2025
Viaarxiv icon

FACTS: Fine-Grained Action Classification for Tactical Sports

Add code
Dec 21, 2024
Viaarxiv icon

SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models

Add code
Oct 11, 2024
Figure 1 for SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
Figure 2 for SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
Figure 3 for SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
Figure 4 for SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
Viaarxiv icon

Sports Intelligence: Assessing the Sports Understanding Capabilities of Language Models through Question Answering from Text to Video

Add code
Jun 21, 2024
Viaarxiv icon

Language and Multimodal Models in Sports: A Survey of Datasets and Applications

Add code
Jun 18, 2024
Figure 1 for Language and Multimodal Models in Sports: A Survey of Datasets and Applications
Figure 2 for Language and Multimodal Models in Sports: A Survey of Datasets and Applications
Viaarxiv icon