Picture for Jiapeng Li

Jiapeng Li

Geo-Embed: Towards Unified Multimodal Embeddings for Urban Understanding

Add code
Aug 04, 2026
Viaarxiv icon

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

Add code
Jul 28, 2026
Viaarxiv icon

Near-field Beam Training under Multi-path Channels: A Hybrid Learning-and-Optimization Approach

Add code
Mar 26, 2026
Viaarxiv icon

StreetTree: A Large-Scale Global Benchmark for Fine-Grained Tree Species Classification

Add code
Feb 22, 2026
Viaarxiv icon

Near-field Target Localization: Effect of Hardware Impairments

Add code
Dec 25, 2025
Figure 1 for Near-field Target Localization: Effect of Hardware Impairments
Figure 2 for Near-field Target Localization: Effect of Hardware Impairments
Figure 3 for Near-field Target Localization: Effect of Hardware Impairments
Figure 4 for Near-field Target Localization: Effect of Hardware Impairments
Viaarxiv icon

R^3-VQA: "Read the Room" by Video Social Reasoning

Add code
May 07, 2025
Figure 1 for R^3-VQA: "Read the Room" by Video Social Reasoning
Figure 2 for R^3-VQA: "Read the Room" by Video Social Reasoning
Figure 3 for R^3-VQA: "Read the Room" by Video Social Reasoning
Figure 4 for R^3-VQA: "Read the Room" by Video Social Reasoning
Viaarxiv icon

Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey

Add code
Nov 14, 2024
Figure 1 for Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
Figure 2 for Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
Figure 3 for Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
Figure 4 for Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
Viaarxiv icon

Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue

Add code
Oct 21, 2024
Figure 1 for Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
Figure 2 for Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
Figure 3 for Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
Figure 4 for Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
Viaarxiv icon

KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Add code
Jul 03, 2024
Viaarxiv icon

GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing

Add code
Jun 30, 2024
Viaarxiv icon