Picture for Yanheng Li

Yanheng Li

From Dense Prediction to Visual Editing: Structured Supervision for Unified Image and Video Creation

Add code
Aug 13, 2026
Viaarxiv icon

Parameter-Efficient Adaptation of SAM3 for Prompt-Driven Surgical Concept Segmentation

Add code
Jul 26, 2026
Viaarxiv icon

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation

Add code
Apr 09, 2026
Viaarxiv icon

Hear You in Silence: Designing for Active Listening in Human Interaction with Conversational Agents Using Context-Aware Pacing

Add code
Feb 05, 2026
Viaarxiv icon

Multi-objective fluorescent molecule design with a data-physics dual-driven generative framework

Add code
Jan 20, 2026
Viaarxiv icon

Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview

Add code
May 04, 2025
Figure 1 for Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview
Figure 2 for Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview
Figure 3 for Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview
Figure 4 for Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview
Viaarxiv icon

ProtTeX: Structure-In-Context Reasoning and Editing of Proteins with Large Language Models

Add code
Mar 13, 2025
Viaarxiv icon

ProTeX: Structure-In-Context Reasoning and Editing of Proteins with Large Language Models

Add code
Mar 11, 2025
Viaarxiv icon

Aspect-Guided Multi-Level Perturbation Analysis of Large Language Models in Automated Peer Review

Add code
Feb 18, 2025
Viaarxiv icon

V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy

Add code
Dec 23, 2024
Figure 1 for V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
Figure 2 for V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
Figure 3 for V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
Figure 4 for V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
Viaarxiv icon