Picture for Zhihao Yuan

Zhihao Yuan

WorldSimProbe: Diagnosing Simulator Faithfulness in Action-Conditioned World Models for Embodied Manipulation

Add code
Aug 10, 2026
Viaarxiv icon

Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models

Add code
Aug 07, 2026
Viaarxiv icon

When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

Add code
Aug 07, 2026
Viaarxiv icon

DataLadder: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid

Add code
Jun 15, 2026
Viaarxiv icon

JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Add code
Apr 22, 2026
Viaarxiv icon

See the Forest and the Trees: A Synergistic Reasoning Framework for Knowledge-Based Visual Question Answering

Add code
Jul 23, 2025
Figure 1 for See the Forest and the Trees: A Synergistic Reasoning Framework for Knowledge-Based Visual Question Answering
Figure 2 for See the Forest and the Trees: A Synergistic Reasoning Framework for Knowledge-Based Visual Question Answering
Figure 3 for See the Forest and the Trees: A Synergistic Reasoning Framework for Knowledge-Based Visual Question Answering
Figure 4 for See the Forest and the Trees: A Synergistic Reasoning Framework for Knowledge-Based Visual Question Answering
Viaarxiv icon

Empowering Large Language Models with 3D Situation Awareness

Add code
Mar 29, 2025
Viaarxiv icon

PiSA: A Self-Augmented Data Engine and Training Strategy for 3D Understanding with Large Models

Add code
Mar 13, 2025
Figure 1 for PiSA: A Self-Augmented Data Engine and Training Strategy for 3D Understanding with Large Models
Figure 2 for PiSA: A Self-Augmented Data Engine and Training Strategy for 3D Understanding with Large Models
Figure 3 for PiSA: A Self-Augmented Data Engine and Training Strategy for 3D Understanding with Large Models
Figure 4 for PiSA: A Self-Augmented Data Engine and Training Strategy for 3D Understanding with Large Models
Viaarxiv icon

Generative Semantic Communication for Text-to-Speech Synthesis

Add code
Oct 04, 2024
Viaarxiv icon

Instance-free Text to Point Cloud Localization with Relative Position Awareness

Add code
Apr 27, 2024
Viaarxiv icon