Picture for Yuanbo Yang

Yuanbo Yang

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning

Add code
Aug 10, 2026
Viaarxiv icon

Beyond Isolated Objects: Relationship-aware Open Vocabulary Scene Understanding via 3D Scene Graph Analysis

Add code
Jul 06, 2026
Viaarxiv icon

The Constant Eye: Benchmarking and Bridging Appearance Robustness in Autonomous Driving

Add code
Feb 13, 2026
Viaarxiv icon

ReRoPE: Repurposing RoPE for Relative Camera Control

Add code
Feb 08, 2026
Viaarxiv icon

Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction

Add code
Jan 07, 2026
Viaarxiv icon

Orientation Matters: Making 3D Generative Models Orientation-Aligned

Add code
Jun 10, 2025
Viaarxiv icon

Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation

Add code
Dec 30, 2024
Viaarxiv icon

Learning Temporally Consistent Video Depth from Video Diffusion Priors

Add code
Jun 04, 2024
Viaarxiv icon

MaPa: Text-driven Photorealistic Material Painting for 3D Shapes

Add code
Apr 26, 2024
Figure 1 for MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
Figure 2 for MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
Figure 3 for MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
Figure 4 for MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
Viaarxiv icon

Learning 3D-Aware GANs from Unposed Images with Template Feature Field

Add code
Apr 08, 2024
Figure 1 for Learning 3D-Aware GANs from Unposed Images with Template Feature Field
Figure 2 for Learning 3D-Aware GANs from Unposed Images with Template Feature Field
Figure 3 for Learning 3D-Aware GANs from Unposed Images with Template Feature Field
Figure 4 for Learning 3D-Aware GANs from Unposed Images with Template Feature Field
Viaarxiv icon