Picture for Wang Zhao

Wang Zhao

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation

Add code
Jun 25, 2026
Viaarxiv icon

MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

Add code
Jun 03, 2026
Viaarxiv icon

Pixal3D: Pixel-Aligned 3D Generation from Images

Add code
May 11, 2026
Viaarxiv icon

StdGEN++: A Comprehensive System for Semantic-Decomposed 3D Character Generation

Add code
Jan 12, 2026
Viaarxiv icon

DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation

Add code
Jul 02, 2025
Figure 1 for DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
Figure 2 for DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
Figure 3 for DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
Figure 4 for DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
Viaarxiv icon

DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation

Add code
Dec 19, 2024
Figure 1 for DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
Figure 2 for DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
Figure 3 for DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
Figure 4 for DI-PCG: Diffusion-based Efficient Inverse Procedural Content Generation for High-quality 3D Asset Creation
Viaarxiv icon

AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos

Add code
Nov 29, 2024
Figure 1 for AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos
Figure 2 for AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos
Figure 3 for AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos
Figure 4 for AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos
Viaarxiv icon

StdGEN: Semantic-Decomposed 3D Character Generation from Single Images

Add code
Nov 08, 2024
Figure 1 for StdGEN: Semantic-Decomposed 3D Character Generation from Single Images
Figure 2 for StdGEN: Semantic-Decomposed 3D Character Generation from Single Images
Figure 3 for StdGEN: Semantic-Decomposed 3D Character Generation from Single Images
Figure 4 for StdGEN: Semantic-Decomposed 3D Character Generation from Single Images
Viaarxiv icon

MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction

Add code
Nov 02, 2024
Figure 1 for MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
Figure 2 for MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
Figure 3 for MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
Figure 4 for MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
Viaarxiv icon

Tracking Everything in Robotic-Assisted Surgery

Add code
Sep 29, 2024
Viaarxiv icon