Touch2Shape: Touch-Conditioned 3D Diffusion for Shape Exploration and Reconstruction

About

Diffusion models have made breakthroughs in 3D generation tasks. Current 3D diffusion models focus on reconstructing target shape from images or a set of partial observations. While excelling in global context understanding, they struggle to capture the local details of complex shapes and limited to the occlusion and lighting conditions. To overcome these limitations, we utilize tactile images to capture the local 3D information and propose a Touch2Shape model, which leverages a touch-conditioned diffusion model to explore and reconstruct the target shape from touch. For shape reconstruction, we have developed a touch embedding module to condition the diffusion model in creating a compact representation and a touch shape fusion module to refine the reconstructed shape. For shape exploration, we combine the diffusion model with reinforcement learning to train a policy. This involves using the generated latent vector from the diffusion model to guide the touch exploration policy training through a novel reward design. Experiments validate the reconstruction quality thorough both qualitatively and quantitative analysis, and our touch exploration policy further boosts reconstruction performance.

Yuanbo Wang, Zhaoxuan Zhang, Jiajin Qiu, Dilong Sun, Zhengyu Meng, Xiaopeng Wei, Xin Yang• 2025

Related benchmarks

Task	Dataset	Result
3D Reconstruction	ShapeNet (test)	EMD0.042	74
3D Reconstruction	Simulation Data objects (test)	EMD0.041	18
3D Reconstruction	ABC	Chamfer Distance1.406	12

Showing 3 of 3 rows

Other info

Follow for update

@wizwand_team Discord