Any6D: Model-free 6D Pose Estimation of Novel Objects
About
We introduce Any6D, a model-free framework for 6D object pose estimation that requires only a single RGB-D anchor image to estimate both the 6D pose and size of unknown objects in novel scenes. Unlike existing methods that rely on textured 3D models or multiple viewpoints, Any6D leverages a joint object alignment process to enhance 2D-3D alignment and metric scale estimation for improved pose accuracy. Our approach integrates a render-and-compare strategy to generate and refine pose hypotheses, enabling robust performance in scenarios with occlusions, non-overlapping views, diverse lighting conditions, and large cross-environment variations. We evaluate our method on five challenging datasets: REAL275, Toyota-Light, HO3D, YCBINEOAT, and LM-O, demonstrating its effectiveness in significantly outperforming state-of-the-art methods for novel object pose estimation. Project page: https://taeyeop.com/any6d
Related benchmarks
| Task | Dataset | Result | Rank | |
|---|---|---|---|---|
| 6-DoF Pose Tracking | YCBInEOAT | ATE86.26 | 28 | |
| 6D Object Pose Estimation | Toyota-Light (TOYL) (test) | AR43.3 | 24 | |
| 6-DoF Object Tracking | YCBInEOAT | ADD-S62.2 | 23 | |
| 6D Object Pose Estimation | LM-O (test) | -- | 22 | |
| 6-DoF Pose Tracking | HO3D | ADD-S75.4 | 20 | |
| Pose Estimation | Toyota-Light 52 | AR43.3 | 14 | |
| Pose Estimation | Real275 23 | AR51 | 14 | |
| 6D Pose Tracking | dataset S4 (held-out synthetic) | Absolute Translation Error (mm)865.4 | 14 | |
| 6D Pose Tracking | HO3D | Absolute Translation Error (mm)64.6 | 14 | |
| HOI Reconstruction | HOI4D (test) | F5 Accuracy71 | 13 |