| Task Name | Dataset Name | SOTA Result | Trend | |
|---|---|---|---|---|
| 3D Object Extraction | LeRF | Accuracy99.8 | 26 | |
| Open-Vocabulary 3D Scene Segmentation | LeRF-mask | Figurines mIoU90.8 | 17 | |
| Open-vocabulary 3D object selection | LERF | Ramen Score61.4 | 16 | |
| 3D Object Localization | LERF | Ramen Success Rate75.6 | 14 | |
| 3D Object Selection | LERF figurines scene | Peak VRAM8 | 14 | |
| 3D Semantic Segmentation | LERF (test) | mIoU62.1 | 13 | |
| 3D Scene Reconstruction | LERF average across four scenes | PSNR24.02 | 12 | |
| Open-vocabulary semantic segmentation | LeRF-OVS | mIoU64.4 | 12 | |
| Open-vocabulary Object Selection | LERF 21 (test) | Mean IoU (Figurines)70.1 | 10 | |
| Novel View Synthesis | LERF 21 | PSNR28.1 | 10 | |
| 3D Semantic Segmentation | LERF | mIoU (Ramen)57.6 | 9 | |
| 3D Open-vocabulary Segmentation | LERF-style Dataset bed scene (test) | mIoU89.5 | 8 | |
| Open-vocabulary 3D Scene Understanding | LERF | Feature Distillation Time (h)1 | 7 | |
| 2D Semantic Segmentation | LERF Overall | mIoU60.02 | 6 | |
| 2D Localization | LERF (Overall) | mAcc84.57 | 6 | |
| 3D Segmentation | LERF-Mask | mIoU (Figurines)93.9 | 5 | |
| Open Vocabulary 3D Semantic Segmentation | LERF (held-out test views) | Ramen mIoU52 | 5 | |
| 3D Instance Retrieval | LERF Small-scale Indoor Scene | Figurines Accuracy77.42 | 4 | |
| Semantic Segmentation | LERF ramen | mIoU95.63 | 4 | |
| Semantic Segmentation | LERF figurines | mIoU93.32 | 4 | |
| Open-vocabulary 3D object selection | Lerf ovs (part) | mIoU44.1 | 4 | |
| Open-vocabulary 3D object retrieval | LERF | Ramen mIoU53.34 | 4 | |
| 3D Object Retrieval | LERF standard scene (~300 images) | Geometry Computation Time (mins)15 | 4 | |
| Open-vocabulary 2D object retrieval and localization | LERF | mIoU (Ramen Scene)63.4 | 4 | |
| Novel View Rendering | LERF Figurines | Speed (FPS)354.72 | 4 |