Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

About

Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinforcement learning approaches map perceptual observations directly to actions, they struggle to model long-horizon dependencies, often leading to suboptimal trajectories. To address this limitation, we propose RoamFlow, a generative navigation framework that leverages MeanFlow to predict the average velocity field for trajectory synthesis, enabling efficient few-step generation and reducing inference latency. We further adopt a two-stage training strategy that combines expert imitation for stable initialization with reinforcement learning for task-specific policy refinement. Extensive experiments in both Habitat simulation and real-world robotic platforms demonstrate that RoamFlow achieves efficient inference while maintaining strong navigation performance under real-time constraints.

Zixuan Zhang, Yuqi Chen, Junjie Gao, Siyuan Song, Yongzhou Pan, Beichen Wang, Mir Feroskhan• 2026

Related benchmarks

TaskDatasetResultRank
Image-Goal NavigationMP3D (test)
Success Rate56.1
32
Image-Goal NavigationGibson unseen (test)
Success Rate (SR)68.7
8
Image-Goal NavigationPhysical Experiments Real-world scenarios
SR100
5
Showing 3 of 3 rows

Other info

Follow for update