arXiv · 2507.20892
PixelNav: Towards Model-based Vision-Only Navigation with Topological Graphs
Abstract
This work proposes a novel hybrid approach for vision-only navigation of mobile robots, which combines advances of both deep learning approaches and classical model-based planning algorithms. Today, purely data-driven end-to-end models are dominant solutions to this problem. Despite advantages such as flexibility and adaptability, the requirement of a large amount of training data and limited interpretability are the main bottlenecks for their practical applications. To address these limitations, we propose a hierarchical system that utilizes recent advances in model predictive control, traversability estimation, visual place recognition, and pose estimation, employing topological graphs as a representation of the target environment. Using such a combination, we provide a scalable system with a higher level of interpretability compared to end-to-end approaches. Extensive real-world experiments show the efficiency of the proposed method.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sergey Bakulin, Timur Akhtyamov, Denis Fatykhov, German Devchich, Gonzalo Ferrer. 2025-07-28. PixelNav: Towards Model-based Vision-Only Navigation with Topological Graphs. https://arxiv.org/abs/2507.20892
Cite the original work for its findings. Save a collection to share your selection of sources.