Skip to content

Learning Agile Navigation in Crowded Environments for Quadruped Robots

Jul 2026 · arXiv.org · Vol abs/2607.15036 · 1 citation · 52 references
Computer Science

TL;DR

This work proposes VOP-Nav, a novel navigation system that combines the geometric safety of VO with the agile adaptability of end-to-end learning and validates the system's robustness and efficiency in complex indoor and outdoor dynamic environments.

Abstract

Navigating dynamic and crowded environments presents significant challenges for quadruped robots due to severe sensor occlusion and unpredictable human motion. Existing approaches face a trade-off: model-based methods, such as Velocity Obstacles (VO), theoretically guarantee safety but rely on accurate obstacle motion estimates that often fail in dense crowds, while end-to-end learning methods offer robustness but lack motion prediction capability of obstacles, leading to collisions or conservative behaviors. To solve this, we propose VOP-Nav, a novel navigation system that combines the geometric safety of VO with the agile adaptability of end-to-end learning. Using only local onboard observations, our system avoids explicit obstacle detection and tracking pipelines. The VOP-Net processes multi-frame LiDAR data to implicitly encode dynamic constraints and predict a safe velocity region derived from Velocity Obstacle theory. Importantly, the VO predictions serve a dual role: they are used as input to the navigation policy during inference and as a reward signal during training to encourage safe motion. Evaluations in Isaac Gym demonstrate that VOP-Nav achieves higher success rates than all baselines while balancing locomotion speed and collision avoidance. Real-world deployment on a Unitree Go2 quadruped robot further validates the system's robustness and efficiency in complex indoor and outdoor dynamic environments.

View source

Similar papers

Preprint Aug 2026

Spatiotemporal Agility: Time-Constrained Reinforcement Learning for Vision-Guided Dynamic Quadrupedal Interception

An integrated framework that combines a vision module for landing point and time prediction with a direct position and time conditioned RL locomotion policy, instead of intermediate velocity commands is proposed, which mitigates perception latency during dynamic interception.

Yi-Dong Zhu, Zibo Dai, Tong-Ning Zhang et al. · 0 citations
Preprint Sep 2026

DODGER: Safety-Guided Reinforcement Learning for Robot Navigation Among Dynamic Obstacles

Robots operating in human-centered environments must safely navigate among multiple dynamic obstacles to avoid collisions with people and surrounding infrastructure. Control barrier functions (CBFs) provide an effective mechanism for safety filtering, and recent CBF-based reinforcement learning (RL) methods embed such...

Sanghyuk Park, Kwan-Woo Lee, Taekyung Kim et al. · 0 citations
Aug 2024

LSTP-Nav: Lightweight Spatiotemporal Policy for Map-Free Multi-Agent Navigation With LiDAR

This paper proposes LSTP-Nav, a lightweight, decentralized navigation framework built on LSTP-Net that maps stacked 2D LiDAR observations, goal information, and velocity feedback directly to action and introduces an HS reward to provide smooth, heading-aware safety feedback, and develops PhysReplay-SimLab to improve tr...

Xingrong Diao, Zhi-Qiang Sun, Jian-Wei Peng et al. · 0 citations
Preprint Sep 2026

Experience-Driven Continual Learning of Terrain Traversability for Quadruped Robots

Safe and efficient quadruped navigation over unfamiliar terrain requires predicting terrain-robot interaction before contact: geometry and visual appearance alone cannot reveal how the robot will slip, load its feet, or expend energy. This paper presents a continual learning pipeline that uses locomotion experience to...

Luca Bricarello, J. C. V. Soares, Alberto Sánchez-Delgado et al. · 0 citations
Preprint Aug 2026

DPNet: Efficient Dead-End Prediction and Avoidance for Vision-Based UAV Navigation

Vision-based Unmanned Aerial Vehicles (UAVs) often suffer from navigation failures in dead ends due to limited sensing accuracy and range. To address this challenge, this paper proposes a systematic solution for efficient dead-end prediction and avoidance. The proposed method introduces a lightweight neural network to...

Rui-Bin Zhang, Lun Pan, Zelong Xia et al. · 0 citations
Preprint Sep 2026

FutureRay: Control-Aligned Future Range for Agile Quadruped Navigation

Moving obstacles can block a previously clear route while a quadruped robot executes a motion command. We investigate whether predicting changing clearance improves navigation when motion selection accounts for the robot footprint and the time needed to react and brake. We present FutureRay, which predicts ranges acros...

Tian-Hao Zang, Shan-Ze Wang, Zi-Qian Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.