A Unified Conditional Policy for Multi-Robot Navigation via LiDAR-to-Vision Distillation
Mobile-robot navigation policies have typically assumed a fixed sensing input and robot platform. In this work, we investigate a teacher–student policy where teachers learn continuous velocity commands from LiDAR-based Twin Delayed Deep Deterministic Policy Gradient, and the student navigates using either LiDAR or came...