RECAST: Recasting Vision-Language Semantics into an Actionable Cost Map for Robot Navigation
Safe and robust robot navigation across diverse environments requires a high-level understanding of complex scenes and the ability to carry it into stable motion. Recent works tackle this with learning-based models trained at scale and with approaches built on vision-language models (VLMs). However, learning-based mode...