AVLF-Nav: Embodied Object Navigation via Active Visual Exploration and Lightweight Vision–Language Fusion
Embodied object navigation must couple partial visual perception with efficient exploration in unseen indoor environments. Existing agents often align language and vision without explicitly deciding which viewpoint is most informative, causing repeated visits and inefficient trajectories. This paper proposes AVLF-Nav,...