Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Jul 2026

Less Repetition, Less Energy Cost: A Reinforcement Learning-Based Multiagent Energy-Saving Autonomous Exploration System.

Multiagent autonomous exploration in unknown environments is both meaningful and challenging. Due to the constraint of a partially observable environment, the collaboration among agents is often inadequate, leading to increased energy consumption. Worse still, a decrease in overall exploration performance may occur due to a single agent failure. To address these issues, we propose a distributed Multiagent Energy-saving Autonomous Exploration System (MEAES) based on reinforcement learning. To accurately evaluate the regional complexity of different branches and further enhance the long-term decision-making capabilities of agents, we introduce the dual-scale clustered observation (DSCO) module. The DSCO generates fine-grained representations based on graph modeling, enabling better characterization of both global and long-term exploration values. Furthermore, we propose an energy-saving action (EA) mechanism, which mitigates redundant exploration and reduces energy consumption by selective waiting actions and independent exploration strategies. Finally, we devise the consumption-exploration-balanced training framework (CEBF), which guides agents to transform from lazy exploration to energy-saving exploration strategies through dynamic reward shaping. Extensive experiments validate the effectiveness of MEAES, demonstrating effective zero-shot transfer performance across unseen environments.

Yang Liu, Peng Zhang, Yanting Li et al. · 0 citations
Jul 2026

One Aligned LLM to Serve Them All: A Transfer Recipe for Training VLMs without Visual-Language Re-Alignment

This work demonstrates that the aligned LLM with a general-purpose vision encoder can effectively enhance downstream VQA performance with task-specific encoders, and investigates several alignment strategies between the aligned LLM and new task-specific encoders.

Jiazuo Yu, Yunzhi Zhuge, Lu Zhang et al. · 0 citations