Brain-Inspired Hierarchical Modularity for General Continual Learning
Across visual recognition, vision-language understanding, ego-exo video understanding, and embodied vision-language-action learning, this method consistently improves learning under online and uncertain data streams, with gains exceeding 50 percentage points over replay-free alternatives in embodied manipulation.