The Alpha-Stabler framework is proposed, a plug-and-play framework with a Predictor that monitors principal-subspace intrusion for early collapse warnings, and a Controller that removes the principal-subspace component of activation gradients during backpropagation while preserving the orthogonal complement.
Yu-Chen Cai, Ding Cao, Qi-Xiang Yin et al.· 1 citation
Large language models are increasingly deployed at scale as API-accessible, tool-augmented agents, forming a heterogeneous, fast-evolving agent ecosystem. A central challenge is query-level identification: selecting the most suitable agent per query from candidates provided as black-box services, where costly input-out...
Jian-Dong Liu, Zi-Chen Zhao, Hao Sun et al.· Proceedings of the 32nd ACM...· 3 citations· ⚡1
This work proposes MAFIS, a novel method that addresses limitations for both online and offline MAIL settings by building upon the single-agent IQ-Learn framework and introducing the value decomposition network to factorize the imitation objective at agent level, thus enabling scalable training for multi-agent systems.
Yichen Li, Zhongxiang Ling, Tao Jiang et al.· Neural Information Processin...· 3 citations
ToMap is introduced, a multi-agent framework that structures proof autoformalization as a Decomposer-Formalizer-Prover pipeline with efficient test-time optimization guided by formal verification and semantic rubrics for proof quality.
This work operationalizes constructive specification with constructive specification, which builds hierarchical capability representations from limited profiling over diverse benchmarks, using an optimism-guided profiler that prioritizes informative regions and prunes low-utility areas with guarantees, and enables plug...
Jian-Dong Liu, Zi-Chen Zhao, Haodong Sun et al.· Proceedings of the 32nd ACM...· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.