Video and audio are perceived together, yet most generative models treat them in isolation. We examine methods that model the two modalities jointly, generate one from the other, or edit them in a coupled manner, organized around a single question: how is the output kept coherent across modalities in time and semantics...
Abhinav Sharma, S. Navuluru, Wang Wei et al.· 0 citations
Existing Large Language Model (LLM) routing methods score LLMs independently to select top-$k$ models. However, this ignores model correlations and enforces a rigid computational budget. Consequently, routers often select redundant models that share failure modes, limiting the overall probability of success. To address...
Wang Wei, Harry Yang, Tiankai Yang et al.· 1 citation
Large language model (LLM) agents increasingly rely on external skills, but routing user requests over large skill registries is difficult because many skills are functionally redundant while complex tasks often require complementary skill sets. Existing skill routers typically rank candidates independently by query re...
Wang Wei, Tiankai Yang, Samyadeep Basu et al.· 2 citations
A common layer equation is introduced that represents covered architectures through seven components: an update domain, channel set, propagation bank, per-channel message maps, channel-fusion operator, ego/residual map, and update map, which exposes the empirical inverse problem of mapping measurable graph and task pro...
S. Navuluru, Siddhartha Shankar Das, B. Ni et al.· 0 citations
VLAFP is the first deep audio fingerprinting model capable of processing audio of variable length, for both training and testing, and outperforms existing state-of-the-arts in live audio identification and audio retrieval across three real-world datasets.
This work introduces the problem of personalized auto-research, which conditions every stage of the research process on a representation of the individual researcher, and proposes a general and flexible framework that threads a graph-grounded researcher context through retrieval, hypothesis search, experimentation, wri...
B. Ni, Franck Dernoncourt, Hongjie Chen et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.