We define OmniVChat (Omni Video Chat) as the task of native audio-visual dialogue between a user and an omni model. In OmniVChat, omni models directly and simultaneously receive audio and video from a user and return text. The user's query is embedded in the audio and video, without a separate text question, external c...
Haolin He, Yunfei Chu, Qi Chen et al.· 0 citations
LLM-based agents rely on heterogeneous interaction capabilities to accomplish complex tasks. Existing approaches often distribute these capabilities across multiple LoRA adapters, which increases adapter storage requirements and introduces routing overhead during inference. A single LoRA avoids this overhead, but learn...
Peng-Yang Zhou, Xiao-Bing Tu, Zheng-Xi Liu et al.· 0 citations
Single-cell RNA sequencing now routinely produces detailed maps of cell types and states, but interpreting a finished project remains harder than it should be. Once the analysis is done, the results are usually handed over as static reports, figure panels and supplementary tables. A biologist who later wants to revisit...
Jin-Bo Zhang, Zheng-Xi Liu, Zhongqi Pu et al.· bioRxiv· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.