World simulation is inherently multisensory, demanding synchronized visual and acoustic dynamics in real time. Yet prevailing interactive world models remain strictly silent, focusing exclusively on visual rendering and control while overlooking the acoustic dimension. We present HelixWorld, a real-time interactive aud...
Lei Ke, Jia-Hao Pan, Ze-Yue Tian et al.· 0 citations
Emotional awareness is a foundational component of mental well-being, yet existing digital tools often reduce emotional experience to labels, metrics, or inferred states. Such approaches overlook how emotional awareness is constructed through embodied, multi-modal experience and reflective sense-making. We present Emot...
Zi-Ying Wang, Yi-Hong Lin, Xiang-Lin Zhao et al.· Proceedings of the 28th Inte...· 0 citations
EmotiStage, an emotion-driven performance system powered by multi-modal generative AI that supports emotional awareness by enabling users to externalize, observe, and reinterpret emotions through music and stage performance, is presented.
Zi-Ying Wang, Yi-Hong Lin, Xiang-Lin Zhao et al.· Proceedings of the 28th Inte...· 0 citations
Symbolic models make melody, harmony, rhythm, and form explicit but typically stop before a finished recording; audio models produce complete songs while leaving composition implicit. We introduce YuE2, which unifies symbolic and audio music generation at frontier quality through symbolic planning. A single AR-NAR Mixt...
Rui-Bin Yuan, Jia-Hao Pan, Jun-Yan Jiang et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.