EMODY Flow: Emotion-Aware Audio-Driven Full-Body Motion Generation
This work presents EMODY Flow, a lightweight flow-matching framework that attaches to a frozen Qwen-3 Omni model and reuses its internal Mimi audio-codecs to condition two parallel DiT generators - one for SMPL-X body pose, one for FLAME facial expressions.
H. Agarwal, Xavier Alameda-Pineda, Olivier Perrotin
· 0 citations