Audio-driven 3D facial animation is essential for advancing immersion and interactivity in virtual experiences. Although recent advances have shown promising capabilities, the training and evaluation of existing methods typically rely on ground-truth-based errors, which fall short of aligning with human preferences. To...
Si-Jing Wu, Yunhao Li, Zhilin Gao et al.· IEEE Transactions on Visuali...· 1 citation
This work extends the previous dataset HVEval with pairwise preference annotations and proposes MoE-Rater, a Mixture-of-Experts (MoE)-inspired and multimodal large language model (MLLM)-based all-in-one method that supports multi-dimensional quality rating, multi-dimensional pairwise comparison, and category-specific q...
Si-Jing Wu, Yun-Hao Li, Hui-Yu Duan et al.· IEEE transactions on circuit...· 4 citations
Recent advances in generative video models have enabled camera-controlled world video generation, allowing models to synthesize videos under user-defined camera trajectories. However, existing video quality assessment (VQA) methods are mainly developed for natural videos and fail to capture the unique perceptual charac...
Yunhe Li, Likun Wu, Sijing Wu et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.