Audio-driven 3D facial animation is essential for advancing immersion and interactivity in virtual experiences. Although recent advances have shown promising capabilities, the training and evaluation of existing methods typically rely on ground-truth-based errors, which fall short of aligning with human preferences. To...
Si-Jing Wu, Yunhao Li, Zhilin Gao et al.· IEEE Transactions on Visuali...· 1 citation
FreeShadow is proposed, a training-free shadow removal method built upon pretrained diffusion models, which exploits diffusion priors for shadow removal without any training or optimization.
Yi-Nan Wang, Yan Huang, Yong Xu et al.· arXiv.org· 0 citations
This work extends the previous dataset HVEval with pairwise preference annotations and proposes MoE-Rater, a Mixture-of-Experts (MoE)-inspired and multimodal large language model (MLLM)-based all-in-one method that supports multi-dimensional quality rating, multi-dimensional pairwise comparison, and category-specific q...
Si-Jing Wu, Yun-Hao Li, Hui-Yu Duan et al.· IEEE transactions on circuit...· 4 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.