This survey reviews core tasks, representative methods, and datasets in multimodal facial state analysis, focusing on facial expression recognition, AU detection, and face-based soft biometric estimation, and emphasizing the unique value of language in providing contextual semantics, enhancing reasoning, and generating...
Xuri Ge, Tian-Shuo Zhang, Ruihan Li et al.· 0 citations
Rec is proposed, an RL-free framework that decouples reasoning from prediction, overcoming the fixed d-dimensional state bottleneck of prior methods, and points to a decoupled, multi-vector recipe that unleashes latent reasoning from the single-state bottleneck of prior methods.
Wenhao Deng, Junchen Fu, Han-Wen Du et al.· arXiv.org· 0 citations
DPQ is introduced, a lightweight pre-quantization recipe family that uses full-precision predictions to construct target-aligned calibration mixtures of high-doubt examples and generic anchors that better preserve broad multiple-choice QA behavior.
Zhen Yang, Sizai Hou, Kai-Wen Zheng et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.