Preprint
Jul 2026
Transferability Between Understanding and Generation in Unified Multimodal Models
This work empirically finds that transferability depends on architecture-models with fully shared transformer backbone and a unified visual encoder exhibit consistent cross-task transfer, while loosely coupled designs show little or none.
Jiwon Kang, Heeji Yoon, Jaewoo Jung et al.
· 1 citation