Quantifying and reducing the modality gap in both contrastive and generative multimodal vision–language models
Sep 2026 · International Journal of Multimedia Information Retrieval · Vol 15 · 0 citations
· 30 references