A novel multimodal entity alignment model- TextFusionEA is proposed, which implements structural embedding based on GraphSAGE, uses MobileNetV3 to obtain visual features, introduces an adaptive fusion mechanism of semantic, structural, and visual features, and designs a semi supervised iterative learning strategy to extend the seed subset to alleviate the problem of data scarcity.
Abstract
The entity alignment task in Multi modal Knowledge Graph (MKG) faces challenges such as strong modal heterogeneity, large semantic gap, and inflexible fusion strategies. Existing methods generally rely on static weighted fusion and large amounts of annotated data, making it difficult to cope with complex multimodal environments. This article proposes a novel multimodal entity alignment model- TextFusionEA, which implements structural embedding based on GraphSAGE, uses MobileNetV3 to obtain visual features, introduces an adaptive fusion mechanism of semantic, structural, and visual features, and designs a semi supervised iterative learning strategy to extend the seed subset to alleviate the problem of data scarcity. Experiments on five datasets have shown that TextFusionEA performs well in MRR Hits@1. It significantly outperforms existing mainstream methods in terms of metrics, verifying its alignment performance and robustness in complex multimodal scenarios.
Multimodal Knowledge Graphs (MMKGs) offer a promising paradigm for integrating heterogeneous sources into a unified, queryable, semantically structured representation. However, existing MMKG construction pipelines remain predominantly text-centric, extracting information from textual passages while leaving much of the...
Busisani Mac Dube, Jean Vincent Fonou Dombeu· Big Data and Cognitive Compu...· 0 citations
With the rapid development of information technology in military and complex system evaluation domains, the issue of "information overload" regarding evaluation data has become increasingly prominent. Traditional recommendation algorithms rely heavily on simple historical interactions and lack the capacity to capture s...
Quan-Dong Wang, Peng-Fei Yang, Qian Huang et al.· 2026 12th International Conf...· 0 citations
To address the challenges of evidence chain breakage of vectors, collaborative constraint between image and structured fields is challenging, and the credibility of the generated results is insufficient with multimodal data, this paper proposes a GraphRAG semantic retrieval model for multimodal data. In order to realiz...
Chun-Jing Liao, Pei-Shan Ye, An-Ni Huang et al.· Discover Artificial Intellig...· 0 citations
An adaptive community search framework ECHO is proposed, which consistently outperforms state-of-the-art methods in terms of community quality while achieving superior search efficiency.
Cheng-Yang Luo, Zi-Xing Ding, Qing Liu et al.· Proceedings of the 32nd ACM...· 0 citations
With ever-growing conception of artificial intelligence and smart algorithms, the progress of complicated relationship modeling and combination technologies of multi-source information has become rapid. Graph matching as a key question in the structured data analysis and inter-modal semantic alignment has been of great...
Rong-Hui Liu· International Conference on...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.