NFDI4DS | UHH-SEMS - Publication Details

Tokenization, Fusion, and Augmentation: Towards Fine-grained Multi-modal Entity Representation

FOS: Computer and information sciences Artificial Intelligence (cs.AI) Computer Science - Artificial Intelligence

DOI: 10.48550/arxiv.2404.09468 Publication Date: 2024-01-01

Abstract Supplemental Material References Cited by

AUTHORS (8)

Zhang, Yichi

Chen, Zhuo

Guo, Lingbing

Xu, Yajing

Hu, Binbin

Liu, Ziqi

Zhang, Wen

Chen, Huajun

ABSTRACT

AAAI 2025; Repo is available at https://github.com/zjukg/MyGO<br/>Multi-modal knowledge graph completion (MMKGC) aims to discover unobserved knowledge from given knowledge graphs, collaboratively leveraging structural information from the triples and multi-modal information of the entities to overcome the inherent incompleteness. Existing MMKGC methods usually extract multi-modal features with pre-trained models, resulting in coarse handling of multi-modal entity information, overlooking the nuanced, fine-grained semantic details and their complex interactions. To tackle this shortfall, we introduce a novel framework MyGO to tokenize, fuse, and augment the fine-grained multi-modal representations of entities and enhance the MMKGC performance. Motivated by the tokenization technology, MyGO tokenizes multi-modal entity information as fine-grained discrete tokens and learns entity representations with a cross-modal entity encoder. To further augment the multi-modal representations, MyGO incorporates fine-grained contrastive learning to highlight the specificity of the entity representations. Experiments on standard MMKGC benchmarks reveal that our method surpasses 19 of the latest models, underlining its superior performance. Code and data can be found in https://github.com/zjukg/MyGO<br/>

SUPPLEMENTAL MATERIAL

Coming soon ....

REFERENCES ()

CITATIONS ()

EXTERNAL LINKS

OPENAIRE - Products

PlumX Metrics

Tokenization, Fusion, and Augmentation: Towards Fine-grained Multi-modal Entity Representation

RECOMMENDATIONS

FAIR ASSESSMENT

Coming soon ....

JUPYTER LAB

Coming soon ....