arXiv cs.CLSeptember 22, 2026
Knowing When to Trust Images: Reliability-Aware Multi-modal Entity Alignment
Excerpt
arXiv:2609.23267v1 Announce Type: new Abstract: The visual modality, i.e., images, plays a key role in multi-modal entity alignment (MMEA). Existing approaches often directly fuse the image with other modalities to align different entities. Although simple, such strategies overlook the potential noise in the images and their semantic misalignment with corresponding entities, resulting in suboptimal fusion and degraded performance. Addressing this, we propose a novel Reliability-Aware framework f