← Back to all articles
arXiv cs.CLSeptember 22, 2026

Knowing When to Trust Images: Reliability-Aware Multi-modal Entity Alignment

Excerpt

arXiv:2609.23267v1 Announce Type: new Abstract: The visual modality, i.e., images, plays a key role in multi-modal entity alignment (MMEA). Existing approaches often directly fuse the image with other modalities to align different entities. Although simple, such strategies overlook the potential noise in the images and their semantic misalignment with corresponding entities, resulting in suboptimal fusion and degraded performance. Addressing this, we propose a novel Reliability-Aware framework f