Knowing When to Trust Images: Reliability-Aware Multi-modal Entity Alignment
The visual modality, i.e., images, plays a key role in multi-modal entity alignment (MMEA). Existing approaches often directly fuse the image with other modalities to align different entities. Although simple, such strategies overlook the potential noise in the images and their semantic misalignment with corresponding entities, resulting in suboptimal fusion and degraded performance. Addressing…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.