CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection

Fanxiao Li; Jiaying Wu; Canyuan He; Wei Zhou

CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection

Fanxiao Li, Jiaying Wu, Canyuan He, Wei Zhou

Abstract

Multimodal large language models (MLLMs) have demonstrated impressive capabilities in visual reasoning and text generation. While previous studies have explored the application of MLLM for detecting out-of-context (OOC) misinformation, our empirical analysis reveals two persisting challenges of this paradigm. Evaluating the representative GPT-4o model on direct reasoning and evidence augmented reasoning, results indicate that MLLM struggle to capture the deeper relationships—specifically, cases in which the image and text are not directly connected but are associated through underlying semantic links. Moreover, noise in the evidence further impairs detection accuracy.To address these challenges, we propose CMIE, a novel OOC misinformation detection framework that incorporates a Coexistence Relationship Generation (CRG) strategy and an Association Scoring (AS) mechanism. CMIE identifies the underlying coexistence relationships between images and text, and selectively utilizes relevant evidence to enhance misinformation detection. Experimental results demonstrate that our approach outperforms existing methods.

Anthology ID:: 2025.findings-acl.487
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 9342–9354
Language:
URL:: https://preview.aclanthology.org/landing_page/2025.findings-acl.487/
DOI:
Bibkey:
Cite (ACL):: Fanxiao Li, Jiaying Wu, Canyuan He, and Wei Zhou. 2025. CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection. In Findings of the Association for Computational Linguistics: ACL 2025, pages 9342–9354, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection (Li et al., Findings 2025)
Copy Citation:
PDF:: https://preview.aclanthology.org/landing_page/2025.findings-acl.487.pdf

PDF Cite Search Fix data