Rafael Frade


2025

pdf bib
ClaimCatchers at SemEval-2025 Task 7: Sentence Transformers for Claim Retrieval
Rrubaa Panchendrarajan | Rafael Frade | Arkaitz Zubiaga
Proceedings of the 19th International Workshop on Semantic Evaluation (SemEval-2025)

Retrieving previously fact-checked claims from verified databases has become a crucial area of research in automated fact-checking, given the impracticality of manual verification of massive online content. To address this challenge, SemEval 2025 Task 7 focuses on multilingual previously fact-checked claim retrieval. This paper presents the experiments conducted for this task, evaluating the effectiveness of various sentence transformer models—ranging from 22M to 9B parameters—in conjunction with retrieval strategies such as nearest neighbor search and reranking techniques. Further, we explore the impact of learning context-specific text representation via finetuning these models. Our results demonstrate that smaller and medium-sized models, when optimized with effective finetuning and reranking, can achieve retrieval accuracy comparable to larger models, highlighting their potential for scalable and efficient misinformation detection.