Arafa is a large-scale Arabic fact-checking dataset created via an automated LLM pipeline. It contains 181,976 claim-evidence pairs categorized as supported, refuted, or insufficient, addressing the scarcity of resources for Arabic NLP.
HOW THIS AFFECTS YOU
●
builderYou can integrate more accurate Arabic fact-checking capabilities into your NLP products.
●
researcherYou can now benchmark Arabic fact-checking models on a large, structured dataset.