A Benchmark Dataset of Check-Worthy Factual Claims

Fatma Arslan, Naeemul Hassan, Chengkai Li, Mark Tremayne

Paper type: Dataset

Keywords: building, claims, communities, elections, humans, sources, traditional

2020-06-09 P4 (23:00-00:00 GMT) [Zoom] [Cal]

Abstract: In this paper we present the ClaimBuster dataset of 23,533 statements extracted from all U.S. general election presidential debates and annotated by human coders. The ClaimBuster dataset can be leveraged in building computational methods to identify claims worth fact-checking from the myriad of sources of digital or traditional media. The ClaimBuster dataset is publicly available to the research community, and it can be found at http://doi.org/10.5281/zenodo.3609356.

Similar Papers

Characterizing the Social Media News Sphere through User Co-Sharing Practices
Mattia Samory , Vartan Kesiz Abnousi , Tanushree Mitra
Modeling and Measuring Expressed (Dis)belief in (Mis)information
Shan Jiang , Miriam Metzger , Andrew Flanagin , Christo Wilson