Skip to main content
zenodoopen

ClaimBuster: A Benchmark Dataset of Check-worthy Factual Claims

<p>The ClaimBuster dataset consists of&nbsp;statements extracted from all U.S. general election presidential debates (1960-2016) along with human-annotated check-worthiness labels where each sentence is categorized into one of the three categories: non-factual statement, unimportant factual statement, and check-worthy factual statement.</p> <p>If you use this dataset, please cite the following paper:</p> <p>@inproceedings{arslan2020claimbuster,<br> &nbsp; &nbsp; title={{A Benchmark Dataset of Check-worthy Factual Claims}},<br> &nbsp; &nbsp; author={Arslan, Fatma and Hassan, Naeemul and Li, Chengkai and Tremayne, Mark },<br> &nbsp; &nbsp; booktitle={14th International AAAI Conference on Web and Social Media},<br> &nbsp; &nbsp; year={2020},<br> &nbsp; &nbsp; organization={AAAI}<br> }</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
8
Engagement
0

Topics