ClaimBuster: A Benchmark Dataset of Check-worthy Factual Claims
<p>The ClaimBuster dataset consists of statements extracted from all U.S. general election presidential debates (1960-2016) along with human-annotated check-worthiness labels where each sentence is categorized into one of the three categories: non-factual statement, unimportant factual statement, and check-worthy factual statement.</p> <p>If you use this dataset, please cite the following paper:</p> <p>@inproceedings{arslan2020claimbuster,<br> title={{A Benchmark Dataset of Check-worthy Factual Claims}},<br> author={Arslan, Fatma and Hassan, Naeemul and Li, Chengkai and Tremayne, Mark },<br> booktitle={14th International AAAI Conference on Web and Social Media},<br> year={2020},<br> organization={AAAI}<br> }</p>
ShareScore
36/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0