Skip to main content
zenodoopen

TIGQA: An Expert-Annotated Question-Answering Dataset in Tigrinya

<div> <h3>What is TIGQA?</h3> <a href="https://github.com/hailaykidu/TigQA-Dataset#what-is-tigqa"></a></div> <p>TigQA is an expert-annotated dataset in Tigrinya, a low-resource language spoken by approximately 10 million speakers in Eritrea and the Tigray region of Ethiopia. Our proposed SQuAD-like dataset contains 2,685 question-answer pairs covering 122 diverse topics such as climate, water, and traffic. These pairs are from 537 context paragraphs in publicly accessible Tigrinya and Biology books, with the answers being provided by teachers from the region.</p> <div>&nbsp;</div> <div>&nbsp;</div>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
8
Access
16
Reuse readiness
8
Engagement
0