zenodoopen
TIGQA: An Expert-Annotated Question-Answering Dataset in Tigrinya
<div> <h3>What is TIGQA?</h3> <a href="https://github.com/hailaykidu/TigQA-Dataset#what-is-tigqa"></a></div> <p>TigQA is an expert-annotated dataset in Tigrinya, a low-resource language spoken by approximately 10 million speakers in Eritrea and the Tigray region of Ethiopia. Our proposed SQuAD-like dataset contains 2,685 question-answer pairs covering 122 diverse topics such as climate, water, and traffic. These pairs are from 537 context paragraphs in publicly accessible Tigrinya and Biology books, with the answers being provided by teachers from the region.</p> <div> </div> <div> </div>
ShareScore
36/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 8
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 0