zenodoopen
SIMPITIKI corpus for simplification in Italian
<p>SIMPITIKI is a Simplification corpus for Italian and it consists of two sets of simplified pairs: the first one is harvested from the Italian Wikipedia in a semi-automatic way; the second one is manually annotated sentence-by-sentence from documents in the administrative domain.</p> <p>For more details, see https://github.com/dhfbk/simpitiki</p>
ShareScore
44/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 4
- Harmonization
- 8
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4