Skip to main content
zenodoopen

SIMPITIKI corpus for simplification in Italian

<p>SIMPITIKI is a Simplification corpus for Italian and it consists of two sets of simplified pairs: the first one is harvested from the Italian Wikipedia in a semi-automatic way; the second one is manually annotated sentence-by-sentence from documents in the administrative domain.</p> <p>For more details, see&nbsp;https://github.com/dhfbk/simpitiki</p>

ShareScore

44/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
8
Access
20
Reuse readiness
8
Engagement
4