Semantic annotation of a part of the Italian Copyright Legislation
<p>The dataset is a structured JSONL file focusing on copyright law. Each entry contains key fields that annotate legal texts, mainly in Italian. These fields include:</p> <p>1. ID: A unique numerical identifier.</p> <p>2. Text: Contains the actual legal provisions.</p> <p>3. Chapter ID & Heading: Identifiers and titles for chapters, categorizing the legal text.</p> <p>4. Article and Paragraph ID: Further break down of the text into articles and paragraphs.</p> <p>5. Insertions: Highlights inserted text fragments in the legal text.</p> <p>6. References: Cites external references with URLs and descriptions.</p> <p>7. Entities: Labels sections of the text, identifying their beginning and ending offsets.</p> <p>8. Relations: Intended to describe relationships between entities, although this field is empty in the sample.</p> <p>9. Comments: A field for comments, also empty in the sample.</p> <p> </p>
ShareScore
44/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4