Skip to main content
zenodoopen

Manual Dependency Annotation of Three German Text Extracts from the Project hermA (Gold Standard Data)

<p>This dataset was created in the digital humanities project hermA (www.herma.uni-hamburg.de) and comprises annotated extracts of the following three texts:</p> <ul> <li>Modern literature (Lit2009): novel <em>Corpus Delicti: Ein Prozess </em>by German author Juli Zeh, published in Frankfurt/Main in 2009.</li> <li>Non-contemporary literature (Lit1850):<em> Eine Frauenfahrt um die Welt </em>(&#39;A woman&rsquo;s journey around the world&#39;) by Austrian author Ida Pfeiffer (1850). Full text available at Deutsches Textarchiv: http://www.deutschestextarchiv.de/pfeiffer_frauenfahrt01_1850/6.</li> <li>Modern academic writing (Aca2009): <em>Stand, M&ouml;glichkeiten und Grenzen der Telemedizin in Deutschland</em> (&#39;Telemedicine in Germany: status, chances and limits&#39;) by R&uuml;diger Klar and Ernst Pelikan, published in Bundesgesundheitsblatt (&#39;Federal Health Gazette&#39;) in 2009. DOI 10.1007/s00103-009-0787-7.</li> </ul> <p>The texts are annotated for part-of-speech and dependency syntax and are made available in CoNLL file format. We describe the annotation process and report inter-annotator agreements in:</p> <p>Adelmann, Benedikt, Melanie Andresen, Wolfgang Menzel &amp; Heike Zinsmeister. 2018. Evaluation of Out-Of Domain Dependency Parsing for its Application in a Digital Humanities Project. <em>Proceedings of the 14th Conference on Natural Language Processing (KONVENS 2018)</em>. Vienna, Austria.</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
8
Access
16
Reuse readiness
8
Engagement
0