Skip to main content
zenodoopen

dataset_eastern_german_crisis_discourse_1976-86

<p>This dataset has been considered&nbsp;suitable for exploring and analysing the Eastern German Crisis Discourse from 1976 to 1986, and consists of important speeches contained in five volumes of the protocol of the party congress of the Socialist Unity Party of Germany (SED). The speeches have been digitized by scanning them into PDF files and then converting them into machine-readable TEXT files using OCR software. These TEXT files have been processed with TagAnt (v.2.0.4&nbsp;Windows 10 64-bit),&nbsp;an annotation software. According to the data returned from&nbsp;AntConc (v.4.0.5&nbsp;Windows 10 64-bit) the&nbsp;dataset contains four corpora: a main corpus with a total of 184,750 tokens and 16,143 types, a &#39;corpus A&#39; with 70,533 tokens and 11,964 types, a &#39;corpus B&#39; with 65,757 tokens and 11,967 types, and a &#39;corpus C&#39; with 48,460 tokens and 8,145 types.</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
4
Harmonization
4
Access
20
Reuse readiness
8
Engagement
4