Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
724
datasets available to search
ShareScore release 0.9.0
Dataset results
724 results for “german”
German Stoneware
_Shown here are some of the distinctive styles that belong to the German stoneware tradition of the sixteenth and seventeenth centuries. The tall "Schnellen" (tankards) of Siegburg, made of white-clay, are characterized by applied round medallions and panels with bilical or mythological subjects. Further south in the Westerwald region the vessels had distinctive ovoid bodies and narrow nects. They were decorated with a regular diaper pattern and stamped leaf and floral motifs, and cobalt blue and manganese purple glazes. A more ambitious use of color was found on the broad tankards made in Kreussen, South Germany._  Source: Objaverse 1.0 / Sketchfab
Kappmmesser М-1937/I (German gravity knife)
German gunsmiths managed to create a simple, compact and reliable knife, which was destined not only to survive the Third Reich, but also to become one of the most famous army knives in the world, which is still in service with the Bundeswehr and some NATO countries. In total, five different modifications of this knife are known, of which the first two were produced during the period of the existence of Nazi Germany, the other three - in the post-war period. The first model of sling cutter "M-1937", or Type I FKm was produced in the period from 1937 to 1941 in Solingen, at the enterprises of the German companies "Paul Weyersberg & Co" and "SMF" (Solinger Metallwaffenfabrik Stoecker & Co).  * [Additional screenshots](http://www.artstation.com/artwork/q9gwVP) * [Second type of knife (M-1937/II)](http://skfb.ly/opEZx) * [English experimental knife](http://skfb.ly/opHt6) Source: Objaverse 1.0 / Sketchfab
Air-Filtration unit from a German bunker WWII
Many German bunkers from WWII had these anti-gas air filtration systems installed that had to manually be operated to pump outside air through an air filter in to the bunker. The round pressure-valves in the walls can be found in most bunkers. These are called "uberdruckventil" and that basically means over-pressure-valve and would automatically let air out in case of overpressure inside the bunker. Source: Objaverse 1.0 / Sketchfab
German Dagger \ Sword
Hi This is a 3D model inspired by a 16th century German Dagger. The model's fully quad based with a clean topology. It is also fully textured with PBR textures. I have included two types of this model: Normal Polycount : Which Contains 4692 Verts, 4592 Faces, 9180 Tries. Lower Poly: Which Contains 2554 Verts, 2498 Faces, 4988 Tries The texture maps are all 2048 By 2048, the included maps are: ___Diffuse Map. ___Metalic Map. ___Roughness Map. ___Normal Map. ___Edge Map. ___Ambient Occlusion Map. Stay safe. Source: Objaverse 1.0 / Sketchfab
German Hand Grenades - Stielhandgranate 24
2 German Hand Grenades -*not totally historically accurate* Based on a Stielhandgranate 24 * **Light Grenade** : 386 vertices * **Heavy Grenade **: 1,454 vertices Source: Objaverse 1.0 / Sketchfab
Paris catacombs - in the German bunker
A portion of the German bunker with some cool murals. Scanned with the iPhone12 Pro and 3dScannerApp. Please feel free to follow my collections of daily scans ([link](https://skfb.ly/6YuwK)) as well as my scans in San Francisco, Paris, or in the catacombs: [link](https://sketchfab.com/edemaistre/collections) Source: Objaverse 1.0 / Sketchfab
1 Euro (German)
Another photogrammetry test. Photos taken with newly built coin holders. Setup: OpenScanner (www.openscan.eu) in a light tent with 3 soft boxes, Sony Alpha 77 II, SAL100M28 (100mm / f2.8 Macro), 3 chunks with 539 RAW images total, processed in lightroom, Agisoft Metashape, minor texture editing in Photoshop. I finally solved the problem with hard chunk borders, therefore now missing some detail und depth in the model. Source: Objaverse 1.0 / Sketchfab
German Parliamentary Speeches
<div> <div>These datasets are part of the thesis entitled "A natural language processing analysis of parliamentary speeches in the German Bundestag from 1949 to 2023 with regards to gender equality and women's politics". The focus of the thesis lies on contrasting the thematic preferences of men and women, as well as their connection to women's political issues based on their speeches and highlighting their historical development. Furthermore, the speakers’ reception in parliament is analyzed and the gender-specific distinguishability of speech style is verified. The data is available in a pandas DataFrame format.</div> <div> </div> <div> <div> <div>The plenary protocol data has been extracted from the following two sources:</div> <div> <ul> <li>Blaette, Andreas (2017): GermaParl. Corpus of Plenary Protocols of the German Bundestag. TEI files, availables at: <a href="https://github.com/PolMine/GermaParlTEI">https://github.com/PolMine/GermaParlTEI</a></li> </ul> </div> <div> <ul> <li>Deutscher Bundestag. Open Data - Plenarprotokolle der 20. Wahlperiode und Stammdaten aller Abgeordneten seit 1949. <a href="https://www.bundestag.de/services/opendata">https://www.bundestag.de/services/opendata</a></li> </ul> </div> <div> </div> <div>The following datasets are contained:</div> <div> <ul> <li><strong>data_merged_revised.pkl:</strong> DataFrame containing the extracted and cleaned data from the sources mentioned above </li> <li><strong>data_topics_revised.pkl</strong>: DataFrame containing the data from the sources mentioned above and the topic distributions after topic modelling</li> <li><strong>data_undersampled_processed.pkl</strong>: Undersampled and preprocessed dataset which can be used for training classification models</li> </ul> </div> </div> </div> </div>
Data Sets for Evaluation of the Psychometric Properties and Validity of the German Version of the Process Model of Emotion Regulation Scale (PMERQ)
<p>Data files relate to an investigation of the psychometric properties of the German Version of the Process Model of Emotion Regulation Scale (PMERQ). Data set 1 (pmerq_1) contains information regarding the age, gender, ethnicity, and educational status of participants. In addition, responses to the 45 items of the initial translation of the 10-scale PMERQ are included. Data set 2 (pmerq_2) contains identical sociodemographic variables and responses to the 45 items of the revised translation of the 10-scale PMERQ. In addition, data set 2 contains responses to the 16-item German Interpersonal Emotion Regulation Questionnaire (IERQ), the 10-item German Emotion Regulation Questionnaire (ERQ), , the German version of the 10-item Big Five Inventory-10 (BFI-10), the 4-item German version of the Patient Health Questionnaire-4 (PHQ-4), the German version of the Satisfaction with Life Scale (SWLS), and the 17-item German Social Desirability Scale-17 (SES-17). Data set 2 (pmerq_2) contains identical sociodemographic variables and responses to the 45 items of the readability-improved translation of the 10-scale PMERQ. In addition, data set 2 contains responses to the German version of the Satisfaction with Life Scale (SWLS) and the German version of the 9-items UCLA Loneliness Scale (UCLA).</p>
The genitive alternation in German (dataset)
<p>An annotated dataset, documentation and an R script for reproducing the analysis reported in <a href="https://doi.org/10.1515/cllt-2024-0017">Kopf, Kristin & Felix Bildhauer. 2024. The genitive alternation in German. Corpus Linguistics and Linguistic Theory. Published online November 13, 2024.</a></p> <h2>Contents</h2> <p><code>genitive_alternation_cllt.tsv</code>: a dataset containing 14,684 instances of nouns with either a genitive modifier or a <em>von</em>-modifier (one per line, with several layers of annotation in individual columns). It was used for analyzing the genitive alternation in German, as reported in <a href="https://doi.org/10.1515/cllt-2024-0017">Kopf & Bildhauer (2024)</a>. Tab-separated values, utf-8.</p> <p><code>cllt.standalone.R</code>: an R script that reproduces the analysis from Kopf & Bildhauer (2024)</p> <p><code>documentation.markdown</code>: dataset documentation</p> <h2>License</h2> <p>The dataset includes data from two different sources, as indicated in the column "License", to which different licences apply:</p> <p>Data from the German Reference Corpus DeReKo are subject to the <a href="https://www2.ids-mannheim.de/cosmas2/projekt/register/license_agreement.html">End User Agreement for the Use of the German Reference Corpus DeReKo</a> (version of 2018-05-24). In particular, commercial use is excluded. Furthermore, this data may not be passed on to third parties or published without the written consent of the Leibniz Institute for the German Language. This does not apply to quotations and excerpts.</p> <p>Data from the DECOW16 web corpus are subject to the <a href="https://www.webcorpora.org/license.php">COW TERMS OF USE</a> (version 2.1, 2014-12-16). In particular, commercial use is excluded.</p> <p>The column "Licence" states the applicable licence for each data point.</p> <p>By downloading the dataset, the user agrees to use the data in accordance with the applicable license.</p>
EncycNet: A Knowledge Graph of Historical German Encyclopedias
<p><strong>EncycNet</strong> is an automatically constructed knowledge graph that uses historical German encyclopedias as a data source. You can read more about the project here: <a href="https://encycnet.github.io/">https://encycnet.github.io/</a></p> <p>The first version of the graph (0.1) contains <em>Meyers Großes Konversations-Lexikon </em>(1905), with 5 more encyclopedias to be added in 2024. Formatted in RDF Turtle.</p> <p>The second version of the graph (0.2) contains 3 encyclopedias in total, meaning triples were changed to quads (see code example at <a href="https://github.com/EncycNet/Encyc-Relations">https://github.com/EncycNet/Encyc-Relations</a>):</p> <ul> <li><em>Meyers Großes Konversations-Lexikon </em>(1905)</li> <li><em>Herders Conversations-Lexikon</em> (1854)</li> <li><em>Brockhaus Bilder-Conversations-Lexikon</em> (1837)</li> </ul> <p>The second version also handled some bug fixes concerning wrongly matched Wikidata entries.</p> <p>The third version (0.3) now additionally contains</p> <ul> <li><em>Brockhaus Conversations-Lexikon oder kurzgefaßtes Handwörterbuch</em> (1809)</li> </ul> <p>It also impoved some namespace issues and introduced <a href="https://www.wikidata.org/wiki/Property:P8371">P8371</a> as encyclopedic refeferences in this case.</p>
Adressbuch/Directory German Migrants in Paris 1854 - full sources
<p>All informations in this collection are extracted from the following book :</p> <p>F-A. Kronauge. Adreßbuch der Deutschen in Paris für das Jahr 1854 oder vollständiges Adreßverzeichniß aller in Paris und seinen Vorstädten wohnenden selbständigen Deutschen in alphabetischer Ordnung. (1854). <a href="https://bibliotheques-specialisees.paris.fr/ark:/73873/pf0000884072">https://bibliotheques-specialisees.paris.fr/ark:/73873/pf0000884072 :</a></p> <ol> <li>HD.zip = pages of the book in high quality</li> <li>SD.zip = pages of the book in standard quality</li> <li>pics_metadata.csv info about the pics size and dpi</li> <li>OCR.zip : text extracted from the pictures</li> <li>adressbuch1854.json : extracted data encoded in JSON</li> <li>The MySQL associated database for the web site : adressbuch1854.sql.zip</li> </ol> <p>The website can be found here: <a href="https://adressbuch1854.dhi-paris.fr/">https://adressbuch1854.dhi-paris.fr/</a>.</p> <p><strong>Update of the database on December 12, 2022 (v2)</strong></p> <p>1) adressbuch1854_v2.json - extracted data encoded in JSON</p> <p>2) adressbuch1854_v2.xml - extracted data encoded in XML</p> <p>3) adressbuch1854_v2.sql - database in MySQL</p> <p>4) adressbuch1854_v2.csv - datasheet with most important data tables - complete version on website</p> <p><strong>Update of the OCR-files on March, 27, 2025 (v2)</strong></p> <p>We have improved the OCR of the scanned images in order to achieve better recognition of the combination of Fraktur and Latin scripts.</p> <p>1) ocr_v2.zip: text extracted from the pictures</p>
Kobalt_RST (RST German Learner Treebank): Die Annotation von rhetorischen Strukturen im Kobalt-DaF-Korpus
<p>Das <a href="https://www.linguistik.hu-berlin.de/de/institut/professuren/korpuslinguistik/forschung/kobalt-daf">Kobalt-DaF-Korpus</a> ist ein systematisch erhobenes und tief annotiertes Deutschlernerkorpus, welches 80 deutschsprachige argumentative Texte von deutschen L1-Sprecher:innen und Deutschlerner:innen unterschiedlicher L1 enthält. Dieses Repositorium stellt eine zusätzliche Annotation des Kobalt-DaF-Korpus bzgl. rhetorischer Strukturen frei zur Verfügung. Folgende Informationen sind hier zu finden: (1) Die Darstellung des Annotationsprozesses (Annotationsframework, -richtlinie, und -verfahren). (2) Die annotierten rs3-Dateien.</p> <p>*Versionshinweise: Bislang sind ausschließlich die Texte der chinesischen Deutschlerner:innen und der deutschen L1-Sprecher:innen (insgesamt 40 Texte) verfügbar. Die Annotation der übrigen Texte folgt demnächst. </p> <p>*Die Annotationsarbeit wurde gefördert durch das Chinese Scholarship Council und die Deutsche Forschungsgemeinschaft (DFG) – SFB 1412, 416591334.</p>
RefWUG: Diachronic Reference Word Usage Graphs for German
<p>This data collection contains diachronic Word Usage Graphs (WUGs) for German created with reference use sampling. Find a description of the data format, code to process the data and further datasets on the <a href="https://www.ims.uni-stuttgart.de/data/wugs">WUGsite</a>.</p> <p>Please find more information on the provided data in the paper referenced below.</p> <p>Version: 1.1.0, 15.12.2021.</p> <p><strong>Reference</strong></p> <p>Dominik Schlechtweg and Sabine Schulte im Walde. submitted. Clustering Word Usage Graphs: A Flexible Framework to Measure Changes in Contextual Word Meaning.</p>
Text of Wikisource pages of German magazine 'Die Gartenlaube'
<p>Text of all Gartenlaube pages transcribed in German Wikisource. Text parsed on 2021-12-13, the output is combinend in separeted json files, each file per volume, starting 1853 and ending 1899. All 47 json files are compressed into one *.tar.xz file.</p> <p>The syntax of the json looks like:</p> <pre><code class="language-json"> [{"pageid" : {PAGEID}, "title" : {PAGETITLE}, "lastrevid" : {REVISIONID}, "proofread" : {{JSON_OBJECT_Proofread_Status}} "html" : {HTML_OUTPUT}, "wikitext": {WIKI_MARKUP}, "plaintxt": {mwparserfromhell(WIKI_MARKUP).strip_code)} }]</code></pre>
Stocks data of Stocks in the German Prime Standard 2018
<p>Data was used in publication:</p> <p>Thrun, M. C.: Exploiting Distance-Based Structures in Data Using an Explainable AI for Stock Picking, Information, MDPI, 2022.</p> <p>Using this data please cite the publication.</p>
Bathymetry data from detonation scars in the Fehmarnbelt, German Baltic Sea.
<p>The bathymetric data were collected on the 27<sup>th</sup> of June 2020 as underway research data on a 1.5 km track during the cruise EMB239 with the German research vessel Elisabeth Mann Borgese. The objective of the data acquisition was to survey seafloor scars resulting from the controlled detonation of ground mines. For data acquisition, the ship’s hull-mounted Sonic 2024 (R2Sonic Inc.) multibeam echosounder was used. The raw sonar data were loaded in Qimera v2.4.3 (Quality Positioning Services B.V.) and automatically processed to compute sounding footprint location under consideration of sound velocity, position, motion, and heading information. To make the data usable without any specific software, the georeferenced soundings were exported without any bathymetric data cleaning as comma-separated ASCII file in the coordinate reference system EPSG: 32632 - WGS84 / UTM zone 32N.</p> <p>For more details please refer to Papenmeier, S., Darr, A., Feldens, P. (in prep): Geomorphological data from detonation craters in the Fehmarnbelt, German Baltic Sea.</p>
PDF and PSD files of DiapixGEtv picture materials – German version adapted to elicit tense vowels
<p>This zipped folder contains PDF and PSD files of our translation of the DiapixUK picture materials by Baker & Hazan (2011) - adapted to elicit tense vowels in German (DiapixGEtv).</p> <p>For our purpose we only modified the written information when making adjustments to the original DiapixUK versions of the files. Our aim was to include as many tense vowels [i:, e:, a:, o:, u:] as possible while remaining subtle enough so as not to alert the participants to our research focus . The additional documentation contains information on the translation process and an overview over the adapted text parts containing target vowels in stressed syllables including the target word with its target vowel and pronunciation as well as its meaning in English.</p> <p>In order to facilitate further adaptation/modification each item in the PSD files is in a separate layer.</p> <p>Funded by the Deutsche Forschungsgemeinschaft (DFG, German Research Foundation) – SFB 1412, 416591334</p>
SCoRe - Student Crowd Research. A german dataset of university students' written reflections on their participation in collaborative research-based learning focused on sustainability topics and using videos as a research tool.
<p>Schriftliche Reflexionen von 57 Studierenden als Prüfungsleistung im Rahmen des forschenden Lernens mit Video zu Nachhaltigkeitsthemen. Die Reflexionen wurden durch vorgegebene Fragen angeleitet. Die hier vorliegenden Texte sind, von den Studierenden selbst niedergeschriebene, Transkripte von Sprechtexten eines Self-Video-Casts. Es handelt sich dabei um eine benotete Prüfungsleistung einer universitätsübergreifenden Wahlpflicht-Lehrveranstaltung mit 1-3 Credit-Points.</p> <p>A dataset containing written reflections of 57 university students from several German universities and fields of study on their participation in collaborative research-based learning focused on sustainability topics and using videos as a research tool. These reflections were guided by given questions. The texts presented here are transcripts of spoken texts of self-video-casts, written down by the students themselves. The texts were graded as part of an inter-university elective course worth 1-3 credit points.</p>
Overview: GERMAN LANGUAGE OER FOR SOCIAL SCIENCE METHODS EDUCATION
<p>This document contains the corpus of identified OER as well as OER-like on social science research methods in the German language and the used codes to classify the identified resources. This document is provided for secondary use as well as expansion and revision.</p> <p>The collection and classification of the corresponding OER and OER-like was last updated in August 2021 and has to be discussed accordingly. OER and OER-like that were created later or were offline during data collection are not included.</p> <p>The codings provided are only covering content-based aspects and not feature e.g. like license framework or opportunity to comment and discuss.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.