Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
724
datasets available to search
ShareScore release 0.9.0
Dataset results
724 results for “german”
Flexibilization or biomethane upgrading? Investment preference of German biogas plant operators for the follow-up of guaranteed feed-in tariffs
<p><span>This dataset reports the results of a discrete choice experiment with 183 German biogas plant operators designed to elicit the respondents' plans for biogas utilization pathways after the end of guaranteed feed-in tariffs. Participants could choose between 'flexibilization' for demand-based electricity generation and conversion to biomethane upgrading for direct feed-in into the natural gas grid. A binomial logit model revealed a 37% probability of switching to biomethane upgrading. These plants are characterized by higher capacities, several involved shareholders, secured succession, costly digestate disposal and belonging to the upper performance quartile. Mixed logit estimations conducted separately for the two investment concepts revealed a very high overall willingness to invest: 71% for flexibilization and 82% for biomethane upgrading. The respondents demand a return on investment of 19% for flexibilization and 26% for biomethane upgrading. Within the flexibilization, twofold overbuilding (installed capacity equals 2 times the rated power) is clearly preferred to fivefold overbuilding. For the biomethane upgrading, private ownership of the upgrading plant is preferred to a joint investment in a central upgrading facility. Limiting the use of energy crops reduces the propensity to invest in both models, while a longer utilization period enhances it. The respondents consider lack of planning reliability as the biggest obstacle to invest, followed by long approval procedures and high investment costs due to restrictive legal requirements.</span></p>
Form and Function of Polysemic Reflexive Markers in German, with Special Regard to Central Bavarian Varieties
<p>This upload includes the data, its analysis and the related questionnaire on the use of reflexive an reciprocal forms in central Bavarian varieties. The questionnaire (n=218) was part of a study that aimed at examining the actual use of reflexive and reciprocal markers in central Bavarian varieties as well as in Standard German.</p>
Speyer um 1200, St German, Entwurf I
Source: Objaverse 1.0 / Sketchfab
Historically Accurate German Messer
Source: Objaverse 1.0 / Sketchfab
pbr ww2 german Grenade
world war 2 german grenade made in suptaince painter Source: Objaverse 1.0 / Sketchfab
Paris catacombs - German bunker (Polycam)
German arrows and signs on the wall in the old German bunker, in the catacombs of Paris. Once can also see "Rauchen Verboten" (Smoking is prohibited) on the top of the scan. Scanned with the iPhone12 Pro and Polycam. Please feel free to follow my collections of daily scans ([link](https://skfb.ly/6YuwK)) as well as my scans in San Francisco, Paris, or in the catacombs: [link](https://sketchfab.com/edemaistre/collections) Source: Objaverse 1.0 / Sketchfab
Pfennig of the German Democratic Republic
Денежная единица: пфенниг. Номинал: 10. Год выпуска: 1971 год. Материал: алюминий. Гурт гладкий. Диаметр: 21 мм. Герб ГДР. Надпись: DEUTSCHE DEMOKRATISCHE REPUBLIK/Monetary unit: pfennig. Denomination: 10. Year of issue: 1971. Material: aluminum. The edge is smooth. Diameter: 21 mm. Coat of arms of the GDR. Caption: DEUTSCHE DEMOKRATISCHE REPUBLIK. A random find in the vicinity of Sterlitamak (Russia, Republic of Bashkortostan). Source: Objaverse 1.0 / Sketchfab
Flexibility Matters: Assessing the Flexibility Impact of Small and Medium Enterprises on the German Energy Transition: Dataset for Reference Scenario
<p>This repo contains the dataset for reference scenario for the paper</p> <p><strong><em>Flexibility Matters: Assessing the Flexibility Impact of Small and Medium Enterprises on the German Energy Transition</em></strong></p> <p>More information can be found here: https://github.com/AnasAbuzayed/SME_Flexibility</p> <p> </p>
GMHP7k: A corpus of german misogynistic hatespeech posts
<p>We provide a german corpus consisting of 7,061 posts authored by users of social media platforms. A group of volunteers annotated each post according to hatespeech and misogynistic/misogynous hatespeech in a binary fashion. The interrater reliability over all annotators according to Fleiss’ Kappa is 0.6409 for hatespeech and 0.8258 for misogynistic hatespeech. Furthermore, baseline measurements with machine learning based text classification with BERT are presented. Initial experiments with the corpus achieve macro average F1-scores up to 0.79 for hatespeech and 0.75 for misogynistic hatespeech. The dataset of the corpus on German Misogynistic Hatespeech Posts (GMHP7k) is publicly available.</p>
Supplementary Material: MyPyPSA-Ger: Introducing CO2 taxes on a multi-regional myopic roadmap of the German electricity system towards achieving the 1.5 °C target by 2050
<p>Supplementary Material to run the open-source model MyPyPSA-Ger</p>
The German research-performing pharmaceutical industry on Twitter
<p>Science communication undergraduate student. </p>
Data of the Shared Task on the Disambiguation of German Verbal Idioms at KONVENS 2021
<p>This dataset was used in the Shared Task on the Disambiguation of German Verbal Idioms (VID) at <a href="https://konvens2021.phil.hhu.de/">KONVENS 2021</a>. For further details, please refer to the description paper of the shared task:</p> <blockquote> <p>Ehren, Rafael, Timm Lichte, Jakub Waszczuk & Laura Kallmeyer. 2021. Shared Task on the Disambiguation of German Verbal Idioms at KONVENS 2021. In Proceedings of the Shared Task on the Disambiguation of German Verbal Idioms at KONVENS 2021. <a href="https://doi.org/10.5281/zenodo.5730322">https://doi.org/10.5281/zenodo.5730322</a>. <a href="https://konvens.org/proceedings/2021/index.html">https://konvens.org/proceedings/2021/index.html</a>.</p> </blockquote> <p><strong>Please cite this paper when using the dataset.</strong></p> <p>The content of the zip file is identical to that of the data directory in the <a href="https://github.com/rafehr/vid-disambiguation-sharedtask/tree/3ceb0bb423fa73e70ac018d4e02063ae449b4542">Github repository of the shared task</a>.</p> <p>The dataset consists of 9901 instances of a German VID type or its literal counterpart in context. The set of VID types was pre-selected, thus it constitutes a lexical sample data set. It is a merger of two datasets:</p> <ul> <li><a href="https://www.aclweb.org/anthology/2020.figlang-1.29.pdf">COLF-VID</a> (instances with T*)</li> <li><a href="https://www.aclweb.org/anthology/S13-2007.pdf">German SemEval-2013 task 5b</a> (instances with S*)</li> </ul> <p>The data comes in tsv files and every line has the following format:</p> <pre><code>Instance_ID \t VID_type \t label \t text</code></pre> <p>Consider this example:</p> <pre><code>T890202.28.4077 in wasser fallen figuratively Der Streit ums Hormonfleisch zwischen USA und EG provozierte den Polizeieinsatz . Aber nicht nur der Steakverkauf , auch die Aktionen gegen den Hormonstand , auf die sich Gruppen der Bauernopposition schon vorbereitet hatten , <b>fielen</b> <b>ins</b> <b>Wasser</b> . Die Fleischexporteure der USA wollten ihrerseits die " Grüne Woche " zur " Aufklärung " nutzen .</code></pre> <p>So the first column contains the ID (T890202.28.4077 in the example), the second the VID type (in wasser fallen), the third the label (figuratively) and the fourth the sentence with either the instance of the VID type or its literal counterpart (and two additional context sentences). The parts of the target expression are marked with the <b> tag (<b>fielen</b> <b>ins</b> <b>Wasser</b>). There are four possible labels:</p> <ul> <li>figuratively</li> <li>literally</li> <li>undecidable</li> <li>both</li> </ul> <p>The first two should be self-explanatory. The label undecidable was used by the annotators if it was not possible to disambiguate an instance given the context. The label both was applied when both the literal and the idiomatic readings were active.</p>
Dataset Compliance study among German blood donors 2020 - Undisclosed sexual risk exposures
<p>The dataset contains data of a nationwide anonymous online survey among German (whole) blood donors, which was performed to quantify and to investigate non-compliance of donors with deferral criteria for sexual risk exposures. It covers anonymous sociodemografic information, sexual risk exposures prior to last donation, and perception of donation-related factors that could potentially be related to compliance (e.g. confidentiality, donor education).</p> <p>The file contains data (sheet "data") as well as the description of variable content and coding (sheet "variables").</p>
Dataset for documents referenced as part of my research on immigration in the German context
<p>A combination of legal documents, surveys and reports (published by organisations affiliated to the EU or the German state and non-governmental agencies/institutes. These are open access documents downloaded either directly from the owner's website or other online platforms.</p> <p><strong>Declaration: I am not the author or owner of any of the documents uploaded here.</strong></p> <p>They have been uploaded here as part of the Marie Curie grant stipulations to make the research data open access.</p> <p>Documents are in German and English.</p> <p>MSCA Project: RE-NUP: Spousal Reunification and Integration Laws in Europe</p> <p>Grant agreement No. 890826</p>
German Argument Similarity dataset (GerArgSimilarity)
<p>This is the German-language argument similarity dataset accompanying our paper titled <em>Argument Similarity Assessment in German for Intelligent Tutoring: Crowdsourced Dataset and First Experiments</em> (to be presented at LREC 2022). It consists of 2940 argumentative pairs of text snippets in German which are annotated for argument similarity on a scale from 0 (<em>no similarity</em>) to 4 (<em>high similarity</em>). The snippets cover arguments on 3 discussion topics. Please refer to the paper for details.</p>
Clean OpenLegalData - German
<p>This dataset of German court cases is obtained from <a href="https://arxiv.org/abs/2005.13342">OpenLegalData</a>. The data received clearly mentioned information such as <em>court</em>, <em>level of appeal</em>, and <em>ECLI</em> (<em>European Case Law Identifier</em>). However, <em>tenor</em>, <em>tatbestand</em>, <em>gründe</em>, and <em>entscheidungsgründe</em> were available only in HTML format. Out of over <em>100000</em> extracted cases, we were able to parse <em>43337</em> HTML only due to structural problems with HTML content.</p> <p>The resulting dataset is approximate <em>~1.1 GBs</em> with <em>43337</em> rows and has the following <em>12</em> features:</p> <table> <tbody> <tr> <td><strong>Feature</strong></td> <td><strong>Total</strong></td> <td><strong>Example content</strong></td> </tr> <tr> <td>id</td> <td>43337</td> <td>127981</td> </tr> <tr> <td>slug</td> <td>43337</td> <td>ag-volklingen-2002-07-10-5c-c-24102</td> </tr> <tr> <td>ecli</td> <td>10831</td> <td>NaN</td> </tr> <tr> <td>date</td> <td>43337</td> <td>2002-07-10</td> </tr> <tr> <td>court</td> <td>43337</td> <td>Amtsgericht Völklingen</td> </tr> <tr> <td>jurisdiction</td> <td>43337</td> <td>Ordentliche Gerichtsbarkeit</td> </tr> <tr> <td>level_of_appeal</td> <td>43337</td> <td>Amtsgericht</td> </tr> <tr> <td>type</td> <td>43337</td> <td>Urteil</td> </tr> <tr> <td>tenor</td> <td>36282</td> <td>1. Die Beklagten werden als Gesamtschuldner verurteilt, an die ...</td> </tr> <tr> <td>tatbestand</td> <td>24243</td> <td>Auf die Darstellung des Tatbestandes wird gemäß § 313 Abs ...</td> </tr> <tr> <td>gründe</td> <td>27144</td> <td>Die Klage ist zulässig und begründet. Die Klägerin kann von de...</td> </tr> <tr> <td>entscheidungsgründe</td> <td>24038</td> <td>Die Klage ist zulässig und begründet. Die Klägerin kann von de...</td> </tr> </tbody> </table> <p><strong>Important</strong>: Make sure to use "<strong>|</strong>"<strong> </strong><em>(pipe-symbol)</em> as CSV separator.</p> <p><strong>Example</strong>:</p> <pre><code class="language-python">data = pd.read_csv('clean_OLD.csv', sep="|")</code></pre> <p> </p>
Role of Social Aggregation in the Fitness Cost of Pyrethroid-resistant German Cockroaches
<p>The data reported in the dataset refer to fitness cost of German cockroach and the effect of group size on biological attributes of field German cockroach </p>
Risk Classification of contaminates sites in Anderstorp using the German Einzelfallbewertung Altlastenstandorte (EB) method from the Hessian Agency for Nature Conservation, Environment and Geology (HLNUG)
<p>This data set includes the documents for risk classifying contaminated sites in Anderstorp, Sweden using the German Einzelfallbewertung Altlastenstandorte (EB) method from the Hessian Agency for Nature Conservation, Environment, and Geology (HLNUG).</p>
Speyer um 800, St. German, Entwurf I
Source: Objaverse 1.0 / Sketchfab
Supplementary material 1 from: Wesener T, Voigtländer K, Decker P, Oeyen JF, Spelda J, Lindner N (2015) First results of the German Barcode of Life (GBOL) – Myriapoda project: Cryptic lineages in German Stenotaenia linearis (Koch, 1835) (Chilopoda, Geophilomorpha). In: Tuf IH, Tajovský K (Eds) Proceedings of the 16th International Congress of Myriapodology, Olomouc, Czech Republic. ZooKeys 510: 15-29. https://doi.org/10.3897/zookeys.510.8852
Table. Estimates of Evolutionary Divergence between Sequences: Explanation note: The number of base differences per site from between sequences are shown. The analysis involved 45 nucleotide sequences. Codon positions included were 1st+2nd+3rd+Noncoding. All ambiguous positions were removed for each sequence pair. There were a total of 658 positions in the final dataset. Evolutionary analyses were conducted in MEGA6.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.