Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
8
datasets available to search
ShareScore release 0.9.0
Dataset results
8 results for “Europeana”
BotanicalEuropeana: Berlin-Dahlem, Europeana collection
Botanic Garden and Botanical Museum Berlin-Dahlem, via the Europeana portal: <p></p>https://www.europeana.eu/portal/en/collections/natural-history?f%5BDATA_PROVIDER%5D%5B%5D=Botanic+Garden+and+Botanical+Museum+Berlin-Dahlem&q=&view=grid<p></p>Botanic Garden and Botanical Museum Berlin-Dahlem, via the Europeana portal: <p></p>https://www.europeana.eu/portal/en/collections/natural-history?f%5BDATA_PROVIDER%5D%5B%5D=Botanic+Garden+and+Botanical+Museum+Berlin-Dahlem&q=&view=grid
Wikidata's linked data for cultural heritage digital resources: an evaluation based on the Europeana Data Model
<p>Wikidata is an open data source with many potential applications. Our study aims to evaluate the usability of Wikidata as a linked data source for acquiring richer descriptions of digital objects within the context of Europeana, a data aggregator from the cultural heritage domain. Specifically, we aim to crawl and convert Wikidata using the standard approaches and operations developed for the (Semantic) Web of Data, i.e. using technologies like linked data consumption and RDF(S)/OWL ontology expression and reasoning. We also seek to re-use existing “semantic” specifications, such as conversions to and from generic data models like Schema.org and SKOS. We have developed an experimental set-up and accompanying software to test the feasibility of this approach. We conclude that Wikidata’s linked data is able to express an interesting level of semantics for cultural heritage, but quality can still be improved and a human operator still must assist linked data applications to interpret Wikidata’s RDF.</p>
Europeana Sounds genres dataset
<p><a href="https://pro.europeana.eu/project/europeana-sounds">Europeana Sounds</a> was a project focused on accessing digital audio files. The current dataset aims to provide semistructured data for learning audio-representations using machine learning techniques.</p> <p>The motivation for this dataset is to allow experimentation with audio content and metadata. The current dataset is a subset of a dataset used in <a href="https://arxiv.org/pdf/2003.12265.pdf">this research paper</a> containing 24k audios belonging to different musical genres. Find more information about the Europeana Sounds project and dataset in <a href="http://www.ifs.tuwien.ac.at/~schindler/eusounds_challenge/">this website</a>, together with raw audio features.</p> <p><br> The audio files can be downloaded using the original media URL. Additional data about the object may be obtained in JSON or in RDF. In JSON, by querying with the ID on the <a href="https://pro.europeana.eu/page/record">Europeana Record API</a>. RDF can be obtained by requesting the URI with HTTP Content Negotiation for a well-known RDF serialization format. The URI also provides access to the object's page at Europeana, if HTTP Content Negotiation is not used.</p> <p>The objects were obtained using the Europeana Search API. More information about Europeana APIs can be found <a href="https://pro.europeana.eu/page/apis">here</a>.</p>
Data set: Europeana Metadata to Analyse The EU's Policy Effectiveness in Digitising European Heritage
<p>This dataset was produced as a part of a research MSc thesis project, “<em>Using Europeana Metadata to Analyse The EU’s Policy Effectiveness in Digitising European Heritage</em><em>” </em>published by the KU Leuven in joint cooperation with Europeana.</p> <p>The dataset contains the data from the Europeana API call, graphs created form the data and the code needed to collect the CH metadata from Europeana digital collections which was used to evaluate the policy aims used in the study. Within this repository one will find the code scripts used to make the API calls, three Raw data files and two cleaned data spreadsheets with graphs used in the published revised research thesis</p> <p>Links to the research produced with this dataset and the Original Thesis published by the KU Leuven can be found below.</p> <p>Research Report: Morgenstern, Paul Simon. (2022). Using Europeana Metadata to Analyse The EU's Policy Effectiveness in Digitising European Heritage. Zenodo. https://doi.org/10.5281/zenodo.7278645</p> <p> </p> <p>Thesis: Morgenstern, P., Truyen, F., & Nyi Nyi Htun, ’. (2022). Using Europeana Metadata to Analyse The EU’s Policy Effectiveness in Digitising European Heritage. KU Leuven. Faculteit Wetenschappen.</p>
University of Tartu Europeana Collection
Countless natural history treasures are deposited in museums across the world, many hidden away beyond easy access. The OpenUp! project creates a free access to these resources, offering over one million items belonging to the world__s biodiversity heritage. The objects made available through OpenUp! consist of high quality images, videos and sounds, as well as natural history artworks and specimens, and also include many items previously inaccessible to visitors. Information provided through OpenUp! is checked by scientists and made available through the Europeana portal at www.europeana.eu.
Automatic translation and multilingual cultural heritage retrieval: a case study with transcriptions in Europeana (dataset)
<p>The dataset contains all the data required to reproduce the experiments done in the paper "Automatic translation and multilingual cultural heritage retrieval: a case study with transcriptions in Europeana", published in the 25th International Conference on Theory and Practice of Digital Libraries (<a href="http://www.tpdl.eu/tpdl2021/">TPDL'21</a>). In that work we run an experiment using the Europeana CH digital library as a use case, and we evaluated the effectiveness of a multilingual information retrieval strategy using machine translations to English as pivot language. We used the CEF translation service (eTranslation) for the translation of queries and content to English (<a href="https://ec.europa.eu/cefdigital/wiki/display/CEFDIGITAL/eTranslation">https://ec.europa.eu/cefdigital/wiki/display/CEFDIGITAL/eTranslation</a>).</p> <p>The dataset is also available at <a href="https://rnd-2.eanadev.org/share/crosslingual-search/">https://rnd-2.eanadev.org/share/crosslingual-search/</a>, and it is organized in four main folders:</p> <ul> <li><strong>queries</strong>: sample of 68 queries and their translations to English. The queries were issued in languages other than English from the Europeana Portal, using the Europeana’s 1914-1918 thematic collection, between January and August 2019.</li> <li><strong>transcriptions</strong>: sample of 18,257 handwriting transcriptions and its translations to English. The transcriptions are taken from the Europeana 1914-1918 thematic collection, and obtained from the Transcribathon crowdsourcing platform (https://europeana.transcribathon.eu/).</li> <li><strong>solr_configuration</strong>: Apache Solr search engine configuration used in the experiments (which replicates the one used in Europeana).</li> <li><strong>results</strong>: manual evaluation of the query translations, and automatic evaluation of the multilingual retrieval.</li> </ul> <p> </p>
Implementation and evaluation of a multilingual search pilot in the Europeana digital library (dataset)
<p>The dataset contains the data required to reproduce the experiments done in the paper "Implementation and evaluation of a multilingual search pilot in the Europeana digital library", published in the 26th International Conference on Theory and Practice of Digital Libraries (<a href="http://tpdl2022.dei.unipd.it/">TPDL'22</a>). In that work we implemented a pilot applying query translation to English from the Spanish version of the website in order to surface results that have English metadata associated with them. The dataset is also available at <a href="https://rnd-2.eanadev.org/share/crosslingual_SpanishPilot/">https://rnd-2.eanadev.org/share/crosslingual_SpanishPilot/</a>, and it is organized in three main folders:</p> <ul> <li><strong>sample</strong>: stratified sample of 300 queries queries issued from the Europeana Spanish portal from 1st<br> December 2020 to 28th February 2021.</li> <li><strong>evaluation.translations: </strong>manual annotation of the quality of the identification of the language of the queries using Google Cloud Translation API, and the quality of the translation obtained using Google plus the CEF translation service (eTranslation).</li> <li><strong>evaluation.search_retrieval: </strong>manual annotation of the relevancy of the (binary) relevance of the documents that are retrieved by one system but not by the other (current monolingual version vs pilot) in their top ten.</li> </ul>
Domain-focused linked data crawling driven by a semantically defined frontier a cultural heritage case study in Europeana
<p>Supporting data referenced in the paper with the same title form ICADL 2020.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.