Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

688

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

688 results for “Abstract”

Learn how ShareScore rates datasets ↗
zenodo52/100

Dataset of "Marcus cross relation in the space of H-atom abstraction reactions boosted through off-diagonal thermodynamics"

<p>Proton-coupled electron transfer (PCET) and hydrogen-atom transfer (HAT) reactions play critical roles in biological processes and modern organic synthesis. The kinetics of these processes can align with the principles described in the renowned Marcus cross relation (MCR), a framework initially formulated to describe electron transfer mechanisms. The MCR provides an outstanding link between the kinetics of PCET/HAT reaction involving two distinct reactants and two related auxiliary self-exchange reactions &ndash; each between a molecule of one of the reactants and its coupled radical. In this study, we investigate the applicability and limitations of the canonical MCR across over 300 PCET and HAT reactions, providing a comprehensive theoretical analysis. Our findings reveal the need for an enhanced framework that incorporates &lsquo;off-diagonal&rsquo; thermodynamic factors&mdash;asynchronicity and frustration. Of these factors, asynchronicity, which quantifies the imbalance between the proton vs. electron transfer components of the reaction, is identified as the dominant contributor to the improved predictive accuracy of the MCR. Notably, the incorporation of off-diagonal thermodynamics yields a more pronounced enhancement for HAT reactions than for PCET reactions. This advancement offers a refined theoretical basis for understanding H-atom abstraction mechanisms and underscores the importance of off-diagonal effects in PCET/HAT chemistry.</p>

opencc-by-4.0Dec 2024View details →
zenodo48/100

Automated Literature Screening for Systematic Reviews: Dataset for Evaluation Against Human Title and Abstract and Full-Text Screening Decisions

<p>This Zenodo entry contains the supplementary material associated with the manuscript titled&nbsp;<em>Automated Literature Screening for Systematic Reviews: A 5-Tier Prompting Approach Meeting Cochrane&rsquo;s Sensitivity Requirement of Greater Than 0.99.</em> The paper will be presented at <a href="https://dbis.rwth-aachen.de/LLMs4MI2024/">LLMsMI 2024</a> in November 2024.</p> <p>A script is provided for replicating the executed experiments, along with a comprehensive evaluation file that reports all the experiment results. Provided data files represent an extension to the original datasets as provided by [1]. For associated systematic review manuscripts and eligibility criteria, please refer to [1] as well.&nbsp;</p> <p>[1] Guo, Eddie; Gupta, Mehul; Deng, Jiawen; Park, Ye-Jean; Paget, Mike; Naugler, Christopher (2023). "Automated Paper Screening for Clinical Reviews Using Large Language Models."&nbsp;<em>Mendeley Data</em>, V1, doi: 10.17632/np79tmhkh5.1. Accessed from: <a href="https://data.mendeley.com/datasets/np79tmhkh5/1" target="_new" rel="noopener">https://data.mendeley.com/datasets/np79tmhkh5/1</a>.</p>

opencc-by-4.0Dec 2023View details →
zenodo44/100

Word embeddings learnt on MEDLINE abstracts

<p>Accompanying a preprint manuscript and code repository, this folder contains both raw text data and learnt word embeddings. The data source is the set of MEDLINE articles published on or after 2000. Preprocessing consists of extraction of each article's title and abstract and some minor text processing. The result is a corpus of 10.5 million documents in a single 14 GB file. </p> <p>word2vec and fastText are used to learn word embeddings on this corpus and three sets of word embeddings are shared here: 1) word2vec skip-gram, 2) word2vec CBOW, and 3) fastText skip-gram. All three sets use the default parameters of the software (e.g. context=5) with the exception of hierarchical softmax optimization and dimension=200.</p> <p>Preprint manuscript: https://arxiv.org/abs/1705.06262<br> GitHub repository: https://github.com/vincentmajor/ctsa_prediction</p>

opencc-by-sa-4.0Jun 2017View details →
zenodo44/100

Data inputs and results from AI-supported title and abstract screening "Lack of evidence regarding markers identifying acute heart failure in patients with COPD: an AI-supported systematic review"

<p>These comma-separated data files were used to conduct the AI supported screening of [Lack of Evidence Regarding Markers Identifying Acute Heart Failure in Patients with COPD: An AI-supported Systematic Review (working title)], following the methodology described in the publication (URL/doi to be uploaded).</p> <p>These files provide insight into the AI-supported screening process and the choices made by the human reviewer.</p>

opencc-by-4.0Jan 2024View details →
zenodo44/100

Abstracts that contain "justice" from AGU Fall Meeting 2014-2024

<p>CSV containing a list of abstracts from AGU Fall Meetings 2014-2024 that contain the word "justice" in the title or text of the abstract. Abstracts were assembled by searching the individual Fall Meeting confex pages and copying the information directly into a CSV.&nbsp;</p> <p>The file contains the conference year, abstract number, full text of the abstract, and a characterization of the sector of the authors, based on listed affiliation in the AGU Fall Meeting system. The sectors include academic, governmental, NGO, commercial, informal education, and cross-sector.</p>

opencc-by-4.0Nov 2024View details →
zenodo44/100

CafeteriaSA corpus: Scientific abstracts annotated across different food semantic resources

<p>In the last decades, a great amount of work has been done in predictive modeling of issues related to human and environmental health. Resolution of issues related to healthcare is made possible by the existence of several biomedical vocabularies and standards, which play a crucial role in understanding health information, together with a large amount of health-related data. However, despite the large number of available resources and work done in the health and environmental domains, there is a lack of semantic resources that can be utilized in the food and nutrition domain, as well as their interconnections. For this purpose, in an European Food Safety Authority-funded project CAFETERIA, we have developed the first annotated corpus of 500 scientific abstracts that consists of 6,407 annotated food entities with regard to Hansard taxonomy, 4,299 for FoodOn, and 3,623 for SNOMED-CT.&nbsp; The CafeteriaSA corpus will enable further development of natural language processing methods for food information extraction from textual data that will allow extracting of food information from scientific textual data.</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

Mars Target Encyclopedia - Labeled LPSC abstracts for four Mars missions

<p>This data set contains annotated text versions of 1635 two-page abstracts published at the Lunar and Planetary Science Conference from 1998&nbsp;to 2020 of relevance to four Mars missions.&nbsp; The annotations were generated using named entity recognition and relation extraction provided by the MTE processing pipeline (available at&nbsp;https://github.com/wkiri/MTE), followed by manual review.&nbsp; Annotated entities include Element, Mineral, Property, and Target.&nbsp; Annotated relations include <strong>Contains</strong>(Target, Element | Mineral) and <strong>HasProperty</strong>(Target, Property).&nbsp; The extracted&nbsp;information (without full texts) is also available as a database (stored in .csv files) at&nbsp;https://pds-geosciences.wustl.edu/missions/mte/mte.htm .&nbsp;The complete annotated texts are provided here as a resource for further research and experimentation on&nbsp;information extraction methods.&nbsp; For more information about the Mars Target Encyclopedia and these annotations, please see:</p> <ul> <li>&quot;<a href="https://www.hou.usra.edu/meetings/lpsc2022/pdf/1231.pdf">Targets from the Spirit Mars Exploration Rover in the Mars Target Encyclopedia</a>&quot;,&nbsp;Kiri L. Wagstaff, Raymond Francis, Matthew Golombek, Steven Lu, Ellen Riloff, Leslie Tamppari, Yuan Zhuang, and Thomas Stein.<br> <em>53rd Lunar and Planetary Science Conference</em>, Abstract #1231, March 2022.</li> <li>&quot;<a href="https://www.hou.usra.edu/meetings/lpsc2021/pdf/1278.pdf">The Mars Target Encyclopedia Now Includes Mars Pathfinder and Mars Phoenix Targets</a>&quot;,&nbsp;Kiri L. Wagstaff, Raymond Francis, Matthew Golombek, Steven Lu, Ellen Riloff, Leslie Tamppari, and Thomas C. Stein.<br> <em>52nd Lunar and Planetary Science Conference</em>, Abstract #1278, March 2021.</li> </ul> <p>The original PDF abstracts are available at:&nbsp;</p> <ul> <li>For years prior to 2000:&nbsp; https://www.lpi.usra.edu/meetings/LPSC${two-digit-year}/pdf/${id}.pdf</li> <li>For year 2000:&nbsp; https://www.lpi.usra.edu/meetings/LPSC${four-digit-year}/pdf/${id}.pdf</li> <li>For years 2001-2017 (note lower-case lpsc):&nbsp; https://www.lpi.usra.edu/meetings/lpsc${four-digit-year}/pdf/${id}.pdf</li> <li>For years 2018-2020:&nbsp; https://www.hou.usra.edu/meetings/lpsc${four-digit-year}/pdf/${id}.pdf</li> </ul> <p>where ${id} is a four-digit abstract number, starting with 1001 (if available).</p> <p>The text files provided in this archive were extracted from the PDF files using the Apache Tika PDF parsing tool.&nbsp; They are named as ${four-digit-year}_${id}.txt.&nbsp; The text is provided here so that the annotations can be viewed in context.&nbsp; The text content remains copyright of the original abstract authors.</p> <p>The annotations (entities and relations) are provided in the format used by the brat annotation tool.&nbsp; They are named as ${four-digit-year}_${id}.ann. To view the annotations in a web-based graphical form, install the brat tool (http://brat.nlplab.org/).&nbsp; These annotations were generated using brat v1.3.&nbsp; The annotation files are also human-readable and can be parsed in to be used directly in code.&nbsp; If the .ann file is empty, then there are no relevant annotations for the associated text file.</p> <p><strong>Contents</strong>:</p> <ul> <li>mpf.zip: 591 abstracts relating to the Mars Pathfinder mission (1998-2020)</li> <li>mer-a.zip: 397 abstracts relating to the MER-A (Spirit) rover mission (2004-2020)</li> <li>mer-b.zip: 256 abstracts relating to the MER-B (Opportunity) rover mission (2005-2020)</li> <li>phx.zip: 391 abstracts relating to the Mars Phoenix Lander mission (2009-2020)</li> </ul> <p>Each directory contains a .txt and .ann file for each abstract.&nbsp; The .ann file is in brat standoff format (http://brat.nlplab.org/standoff.html).&nbsp; Additional .conf files are provided to generate color highlighting and keyboard shortcuts.&nbsp; These are used by the brat tool.</p> <p>Note: the same abstract may appear in more than one mission directory, if it discusses targets from more than one mission.&nbsp; It will have a different .ann file for each such appearance.&nbsp; Within each directory, a&nbsp;&quot;Target&quot; annotation is understood to refer to a target of the relevant mission.</p> <p><strong>Attribution</strong>:</p> <p>If you use this data set in your own work, please cite it as follows:</p> <p>Kiri L. Wagstaff, Raymond Francis, Matthew Golombek, Leslie Tamppari, and Steven Lu. (2022).&nbsp;Mars Target Encyclopedia - Labeled LPSC abstracts for four Mars missions&nbsp;(1.0.0.0) [Data set]. Zenodo. DOI: 10.5281/zenodo.7066107</p>

opencc-by-4.0Sep 2022View details →
zenodo44/100

Graphical abstract_CoBiol_260922

<p>With the static&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=global&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#global</a>&nbsp;growth 🌏 of the&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=biomass&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#biomass</a>-derived&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=fuel&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#fuel</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=supplychain&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#supplychain</a>, coupled with their inclusion in the long-term&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=decarbonization&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#decarbonization</a>&nbsp;of many transport enterprises, the sector still addresses the following&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=questions&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#questions</a>: are the emerging technologies assessing the&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=sustainability&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#sustainability</a>&nbsp;and&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=scalability&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#scalability</a>&nbsp;of biofuels? and if so, what are the&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=challenges&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#challenges</a>&nbsp;in the implementation of the future global&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=energy&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#energy</a>&ndash;&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=water&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#water</a>&ndash;<a href="https://www.linkedin.com/feed/hashtag/?keywords=climate&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#climate</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=nexus&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#nexus</a>?<br> <br> Despite&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=hydrothermal&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#hydrothermal</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=liquefaction&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#liquefaction</a>&nbsp;(HTL) being a competitive technology and apart from&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=crudeoil&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#crudeoil</a>&nbsp;as the target product, the major challenge limiting the economic viability and technical scalability of the HTL-related process is the safe disposal of generated by-products, including nearly 25 wt.% to 50 wt.% post-hydrothermal&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=aqueous&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#aqueous</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=phase&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#phase</a>&nbsp;(HTL-AP 💧) and 5 wt.% to 20 wt.% solid&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=hydrochar&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#hydrochar</a>&nbsp;residues (HCs).<br> <br> <a href="https://www.linkedin.com/feed/hashtag/?keywords=integrating&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#Integrating</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=adsorbents&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#adsorbents</a>&nbsp;such as commercially&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=activatedcarbon&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#activatedcarbon</a>&nbsp;and&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=zeolites&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#zeolites</a>&nbsp;can be a potential step to the removal of xenobiotic and recalcitrant&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=organic&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#organic</a>&nbsp;and&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=nutrient&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#nutrient</a>&nbsp;compounds from HTL-AP. However, these expensive adsorbents require a substitute for adsorption sustainability. Solid HCs, which are generated from HTL with carbon as the major element, can be used as low-cost adsorbents. However, its&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=porosity&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#porosity</a>&nbsp;of 0.058&ndash;0.082 cm3/g and BET-specific surface area of 1.56&ndash;17 m2/g remain comparatively low because of the formation and condensation of hydrocarbons on the surface, thereby clogging pores and reducing its&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=adsorption&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#adsorption</a>&nbsp;capacity. To the best of our knowledge, existing HTL studies have well-characterized crude bio-oils but have not focused on hydrochar except for primary details such as yield and ultimate analysis. Furthermore, the use of HTL hydrochar as an alternative low cost for&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=commercial&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#commercial</a>&nbsp;adsorbents and&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=wastewater&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#wastewater</a>&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=remediation&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#remediation</a>&nbsp;purposes is very less explored in comparison to pyrolytic char. Hence, the focus of our just available online accepted article is to contribute to the knowledge development of these&nbsp;<a href="https://www.linkedin.com/feed/hashtag/?keywords=gaps&amp;highlightedUpdateUrns=urn%3Ali%3Aactivity%3A6980062930067189760">#gaps</a>.</p>

opencc-by-4.0Sep 2022View details →
zenodo44/100

Mars Target Encyclopedia - LPSC abstracts labeled data set

<p>This data set contains annotated text versions of 2-page abstracts published at the Lunar and Planetary Science Conference in 2015 and 2016.</p> <p>The original PDF abstracts are available at:</p> <ul> <li>https://www.hou.usra.edu/meetings/lpsc2015/programAbstracts/view/</li> <li>https://www.hou.usra.edu/meetings/lpsc2016/programAbstracts/view/</li> </ul> <p>The text files in this archive were extracted using the Apache Tika PDF parsing tool.  The text is provided here so that the annotations can be viewed.  The text content remains copyright of the original abstract authors.</p> <p>The annotations (entities and relations) are provided in the format used by the brat annotation tool.  To view the annotations in a web-based graphical form, install the brat tool (http://brat.nlplab.org/).  These annotations were generated using brat v1.3.  The annotation files are also human-readable and can be parsed in to be used directly in code.</p> <p><strong>Contents</strong>:</p> <ul> <li>lpsc15/: 62 abstracts</li> <li>lpsc16/: 55 abstracts</li> </ul> <p>Each directory contains a .txt and .ann file for each abstract.  The .ann file is in brat standoff format (http://brat.nlplab.org/standoff.html).</p> <p>Additional .conf files are provided to generate color highlighting and keyboard shortcuts.  These are used by the brat tool.</p> <p><strong>Attribution</strong>:</p> <p>If you use this data set in your own work, please cite this DOI:</p> <p>10.5281/zenodo.1048419</p> <p>Please also cite this paper, which provides additional details about the data set.</p> <p>Kiri L. Wagstaff, Raymond Francis, Thamme Gowda, You Lu, Ellen Riloff, Karanjeet Singh, and Nina Lanza. "Mars Target Encyclopedia: Rock and Soil Composition Extracted from the Literature."  <em>Proceedings of the Thirtieth Annual Conference on Innovative Applications of Artificial Intelligence</em>, 2018.</p>

opencc-by-sa-4.0Nov 2017View details →
zenodo44/100

myExperiment Workflows, "abstracted" (all non-analytical nodes removed)

<p>To do the SCOFF analysis (detecting highly similar workflow fragments) we took all bioinformatics-related workflows from myExperiment and removed all non-analytical nodes.&nbsp; These included nodes referred-to as &quot;shims&quot; - those that do data structure/type transformations, but not any &quot;semantic&quot; transformation.&nbsp; This deposit contains all such abstracted workflows.</p>

opencc-by-4.0Jan 2018View details →
zenodo44/100

Abstracts from the Digital Humanities Conference 2005-2018

<p>Plain-text versions of the <a href="http://adho.org/conference">Digital Humanities Conference</a>s&#39;&nbsp;books of abstracts&nbsp;(2005-2018).<br> &nbsp;</p>

opencc-by-nc-4.0Aug 2018View details →
zenodo44/100

Visual abstract for SAMPL6 logP Challenge

<p>This figure was created as a visual abstract for the SAMPL6 Part II logP Challenge, which was a blind computational prediction challenge for predicting octanol-water partition coefficients of kinase inhibitor fragment-like small molecules. This figure&nbsp;can possibly be used as a cover art for the special journal issue organized for this&nbsp;challenge.&nbsp;</p>

opencc-by-4.0Nov 2019View details →
zenodo44/100

Map of the archaeological sites mentionned in the paper "Abstraction in Archaeological Stratigraphy: a Pyrenean Lineage of Innovation (late 19th–early 21th century)"

<p>Projection: WGS 84. QGIS 3.14.16</p> <p>Sources:</p> <ul> <li>DEM: GEBCO (<a href="https://doi.org/10.5285/A29C5465-B138-234D-E053-6C86ABC040B9">https://doi.org/10.5285/A29C5465-B138-234D-E053-6C86ABC040B9</a>)</li> <li>Borders:<em> L&iacute;mites municipales, provinciales y auton&oacute;micosRecintos municipales y l&iacute;neas l&iacute;mite (municipales, provinciales y auton&oacute;micos)</em>. BDLJE CC-BY 4.0.</li> </ul>

opencc-by-4.0May 2021View details →
zenodo44/100

Tan-Tári Bricks graphical abstract in iGEM Design League 2022

<p>Graphical abstract of Tan-T&aacute;ri Bricks Team from the University of Chile participating in iGEM Design League 2022.</p>

opencc-by-4.0Nov 2022View details →
zenodo44/100

arXiv abstracts and titles from 1,469 single-authored papers (100 unique authors) in computer science

<p>This dataset is meant to be used for experiments of Authorship Analysis. The dataset&nbsp;consists of abstracts of single-author papers from arXiv crawled using the arXiv&#39;s API by querying a list of computer-science-related keywords (&quot;deep learning&quot;, &quot;machine learning&quot;, &quot;information retrieval&quot;, &quot;computer science&quot;, &quot;data mining&quot;, &quot;support vector&quot;, &quot;logistic regression&quot;, &quot;artificial intelligence&quot;, &quot;supervised learning&quot;&#39;).&nbsp;The corpus somehow follows a power-law distribution, with few prolific authors and many authors accounting for very few papers each: we retained authors with at&nbsp;least 10 papers, resulting in a total of 1,469 documents from 100 authors. The most prolific authors (Peter D. Turney and Subhash Kak) have 34 abstracts to their names, the 10 most prolific authors have written 22 or more articles, while 50% of the authors have no more than 12 abstracts to their names. In order to divide the corpus into a training set and a test set we perform a stratified split, with the production of each author being split into a training set (70%) and a test set (30%). We use these documents as examples of &quot;scientific communication&quot;, characterised by a precise and compact style, with an abundance of technical terminology.</p>

opencc-by-4.0Dec 2022View details →
zenodo44/100

Dataset for Evaluating Abstractive Summaries of Crisis-Related Social Media

<p>The dataset created for evaluation of summaries generated from social media posted during five natural disasters.</p> <p>The dataset contains:</p> <ul> <li> <p>ground truth reports created by human assessor based on ERCC Echo Flash reports (5 events);</p> </li> <li> <p>summaries generated by extractive and abstractive state-of-the-art with manual annotation of category-relevant and crisis-relevant claim.</p> </li> </ul> <p>The dataset with annotated social media postings &mdash; <a href="https://doi.org/10.5281/zenodo.7714014">link</a></p> <p>The full description of used summarization methods, sources of data and evaluation metrics available in the paper &mdash; <a href="https://dl.acm.org/doi/abs/10.1145/3511095.3531279">link</a></p>

opencc-by-4.0Mar 2023View details →
zenodo44/100

Traces for studying Datacenter Scheduler Programming Abstractions

<p>Traces for the experiments for the research work that&nbsp;investigates the performance impact of various datacenter scheduler programming abstractions.</p>

opencc-by-4.0May 2023View details →
zenodo44/100

ROADMAP Technical Abstract Video: The Living Labs

<p>The Living Labs are one of the pillars of the ROADMAP project, which is focused on the responsible use and reduction of antimicrobials in livestock production. In this video, you will have the chance to get to know better how a Living Lab works, its importance in the process of getting awareness of AMU (antimicrobial use) and how tools are given to farmers in order to implement good practices, recommendations and acknowledging other ways of improving animal health and welfare.</p> <p>Watch the video here:&nbsp;https://youtu.be/afXe1eHX2lM</p>

opencc-by-4.0May 2023View details →
zenodo44/100

Analysis of interventions of practice abstracts and videos

<p>This dataset containts the underlying data that was used to analyse the SmartCulTour identified interventions that were described under the production of abstracts and practice videos of D6.2. The data collection established:</p> <ul> <li>Context &amp; background</li> <li>Reason why of the intervention</li> <li>Resources and tools</li> <li>Expected economic impact</li> <li>Expected social impact</li> <li>Expected cultural impact</li> <li>Expected environmental impact</li> <li>Success conditions</li> </ul> <p>These criteria were analysed by the SmartCulTour local experts and meant to feed into the updated taxonomy on cultural tourism interventions of D3.4. The interventions that are described are linked to the practice videos that can be found under the Living Labs section of the SmartCulTour website (<a href="http://www.smartcultour.eu">www.smartcultour.eu</a>) and pertain to:</p> <ul> <li>Rotterdam Living Lab: Planning for the future of Hoek van Holland &amp; Bospolder-Tussendijken</li> <li>Scheldeland Living Lab: Hof van Coolhem: social employment and care project in tourism</li> <li>Scheldeland Living Lab: Bornem Castle: upgrades historical exhibitions &amp; creates visitor centre</li> <li>Scheldeland Living Lab: Steam train Dendermonde-Puurs: volunteers protecting industrial heritage</li> <li>Utsjoki Living Lab: Traces in Utsjoki: inspiring respectful visitor behaviour in nature areas</li> <li>Utsjoki Living Lab: Placemaking as a technique to support meaningful visitor experiences</li> <li>Huesca Living Lab: The Somontano Wine Route: a resilient strategy for Huesca</li> <li>Huesca Living Lab: The R&iacute;o Vero Cultural Park. From Palaeolithic human history to the present</li> <li>Split Living Lab: Making traditional Easter bread-Sirnica in Solin</li> <li>Split Living Lab: The cultural heritage of Sinj: the story of Alka</li> <li>Vicenza Living Lab: Vicenza: the city of Palladio</li> <li>Vicenza Living Lab: The international library &quot;La Vigna&quot; becomes an open innovation Living Lab</li> </ul>

opencc-by-4.0Jun 2023View details →
zenodo44/100

Most common terms found within the abstracts

<p>Most common terms found within the abstracts.&nbsp;Part of the study &quot;What do we mean by GenAI?&quot;</p>

opencc-by-4.0Jul 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record