Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

311

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

311 results for “new dataset”

Learn how ShareScore rates datasets ↗
zenodo40/100

Dataset: New Providence Acquisition Corp. II (NPABW) Stock Performance

This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.

opencc-zeroJun 2024View details →
zenodo40/100

Dataset: New Providence Acquisition Corp. II (NPABU) Stock Performance

This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.

opencc-zeroJun 2024View details →
zenodo40/100

Dataset: New Fortress Energy Inc. (NFE) Stock Performance

This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.

opencc-zeroJun 2024View details →
zenodo40/100

Dataset: New Horizon Aircraft Ltd. (HOVRW) Stock Performance

This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.

opencc-zeroJun 2024View details →
zenodo40/100

Dataset: New Horizon Aircraft Ltd. (HOVR) Stock Performance

This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.

opencc-zeroJun 2024View details →
zenodo40/100

Dataset: Towards a two-step assessment of the chloride ingress behaviour of new cementitious binders

<p>Data and results to accompany the publication:</p> <p>&nbsp;</p> <p><strong>Towards a two-step assessment of the chloride ingress behaviour of new cementitious binders</strong></p> <p><strong>William Wilson<sup>a,b,</sup><sup>⁎</sup>, Fabien Georget<sup>a,c</sup>, Karen L. Scrivener<sup>a</sup></strong></p> <p><strong>Cement and Concrete Research, Volume 184, July 2024, 107594</strong></p> <p>&nbsp;</p> <p><sup>a</sup>Laboratory of Construction Materials, EPFL, Lausanne, Switzerland</p> <p><sup>b</sup>Universit&eacute; de Sherbrooke, Sherbrooke, Canada</p> <p><sup>c</sup>Institute of Building Materials Research, RWTH Aachen, Aachen, Germany</p> <p>&nbsp;</p> <p>*Corresponding author: william.wilson@usherbrooke.ca</p>

opencc-by-4.0Jul 2024View details →
zenodo40/100

Embedded HuffPost New Category Dataset

<p>This dataset was created by embedding the concatenation of title and short description of each entry in the <a href="https://arxiv.org/pdf/2209.11429">HuffPost news category</a> dataset, ordered by timestamp, using OpenAI's t<a href="https://openai.com/index/new-embedding-models-and-api-updates/">ext-embedding-3-small embedding</a>. Usage is subject to Huffington Post's <a href="https://www.huffpost.com/static/user-agreement">user agreement</a> and OpenAI's <a href="https://openai.com/policies/row-terms-of-use/">term of use</a>.</p>

opencc-by-4.0Jul 2024View details →
zenodo40/100

Dataset for New Results on Tripod Packings

<p>The files contain tripod packings obtained in the study &quot;P.R.J. &Ouml;sterg&aring;rd &amp; A. P&ouml;ll&auml;nen, New results on tripod packings&quot;. More specifically, they prove the lower bounds in Table 5 of the paper. The two files contain sets of orbit representatives of codes and explicit packings in matrix form. For each type of symmetry, the results presented are for <span class="math-tex">\(n = 3,4,\ldots\)</span>. Codewords are of the form <span class="math-tex">\((x,y,z)\)</span>, where <span class="math-tex">\(x,y,z \in \{0,1,\ldots ,n-1\}\)</span>. The symmetries are</p> <ul> <li><span class="math-tex">\(r(x,y,z) = (z,x,y)\)</span>,</li> <li><span class="math-tex">\(p(x,y,z) = (y,x,z)\)</span>,</li> <li><span class="math-tex">\(o(x,y,z) = (n-1-x,n-1-y,n-1-z)\)</span>.</li> </ul> <p>&nbsp;</p>

opencc-by-4.0Apr 2018View details →
zenodo40/100

Dataset: Phylogenetic matrix: A new myromecophilic spider from the Chihuahuan desert

<p>Dataset: Phylogenetic matrix: A new myromecophilic spider from the Chihuahuan desert</p> <p>combined.nex.txt: MrBayes Nexus data matrix&nbsp;of DNA and morphology combined.</p> <p>morphology.tnt.txt: TNT data matrix of the morphology partition.</p>

opencc-by-4.0Apr 2019View details →
zenodo40/100

#nowplaying-RS: A New Benchmark Dataset for Building Context-Aware Music Recommender Systems

<p>Music recommender systems can offer users personalized and contextualized recommendation and are therefore important for music information retrieval. An increasing number of datasets have been compiled to facilitate research on different topics, such as content-based, context-based or next-song recommendation. However, these topics are usually addressed separately using different datasets, due to the lack of a unified dataset that contains a large variety of feature types such as item features, user contexts, and timestamps. To address this issue, we propose a large-scale benchmark dataset called #nowplaying-RS, which contains 11.6 million music listening events (LEs) of 139K users and 346K tracks collected from Twitter. The dataset comes with a rich set of item content features and user context features, and the timestamps of the LEs. Moreover, some of the user context features imply the cultural origin of the users, and some others&mdash;like hashtags&mdash;give clues to the emotional state of a user underlying an LE. In this paper, we provide some statistics to give insight into the dataset, and some directions in which the dataset can be used for making music recommendation. We also provide standardized training and test sets for experimentation, and some baseline results obtained by using factorization machines.</p> <p>The dataset contains three files:</p> <ul> <li>user_track_hashtag_timestamp.csv contains basic information about each listening event. For each listening event, we provide an id, the user_id, track_id, hashtag, created_at&nbsp;</li> <li>context_content_features.csv: contains all context and content features. For each listening event, we provide the id of the event, user_id, track_id, artist_id, content features regarding the track mentioned in the event (instrumentalness, liveness, speechiness, danceability, valence, loudness, tempo, acousticness, energy, mode, key) and context features regarding the listening event (coordinates (as geoJSON), place (as geoJSON), geo (as geoJSON), tweet_language, created_at, user_lang, time_zone, entities contained in the tweet).</li> <li>sentiment_values.csv contains sentiment information for hashtags. It contains the hashtag itself and the sentiment values gathered via four different sentiment dictionaries: AFINN, Opinion Lexicon, Sentistrength Lexicon and vader. For each of these dictionaries we list the minimum, maximum, sum and average of all&nbsp;sentiments of the tokens of the hashtag (if available, else we list empty values). However, as most hashtags only consist of a single token, these&nbsp;values are equal in most cases. Please note that the lexica are rather diverse and therefore, are able to resolve very different terms against a score. Hence,&nbsp;the resulting csv is rather sparse. The file contains the following comma-separated values: &lt;hashtag, vader_min, vader_max, vader_sum,vader_avg, &nbsp;afinn_min, afinn_max,&nbsp;afinn_sum, afinn_avg, ol_min, ol_max, ol_sum, ol_avg, ss_min, ss_max, ss_sum, ss_avg &gt;, where we abbreviate all scores gathered over the Opinion Lexicon with the&nbsp;prefix &#39;ol&#39;. Similarly, &#39;ss&#39; stands for SentiStrength.&nbsp;</li> </ul> <p>Please also find the training and test-splits for the dataset in this repo. Also, prototypical implementations of a context-aware recommender system based on the dataset can be found at&nbsp; <a href="https://github.com/asmitapoddar/nowplaying-RS-Music-Reco-FM">https://github.com/asmitapoddar/nowplaying-RS-Music-Reco-FM</a>.</p> <p>If you make use of this dataset, please cite the following paper where we describe and experiment with the dataset:</p> <p>@inproceedings{smc18,<br> title = {#nowplaying-RS: A New Benchmark Dataset for Building Context-Aware Music Recommender Systems},<br> author = {Asmita Poddar and Eva Zangerle and Yi-Hsuan Yang},<br> url = {http://mac.citi.sinica.edu.tw/~yang/pub/poddar18smc.pdf},<br> year = {2018},<br> date = {2018-07-04},<br> booktitle = {Proceedings of the 15th Sound &amp; Music Computing Conference},<br> address = {Limassol, Cyprus},<br> note = {code at https://github.com/asmitapoddar/nowplaying-RS-Music-Reco-FM},<br> tppubtype = {inproceedings}<br> }</p>

opencc-by-4.0Jul 2018View details →
zenodo40/100

Dataset related to article "Glia-to-neuron transfer of miRNAs via extracellular vesicles: a new mechanism underlying inflammation-induced synaptic alterations"

<p>This record contains raw data related to article &quot;Glia-to-neuron transfer of miRNAs via extracellular vesicles: a new mechanism underlying inflammation-induced synaptic alterations&quot;</p> <p>Recent evidence indicates synaptic dysfunction as an early mechanism affected in neuroinflammatory diseases, such as multiple sclerosis, which are characterized by chronic microglia activation. However, the mode(s) of action of reactive microglia in causing synaptic defects are not fully understood. In this study, we show that inflammatory microglia produce extracellular vesicles (EVs) which are enriched in a set of miRNAs that regulate the expression of key synaptic proteins. Among them, miR-146a-5p, a microglia-specific miRNA not present in hippocampal neurons, controls the expression of presynaptic synaptotagmin1 (Syt1) and postsynaptic neuroligin1 (Nlg1), an adhesion protein which play a crucial role in dendritic spine formation and synaptic stability. Using a Renilla-based sensor, we provide formal proof that inflammatory EVs transfer their miR-146a-5p cargo to neuron. By western blot and immunofluorescence analysis we show that vesicular miR-146a-5p suppresses Syt1 and Nlg1 expression in receiving neurons. Microglia-to-neuron miR-146a-5p transfer and Syt1 and Nlg1 downregulation do not occur when EV-neuron contact is inhibited by cloaking vesicular phosphatidylserine residues and when neurons are exposed to EVs either depleted of miR-146a-5p, produced by pro-regenerative microglia, or storing inactive miR-146a-5p, produced by cells transfected with an anti-miR-146a-5p. Morphological analysis reveals that prolonged exposure to inflammatory EVs leads to significant decrease in dendritic spine density in hippocampal neurons in vivo and in primary culture, which is rescued in vitro by transfection of a miR-insensitive Nlg1 form. Dendritic spine loss is accompanied by a decrease in the density and strength of excitatory synapses, as indicated by reduced mEPSC frequency and amplitude. These findings link inflammatory microglia and enhanced EV production to loss of excitatory synapses, uncovering a previously unrecognized role for microglia-enriched miRNAs, released in association to EVs, in silencing of key synaptic genes.</p>

opencc-by-4.0Sep 2019View details →
zenodo40/100

Efficiency and heat transport processes of low-temperature aquifer thermal energy storage systems: new insights from global sensitivity analyses - Supporting Dataset

<p>This dataset contains the files used to substantiate the outcomes of the publication <em>"Efficiency and heat transport processes of low-temperature aquifer thermal energy storage systems: new insights from global sensitivity analyses"</em>.&nbsp;</p> <p>It includes the output of 250 random model realizations of an aquifer thermal energy storage system in a thick productive aquifer (Case 1). It also includes the output of 500 random model realizations of an aquifer thermal energy storage system in a shallow alluvial aquifer (Case 2 part 1 and part 2).</p> <p>If there is interest in generating new output, the datset also includes the model input files for both cases.</p> <p>(Scripts to process the output data or to generate new output data can be found in the corresponding GitHub repository: https://github.com/lukatas/ATES_SensitivityAnalyses.git )</p>

opencc-by-4.0Aug 2024View details →
zenodo40/100

Dataset for "On the potential of the Cluster Ion Counter (CIC) to observe local new particle formation, condensation sink and growth rate of newly formed particles"

<p>Data for Kulmala et al. (2024 )"On the potential of the Cluster Ion Counter (CIC) to observe local new particle formation, condensation sink and growth rate of newly formed particles" (https://doi.org/10.5194/ar-2024-14).</p> <p>Included in the file are number concentrations of sub-2 nm ions and 2-2.3 nm ions measured with&nbsp; Cluster Ion Counter (CIC) and Neutral cluster and&nbsp; Air Ion Spectrometer (NAIS) at&nbsp; SMEAR II station in Hyyti&auml;l&auml;, Finland. Concentrations of 1-2 nm ions measured with the NAIS are also included. Sub-2 nm (2-2.3 nm) ion concentrations measured with CIC are refered as Channel 1 (Channel 2-Channel 3) in the .csv file.</p> <p>Contact Santeri Tuovinen (santeri.tuovinen@helsinki.fi) for more details.</p>

opencc-by-4.0Oct 2024View details →
zenodo40/100

Linked collectors and determiners for: A Distribution and Taxonomic Reference Dataset of Geranium (Geraniaceae) in the New World.

Natural history specimen data linked to collectors and determiners held within, "A Distribution and Taxonomic Reference Dataset of Geranium (Geraniaceae) in the New World". Claims or attributions were made on Bionomia by volunteer Scribes, <a href="https://bionomia.net/dataset/26d72d3b-4544-4645-aa56-27aa8a669c6f">https://bionomia.net/dataset/26d72d3b-4544-4645-aa56-27aa8a669c6f</a> using specimen data from the dataset aggregated by the Global Biodiversity Information Facility, <a href="https://gbif.org/dataset/26d72d3b-4544-4645-aa56-27aa8a669c6f">https://gbif.org/dataset/26d72d3b-4544-4645-aa56-27aa8a669c6f</a>. Formatted as a Frictionless Data package.

opencc-zeroJan 2024View details →
zenodo40/100

Datasets underlying the publication "A new lineage nomenclature to aid genomic surveillance of dengue virus"

<p>These datasets are underlying the scientific publication titled "A new lineage nomenclature to aid genomic surveillance of dengue virus", published in the <a href="https://journals.plos.org/plosbiology/article?id=10.1371/journal.pbio.3002834#abstract0">PLOS Biology</a>&nbsp;journal.&nbsp;</p> <p>All sequences used to design the lineage system are from Genbank and GISAID, with accession numbers listed in the tables. Custom scripts and alignments of representative sequences from Genbank can be found on the github of the publication authors (<a href="https://github.com/DENV-lineages/lineages-paper">https://github.com/DENV-lineages/lineages-paper</a>).</p> <p>Sequences for the Vietnam case study can be found on Genbank under accession numbers PP269455-PP270050, in bioproject PRJNA1072696. For the case study from Tanzania, sequences can be found on Genbank under accession numbers OM920035-OM920066 for DENV-3 and OM920075-OM920415 for DENV-1. Sequences for the Brazil case study can be found on GISAID under accession numbers EPI_ISL_17733558 ‐ EPI_ISL_191469691.<br><br>The provided information in the datasets are further discussed and interpreted in detail, as well as their subsequent results, in the scientific publication.</p> <p>VIRTIGATION partner EMWEB contributed to this publication with findings from the VIRTIGATION project, a project which is part of the EU Open Research Data pilot. This project has received funding from the European Union's Horizon 2020 research and innovation program under grant agreement No. 101000570.</p>

opencc-by-4.0Sep 2024View details →
zenodo40/100

Time Series Comparisons, Model Code, and a Demo Dataset for SIBaR: A New Method for Background Quantification and Removal from Mobile Air Pollution Measurements

<p>Time series comparisons between SIBaR, Brantley, and Apte background signals for all 312 time series in the Houston mobile monitoring campaign. Additionally, a R script demo (DemoData.R) of the SIBaR partitioning step on the demo datatset (DemoData.csv).</p>

opencc-by-4.0Jun 2021View details →
zenodo40/100

Fig. 1. Maximum likelihood tree inferred from the COI dataset with 1000 in Seven new giant pill-millipede species and numerous new records of the genus Zoosphaerium from Madagascar (Diplopoda, Sphaerotheriida, Arthrosphaeridae)

Fig. 1. Maximum likelihood tree inferred from the COI dataset with 1000 bootstrap pseudoreplicates implementing the GTR + I + G model. Colors representing newly described species of Zoosphaerium: orange = Z. nigrum sp. nov.; green = Z. silens sp. nov.; red = Z. ambatovaky sp. nov.; yellow = Z. beanka sp. nov.; purple = Z. voahangy sp. nov.; blue = Z. masoala sp. nov. Round-cornered rectangles indicate well-supported sister group relationships.

opencc-by-4.0Jul 2021View details →
zenodo40/100

Dataset for submitted manuscript "Temperatures and cooling rates recorded by the New Caledonia ophiolite: implications for cooling mechanisms in young forearc sequence"

<p>original dataset for the manuscript &quot;Temperatures and cooling rates recorded by the New Caledonia&nbsp;ophiolite: implications for cooling mechanisms in young forearc&nbsp;sequences&quot; submitted to the journal &quot;G3&nbsp;Geophysics, Geochemistry, Geosystems&quot;</p>

opencc-by-4.0Apr 2021View details →
zenodo40/100

New estimation of the NOx snow-source on the Antarctic Plateau - repository dataset

<p>Notebook and data set used to present the results.</p> <p>For more information, please, do not hesitate to contact the corresponding author to this study.</p>

opencc-by-4.0Aug 2021View details →
zenodo40/100

FIGURE 2 in A new morphological dataset reveals a novel relationship for the adzebills of New Zealand (Aptornis) and provides a foundation for total evidence neoavian phylogenetics

FIGURE 2. Majority-rule cladogram of nine most parsimonious trees (length: 2038, CI: 0.2498, RI: 0.5337, RC: 0.1333, HI: 0.7502) from analysis of our new dataset of 40 taxa and 368 osteological characters. All trees show optimization of an Aptornis defossor+Psophia obscura sister group. Synapomorphies are detailed in table 2. Extinct taxa are denoted with daggers. Majority-rule percentages are annotated above branches, followed by bootstrap support values greater than 50% in parentheses. Branch length ranges are below branches.

opencc-by-4.0May 2019View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record