Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
612
datasets available to search
ShareScore release 0.7.1
Dataset results
612 results for “CITES”
PBMC CITE-seq, 10x Multiome, and TEA-seq multiomic datasets from Swanson, et al. eLife (2021).
<p>Assembled multiomic datasets from Swanson, et al. <em>Simultaneous trimodal single-cell measurement of transcripts, epitopes, and chromatin accessibility using TEA-seq</em>. eLife 2021;10:e63632 DOI: <a href="https://doi.org/10.7554/eLife.63632">10.7554/eLife.63632</a> </p> <p>Datasets were assembled from the GEO repository: <a href="https://www.ncbi.nlm.nih.gov/geo/query/acc.cgi?acc=GSE158013">GSE158013</a> </p> <p>Data are provided as SeuratObjects stored in separate .rds files using the saveRDS() function in R.</p> <p>MuData files for use with Python tools (muon, scanpy, scvitools, et al.) will be added soon.</p> <p>Contact Lucas Graybuck (lucasg at alleninstitute dot org) if there are problems with these datasets.</p>
Raw and aggregated data for the study introduced in the paper "The way we cite: common metadata used across disciplines for defining bibliographic references"
<p>These data have been gathered in the context of a study aiming to investigate citation practices for referencing different types of entities and, in particular, for understanding the most used metadata in bibliographic references. The data are stored in two documents in XLSX format:</p> <ul> <li>file "links-intext-pointers-and-cited-entity-types.xlsx" - it contains information about whether the in-text reference pointers of the various PDF articles of the corpus have specified hypertextual links from the in-text reference pointers to the denoted bibliographic reference, plus information about the types of all the entities cited by each article in the corpus;</li> <li>file "metadata-bibliographic-references.xlsm" - it contains information about the metadata used to identify the various descriptive elements of all the bibliographic references defined in the article of the corpus.</li> </ul> <p>The methodology used to gather all these data is described in:</p> <blockquote> <p>Santos, E. A. d., Peroni, S., Mucheroni, M. L.: Workflow for retrieving all the data of the analysis introduced in the article "Citing and referencing habits in Medicine and Social Sciences journals in 2019". (2020), <a href="https://doi.org/10.17504/protocols.io.bbifikbn">https://doi.org/10.17504/protocols.io.bbifikbn</a></p> </blockquote>
CITES species records: cites_taxa.tar.gz
<p></p>https://eol-jira.bibalex.org/browse/DATA-1790<p></p>CITES (the Convention on International Trade in Endangered Species of Wild Fauna and Flora) is an international agreement between governments. Its aim is to ensure that international trade in specimens of wild animals and plants does not threaten their survival. <p></p>https://www.cites.org
Raw and aggregated data for the study introduced in the article "An analysis of citing and referencing habits across all scholarly disciplines: approaches and trends in bibliographic metadata errors"
<p>This dataset contains all the raw data and aggregated data subject of the study introduced in the article "An analysis of citing and referencing habits across all scholarly disciplines: approaches and trends in bibliographic metadata errors". The study is based on the bibliographic and citation data contained in 729 articles published in 147 journals in 27 subject areas. The articles contained a total amount of 34,140 bibliographic references and 55,100 mentions and quotations overall.</p> <p>The dataset is composed of a series of files:</p> <ul> <li>the files "subject_area_<discipline-name>.csv" contain the raw data of the articles published in the journals of all the disciplines considered in the study;</li> <li>the file "article_data_summary.csv" contains the aggregated data created considering the raw data in the previous files, which have been used to creating all the tables and figures in the article;</li> <li>the file "starred_metadata_set.csv" contains information about the most used subset of bibliographic metadata;</li> <li>the file "journals_selection.csv" contains information about all the journals selected for the study.</li> </ul>
Articles citing HIMF as marker for alternatively activated macrophages
<p>These articles offer good evidence that HIMF is a marker for alternatively activated macrophages.</p>
How to cite & reference in Chicago (footnote) style
<p>This video shows how to cite and reference a book and a journal article in Chicago style, using style guide Cite Them Right.</p>
Dataset: Cartica Acquisition Corp (CITE) Stock Performance
This dataset provides historical stock market performance data for specific companies. It enables users to analyze and understand the past trends and fluctuations in stock prices over time. This information can be utilized for various purposes such as investment analysis, financial research, and market trend forecasting.
Bibliometric dataset: list of highly cited papers in bibliometric
<p><strong>Motivation</strong></p> <p>My motivation in providing this dataset is to invite more interests from Indonesia's librarian to understand their diverse field of study. </p> <p><strong>Method</strong></p> <p>This dataset is harvested in 19 January 2019 from Scopus database provided by The University of Sydney. I used the keyword "bibliometric" in title, sort the search results by total citation, then download the first 2000 papers as RIS file. This file can be converted to other formats like bibtex or csv using available reference manager, like Zotero.</p> <p><strong>Visualisations</strong></p> <p>I did two small visualisations using the following options:</p> <ol> <li>"create a map based on bibliographic data"</li> <li>"create a map based on text data"</li> </ol> <p>Both mappings are done using <a href="http://vosviewer.com">VosViewer </a>open source app from <a href="https://www.cwts.nl/">CWTS Leiden University</a>.</p>
Text-fig. 2. Map showing the distribution of localities included in the treatment of Brown (1962), based on the coordinates presented in Appendix Table 1. Localities are distributed across New Mexico (NM), Colorado (CO), Wyoming (WY), Montana (MT), South Dakota (SD) and North Dakota (ND). Also plotted are cited Canadian localities: Joffre Bridge, in Alberta (AB), and Ravenscrag, in Saskatchewan (SK). Map generated with Map-It (2014) website. in Revisions To Roland Brown'S North American Paleocene Flora
Text-fig. 2. Map showing the distribution of localities included in the treatment of Brown (1962), based on the coordinates presented in Appendix Table 1. Localities are distributed across New Mexico (NM), Colorado (CO), Wyoming (WY), Montana (MT), South Dakota (SD) and North Dakota (ND). Also plotted are cited Canadian localities: Joffre Bridge, in Alberta (AB), and Ravenscrag, in Saskatchewan (SK). Map generated with Map-It (2014) website.
A Dataset of Metadata of Articles Citing Retracted Articles
<p>This dataset comprises of metada of articles citing retracted publications. Originally, we obtained the DOIs from the<a href="https://www.irit.fr/~Guillaume.Cabanac/problematic-paper-screener"> Feet of Clay Detector of the Problematic Paper Screener (PPS - FoCD)</a>. Additional columns that were not provided in PPS were added using Crossref & Retraction Watch Database (CRxRW) and Dimensions API services. This detector flags publications that cite retracted articles with additional metadata. </p> <p>By querying the Dimensions API with the DOIs of the FoC articles, we acquired information such as more detailed document types (editorial, review article, research article), open access status (we only kept open access FoC articles in the dataset since we want to access the full-texts in the future), and research fields (classified according to the Australian and New Zealand Standard Research Classification (ANZSRC) Fields of Research (FoR), comprising of 23 main fields such as biological sciences, education.</p> <p>To get further information about the cited retracted articles in the dataset, we used the joint release of CRxRW. Using this dataset, we added the retraction reasons and retraction years. </p> <p>The original dataset was obtained from the PPS FoCD in December 2023. At this time there were 22558 total articles flagged in FoCD. Using the data filtering feature in PPS, we had a preliminary selection before downloading the first version of the dataset. We applied a filter to obtain:</p> <ul> <li>non-retracted citing articles at the time of data curation*</li> <li>open-access citing articles since we need the whole text to go forward with natural language processing tasks</li> <li>cited retracted articles with at least one scientific content related reason of retraction </li> <li>only articles (not monographs, chapters) to retain a unified text type</li> </ul> <p>More information about the usage of this dataset will be updated. </p> <p>*Current retraction status of the citing articles can be different since this is a static dataset and scientific literature is dynamic. </p>
Linked collectors and determiners for: Espèces vasculaires endémiques et orchidées (CITES).
Natural history specimen data linked to collectors and determiners held within, "Espèces vasculaires endémiques et orchidées (CITES)". Claims or attributions were made on Bionomia by volunteer Scribes, <a href="https://bionomia.net/dataset/912de54b-1fbe-4ef1-a283-9f7329238022">https://bionomia.net/dataset/912de54b-1fbe-4ef1-a283-9f7329238022</a> using specimen data from the dataset aggregated by the Global Biodiversity Information Facility, <a href="https://gbif.org/dataset/912de54b-1fbe-4ef1-a283-9f7329238022">https://gbif.org/dataset/912de54b-1fbe-4ef1-a283-9f7329238022</a>. Formatted as a Frictionless Data package.
Cite & reference to avoid plagiarism
<p>This video explains the importance of citing and referencing, and shows what name/date and footnote citation style formats look like.</p>
How to cite & reference in IHS style
<p>This video shows you how to cite & reference a book, a chapter in an edited book, and a journal article in IHS (Irish Historical Studies) style.</p>
How to cite & reference in MHRA style (author-date system)
<p>This video explains how to cite & reference a chapter in an edited book in MHRA style using the author-date system.</p>
How to cite & reference in MHRA style (with footnotes)
<p>This video shows how to cite and reference a printed book and a journal article in MHRA (footnote) style, using Cite Them Right as style guide.</p>
Highly cited tropical medicine articles in the early COVID 19 pandemic. Original Excel data on citation and subjects
<p><strong>ORIGINAL DATA SET FOR STUDY OF PUBLICATION AND CITATION TRENDS, TROPICAL MEDICINE, EARLY COVID 19 PANDEMIC. Background: </strong>An adequate response to health needs includes the identification of research patterns about the large number of people living in the tropics and subjected to tropical diseases. Studies have shown that research does not always match the real needs of those populations, and that citation reflects mostly the amount of money behind particular publications. Here we test the hypothesis that research from richer institutions is published in better-indexed journals, and thus has greater citation rates.</p> <p><strong>Methods:</strong> The data in this study was extracted from the Science Citation Index Expanded database; the 2020 journal Impact Factor (<em>IF</em><sub>2020</sub>) was updated to 30 June 2021. We considered places, subjects, institutions and journals.</p> <p><strong>Results:</strong> We identified 1 041 highly cited articles with 100 citations or more in the category of tropical medicine. About a decade is needed for an article to reach peak citation. Only two Covid-19 related were highly cited in the last three years. Most cited articles were published by the journals <em>Memorias Do Instituto Oswaldo Cruz</em> (Brazil), <em>Acta Tropica</em> (Switzerland), and <em>PLoS Neglected Tropical Diseases</em> (USA). The USA dominated five of the six publication indicators. International collaboration articles had more citations than single-country articles. The UK, South Africa, and Switzerland had high citation rates, as did the London School of Hygiene and Tropical Medicine in the UK, the Centers for Disease Control and Prevention in the USA, and the WHO in Switzerland.</p> <p><strong>Conclusions:</strong> About ten years of accumulated citations are needed to get 100 citations or more as highly cited articles in the Web of Science category of tropical medicine. Six publication and citation indicators, including authors’ publication potential and characteristics evaluated by <em>Y</em>-index, indicate that the currently available indexing system places tropical researchers at a disadvantage against their colleagues in temperate countries, and suggest that, to progress towards better control of tropical diseases, international collaboration should increase, and other tropical countries should follow the example of Brazil, which provides significant financing to its scientific community.Julián Monge-Nájera<sup>1</sup>, and Yuh-Shan Ho<sup>2</sup>*</p> <p><sup>1</sup>Laboratorio de Ecología Urbana, Vicerrectoría de Investigación, Universidad Estatal a Distancia, 2050 San José, Costa Rica; <a href="mailto:julianmonge@gmail.com"><em>julianmonge@gmail.com</em></a> (https://orcid.org/0000-0001-7764-2966)</p> <p>*Corresponding author: Trend Research Centre, Asia University, No. 500 Lioufeng Road, Wufeng, Taichung 41354, Taiwan; <a href="mailto:ysho@asia.edu.tw">ysho@asia.edu.tw</a> (<em>https://orcid.org/0000-0002-2557-8736</em>)</p>
Characterizing Highly Cited Papers in Mass Cytometry through H-Classics: WoS dataset and citation report
<p>Dataset and citation report extracted from Web of Science (WoS) used to characterize highly cited papers in mass cytometry research field from 2010 to 2019.</p>
Genus Aconitum in Slovakia: a phenetic approach. S30: Database of specimens from Slovakia cited in published sources
<p>This is database generated from published sources. It contains data about Aconitum specimens mentioned in other publications and preserved in herbaria outside of Slovakia.</p>
10K-Cell Subset of PBMC CITE-Seq Dataset for CITEViz
<p>This repository contains an example CITE-Seq data (10K peripheral blood mononuclear cells) to test the CITEViz program. The CITEViz preprint is available <a href="https://www.biorxiv.org/content/10.1101/2022.05.15.491411v1">here</a>, and the documentation website is located <a href="https://maxsonbraunlab.github.io/CITEViz/">here</a>. The original data underlying this article are available in GEO (Gene Expression Omnibus) at <a href="https://www.ncbi.nlm.nih.gov/geo/">https://www.ncbi.nlm.nih.gov/geo/</a>, and can be accessed with GSE164378. </p>
How to cite & reference using OSCOLA
<p>This video shows how to cite and reference using OSCOLA and style guide Cite Them Right.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.