Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

9

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

9 results for “Bibliographic references”

Learn how ShareScore rates datasets ↗
zenodo44/100

Raw and aggregated data for the study introduced in the paper "The way we cite: common metadata used across disciplines for defining bibliographic references"

<p>These data have been gathered in the context of a study aiming to investigate citation practices for referencing different types of entities and, in particular, for understanding the most used metadata in bibliographic references. The data are stored in two documents in XLSX format:</p> <ul> <li>file &quot;links-intext-pointers-and-cited-entity-types.xlsx&quot; - it contains information about whether the in-text reference pointers of the various PDF articles of the corpus have specified hypertextual links from the in-text reference pointers to the denoted bibliographic reference, plus information about the types of all the entities cited by each article in the corpus;</li> <li>file &quot;metadata-bibliographic-references.xlsm&quot; - it contains information about the metadata used to identify the various descriptive elements of all the bibliographic references defined in the article of the corpus.</li> </ul> <p>The methodology used to gather all these data is described in:</p> <blockquote> <p>Santos, E. A. d., Peroni, S., Mucheroni, M. L.: Workflow for retrieving all the data of the analysis introduced in the article &quot;Citing and referencing habits in Medicine and Social Sciences journals in 2019&quot;. (2020), <a href="https://doi.org/10.17504/protocols.io.bbifikbn">https://doi.org/10.17504/protocols.io.bbifikbn</a></p> </blockquote>

opencc-by-4.0May 2022View details →
zenodo40/100

Dataset of first appearances of the scholarly bibliographic references on English Wikipedia articles as of 1 March 2017 and as of 1 October 2021

<p><strong>Abstract</strong></p> <p>We developed a methodology to detect the oldest scholarly reference added to Wikipedia articles by which a certain paper is uniquely identifiable as the &quot;first appearance of the scholarly reference.&quot; We identified the first appearances of 923,894 scholarly references (611,119 unique DOIs) in180,795 unique pages on English Wikipedia as of March 1, 2017, and stored them in the dataset. Moreover, we assessed the precision of the dataset, which was and it was a high precision regardless of the research field. In this version,&nbsp;it is available not only the dataset of English Wikipedia as of March 1, 2017, but also English Wikipedia as of October 1, 2021, generated by using&nbsp;the same methodology.</p> <p>&nbsp;</p> <p><strong>Data Records</strong></p> <p>The data format of the dataset is JSON lines, where each line is a single record. In this dataset, we detected the first appearance of each scholarly reference added to Wikipedia articles. If there are multiple references corresponding to the same paper on the same page, only the oldest one is collected. Sample of the record is the following.</p> <ul> <li>doi -- DOI corresponding to the paper (String), e.g., &quot;10.1006/anbe.1996.0497&quot;</li> <li>paper_type -- Document type of the paper (String), e.g., &quot;journal-article&quot;</li> <li>paper_container_title -- Journal title, book title, or proceedings title (Array of String), e.g., [&quot;Animal Behaviour&quot;]</li> <li>paper_publisher -- Publisher name (String), e.g., &quot;Elsevier BV&quot;</li> <li>paper_title -- Paper title (Array of String), e.g., [&quot;Push or pull: an experimental study on imitation in marmosets&quot;]</li> <li>paper_published_year -- Published year (String), e.g., &quot;1997&quot;</li> <li>paper_issue -- Issue number (String), e.g., &quot;4&quot;</li> <li>paper_volume -- Volume number (String), e.g., &quot;54&quot;</li> <li>paper_page -- Page numbers (String), e.g., &quot;817-831&quot;</li> <li>paper_author -- Authors information consisted of given and family names, sequences (order in author names), and affiliations (Array of JSON), e.g., [{&quot;given&quot;:&quot;THOMAS&quot;, &quot;family&quot;:&quot;BUGNYAR&quot;, &quot;sequence&quot;:&quot;first&quot;, &quot;affiliation&quot;:[]}, {&quot;given&quot;:&quot;LUDWIG&quot;, &quot;family&quot;:&quot;HUBER&quot;, &quot;sequence&quot;:&quot;additional&quot;, &quot;affiliation&quot;:[]}]</li> <li>issn -- ISSN related to the paper (Array of String), e.g., [&quot;0003-3472&quot;]</li> <li>research_field -- Research fields from ESI categories (Array of String), e.g., [&quot;PLANT &amp; ANIMAL SCIENCE&quot;]</li> <li>page_id -- Page id (String), e.g., &quot;577858&quot;</li> <li>page_title -- Page title (String), e.g., &quot;Imitation&quot;</li> <li>revision_id -- Revision id (String), e.g., &quot;203309031&quot;</li> <li>revision_timestamp -- Revision timestamp (String), e.g., &quot;2008-04-04 15:54:09 UTC&quot;</li> <li>revision_comment -- Revision comment (edit summary) (String), e.g., &quot;/* Animal Behaviour */&quot;</li> <li>editor_name -- Wikipedia editor&#39;s name (String), e.g., &quot;Nicemr&quot;</li> <li>editor_type -- Type of the editor (String), e.g., &quot;User&quot;</li> </ul> <p><strong>References</strong></p> <ul> <li>Kikkawa, J., Takaku, M. &amp; Yoshikane, F. &quot;Dataset of first appearances of the scholarly bibliographic references on Wikipedia articles&quot;, Scientific Data, Vol. 9, Article number 85, pp. 1-11, 2022. <a href="https://doi.org/10.1038/s41597-022-01190-z">https://doi.org/10.1038/s41597-022-01190-z</a>.</li> </ul> <p><strong>FUNDING</strong></p> <ul> <li>JSPS KAKENHI Grant Number <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-20K12543">JP20K12543</a></li> <li>JSPS KAKENHI Grant Number <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-21K21303">JP21K21303</a></li> </ul>

opencc-by-sa-4.0Oct 2021View details →
zenodo40/100

Telescopus finkeldeyi Haacke, 2013 DAMARA TIGER SNAKE Telescopus finkeldeyi Haacke 2013:281. Holotype: TM 53542 (collector J.A. van Rooyen). Type locality: "Rössing Uranium mine area, Swako- mund [sic] district (2214Db) Namibia." Global conservation status (IUCN): Not Evaluated. Global distribution: The species is known from Angola and Namibia. Ocurrences in Angola (Map 364): The species occurs in southwestern Angola. Namibe: "5 km north Namibé" [-15.20000, 12.15000] (Haacke 2013:285). Taxonomic and distributional notes: Some earlier records of T. semiannulatus polystictus in Namibia actually refer to this recently described species. MAP 364. Distribution of Telescopus finkeldeyi in Angola. in Diversity and Distribution of the Amphibians and Terrestrial Reptiles of Angola Atlas of Historical and Bibliographic Records (1840-2017)

Telescopus finkeldeyi Haacke, 2013 DAMARA TIGER SNAKE Telescopus finkeldeyi Haacke 2013:281. Holotype: TM 53542 (collector J.A. van Rooyen). Type locality: "Rössing Uranium mine area, Swako- mund [sic] district (2214Db) Namibia." Global conservation status (IUCN): Not Evaluated. Global distribution: The species is known from Angola and Namibia. Ocurrences in Angola (Map 364): The species occurs in southwestern Angola. Namibe: "5 km north Namibé" [-15.20000, 12.15000] (Haacke 2013:285). Taxonomic and distributional notes: Some earlier records of T. semiannulatus polystictus in Namibia actually refer to this recently described species. MAP 364. Distribution of Telescopus finkeldeyi in Angola.

opencc-by-4.0Sep 2018View details →
zenodo40/100

ly a misidentification as today the recognized distribution of this species is restricted to western Africa. It is possible that the Cabinda frog may be referable to Phrynobatrachus auritus Boulenger, 1900. MAP 96. Distribution of Phrynobatrachus plicatus in Angola. in Diversity and Distribution of the Amphibians and Terrestrial Reptiles of Angola Atlas of Historical and Bibliographic Records (1840-2017)

ly a misidentification as today the recognized distribution of this species is restricted to western Africa. It is possible that the Cabinda frog may be referable to Phrynobatrachus auritus Boulenger, 1900. MAP 96. Distribution of Phrynobatrachus plicatus in Angola.

opencc-by-4.0Sep 2018View details →
zenodo36/100

YA Domain Dataset: Dataset of scholarly bibliographic references on YouTube videos

<p><strong>Abstract</strong></p> <p>Scholarly communication through YouTube videos has been increasing. Although Altmetric (<a href="https://altmetric.com/">https://altmetric.com/</a>) provides the dataset on such references, its coverage is unclear, and it does not contain the original external links in each video. Considering this background, we built and published a dataset of scholarly bibliographic references on YouTube videos by using YouTube Data API v3, targeting six types of domain names: "doi.org," "ncbi.nlm.nih.gov," ieeexplore.ieee.org," "link.springer.com," "onlinelibrary.wiley.com," and "sciencedirect.com." As a result, we identified approximately 480,000 references associated with Crossref DOIs among 230,000 videos published by December 31, 2023, posted on 55,000 channels. Notably, over half of these references were not covered by the Altmetric dataset, resulting in a 150% increase in the number of references when combining the dataset constructed by the proposed method with the Altmetric dataset, compared to the Altmetric dataset alone. Regarding external links, PubMed and DOI links were prominent; however, a substantial number of direct links to publisher platforms were observed. Most channels and videos contained external links to a single platform, scattered across each platform. This dataset is helpful for identifying and analyzing scholarly references on YouTube.<br>As for the original paper related to this dataset, please refer to the references section.</p> <p>&nbsp;</p> <p><strong>Data Records</strong></p> <p>The data format of the dataset is JSON lines, where each line is a single record. The data is split into files by DOI Registration Agencies. A sample of the record is as follows:</p> <table> <tbody> <tr> <td>{<br>&nbsp; &nbsp; "channel_id": "UCEfEi-IMiB87UsxY3765P6w",<br>&nbsp; &nbsp; "video_id": "e7YmyVd4uOE",<br>&nbsp; &nbsp; "is_covered_by_altmetric_com": false,<br>&nbsp; &nbsp; "youtube_data_api_search": [<br>&nbsp; &nbsp; &nbsp; &nbsp; {<br>&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; "query": "doi.org",<br>&nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; "uri": "http://dx.doi.org/10.1145/2807442.2814654"<br>&nbsp; &nbsp; &nbsp; &nbsp; }<br>&nbsp; &nbsp; ],<br>&nbsp; &nbsp; "doi": "10.1145/2807442.2814654",<br>&nbsp; &nbsp; "doiRA": "Crossref"<br>}</td> </tr> </tbody> </table> <ul> <li>channel_id (String) -- Channel ID of the YouTube channel that uploaded the video.</li> <li>video_id (String) -- Video ID.</li> <li>is_covered_by_altmetric_com (Boolean) -- Whether this reference is covered by altmetric.com or not.</li> <li>youtube_data_api_search (Array) <ul> <li>&nbsp; query (String) -- The query used in the search:list of YouTube Data API v3. (<a href="https://developers.google.com/youtube/v3/docs/search/list?hl=en">https://developers.google.com/youtube/v3/docs/search/list?hl=en</a>)</li> <li>&nbsp; uri (String)-- The original external links written in the description text or video title in each video.</li> </ul> </li> <li>doi (String) -- DOI corresponding to the bibliographic reference in the video.</li> <li>doiRA (String) -- DOI registration agency for the DOI. We obtained this data using the WhichRA? API (<a href="https://www.doi.org/the-identifier/resources/factsheets/doi-resolution-documentation#4-which-ra">https://www.doi.org/the-identifier/resources/factsheets/doi-resolution-documentation#4-which-ra</a>).</li> </ul> <p>We note that the altmetric dataset obtained from Altmetric Explorer in this study is not included in this dataset.</p> <p><strong>References</strong></p> <ul> <li>Kikkawa, Jiro; Takaku, Masao; Yoshikane, Fuyuki: "Enhancing Identification of Scholarly Reference on YouTube: Method Development and Analysis of External Link Characteristics", <em>Proceedings of the 28th International Conference on Theory and Practice of Digital Libraries (<a href="https://tpdl2024.nuk.si/">TPDL 2024</a>)</em>, Ljubljana, Slovenia, Lecture Notes in Computer Science (LNCS), Vol.15178, 2024.09. (in press).</li> </ul> <p><strong>Fundings</strong></p> <p>JSPS KAKENHI Grant Numbers <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-22K18147/">JP22K18147</a>, <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-23K11761">JP23K11761</a>, and <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-24K15652">JP24K15652</a>.</p>

opencc-by-4.0Jul 2024View details →
zenodo36/100

Bibliographic references extracted from theses and dissertations in the Electronic Theses and Dissertations Library of the Federal University of Paraná, Brazil

<p>Bibliographic references obtained from a <a href="https://doi.org/10.5281/zenodo.1169586">subset of PDF files in the Electronic Theses and Dissertations Library of the Federal University of Paran&aacute;, Brazil</a>, extracted through ParsCit.</p>

openother-pdFeb 2018View details →
zenodo36/100

PubMed inner references obtained from five freely available bibliographic data sources

<p>This dataset contains PMID-to-PMID citations of&nbsp;PubMed 2020 Baseline extracted from five freely available bibliographic data sources (COCI, Dimensions, MAG, NIH-OCC, and S2ORC).</p> <p>Each line contains one citing PubMed document and its cited references. The citing and cited documents are separated by a tab (\t) and the cited references are separated by a semicolon (;).</p>

opencc-by-4.0Aug 2021View details →
zenodo32/100

Table A.12: Bibliographic references and abbreviations in full table

Open the record for dataset details and reuse information.

opencc-by-4.0Nov 2024View details →
zenodo28/100

The Brill Knowledge Graph: A Database of Bibliographic References and Index Terms extracted from Books in Humanities and Social Sciences

<p>We present a complete dataset of linked bibliography and index data, partially disambiguated and augmented with references to external resources, extracted from the Brill&rsquo;s archive in the field of Classics. Processed book identifiers are listed in a separate&nbsp;text file. Text fragments extracted from different books via this process are then parsed and compared using a string-based similarity metric to form clusters of bibliographic references to the same published work or (variants of) the same subjects discussed in these books. The entire set of references was then disambiguated using Google Books and Crossref APIs.</p> <p><a href="https://jdmdh.episciences.org/11062">Paper about extraction pipeline</a></p> <p><a href="https://www.nkokash.com/documents/KIEM-RDJ.pdf">Paper about extracted KG</a></p> <p>&nbsp;</p>

opencc-by-4.0Mar 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record