Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
9
datasets available to search
ShareScore release 0.9.0
Dataset results
9 results for “Bibliographic references”
Raw and aggregated data for the study introduced in the paper "The way we cite: common metadata used across disciplines for defining bibliographic references"
<p>These data have been gathered in the context of a study aiming to investigate citation practices for referencing different types of entities and, in particular, for understanding the most used metadata in bibliographic references. The data are stored in two documents in XLSX format:</p> <ul> <li>file "links-intext-pointers-and-cited-entity-types.xlsx" - it contains information about whether the in-text reference pointers of the various PDF articles of the corpus have specified hypertextual links from the in-text reference pointers to the denoted bibliographic reference, plus information about the types of all the entities cited by each article in the corpus;</li> <li>file "metadata-bibliographic-references.xlsm" - it contains information about the metadata used to identify the various descriptive elements of all the bibliographic references defined in the article of the corpus.</li> </ul> <p>The methodology used to gather all these data is described in:</p> <blockquote> <p>Santos, E. A. d., Peroni, S., Mucheroni, M. L.: Workflow for retrieving all the data of the analysis introduced in the article "Citing and referencing habits in Medicine and Social Sciences journals in 2019". (2020), <a href="https://doi.org/10.17504/protocols.io.bbifikbn">https://doi.org/10.17504/protocols.io.bbifikbn</a></p> </blockquote>
Dataset of first appearances of the scholarly bibliographic references on English Wikipedia articles as of 1 March 2017 and as of 1 October 2021
<p><strong>Abstract</strong></p> <p>We developed a methodology to detect the oldest scholarly reference added to Wikipedia articles by which a certain paper is uniquely identifiable as the "first appearance of the scholarly reference." We identified the first appearances of 923,894 scholarly references (611,119 unique DOIs) in180,795 unique pages on English Wikipedia as of March 1, 2017, and stored them in the dataset. Moreover, we assessed the precision of the dataset, which was and it was a high precision regardless of the research field. In this version, it is available not only the dataset of English Wikipedia as of March 1, 2017, but also English Wikipedia as of October 1, 2021, generated by using the same methodology.</p> <p> </p> <p><strong>Data Records</strong></p> <p>The data format of the dataset is JSON lines, where each line is a single record. In this dataset, we detected the first appearance of each scholarly reference added to Wikipedia articles. If there are multiple references corresponding to the same paper on the same page, only the oldest one is collected. Sample of the record is the following.</p> <ul> <li>doi -- DOI corresponding to the paper (String), e.g., "10.1006/anbe.1996.0497"</li> <li>paper_type -- Document type of the paper (String), e.g., "journal-article"</li> <li>paper_container_title -- Journal title, book title, or proceedings title (Array of String), e.g., ["Animal Behaviour"]</li> <li>paper_publisher -- Publisher name (String), e.g., "Elsevier BV"</li> <li>paper_title -- Paper title (Array of String), e.g., ["Push or pull: an experimental study on imitation in marmosets"]</li> <li>paper_published_year -- Published year (String), e.g., "1997"</li> <li>paper_issue -- Issue number (String), e.g., "4"</li> <li>paper_volume -- Volume number (String), e.g., "54"</li> <li>paper_page -- Page numbers (String), e.g., "817-831"</li> <li>paper_author -- Authors information consisted of given and family names, sequences (order in author names), and affiliations (Array of JSON), e.g., [{"given":"THOMAS", "family":"BUGNYAR", "sequence":"first", "affiliation":[]}, {"given":"LUDWIG", "family":"HUBER", "sequence":"additional", "affiliation":[]}]</li> <li>issn -- ISSN related to the paper (Array of String), e.g., ["0003-3472"]</li> <li>research_field -- Research fields from ESI categories (Array of String), e.g., ["PLANT & ANIMAL SCIENCE"]</li> <li>page_id -- Page id (String), e.g., "577858"</li> <li>page_title -- Page title (String), e.g., "Imitation"</li> <li>revision_id -- Revision id (String), e.g., "203309031"</li> <li>revision_timestamp -- Revision timestamp (String), e.g., "2008-04-04 15:54:09 UTC"</li> <li>revision_comment -- Revision comment (edit summary) (String), e.g., "/* Animal Behaviour */"</li> <li>editor_name -- Wikipedia editor's name (String), e.g., "Nicemr"</li> <li>editor_type -- Type of the editor (String), e.g., "User"</li> </ul> <p><strong>References</strong></p> <ul> <li>Kikkawa, J., Takaku, M. & Yoshikane, F. "Dataset of first appearances of the scholarly bibliographic references on Wikipedia articles", Scientific Data, Vol. 9, Article number 85, pp. 1-11, 2022. <a href="https://doi.org/10.1038/s41597-022-01190-z">https://doi.org/10.1038/s41597-022-01190-z</a>.</li> </ul> <p><strong>FUNDING</strong></p> <ul> <li>JSPS KAKENHI Grant Number <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-20K12543">JP20K12543</a></li> <li>JSPS KAKENHI Grant Number <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-21K21303">JP21K21303</a></li> </ul>
Telescopus finkeldeyi Haacke, 2013 DAMARA TIGER SNAKE Telescopus finkeldeyi Haacke 2013:281. Holotype: TM 53542 (collector J.A. van Rooyen). Type locality: "Rössing Uranium mine area, Swako- mund [sic] district (2214Db) Namibia." Global conservation status (IUCN): Not Evaluated. Global distribution: The species is known from Angola and Namibia. Ocurrences in Angola (Map 364): The species occurs in southwestern Angola. Namibe: "5 km north Namibé" [-15.20000, 12.15000] (Haacke 2013:285). Taxonomic and distributional notes: Some earlier records of T. semiannulatus polystictus in Namibia actually refer to this recently described species. MAP 364. Distribution of Telescopus finkeldeyi in Angola. in Diversity and Distribution of the Amphibians and Terrestrial Reptiles of Angola Atlas of Historical and Bibliographic Records (1840-2017)
Telescopus finkeldeyi Haacke, 2013 DAMARA TIGER SNAKE Telescopus finkeldeyi Haacke 2013:281. Holotype: TM 53542 (collector J.A. van Rooyen). Type locality: "Rössing Uranium mine area, Swako- mund [sic] district (2214Db) Namibia." Global conservation status (IUCN): Not Evaluated. Global distribution: The species is known from Angola and Namibia. Ocurrences in Angola (Map 364): The species occurs in southwestern Angola. Namibe: "5 km north Namibé" [-15.20000, 12.15000] (Haacke 2013:285). Taxonomic and distributional notes: Some earlier records of T. semiannulatus polystictus in Namibia actually refer to this recently described species. MAP 364. Distribution of Telescopus finkeldeyi in Angola.
ly a misidentification as today the recognized distribution of this species is restricted to western Africa. It is possible that the Cabinda frog may be referable to Phrynobatrachus auritus Boulenger, 1900. MAP 96. Distribution of Phrynobatrachus plicatus in Angola. in Diversity and Distribution of the Amphibians and Terrestrial Reptiles of Angola Atlas of Historical and Bibliographic Records (1840-2017)
ly a misidentification as today the recognized distribution of this species is restricted to western Africa. It is possible that the Cabinda frog may be referable to Phrynobatrachus auritus Boulenger, 1900. MAP 96. Distribution of Phrynobatrachus plicatus in Angola.
YA Domain Dataset: Dataset of scholarly bibliographic references on YouTube videos
<p><strong>Abstract</strong></p> <p>Scholarly communication through YouTube videos has been increasing. Although Altmetric (<a href="https://altmetric.com/">https://altmetric.com/</a>) provides the dataset on such references, its coverage is unclear, and it does not contain the original external links in each video. Considering this background, we built and published a dataset of scholarly bibliographic references on YouTube videos by using YouTube Data API v3, targeting six types of domain names: "doi.org," "ncbi.nlm.nih.gov," ieeexplore.ieee.org," "link.springer.com," "onlinelibrary.wiley.com," and "sciencedirect.com." As a result, we identified approximately 480,000 references associated with Crossref DOIs among 230,000 videos published by December 31, 2023, posted on 55,000 channels. Notably, over half of these references were not covered by the Altmetric dataset, resulting in a 150% increase in the number of references when combining the dataset constructed by the proposed method with the Altmetric dataset, compared to the Altmetric dataset alone. Regarding external links, PubMed and DOI links were prominent; however, a substantial number of direct links to publisher platforms were observed. Most channels and videos contained external links to a single platform, scattered across each platform. This dataset is helpful for identifying and analyzing scholarly references on YouTube.<br>As for the original paper related to this dataset, please refer to the references section.</p> <p> </p> <p><strong>Data Records</strong></p> <p>The data format of the dataset is JSON lines, where each line is a single record. The data is split into files by DOI Registration Agencies. A sample of the record is as follows:</p> <table> <tbody> <tr> <td>{<br> "channel_id": "UCEfEi-IMiB87UsxY3765P6w",<br> "video_id": "e7YmyVd4uOE",<br> "is_covered_by_altmetric_com": false,<br> "youtube_data_api_search": [<br> {<br> "query": "doi.org",<br> "uri": "http://dx.doi.org/10.1145/2807442.2814654"<br> }<br> ],<br> "doi": "10.1145/2807442.2814654",<br> "doiRA": "Crossref"<br>}</td> </tr> </tbody> </table> <ul> <li>channel_id (String) -- Channel ID of the YouTube channel that uploaded the video.</li> <li>video_id (String) -- Video ID.</li> <li>is_covered_by_altmetric_com (Boolean) -- Whether this reference is covered by altmetric.com or not.</li> <li>youtube_data_api_search (Array) <ul> <li> query (String) -- The query used in the search:list of YouTube Data API v3. (<a href="https://developers.google.com/youtube/v3/docs/search/list?hl=en">https://developers.google.com/youtube/v3/docs/search/list?hl=en</a>)</li> <li> uri (String)-- The original external links written in the description text or video title in each video.</li> </ul> </li> <li>doi (String) -- DOI corresponding to the bibliographic reference in the video.</li> <li>doiRA (String) -- DOI registration agency for the DOI. We obtained this data using the WhichRA? API (<a href="https://www.doi.org/the-identifier/resources/factsheets/doi-resolution-documentation#4-which-ra">https://www.doi.org/the-identifier/resources/factsheets/doi-resolution-documentation#4-which-ra</a>).</li> </ul> <p>We note that the altmetric dataset obtained from Altmetric Explorer in this study is not included in this dataset.</p> <p><strong>References</strong></p> <ul> <li>Kikkawa, Jiro; Takaku, Masao; Yoshikane, Fuyuki: "Enhancing Identification of Scholarly Reference on YouTube: Method Development and Analysis of External Link Characteristics", <em>Proceedings of the 28th International Conference on Theory and Practice of Digital Libraries (<a href="https://tpdl2024.nuk.si/">TPDL 2024</a>)</em>, Ljubljana, Slovenia, Lecture Notes in Computer Science (LNCS), Vol.15178, 2024.09. (in press).</li> </ul> <p><strong>Fundings</strong></p> <p>JSPS KAKENHI Grant Numbers <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-22K18147/">JP22K18147</a>, <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-23K11761">JP23K11761</a>, and <a href="https://kaken.nii.ac.jp/en/grant/KAKENHI-PROJECT-24K15652">JP24K15652</a>.</p>
Bibliographic references extracted from theses and dissertations in the Electronic Theses and Dissertations Library of the Federal University of Paraná, Brazil
<p>Bibliographic references obtained from a <a href="https://doi.org/10.5281/zenodo.1169586">subset of PDF files in the Electronic Theses and Dissertations Library of the Federal University of Paraná, Brazil</a>, extracted through ParsCit.</p>
PubMed inner references obtained from five freely available bibliographic data sources
<p>This dataset contains PMID-to-PMID citations of PubMed 2020 Baseline extracted from five freely available bibliographic data sources (COCI, Dimensions, MAG, NIH-OCC, and S2ORC).</p> <p>Each line contains one citing PubMed document and its cited references. The citing and cited documents are separated by a tab (\t) and the cited references are separated by a semicolon (;).</p>
Table A.12: Bibliographic references and abbreviations in full table
Open the record for dataset details and reuse information.
The Brill Knowledge Graph: A Database of Bibliographic References and Index Terms extracted from Books in Humanities and Social Sciences
<p>We present a complete dataset of linked bibliography and index data, partially disambiguated and augmented with references to external resources, extracted from the Brill’s archive in the field of Classics. Processed book identifiers are listed in a separate text file. Text fragments extracted from different books via this process are then parsed and compared using a string-based similarity metric to form clusters of bibliographic references to the same published work or (variants of) the same subjects discussed in these books. The entire set of references was then disambiguated using Google Books and Crossref APIs.</p> <p><a href="https://jdmdh.episciences.org/11062">Paper about extraction pipeline</a></p> <p><a href="https://www.nkokash.com/documents/KIEM-RDJ.pdf">Paper about extracted KG</a></p> <p> </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.