Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

214

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

214 results for “citation”

Learn how ShareScore rates datasets ↗
zenodo40/100

Figs 1-6 in New citations of Alysiini from Spain, with a description of Dinotrema mediocornis hispanicum nov.ssp. and of the females of Aspilota inflatinervis and Synaldis azorica (Hymenoptera, Braconidae, Alysiinae)

Figs 1-6: Aspilota inflatinervis FISCHER: (1) head in fronto-lateral view with detail of antennae. (2) right anterior wing. Aspilota mediocornis hispanicum ssp. n.: (3) head in lateral view with detail of antennae. (4) right anterior wing. Aspilota azorica FISCHER: (5) right mandible. (6) mesosoma with detail of propleuron, mesopleuron, and propodeum.

opencc-by-4.0Dec 2008View details →
zenodo40/100

Citation Location, Elements and Purpose of ICPSR Research Data Citations

<p>The dataset contains an analysis of a randomly chosen sample of 1,073 publications that cite ICPSR research datasets. The collected research data (re)use indications were analyzed according to their location in the full-text, their metadata elements, and their citation purpose.</p> <p>The data was collected and analyzed in 2020 for a PhD thesis on research data and software (re)use indications in scholarly works.</p>

opencc-by-4.0Jun 2021View details →
zenodo40/100

COVID-19++: A Citation-Aware Covid-19 Dataset for the Analysis of Research Dynamics

<p>COVID-19++ is a citation-aware COVID-19 dataset for the analysis of research dynamics. In addition to primary COVID-19 related articles and preprints from 2020, it includes citations and the metadata of first-order cited work. All publications are annotated with MeSH terms, either from the ground truth, or via ConceptMapper, if no ground truth was available.&nbsp;</p> <p>The data is organized in CSV files</p> <p>- Paper metadata (paper_id, publdate, title, data_source): paper.csv</p> <p>- Annotation data, mapping paper_id to MeSH terms: annotation.csv&nbsp;</p> <p>- Authorship data, mapping paper_id to author, optionally with ORCID: authorship.csv<br> - Paired DOIs of citing and cited papers: references.csv</p> <p>The column data source within the paper metadata has the value KE (for metadata from ZB MED KE), PP (for preprints) or CR (for cited resources from CrossRef)<br> &nbsp;</p> <p>This work was supported by BMBF within the programme ``Quantitative Wissenschaftsforschung&#39;&#39; under grant numbers 01PU17013A, 01PU17013B, 01PU17013C.<br> &nbsp;</p>

opencc-by-4.0Sep 2021View details →
zenodo40/100

DATABASE: Electric Vehicle, Battery and Smart Grid patent citation networks and main paths.

<p>This dataset comprises the original patent citation networks that were created to calculate&nbsp;the main citation paths for the technologies of Electric Vehicle, Battery and Smart Grid.</p> <p>For each technology (1- Electric Vehicle, 2- Battery, 3- Smart Grid), four outputs are provided:</p> <p>a- Patent extraction: USPTO patents filtered by IPC or CPC and found in the Triadic Patent Families database (OECD, 2021)&nbsp;&nbsp;</p> <p>b- Full nodes and links reconstructed by following patent citations through a snowball method&nbsp;(until no further patents found)</p> <p>c- Filtered nodes and links according to keywords</p> <p>d- Main path nodes and links (with citation weights).</p> <p>For a detailed explanation of the methodology please refer to the submitted paper:</p> <p><strong>Transitions as a coevolutionary process: the urban emergence of electric vehicle inventions</strong></p>

opencc-by-4.0Sep 2021View details →
zenodo40/100

Use and sharing of raw data in the Journal Citation Reports' Emergency Medicine Category: Metrics and Journals including supplementary material classification sorted by quartile of the JCR emergency medicine category.

<p>Raw data belonged to the study of use and sharing of raw research data in the Journal Citation Reports&#39; Emergency Medicine Category.</p>

opencc-by-4.0Dec 2018View details →
zenodo40/100

Data for "Measuring Back: Bibliodiversity and the Journal Impact Factor brand. A Case study of IF-journals included in the 2021 Journal Citations Report."

<p>This is the open data for the preprint &quot;Measuring Back: Bibliodiversity and the Journal Impact Factor brand.&nbsp; A Case study of IF-journals included in the 2021 Journal Citations Report.&quot;</p>

opencc-by-4.0Feb 2023View details →
zenodo40/100

Bibliographic data on datasets affiliated to Poznan University of Technology and indexed in Data Citation Index (retrieved by Web of Science service in January 2023))

<p>The file contains the number of datasets published by the researchers affiliated to Poznan University of Technology and indexed in Data Citation Index provided by Web of Science (database updated 10.01.2023). The Search was performed using the name of institution in the &#39;Affiliation&#39; field. Dataset contains two files in two diffrent formats: plain text and xls.</p>

opencc-byJan 2023View details →
zenodo40/100

Highly cited tropical medicine articles in the early COVID 19 pandemic. Original Excel data on citation and subjects

<p><strong>ORIGINAL DATA SET FOR STUDY OF PUBLICATION AND CITATION TRENDS, TROPICAL MEDICINE, EARLY COVID 19 PANDEMIC. Background: </strong>An adequate response to health needs includes the identification of research patterns about the large number of people living in the tropics and subjected to tropical diseases. Studies have shown that research does not always match the real needs of those populations, and that citation reflects mostly the amount of money behind particular publications. Here we test the hypothesis that research from richer institutions is published in better-indexed journals, and thus has greater citation rates.</p> <p><strong>Methods:</strong> The data in this study was extracted from the Science Citation Index Expanded database; the 2020 journal Impact Factor (<em>IF</em><sub>2020</sub>) was updated to 30 June 2021. We considered places, subjects, institutions and journals.</p> <p><strong>Results:</strong> We identified 1 041 highly cited articles with 100 citations or more in the category of tropical medicine. About a decade is needed for an article to reach peak citation. Only two Covid-19 related were highly cited in the last three years. Most cited articles were published by the journals <em>Memorias Do Instituto Oswaldo Cruz</em> (Brazil), <em>Acta Tropica</em> (Switzerland), and <em>PLoS Neglected Tropical Diseases</em> (USA). The USA dominated five of the six publication indicators. International collaboration articles had more citations than single-country articles. The UK, South Africa, and Switzerland had high citation rates, as did the London School of Hygiene and Tropical Medicine in the UK, the Centers for Disease Control and Prevention in the USA, and the WHO in Switzerland.</p> <p><strong>Conclusions:</strong> About ten years of accumulated citations are needed to get 100 citations or more as highly cited articles in the Web of Science category of tropical medicine. Six publication and citation indicators, including authors&rsquo; publication potential and characteristics evaluated by <em>Y</em>-index, indicate that the currently available indexing system places tropical researchers at a disadvantage against their colleagues in temperate countries, and suggest that, to progress towards better control of tropical diseases, international collaboration should increase, and other tropical countries should follow the example of Brazil, which provides significant financing to its scientific community.Juli&aacute;n Monge-N&aacute;jera<sup>1</sup>, and Yuh-Shan Ho<sup>2</sup>*</p> <p><sup>1</sup>Laboratorio de Ecolog&iacute;a Urbana, Vicerrector&iacute;a de Investigaci&oacute;n, Universidad Estatal a Distancia, 2050 San Jos&eacute;, Costa Rica; <a href="mailto:julianmonge@gmail.com"><em>julianmonge@gmail.com</em></a> (https://orcid.org/0000-0001-7764-2966)</p> <p>*Corresponding author: Trend Research Centre, Asia University, No. 500 Lioufeng Road, Wufeng, Taichung 41354, Taiwan; <a href="mailto:ysho@asia.edu.tw">ysho@asia.edu.tw</a> (<em>https://orcid.org/0000-0002-2557-8736</em>)</p>

opencc-by-4.0Apr 2023View details →
zenodo40/100

Signed Citation of Provenance of GBIF Occurrence Downloads referenced in Chesshire et al. 2023 doi:10.1111/ecog.06584 hash://sha256/9e3ca96d94229e20f47c14efaa59f793845aa37d9f6c698d2dd35876705e9feb hash://md5/43652e3d26989008026e092e3f04b04d

<p>Chesshire et al. 2023. scientific publication [1] used and referenced&nbsp; three GBIF mediated occurrence download queries [2,3,4]&nbsp;and associated data. However, in their GBIF records indicate that the data associated with the three download queries are slated for removal at any point after&nbsp;2021-08-03 . This publication explicitly references the DOIs associated with [2,3,4] and documents the provenance of their associated meta-data records. The provenance was captured using Preston [5,6], a biodiversity dataset tracker.&nbsp;</p> <p>The signed citation of this provenance publication can be derived from:</p> <pre><code class="language-bash">preston history\ --anchor hash://sha256/9e3ca96d94229e20f47c14efaa59f793845aa37d9f6c698d2dd35876705e9feb\ --remote https://zenodo.org/record/7849559/files</code></pre> <pre><code>&lt;hash://sha256/f2d8bdaec7a416a0039e9398cf07c6fa69083f64a6f22de3f252ebb5dd4fd412&gt; &lt;http://www.w3.org/ns/prov#wasDerivedFrom&gt; &lt;hash://sha256/c457565ea0cec7b0392f1271fcda08440919f03bbf29bb8df1eb926c78a972cc&gt; . &lt;urn:uuid:0659a54f-b713-4f86-a917-5be166a14110&gt; &lt;http://purl.org/pav/hasVersion&gt; &lt;hash://sha256/c457565ea0cec7b0392f1271fcda08440919f03bbf29bb8df1eb926c78a972cc&gt; .</code></pre> <p>And their tracked content include, as obtained via&nbsp;</p> <pre><code class="language-bash">preston alias\ --anchor hash://sha256/9e3ca96d94229e20f47c14efaa59f793845aa37d9f6c698d2dd35876705e9feb\ --remote https://zenodo.org/record/7849559/files</code></pre> <table> <caption>Tracked content associated with hash://sha256/9e3ca96d94229e20f47c14efaa59f793845aa37d9f6c698d2dd35876705e9feb</caption> <tbody> <tr> <td>content location</td> <td>content relation</td> <td>content id</td> </tr> <tr> <td>https://doi.org/10.15468/dl.6cxfsw</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/6b8b5f79af53dee98c3654b945389628194c9e9f0ad610852327574b3f99ff7a</td> </tr> <tr> <td>https://doi.org/10.15468/dl.b9rfa7</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/a74cfe8a6c6b7d2361f41cc04979c262b5fbba60c0992253ae78fe6d31f414bb</td> </tr> <tr> <td>https://doi.org/10.15468/dl.w2nndm</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/6a587a219e78ff2674fbb54d99fed48c21b77bc46608d8f991e79ee06a547fac</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182006-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/1c5d8a7399793a634a0dde32f3a94ccf64199f010d7f93baa422c2e1dbb98b2f</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182032-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/23d7c875420bea71d24c1ec3ba127f91eff5b368744de14824de0fc4fc090bb2</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182076-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/6555d581e0ce75c77740811e547da726297d02369b149893faf531f132a2aff0</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182006-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/2c4c4f4cd1151bc65394466416b066c19422fe22b8eb64c5c144fb7889ea2f16</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182032-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/e4e9742259e9232c773ab157e34af1cfebfd09050effb49c15db032057fc5750</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/0182076-200613084148143</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/20915d475c63fa6f96ab127ff5efb5554df40208596244349d110432b478168b</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/request/0182006-200613084148143.zip</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/d14a14e549e3caa8965daecad6fcb0cfddd4be12fb78a495b248c380df41db9b</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/request/0182032-200613084148143.zip</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/3fc1b6491813f5d7e2d32b7c6cadb1ae60558f31a4489e23735c43bd74ed4db6</td> </tr> <tr> <td>https://api.gbif.org/v1/occurrence/download/request/0182076-200613084148143.zip</td> <td>http://purl.org/pav/hasVersion</td> <td>hash://sha256/7ddea84a67329ec8eea389d09798e5b6d60d86c39b975590f117679cdbbe8e20</td> </tr> </tbody> </table> <p>This data publication, and associated tracked content, can be cloned using:</p> <pre><code class="language-bash">preston clone https://zenodo.org/record/7849559/files</code></pre> <p><br> Note that the original publication dated 2023-04-03 did *not* include the associated tracked data retrieved from https://api.gbif.org/v1/occurrence/download/request/0182006-200613084148143.zip, https://api.gbif.org/v1/occurrence/download/request/0182032-200613084148143.zip, https://api.gbif.org/v1/occurrence/download/request/0182076-200613084148143.zip. However, on 2023-03-17, the data associated with [2], [3], [4] were still marked for deletion in the GBIF ecosystem,&nbsp;two weeks after the respective DOIs were first cited in the v0.1 of&nbsp;this data publication. This 2023-03-17 publication includes tracked content that shows the associated data is marked for deletion, and contains the associated data archives. &nbsp;</p> <p>The example below shows a tracked versions of the metadata associated with the download request/query doi:10.15468/dl.w2nndm &nbsp;[4] indicates that the associated data is scheduled to be &quot;eraseAfter&quot; &nbsp;&quot;2021-08-03T19:18:46.611+00:00&quot;.&nbsp;</p> <pre><code class="language-bash">preston cat\ --remote https://zenodo.org/record/7837572/files\ hash://sha256/6555d581e0ce75c77740811e547da726297d02369b149893faf531f132a2aff0\ | jq .</code></pre> <p>&nbsp;</p> <pre><code class="language-json">{ "key": "0182076-200613084148143", "doi": "10.15468/dl.w2nndm", "license": "http://creativecommons.org/licenses/by-nc/4.0/legalcode", "request": { "predicate": { "type": "and", "predicates": [ { "type": "equals", "key": "DATASET_KEY", "value": "e05f6e7d-418e-4407-8e0f-7b8ccf21109e", "matchCase": false }, { "type": "or", "predicates": [ { "type": "equals", "key": "TAXON_KEY", "value": "4334", "matchCase": false }, { "type": "equals", "key": "TAXON_KEY", "value": "4345", "matchCase": false }, { "type": "equals", "key": "TAXON_KEY", "value": "7911", "matchCase": false }, { "type": "equals", "key": "TAXON_KEY", "value": "7908", "matchCase": false }, { "type": "equals", "key": "TAXON_KEY", "value": "7901", "matchCase": false }, { "type": "equals", "key": "TAXON_KEY", "value": "7905", "matchCase": false } ] } ] }, "sendNotification": true, "format": "DWCA", "type": "OCCURRENCE", "verbatimExtensions": [] }, "created": "2021-02-03T19:18:46.687+00:00", "modified": "2021-02-03T19:20:03.899+00:00", "eraseAfter": "2021-08-03T19:18:46.611+00:00", "status": "SUCCEEDED", "downloadLink": "https://api.gbif.org/v1/occurrence/download/request/0182076-200613084148143.zip", "size": 2624689, "totalRecords": 11654, "numberDatasets": 1 }</code></pre> <p>Also, on after (re-)running</p> <pre><code class="language-bash">preston track\ https://doi.org/10.15468/dl.6cxfsw\ https://doi.org/10.15468/dl.b9rfa7\ https://doi.org/10.15468/dl.w2nndm</code></pre> <p>on 2023-04-20, the download record metadata retrieved from&nbsp;&nbsp;https://api.gbif.org/v1/occurrence/download/0182006-200613084148143 and associated with&nbsp;https://doi.org/10.15468/dl.6cxfsw appeared to no longer be marked for deletion, as shown by the difference between a pre-2023-04-20 version (i.e.&nbsp;hash://sha256/1c5d8a7399793a634a0dde32f3a94ccf64199f010d7f93baa422c2e1dbb98b2f) with the newly retrieved response on 2023-04-20 (i.e.,&nbsp;hash://sha256/2c4c4f4cd1151bc65394466416b066c19422fe22b8eb64c5c144fb7889ea2f16).</p> <p>The difference is highlighted below using the diff and preston tools via</p> <pre><code class="language-bash">diff\ &lt;(preston cat hash://sha256/2c4c4f4cd1151bc65394466416b066c19422fe22b8eb64c5c144fb7889ea2f16 | jq .)\ &lt;(preston cat hash://sha256/1c5d8a7399793a634a0dde32f3a94ccf64199f010d7f93baa422c2e1dbb98b2f | jq .)</code></pre> <p>yielding:</p> <pre><code class="language-diff">116c116,117 &lt; "modified": "2023-04-18T08:09:09.757+00:00", --- &gt; "modified": "2021-02-03T18:00:50.416+00:00", &gt; "eraseAfter": "2021-08-03T17:50:18.453+00:00",</code></pre> <p>&nbsp;This observation is consistent with the 2023-04-18 claim by Daniel Noesgaard [7]&nbsp;that associated download records are no longer marked for deletion.</p> <p><strong>References&nbsp;</strong></p> <p>[1]&nbsp;Chesshire, P.R., Fischer, E.E., Dowdy, N.J., Griswold, T.L., Hughes, A.C., Orr, M.C., Ascher, J.S., Guzman, L.M., Hung, K.-L.J., Cobb, N.S. and McCabe, L.M. (2023), Completeness analysis for over 3000 United States bee species identifies persistent data gap. Ecography e06584.&nbsp;<a href="https://doi.org/10.1111/ecog.06584">https://doi.org/10.1111/ecog.06584</a></p> <p>[2]&nbsp;GBIF.org (3 February 2021) GBIF Occurrence Download&nbsp;<a href="https://doi.org/10.15468/dl.6cxfsw">https://doi.org/10.15468/dl.6cxfsw</a></p> <p>[3]&nbsp;GBIF.org (3 February 2021) GBIF Occurrence Download&nbsp;<a href="https://doi.org/10.15468/dl.b9rfa7">https://doi.org/10.15468/dl.b9rfa7</a></p> <p>[4]&nbsp;GBIF.org (3 February 2021) GBIF Occurrence Download&nbsp;<a href="https://doi.org/10.15468/dl.w2nndm">https://doi.org/10.15468/dl.w2nndm</a></p> <p>[5]&nbsp;MJ Elliott, JH Poelen, JAB Fortes (2020). Toward Reliable Biodiversity Dataset References. Ecological Informatics.&nbsp;<a href="https://doi.org/10.1016/j.ecoinf.2020.101132">https://doi.org/10.1016/j.ecoinf.2020.101132</a></p> <p>[6] Elliott, M. J., Poelen, J. H., &amp; Fortes, J. (2022, August 29, accepted with minor revisions). Signed Citations: Making Persistent and Verifiable Citations of Digital Scientific Content.&nbsp;<a href="https://doi.org/10.31222/osf.io/wycjn">https://doi.org/10.31222/osf.io/wycjn</a></p> <p>[7] Noesgaard, D. 2023. https://discourse.gbif.org/t/data-queries-doi-10-15468-dl-6cxfsw-doi-10-15468-dl-b9rfa7-doi-10-15468-dl-w2nndm-used-in-chesshire-et-al-2023-were-cited-but-remain-marked-for-deletion/3915/2 accessed at 2023-04-20 .</p>

opencc-zeroApr 2023View details →
zenodo40/100

Exploring the Impact of Negative Sampling on Patent Citation Recommendation

<ul> <li> <p><strong>pcr_patents.csv </strong>is the dataset which is generated by collecting samples randomly from Google Patents by exploiting a <a href="https://pypi.org/project/google-patent-scraper/">Python library</a>. The dataset comprises around 250,000 US patents and their titles, abstracts, and citations.&nbsp; Each patent has roughly on average 27 citations.</p> </li> </ul> <p>The zip file contains 3 different datasets for training and testing patent citation recommendation systems. These datasets were generated by utilizing the main dataset. They consist of around 1 million instances which are positive as well as negative samples.&nbsp;&nbsp;</p> <ul> <li> <p><strong>pcr_cpc_negative_sample_data.csv</strong> &nbsp;consists of negative samples that were generated based on CPC subclass codes.&nbsp;</p> </li> <li> <p><strong>pcr_random_negative_sample_data.csv</strong> consists of negative samples that were generated randomly.&nbsp;</p> </li> <li> <p><strong>pcr_sem_sim_negative_sample_data_2.csv</strong> consists of negative samples that were generated based on nearest neighbor relation.</p> </li> </ul>

opencc-by-4.0Apr 2023View details →
zenodo40/100

Using Open Citation Databases for Snowballing in Software Engineering Research

<p>Dataset for our study on the coverage of software engineering articles in open citation databases:</p> <ul> <li>a list of the 23 sampled venues with their respective CORE ranks and publishers, <ul> <li>01-venues.csv,</li> </ul> </li> <li>a list of the 204 sampled articles with their respective number of references/citations per citation database, <ul> <li>02-articles.csv (articles with publication information),</li> <li>03-references-absolute.csv (number of references in published PDF &amp; absolute numbers for reference coverage in databases),</li> <li>04-references-relative.csv (relative numbers for reference coverage in databases),</li> <li>05-citations-absolute.csv (absolute numbers for citation coverage in databases),</li> <li>06-citations relative.csv (relative numbers for citation coverage in databases),</li> </ul> </li> <li>a list of the 8 articles analyzed in more detail with complete references data from the citation databases, <ul> <li>07-selected-articles.csv (articles with publication information),</li> <li>08A&ndash;08H (comparison of references found in databases for each article),</li> </ul> </li> <li>and additional statistical measures and plots <ul> <li>09-Statistics.{pdf,xlsx} (statistical measures &ndash; i.e., minimum, maximum, median, average, variance &ndash; for the whole dataset and for subsets by publisher, CORE rank, or year of publication),</li> <li>10-Figures.zip (figures for references as shown in the study and additional figures for citations &ndash; each in EPS and PNG format).</li> </ul> </li> </ul>

opencc-by-4.0May 2023View details →
zenodo40/100

Uncovering the Citation Landscape: Exploring OpenCitations COCI, OpenCitations Meta, and ERIH-PLUS in Social Sciences and Humanities Journals - DATA PRODUCED

<p>This zipped folders contain all the data produced for the research &quot;Uncovering the Citation Landscape: Exploring OpenCitations COCI, OpenCitations Meta, and ERIH-PLUS in Social Sciences and Humanities Journals&quot;: the results datasets (dataset_map_disciplines, dataset_no_SSH, dataset_SSH, erih_meta_with_disciplines and erih_meta_without_disciplines).</p> <ul> <li> <p><strong>dataset_map_disciplines.zip </strong>contains CSV files with four columns (&quot;id&quot;, &quot;citing&quot;, &quot;cited&quot;, &quot;disciplines&quot;) giving information about publications stored in OpenCitations META (version 3 released on February 2023) and&nbsp; part of&nbsp; SSH journals, according to ERIH PLUS (version downloaded on 2023-04-27), specifying the disciplines associated to them and a boolean value stating if they cite or are cited, according to the OpenCitations COCI dataset (version 19 released on January 2023).</p> </li> <li> <p><strong>dataset_no_SSH.zip </strong>and <strong>dataset_SSH.zip</strong> contain CSV files with the same structure. Each dataset has four columns: &quot;citing&quot;, &quot;is_citing_SSH&quot;, &quot;cited&quot;, and &quot;is_cited_SSH&quot;. &rdquo;Citing&rdquo; and &ldquo;cited&rdquo; columns are filled with DOIs of publications stored in OpenCitations META that according to OpenCitations COCI are involved in a citation. The &quot;is_citing_SSH&quot; and &quot;is_cited_SSH&quot; columns contain boolean values: &quot;True&quot; if the corresponding publication is associated with a SSH (Social Sciences and Humanities) discipline, according to ERIH PLUS,&nbsp; and &quot;False&quot; otherwise. The two datasets are built starting from the two different subsets obtained as a result of&nbsp; the union between OpenCitations META and ERIH PLUS: dataset_SSH comes from erih_meta_with_disciplines and dataset_no_SSH from <strong>erih_meta_without_disciplines. </strong>dataset_no_SSH comes from <strong>erih_meta_with_disciplines.zip</strong> and erih_meta_without_disciplines.zip, as explained before, contain CSV files originating from ERIH PLUS and META. erih_meta_without_disciplines has just one column &ldquo;id&rdquo; and contains the DOIs of all the publications in META that do not have any discipline associated, that is, have not been published on a SSH journal, while erih_meta_with_disciplines derives from all the publications in META that have at least one linked discipline and has two columns: &ldquo;id&rdquo; and &ldquo;erih_disciplines&rdquo;, containing a string with all the disciplines linked to that publication like &quot;History, Interdisciplinary research in the Humanities, Interdisciplinary research in the Social Sciences, Sociology&quot;.</p> </li> </ul> <p>Software:&nbsp;https://doi.org/10.5281/zenodo.8326023</p> <p>Data preprocessed:&nbsp;https://doi.org/10.5281/zenodo.7973159</p> <p>Article:&nbsp;https://zenodo.org/record/8326044</p> <p>DMP:&nbsp;https://zenodo.org/record/8324973</p> <p>Protocol:&nbsp;https://doi.org/10.17504/protocols.io.n92ldpeenl5b/v5</p>

opencc-by-4.0Sep 2023View details →
zenodo40/100

Uncovering the Citation Landscape: Exploring OpenCitations COCI, OpenCitations Meta, and ERIH-PLUS in Social Sciences and Humanities Journals - DATA PREPROCESSED

<p>This zipped folders contain all the data preprocessed for the research &quot;Uncovering the Citation Landscape: Exploring OpenCitations COCI, OpenCitations Meta, and ERIH-PLUS in Social Sciences and Humanities Journals&quot;: the cleaned datasets (coci_preprocessed, meta_preprocessed, erih_preprocessed and erih_meta).</p> <ul> <li> <p><strong>coci_preprocessed.zip</strong>: this archive contains CSVs with two columns &ldquo;citing&rdquo; and &ldquo;cited&rdquo;, giving information about publications involved in citations according to the OpenCitations COCI dataset (version 19 released on January 2023), and that are entirely contained in OpenCitations META (version 3 released on February 2023). This means that the citations which have either the citing or the cited entity (or both) not contained in META are excluded from coci_preprocessed dataset.</p> </li> <li> <p><strong>meta_preprocessed.zip</strong>: all the original columns of OpenCitations META are maintained in this dataset, so the CSVs have the columns: &ldquo;id&rdquo;, &ldquo;title&rdquo;, &ldquo;author&rdquo;, &ldquo;issue&rdquo;, &ldquo;volume&rdquo;, &ldquo;venue&rdquo;, &ldquo;page&rdquo;, &ldquo;pub_date&rdquo;, &ldquo;type&rdquo;, &ldquo;publisher&rdquo; and &ldquo;editor&rdquo;. The only difference with the original dataset is that meta_preprocessed in the columns &ldquo;id&rdquo; and &ldquo;venue&rdquo; has respectively just the DOIs and the ISSNs, without all the other identifiers specified for each entity in META.</p> </li> <li> <p><strong>erih_preprocessed.zip</strong>: it contains a&nbsp; CSV file with two columns &quot;venue_id&quot; and &quot;ERIH_disciplines&quot;. &quot;venue_id&quot; is the union of the original columns &quot;Online ISSN&quot; and &quot;Print ISSN&quot; of ERIH_PLUS (version downloaded on 2023-04-27).</p> </li> <li> <p><strong>erih_meta.zip</strong>: it contains CSV files obtained from the union of meta_preprocessed and erih_preprocessed, they have all the columns of meta_preprocessed plus a new column &ldquo;erih_disciplines&rdquo; containing all the disciplines linked to a venue (identified by an ISSN).</p> </li> </ul> <p>&nbsp;</p> <p>Software:&nbsp;https://doi.org/10.5281/zenodo.8326023</p> <p>Data produced:&nbsp;https://doi.org/10.5281/zenodo.7974816</p> <p>Article:&nbsp;https://zenodo.org/record/8326044</p> <p>DMP:&nbsp;https://zenodo.org/record/8324973</p> <p>Protocol:&nbsp;https://doi.org/10.17504/protocols.io.n92ldpeenl5b/v5</p>

opencc-by-4.0Sep 2023View details →
zenodo40/100

Don't mention it: challenges to using software mentions to investigate citation and discoverability - Data and Notebooks

<p>This deposit contains the data and Jupyter notebooks used for sampling and annotation analysis of our submission to the PeerJ Computer Science special issue <em>Software Citation, Indexing, and Discoverability </em>(<a href="https://peerj.com/collections/84-software">https://peerj.com/collections/84-software</a>):</p> <blockquote> <p>Stephan Druskat, Neil P. Chue Hong, Sammie Buzzard, Olexandr Konovalov, and Patrick Kornek. Don&rsquo;t mention it: challenges to using software mentions to investigate citation and discoverability.</p> </blockquote> <p>See the README in this deposit.</p> <p>The contents of this deposit can also be browsed at <a href="https://github.com/softwaresaved/habeas-corpus/tree/main/replication-package">https://github.com/softwaresaved/habeas-corpus/tree/main/replication-package</a>.</p>

opencc-by-4.0Sep 2021View details →
zenodo40/100

Citation graph of the Journal of Law and Society (1974-2022)

<p>This is a dump of Neo4J v4.4 graph data containing the citation graph of the Journal of Law and Society (1974-2022). It was generated by merging data from OpenAlex.org with data obtained by citation-mining. The graph model can be described as follows:</p> <pre><code>(:Author)-[:CREATOR_OF]-&gt;(:Work) (:Author)-[:AFFILIATED_WITH]-(:Institution) (:Work)-[:CITES]-&gt;(:Work) (:Work)-[:PUBLISHED_IN]-&gt;(:Venue)</code></pre> <p>Data on institutional affiliation is incomplete.</p>

opencc-by-4.0Sep 2023View details →
zenodo40/100

Factors Associated with Scientific Production Citations in Dentistry: Zero-inflated Negative Binomial Regression and Hurdle Modelling

<p><strong>Abstract:</strong> The global scientific literature in dentistry has shown important advances in the field, with major contributions ranging from the analysis of the basic epidemiological aspects of prevention to specialised results in the field of dental treatments. The present investigation aims to analyse the current state of the scientific literature on dentistry hosted in the Web of Science database. The methodology includes two phases in the analysis of articles and indexed reviews in all thematic areas. During the first phase, the following variables are analysed: scientific production by the publisher, the evolution of scientific output published by publishers, the factors associated with the impact of scientific production, and the modelling of the impact of scientific production on dentistry. During the second phase, associations, evolutions, and trends in the use of main keywords in the scientific literature in dentistry are analysed. In conclusion, the study shows that the most studied topics include the association of dental education and the curriculum, the association of pediatric dentistry with oral health, and dental care. The findings show that more recently emphasised topics also stand out, such as evidence-based dentistry, the pandemic, infection control, and endodontics, as well as the need for future research to expand current knowledge based on emerging topics in the scientific literature on dentistry.</p>

opencc-by-4.0Sep 2023View details →
zenodo36/100

Coronavirus Open Citations Dataset

<p>The Coronavirus Open Citations Dataset curated by OpenCitations currently contains (as of 16 May 2020) information about 189,697 citations and about the 49,719 citing or cited articles involved in these citations. A subset of these data (stored in the &quot;_partial.json&quot; files introduced below) is used for creating the visualization available at <a href="https://opencitations.github.io/coronavirus/">https://opencitations.github.io/coronavirus/</a>.</p> <p>Each item in both the JSON files storing citations (&#39;citations_full.json&#39; and &#39;citations_partial.json&#39;) contains the following fields:</p> <ul> <li>&quot;id&quot;, a numeric identifier of the citation;</li> <li>&quot;source&quot;, the citing entity;</li> <li>&quot;target&quot;, the cited entity.</li> </ul> <p>Each item in both the JSON files storing article metadata (&#39;metadata_full.json&#39; and &#39;metadata_partial.json&#39;) contains the following fields:</p> <ul> <li>&quot;id&quot;, the DOI of the article;</li> <li>&quot;author&quot;, the surname of all the authors of the article;</li> <li>&quot;year&quot;, the year of publication;</li> <li>&quot;title&quot;, the title of the article;</li> <li>&quot;source_title&quot;, the title of the venue where the article has been published.</li> </ul> <p>In addition, any item in the file &#39;metadata_partial.json&#39; contains another field:</p> <ul> <li>&quot;count&quot;: the overall number of citations that the article received.</li> </ul>

opencc-zeroApr 2020View details →
zenodo36/100

Topics in Research on International Relations 2011-2015: Results from Clustering of Citation Links

<p>Supplementary information for the paper</p> <p><strong>IR Theory </strong><strong>and the Core-Periphery Structure of Global IR - Lessons from Citation Analysis</strong></p> <p>by Thomas Risse, Wiebke Wemheuer-Vogelaar, and Frank Havemann (<em>International Studies Review</em>, 2022, Vol. 4, Iss. 3, <a href="https://doi.org/10.1093/isr/viac029">doi.org/10.1093/isr/viac029</a>)</p> <p>&nbsp;</p> <p>Description of Files:</p> <p>File 1: Top-300 Cited Sources and Citing IR Journals<br> file name: Havemann2020top-300-sources.pdf (fourth file in list)<br> content: List of top-300 cited sources and a list of&nbsp; IR journals used for downloads from Web of Science</p> <p>File 2: Cited Sources in Clusters<br> file name: Havemann2020clusters.pdf (third file in list)<br> content: Lists of cited sources in link clusters and graphs of these clusters</p> <p>File 3: Citing Journals of Clusters<br> file name: citing.journals.of.clusters.pdf (second file in list)<br> content: Rank lists of publication channels with regard to numbers of papers citing papers in clusters</p> <p>File 4: Rank Lists of Cited Sources in Seven Journals Outside Web of Science<br> file name: cited-sources-in-7-journals.zip (first file in list)<br> content: Seven ASCII files</p> <p>&nbsp;</p> <p>A preprint of the ISR paper by Risse, Wemheuer-Vogelaar, and Havemann was published with the title &quot;Theory Makes Global IR Hang Together. Lessons from Citation Analysis&quot; and is available at: <a href="https://refubium.fu-berlin.de/handle/fub188/2/browse?type=affiliation&amp;value=Otto-Suhr-Institut+f%C3%BCr+Politikwissenschaft+%2F+Arbeitsstelle+Transnationale+Beziehungen%2C+Au%C3%9Fen-+und+Sicherheitspolitik">Refubium-OSI-IR-group</a></p>

opencc-by-4.0Nov 2020View details →
zenodo36/100

Characterizing Highly Cited Papers in Mass Cytometry through H-Classics: WoS dataset and citation report

<p>Dataset and citation report extracted from Web of Science (WoS) used to characterize highly cited papers in mass cytometry research field from 2010 to 2019.</p>

opencc-by-4.0Jan 2021View details →
zenodo36/100

Citations in the German Wikipedia

<p>Using https://github.com/halfak/Extract-scholarly-article-citations-from-Wikipedia we extracted citations from the February 2015 dump of the German Wikipedia based on DOIs, PubMed IDs, and ISBNs.</p>

opencc-zeroMar 2015View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record