Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

345

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

345 results for “Open Access”

Learn how ShareScore rates datasets ↗
zenodo44/100

DOIs linked by the English Wikipedia which could be made available in green Open Access

<p>List of citations from the English Wikipedia articles extracted from the enwiki-20170720-pages-articles XML dump via https://pypi.org/project/mwcites/ , DOIs cleaned with custom regular expressions.</p> <p>The list of scholarly publications identified by the DOIs has been filtered to exclude those which are already available in Open Access and those which may not be depositable according to SHERPA/RoMEO policy summaries, first via the Dissemin API and then by the oaDOI API, with the attached Python script (https://github.com/nemobis/bots/blob/master/doi-doai-openaccess.py ).</p> <p>This produced a list of 194913 DOIs available in open access and 430230 DOIs unavailable and depositable (as of 2017-08-22 data, which for oaDOI was partly v1 and partly v2).</p>

opencc-by-4.0Sep 2017View details →
zenodo44/100

Supplementary data to `Do science maps from open access literature capture the overall topic structure of an academic field?`

<p>The dataset contains the 8,528 academic articles records related to Sustainable Food research sourced with the query `TS=("sustainab*" NEAR/2 "food*")` .</p> <p>They are the records present in the largest component of the citation network, as specified in the manuscript. &nbsp;</p> <p>The dataset was sourced from OpenAlex based on the original data used in the manuscript and it is composed of the following columns:</p> <table> <tbody> <tr> <td><em><strong>Column</strong></em></td> <td><em><strong>Description</strong></em></td> </tr> <tr> <td>Id</td> <td>OpenAlex ID</td> </tr> <tr> <td>DOI</td> <td>Document Object Identifier</td> </tr> <tr> <td>display_name</td> <td>The article title</td> </tr> <tr> <td>publication_year</td> <td>The publication year of the article</td> </tr> <tr> <td>open_access</td> <td>An object with details of the open access status of the article</td> </tr> </tbody> </table> <p>We choose the `.rdata` format for easy loading in R. Use the function `load()` to add the data frame to the enviroment.&nbsp;</p>

opencc-by-4.0Apr 2024View details →
zenodo44/100

2023 Utrecht University Open Access Monitor (peer reviewed journal articles)

<p>Results of the OA monitor of Utrecht University (UU) and University Medical Center Utrecht(UMCU) for the year 2023. It lists the open access availability of all peer reviewed journal articles registered in the CRIS (Pure) of Utrecht University and/or University Medical Center Utrecht.&nbsp;</p>

opencc-zeroJun 2024View details →
zenodo44/100

Interviews with editors of library science journals on transitioning to open access

<p>These three files are related to qualitative, semi-structured interviews conducted in Fall 2023 with editors of Library and Information Science (LIS) journals on transitioning to open access. One subgroup consisted of participants who were editors at the time of an LIS journal when it transitioned (or flipped) to an open access model that does not charge a fee to either readers or authors (which this study refers to as equitable open access), and the other subgroup consisted of current editors (at the time) of LIS journals that have not yet transitioned (or unflipped) to an equitable open access model. Two of the files are the interview protocols for each group of flipped and unflipped editors, and the third file is the codebook the researchers used to analyze the interview transcripts. Interview transcripts are not being publicly shared to ensure confidentiality for interview participants.</p> <p>The interview protocols were created based on the findings of a prior research study:</p> <p>Borchardt, R., Dawson, D., &amp; Schultz, T. (2024). Financial and other perceived barriers to transitioning to an equitable no-publishing fee open access model: A survey of LIS journal editors. College &amp; Research Libraries, 85(1). <a href="https://doi.org/10.5860/crl.85.1.96">https://doi.org/10.5860/crl.85.1.96</a></p> <p>The codebook was created iteratively based on the researchers' review and analysis of the interview transcripts.</p>

opencc-by-4.0Jul 2024View details →
zenodo44/100

Data and Statistical analysis for: "Predator in the pool? A quantitative evaluation of non-indexed open access journals in aquaculture research"

<p>Data and Statistical analysis for: &quot;Predator in the pool? A quantitative evaluation of non-indexed open access journals in aquaculture research&quot; published in&nbsp;<em>Frontiers in Marine Science</em></p>

openmit-licenseMar 2018View details →
zenodo44/100

Open Access levels of Dutch universities' output 2016-2017 (articles & reviews): green, gold, hybrid and bronze - May 2018

<p>Using Web of Science and Unpaywall data, we here provide an update of Open Access (OA) levels of Dutch universities, for 2016 and 2017.</p> <p>Our previous analysis&nbsp;&nbsp;(<a href="http://doi.org/10.5281/zenodo.1133759">10.5281/zenodo.1133759</a> and <a href="http://doi.org/10.7287/peerj.preprints.3520v1">10.7287/peerj.preprints.3520v1</a>)&nbsp;looked at OA classification as included in Web of Science (gold and green OA, based on Unpaywall data), and supplemented that with a breakdown of gold OA into pure gold, hybrid and bronze, taken from Unpaywall data (formerly OADOI) directly. Here, we improve on this by running all DOIs retrieved from WoS through Unpaywall data (using their web interface that allows batch checking of up to 10,000 DOIs at a time). Unlike WoS, Unpaywall data itself includes author-submitted versions in their green OA classification, resulting in more complete green OA levels.&nbsp;</p> <p>In addition, since our initial analysis of December 2017, Unpaywall data has considerably expanded its coverage of institutional repositories&nbsp; (IRs) (see <a href="https://unpaywall.org/sources">https://unpaywall.org/sources</a>). This now includes coverage of the IRs from all Dutch universities.&nbsp;</p> <p>Taken together, the current data show higher levels of green open access, including author-submitted versions, compared to our previous analysis.&nbsp;</p> <p>In this update, we include output (articles and reviews) from 2016 and 2017 for all 14 universities in the Netherlands.&nbsp;</p> <p>The following categories are distinguished (description taken from&nbsp;Piwowar at al., 2018, doi:&nbsp;<a href="https://doi.org/10.7717/peerj.4375">10.7717/peerj.4375</a>)</p> <ul> <li><strong>Pure gold</strong>: Published in an open-access journal (as defined by the DOAJ)</li> <li><strong>Hybrid</strong>: Free under an open license in a toll-access journal</li> <li><strong>Bronze</strong>: Free to read on the publisher page, but without a license</li> <li><strong>Green:&nbsp;</strong>Available from an institutional or disciplinary repository (including PubMedCentral)</li> </ul> <p>Data for Dutch universities were collected from Web of Science using the organization-enhanced field. Only articles and reviews were included. DOIs were extracted from the Web of Science export, run through the Unpaywall data <a href="https://unpaywall.org/products/simple-query-tool">Simple Query Tool</a>. From the resulting data from Unpaywall, OA classification was done using a simple formula in Excel (to be replaced by an R script in a future update). The Excel template used is included in this dataset, as is the OADOI API output for each Dutch university&#39;s article subset, and the lists of DOIs derived from Web of Science. The dataset also includes summarized data and three charts generated from these data, showing levels of different types of OA for 2016, 2017 and the two years compared.&nbsp;&nbsp;</p> <p>----------------------------------------------------------------------------------------------------------------------------------------------------------------------------</p>

opencc-by-4.0May 2018View details →
zenodo44/100

Method Classification of Open Access INTACT Molecular Interaction data.

<p>Simple&nbsp;classification data derived from open access papers indexed in&nbsp;the INTACT database (https://www.ebi.ac.uk/intact/downloads) based on PSI-MI25 codes for interaction detection methods&nbsp;or participant detection methods based on the subfigure caption text.&nbsp;<br> <br> intact_records_and_captions_complete.tsv - This file links available text of subfigure captions to PSI-MI25 codes for the interaction detection method and participant detection method.&nbsp;&nbsp;</p> <p>evidx_run_file.txt - This file provides execution codes for the &#39;EvidX&#39; machine learning text&nbsp;classifier (https://github.com/SciKnowEngine/evidX/releases/tag/v0.1.0)</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Jul 2018View details →
zenodo44/100

Toolkit on Open Access for Research Project Coordinators

<p>The materials in this toolkit were created by Romain F&eacute;ret as a resource for training on how to help project coordinators to comply with their open access requirements. The slides of the training are available on Zenodo at&nbsp;10.5281/zenodo.3381783. This training day took place on Wednesday the 5th of June 2019, at the University of Lille. It was organized with the support of Couperin as a part of its activities in the project OpenAIRE-Advanced.</p> <p>The tutorials are divided into two folders. The &lsquo;Coordinator&rsquo; folder contains documents that can be sent directly to the researchers, while the &lsquo;Support staff&rsquo; folder contains tutorials for support staff (librarians, project managers) who help the coordinators to manage their project. Each tutorial is in .pdf and .docx format for easy reuse and modification. Each document is available in French and in English.</p>

opencc-by-4.0Jun 2019View details →
zenodo44/100

La piattaforma di riviste Open Access dell'Università degli Studi di Milano. [Video]

<p>L&rsquo;intervento percorre i passi fondamentali della creazione della piattaforma di riviste open access dell&rsquo;Universit&agrave; degli Studi di Milano, come modello in cui una istituzione che produce conoscenza decide di assumersi la responsabilit&agrave; di validare e diffondere questa conoscenza ad un pubblico che sia il pi&ugrave; ampio possibile, liberandosi da scelte e vincoli imposti dagli editori commerciali e dalle logiche editoriali e riportando nelle mani dei ricercatori le attivit&agrave; che da tempo erano state consegnate agli editori.</p>

opencc-by-sa-4.0Feb 2017View details →
zenodo44/100

Monitoring and evaluation of UKRI's Open Access Policy: Exploring the use of open data sources to inform baseline values - Dataset

<p>This dataset accompanies the report <em>"Monitoring and evaluation of UKRI's Open Access Policy: Exploring the use of open data sources to inform baseline values"</em>, which is available via Zenodo.<br><br>It provides record-level data of UKRI-funded and UK-affiliated research output (limited to journal articles with Crossref DOIs) published between 2012 and 2022 - including bibliographic metadata as well as data on open access availability, publisher, national and international collaborations, citations, views and downloads, altmetrics and subjects (fields).&nbsp;All variables are documented in the data dictionary included in this Zenodo record.</p> <p>The code used to generate the dataset from open data sources is available on GitHub.&nbsp;</p> <p>The following data sources were used:</p> <ul> <li> <p>Gateway to Research (records downloaded between 2023-11-05 and 2023-11-13)</p> </li> <li> <p>Crossref (Metadata Plus snaphot 2023-10-31, Crossref member route API 2024-01-23)</p> </li> <li> <p>OpenAlex (data snapshot 2023-10-18)</p> </li> <li> <p>Unpaywall (data snapshot 2023-11-27)</p> </li> <li> <p>IRUS UK (2024-04-03)</p> </li> <li> <p>Crossref Event Data (2023-04-01)</p> </li> </ul> <p><strong></strong><br><br>The project made use of Curtin Open Knowledge Initiative (COKI) infrastructure, which is documented on GitHub: <a href="https://github.com/The-Academic-Observatory">https://github.com/The-Academic-Observatory</a>.&nbsp;</p>

opencc-zeroSep 2024View details →
zenodo44/100

Open-Access Data for "Received SignalStrength Measurements with BLE Signals for Contact Tracing and Proximity Detection"

<p>This archive contains three folders which are supplementary material for the paper accepted for publishing in IEEE Sensors Journal.</p> <p><strong>Contents:</strong></p> <ul> <li>&nbsp;The folder `open-access-data/upb/` contains the measurements acquired at UPB. The subfolders are named as `upb_ble_*`, where an asterisk masks&nbsp;the directory number. Whenever UPB is specified, use the data sets from the corresponding directory.</li> <li>The folder `open-access-data/tau/` contains the measurements acquired at TAU. The subfolders are named as `tau_ble_*`, where an asterisk masks the directory number. Whenever TAU is specified, use the data sets from the corresponding directory.</li> <li>The folder `open-access-data/wifi-on-off/` contains a sample code to read the files and plot the data from Fig. 14 in `open-access-data/wifi-on-off/wifi_on_off_read_plot.py` and Fig. 15 in `open-access-data/wifi-on-off/wifi_on_off_read_plot.ipynb`.</li> </ul> <p><strong>Results based on the data have been presented in the paper:</strong><br> Flueratoru, L., Shubina, V., Niculescu, D., Lohan, E.S. (2021). On the High Fluctuations of Received Signal Strength Measurements with BLE Signals for Contact Tracing and Proximity Detection, IEEE Sensors, Special Issue on Advanced Sensors and Sensing Technologies for Indoor Positioning and Navigation</p> <p><strong>To cite these data sets please use the following:</strong><br> Laura Flueratoru, Viktoriia Shubina, Dragoș Niculescu, &amp; Elena Simona Lohan. (2021). Open Access Data for &quot;Received SignalStrength Measurements with BLE Signals for Contact Tracing and Proximity Detection&quot; [Data set]. Zenodo. http://doi.org/10.5281/zenodo.4643668</p>

opencc-by-4.0Jul 2021View details →
zenodo44/100

The Unofficial Guide on applying NCN Open Access rules to GitHub repositories.

<p><b>The Unofficial Guide on applying NCN Open Access rules to GitHub repositories.</b> <i>Some</i> HTML code can be used here.</p>

opencc-zeroDec 2022View details →
zenodo44/100

Data from the OPERAS business models survey on open access books

<p>OPERAS (the European Research Infrastructure for the development of open scholarly communication in the social sciences and humanities) has conducted a survey of publishing organisations throughout Europe to identify and better understand existing and potential business models to support the Open Access publication of research monographs. The results of the survey are used to inform the formulation of recommendations about how to create a sustainable open access book publishing ecosystem within Europe.</p> <p>The survey was designed to serve two core aims:&nbsp;<br> 1. To further, better or improve our understanding of the scholarly publishing landscape and of the challenges that publishers face in the context of publishing OA monographs;<br> 2. To identify main trends (including opportunities and challenges) and the knowledge of collaborative funding and infrastructure models in OA publishing in SSH.&nbsp;</p> <p>The survey was open&nbsp;between 16 February and 14 April 2021.</p> <p>The results are presented in two versions of the white paper&nbsp;of the Open Access Business Models Special Interest Group:&nbsp;&nbsp;Stone, Graham, Błaszczyńska, Marta, Lebon, Chlo&eacute;, Morka, Agata, Mosterd, Tom, Mounier, Pierre, Proudman, Vanessa, Speicher, Lara, &amp; Melin&scaron;čak Zlodi, Iva. (2021). Collaborative models for OA book publishers (1.0). Zenodo. https://doi.org/10.5281/zenodo.5494731 and the second version to be published in Spring 2023.</p>

opencc-by-4.0Feb 2023View details →
zenodo44/100

AgrImOnIA: Open Access dataset correlating livestock and air quality in the Lombardy region, Italy

<p>The AgrImOnIA dataset is a comprehensive dataset relating air quality and livestock (expressed as the&nbsp;density of bovines and swine bred) along with weather and other variables. The AgrImOnIA Dataset represents the first step of the <a href="http://www.agrimonia.net">AgrImOnIA project</a>. The purpose of this dataset is to give the opportunity to assess the impact of agriculture on air quality in Lombardy through statistical techniques capable of highlighting the relationship between the livestock sector and air pollutants concentrations.</p> <p>The building process of the dataset is detailed in the <strong>companion paper:</strong></p> <p>A. Fass&ograve;, J. Rodeschini, A. Fusta Moro, Q. Shaboviq, P. Maranzano, M. Cameletti, F. Finazzi, N. Golini, R. Ignaccolo, and P. Otto&nbsp;(2023). Agrimonia: a dataset on livestock, meteorology and air quality in the Lombardy region, Italy.&nbsp;<em>SCIENTIFIC DATA</em>, 1-19.</p> <p>available <a href="https://rdcu.be/c7T9H">here</a>.</p> <p>This dataset is a collection of estimated daily values for a range of measurements of different dimensions as: air quality, meteorology, emissions, livestock animals and land use. Data are related to Lombardy and the surrounding area for&nbsp;2016-2021, inclusive. The surrounding area is obtained by applying a 0.3&deg; buffer on Lombardy borders.</p> <p>The data uses several aggregation and interpolation methods to estimate the measurement for all days.</p> <p>The files in the record, renamed according to their version (es. .._v_3_0_0),&nbsp;are:</p> <ul> <li> <p>Agrimonia_Dataset.csv(.mat and .Rdata) which is built by joining the daily time series related to the AQ, WE, EM, LI and LA variables. In order to simplify access to variables in the Agrimonia dataset, the variable name starts with the dimension of the variable, i.e., the name of the variables related to the AQ dimension start with &#39;AQ_&#39;. This file is archived also in the&nbsp;format for MATLAB and R software.&nbsp;</p> </li> <li> <p>Metadata_Agrimonia.csv which provides further information about the Agrimonia variables: e.g. sources used, original names of the variables imported, transformations applied.</p> </li> <li> <p>Metadata_AQ_imputation_uncertainty.csv which contains the daily uncertainty estimate of the imputed observation for the AQ to mitigate missing data in the hourly time series.&nbsp;&nbsp;</p> </li> <li> <p>Metadata_LA_CORINE_labels.csv which contains the label and the description associated with the CLC class.&nbsp;&nbsp;</p> </li> <li> <p>Metadata_monitoring_network_registry.csv which contains all details about the AQ monitoring station used to build the dataset. Information about air quality monitoring stations include: station type, municipality code, environment type, altitude, pollutants sampled and other. Each row represents a single sensor.</p> </li> <li> <p>Metadata_LA_SIARL_labels.csv which contains the label and the description associated with the SIARL class.</p> </li> <li> <p>AGC_Dataset.csv(.mat and .Rdata)&nbsp;that&nbsp;includes daily data of almost all variables available in&nbsp;the Agrimonia&nbsp;Dataset (excluding AQ variables)&nbsp;on an&nbsp;equidistant grid covering the Lombardy region and its surrounding area.&nbsp;</p> </li> </ul> <p>The Agrimonia dataset can be reproduced using&nbsp;the code available at the GitHub page: <a href="https://github.com/AgrImOnIA-project/AgrImOnIA_Data">https://github.com/AgrImOnIA-project/AgrImOnIA_Data</a></p> <p><strong>UPDATE 31/05/2023</strong>&nbsp;<strong>- NEW RELEASE - V 3.0.0</strong></p> <p>A new version of the dataset is released:&nbsp;Agrimonia_Dataset_v_3_0_0.csv (.Rdata and .mat), where variable&nbsp;<em>WE_rh_min, WE_rh_mean and WE_rh_max&nbsp;</em>have been recomputed due to some bugs<em>.</em></p> <p>In addition, two new columns are added, they are&nbsp;<em>LI_pigs_v2 and LI_bovine_v2&nbsp;</em>and represents the density of the pigs and bovine (expressed as animals per kilometer squared) of a square of size ~ 10 x 10 km centered at the station localisation.</p> <p>A new dataset is released: the Agrimonia Grid Covariates (AGC) that includes daily information for the period from 2016 to 2020 of almost all variables within the Agrimonia Dataset on a equidistant grid containing the Lombardy region and its surrounding area. The AGC does not include AQ variables as they come from&nbsp;the monitoring stations that are irregularly spread over the area considered.</p> <p><strong>UPDATE 11/03/2023</strong>&nbsp;<strong>- NEW RELEASE - V 2.0.2</strong></p> <p>A new version of the dataset is released:&nbsp;Agrimonia_Dataset_v_2_0_2.csv (.Rdata), where variable&nbsp;<em>WE_tot_precipitation&nbsp;</em>have been recomputed due to some bugs<em>.</em></p> <p>A new version of the metadata is available:&nbsp;Metadata_Agrimonia_v_2_0_2.csv where the spatial resolution of the variable <em>WE_precipitation_t&nbsp;</em>is corrected.</p> <ul> </ul> <p><strong>UPDATE 24/01/2023</strong>&nbsp;<strong>- NEW RELEASE - V 2.0.1</strong></p> <p>minor bug fixed</p> <p><strong>UPDATE 16/01/2023</strong>&nbsp;<strong>- NEW RELEASE - V 2.0.0</strong></p> <p>A new version of the dataset is released, Agrimonia_Dataset_v_2_0_0.csv (.Rdata) and Metadata_monitoring_network_registry_v_2_0_0.csv.&nbsp;Some minor points have been addressed:</p> <ul> <li>Added&nbsp;values for <em>LA_land_use</em> variable for Switzerland stations (in Agrimonia Dataset_v_2_0_0.csv)</li> <li>Deleted&nbsp;incorrect values for <em>LA_soil_use</em> variable for stations outside Lombardy region during 2018 (in Agrimonia Dataset_v_2_0_0.csv)</li> <li>Fixed duplicate&nbsp;sensors corresponding&nbsp;to the same pollutant within the same&nbsp;station<em>&nbsp;</em>(in&nbsp;Metadata_monitoring_network_registry_v_2_0_0.csv)</li> </ul>

opencc-by-4.0Sep 2022View details →
zenodo44/100

zbMATH Open Access Subset

<p>The dataset contains two tables as csv files.</p> <p>1) documents_in_oa_series</p> <p>is a list of zbMath Open documents in serials where the description contains the word &quot;Open Access&quot;. Note that some documents did not appear yet or might have been retracted. Thus when fetching information from the oai-pmh API or the website, be prepared to handle non-existing documents.</p> <p>2) oa_links</p> <p>Lists all links from zbMATH Open documents to fulltext matched by either unpaywall or arxiv.</p> <p>The meaning of the fields is:</p> <ul> <li><strong>zbmath_id</strong> Unique identifier from zbMATH Open. Prefix with <code>https://zbmath.org/</code> to visit additional information on the article. For example, <code>5635019</code> is associated with <a href="https://zbmath.org/5635019">https://zbmath.org/5635019</a></li> <li><strong>link</strong> link to the fulltext. Note that those links are not fully reliable. We estimate a success rate of 90%</li> </ul> <p>Note that there is an overlap between 1 and 2. So some documents are published in OA serials and have arxiv or unpaywall links at the same time.</p> <p>&nbsp;</p>

opencc-by-4.0Jun 2023View details →
zenodo44/100

Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - RESULTS DATASET (with Mega Journals)

<p>The dataset contains all the data produced running the research software for the study:&quot;Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta&quot;.</p> <p>Disclaimer: these results are not considered to be representative, because we have fount that Mega Journals skewed significantly some of the data. The result datasets without Mega Journals are published <a href="https://zenodo.org/record/8249907">here</a>.</p> <p>Description of datasets:</p> <ul> <li><strong>SSH_Publications_in_OC_Meta_and_Open_Access_status.csv:&nbsp;</strong>containing information about OpenCitations Meta coverage of ERIH PLUS Journals as well as their Open Access availability. In this dataset, every row holds data for a Journal of ERIH PLUS also covered by OpenCitations Meta database. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> <li><strong>SSH_Publications_by_Discipline.csv:</strong>&nbsp;containing information about number of publications per&nbsp;discipline&nbsp;(in addition, number of journals&nbsp;per discipline are also included). The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>SSH_Publications_and_Journals_by_Country:</strong>&nbsp;containing information about number of publications and journals per&nbsp;country.&nbsp;The dataset has three columns, the first, labeled <strong>&quot;Country&quot;,</strong>&nbsp;contains single countries of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>result_disciplines.json:</strong> the dictionary containing all disciplines as key and a list of&nbsp;related ERIH PLUS venue identifiers as value.</li> <li><strong>result_countries.json:</strong>&nbsp;the dictionary containing all countries as key and a list of related ERIH PLUS venue identifiers as value.</li> <li><strong>duplicate_omids.csv: </strong>a dataset containing the duplicated Journal entries in OpenCitations Meta, structured with two columns: &quot;<strong>OC_omid&quot;</strong>, the internal OC Meta identifier; &quot;<strong>issn&quot;,&nbsp;</strong>the issn values associated to that identifier</li> <li><strong>eu_data.csv: </strong>contains the data specific for&nbsp;European countries&#39; SSH Journals&nbsp;covered in OCMeta. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;Original_Title&quot;</strong>,<strong> &quot;Country_of_Publication&quot;</strong>,<strong>&quot;ERIH_PLUS_Disciplines&quot;</strong>, <strong>&quot;disc_count&quot;</strong>, the number of disciplines per Journal.</li> <li><strong>eu_disciplines_count.csv:&nbsp;</strong>containing information about number of publications per&nbsp;discipline and number of journals&nbsp;per discipline of european countries. The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>meta_coverage_eu.csv:&nbsp;</strong>contains the data specific for&nbsp;European countries&#39; SSH Journals&nbsp;covered in OCMeta.&nbsp;It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> <li><strong>us_data.csv:&nbsp;</strong>contains the data specific for the&nbsp;United States&#39; SSH Journals&nbsp;covered in OCMeta. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;Original_Title&quot;</strong>,<strong> &quot;Country_of_Publication&quot;</strong>,<strong>&quot;ERIH_PLUS_Disciplines&quot;</strong>, <strong>&quot;disc_count&quot;</strong>, the number of disciplines per Journal.</li> <li><strong>us_disciplines_count.csv:&nbsp;</strong>containing information about number of publications per&nbsp;discipline and number of journals&nbsp;per discipline of the United States. The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>meta_coverage_us.csv:&nbsp;</strong>contains the data specific for the United States&#39; SSH Journals&nbsp;covered in OCMeta.&nbsp;It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> </ul> <p>&nbsp;</p> <p><strong>Abstract of the research:&nbsp;</strong></p> <p><strong>Purpose:</strong>&nbsp;this study aims to investigate the representation and distribution of Social Science and Humanities (SSH) journals within the OpenCitations Meta database, with a particular emphasis on their Open Access (OA) status, as well as their spread across different disciplines and countries. The underlying premise is that open infrastructures play a pivotal role in promoting transparency, reproducibility, and trust in scientific research.<br> <strong>Study Design and Methodology:</strong>&nbsp;the study is grounded on the premise that open infrastructures are crucial for ensuring transparency, reproducibility, and fostering trust in scientific research. The research methodology involved the use of secondary data sources, namely the OpenCitations Meta database, the ERIH PLUS bibliographic index, and the DOAJ index. A custom research software was developed in Python to facilitate the processing and analysis of the data.<br> <strong>Findings:</strong>&nbsp;the results reveal that 78.1% of SSH journals listed in the European Reference Index for the Humanities (ERIH-PLUS) are included in the OpenCitations Meta database. The discipline of Psychology has the highest number of publications. The United States and the United Kingdom are the leading contributors in terms of the number of publications. However, the study also uncovers that only 38% of the SSH journals in the OpenCitations Meta database are OA.<br> <strong>Originality:</strong>&nbsp;this research adds to the existing body of knowledge by providing insights into the representation of SSH in open bibliographic databases and the role of open access in this domain. The study highlights the necessity for advocating OA practices within SSH and the significance of open data for bibliometric studies. It further encourages additional research into the impact of OA on various facets of citation patterns and the factors leading to disparity across disciplinary representation.</p> <p><strong>Related resources:</strong></p> <p>Ghasempouri S., Ghiotto M., &amp; Giacomini S. (2023). Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - RESEARCH ARTICLE.&nbsp;<a href="https://doi.org/10.5281/zenodo.8263908">https://doi.org/10.5281/zenodo.8263908</a></p> <p>Ghasempouri, S.,&nbsp;Ghiotto, M., Giacomini, S., (2023).&nbsp; Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - DATA MANAGEMENT PLAN (Version 4). Zenodo.&nbsp;<a href="https://doi.org/10.5281/zenodo.8174644">https://doi.org/10.5281/zenodo.8174644</a></p> <p>Ghasempouri, S., Ghiotto, M., Giacomini, S. (2023e). Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - PROTOCOL. V.5. (<a href="https://dx.doi.org/10.17504/protocols.io.5jyl8jo1rg2w/v5">https://dx.doi.org/10.17504/protocols.io.5jyl8jo1rg2w/v5</a>)</p>

opencc-byMay 2023View details →
zenodo44/100

Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - RESULTS DATASET (without Mega Journals)

<p>The dataset contains all the data produced running the research software for the study <em>Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta</em>, a research carried out in the contest of the Open Science course 22/23 at the University of Bologna.</p> <p>Mega Journals have been excluded form the datasets, since we found they were significantly skewing the results, the only datasets not interested by this exclusion are&nbsp;<strong>SSH_Publications_in_OC_Meta_and_Open_Access_status </strong>and<strong>&nbsp;duplicate_omids.</strong>&nbsp;The result datasets with Mega Journals included are published <a href="https://doi.org/10.5281/zenodo.8250858">here</a><br> The Journals excluded from the results are: PLOS ONE (issn:1932-6203), PNAS (issn:1091-6490), Science (issn:1095-9203), Nature(issn:0028-0836).</p> <p>Description of datasets:</p> <ul> <li><strong>SSH_Publications_in_OC_Meta_and_Open_Access_status.csv:&nbsp;</strong>containing information about OpenCitations Meta coverage of ERIH PLUS Journals as well as their Open Access availability. In this dataset, every row holds data for a Journal of ERIH PLUS also covered by OpenCitations Meta database. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> <li><strong>SSH_Publications_by_Discipline.csv:</strong>&nbsp;containing information about number of publications per&nbsp;discipline&nbsp;(in addition, number of journals&nbsp;per discipline are also included). The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>SSH_Publications_and_Journals_by_Country:</strong>&nbsp;containing information about number of publications and journals per&nbsp;country.&nbsp;The dataset has three columns, the first, labeled <strong>&quot;Country&quot;,</strong>&nbsp;contains single countries of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>result_disciplines.json:</strong> the dictionary containing all disciplines as key and a list of&nbsp;related ERIH PLUS venue identifiers as value.</li> <li><strong>result_countries.json:</strong>&nbsp;the dictionary containing all countries as key and a list of related ERIH PLUS venue identifiers as value.</li> <li><strong>duplicate_omids.csv: </strong>a dataset containing the duplicated Journal entries in OpenCitations Meta, structured with two columns: &quot;<strong>OC_omid&quot;</strong>, the internal OC Meta identifier; &quot;<strong>issn&quot;,&nbsp;</strong>the issn values associated to that identifier</li> <li><strong>eu_data.csv: </strong>contains the data specific for&nbsp;European countries&#39; SSH Journals&nbsp;covered in OCMeta. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;Original_Title&quot;</strong>,<strong> &quot;Country_of_Publication&quot;</strong>,<strong>&quot;ERIH_PLUS_Disciplines&quot;</strong>, <strong>&quot;disc_count&quot;</strong>, the number of disciplines per Journal.</li> <li><strong>eu_disciplines_count.csv:&nbsp;</strong>containing information about number of publications per&nbsp;discipline and number of journals&nbsp;per discipline of european countries. The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>meta_coverage_eu.csv:&nbsp;</strong>contains the data specific for&nbsp;European countries&#39; SSH Journals&nbsp;covered in OCMeta.&nbsp;It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> <li><strong>us_data.csv:&nbsp;</strong>contains the data specific for the&nbsp;United States&#39; SSH Journals&nbsp;covered in OCMeta. It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;Original_Title&quot;</strong>,<strong> &quot;Country_of_Publication&quot;</strong>,<strong>&quot;ERIH_PLUS_Disciplines&quot;</strong>, <strong>&quot;disc_count&quot;</strong>, the number of disciplines per Journal.</li> <li><strong>us_disciplines_count.csv:&nbsp;</strong>containing information about number of publications per&nbsp;discipline and number of journals&nbsp;per discipline of the United States. The dataset has three columns, the first, labeled <strong>&quot;Discipline&quot;,</strong>&nbsp;contains single disciplines of the ERIH classificaton, the second and the third, labeled <strong>&quot;Journal_count&quot;&nbsp;</strong>and <strong>&quot;Publication_count&quot;,&nbsp;</strong>respectively, the number of Journals and the number of Publications counted for each discipline.</li> <li><strong>meta_coverage_us.csv:&nbsp;</strong>contains the data specific for the United States&#39; SSH Journals&nbsp;covered in OCMeta.&nbsp;It is structured with the following columns:&nbsp; &quot;<strong>EP_id&quot;, </strong>the internal ERIH PLUS identifier; <strong>&quot;Publications_in_venue&quot;, </strong>the<strong>&nbsp;</strong>numbers of Publications counted in each venue; <strong>&quot;</strong><strong>OC_omid&quot;, </strong>the internal OpenCitations Meta identifier for the venue;<strong>&nbsp;&quot;issn&quot;,</strong> numbers of publications in each venue;<strong>&nbsp;&quot;Open Access&quot;,</strong> a value to represent if the journal is OA or not, either &quot;True&quot; or &quot;Unknown&quot;.</li> </ul> <p>&nbsp;</p> <p><strong>Abstract of the research:&nbsp;</strong></p> <p><strong>Purpose:</strong>&nbsp;this study aims to investigate the representation and distribution of Social Science and Humanities (SSH) journals within the OpenCitations Meta database, with a particular emphasis on their Open Access (OA) status, as well as their spread across different disciplines and countries. The underlying premise is that open infrastructures play a pivotal role in promoting transparency, reproducibility, and trust in scientific research.<br> <strong>Study Design and Methodology:</strong>&nbsp;the study is grounded on the premise that open infrastructures are crucial for ensuring transparency, reproducibility, and fostering trust in scientific research. The research methodology involved the use of secondary data sources, namely the OpenCitations Meta database, the ERIH PLUS bibliographic index, and the DOAJ index. A custom research software was developed in Python to facilitate the processing and analysis of the data.<br> <strong>Findings:</strong>&nbsp;the results reveal that 78.1% of SSH journals listed in the European Reference Index for the Humanities (ERIH-PLUS) are included in the OpenCitations Meta database. The discipline of Psychology has the highest number of publications. The United States and the United Kingdom are the leading contributors in terms of the number of publications. However, the study also uncovers that only 38% of the SSH journals in the OpenCitations Meta database are OA.<br> <strong>Originality:</strong>&nbsp;this research adds to the existing body of knowledge by providing insights into the representation of SSH in open bibliographic databases and the role of open access in this domain. The study highlights the necessity for advocating OA practices within SSH and the significance of open data for bibliometric studies. It further encourages additional research into the impact of OA on various facets of citation patterns and the factors leading to disparity across disciplinary representation.</p> <p><strong>Related resources:</strong></p> <p>Ghasempouri S., Ghiotto M., &amp; Giacomini S. (2023). Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - RESEARCH ARTICLE.&nbsp;<a href="https://doi.org/10.5281/zenodo.8263908">https://doi.org/10.5281/zenodo.8263908</a></p> <p>Ghasempouri, S.,&nbsp;Ghiotto, M., Giacomini, S., (2023).&nbsp; Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - DATA MANAGEMENT PLAN (Version 4). Zenodo.&nbsp;<a href="https://doi.org/10.5281/zenodo.8174644">https://doi.org/10.5281/zenodo.8174644</a></p> <p>Ghasempouri, S., Ghiotto, M., Giacomini, S. (2023e). Open Science for Social Sciences and Humanities: Open Access availability and distribution across disciplines and Countries in OpenCitations Meta - PROTOCOL. V.5. (<a href="https://dx.doi.org/10.17504/protocols.io.5jyl8jo1rg2w/v5">https://dx.doi.org/10.17504/protocols.io.5jyl8jo1rg2w/v5</a>)</p>

opencc-byMay 2023View details →
zenodo44/100

A Shortlist of Diamond Open Access Journals for the Faculty of Science at Utrecht University

<p><strong>Context</strong></p> <p>The following shortlist of diamond open-access journals was compiled to increase awareness of alternative scholarly publication models among the six departments of the <a href="https://www.uu.nl/en/organisation/faculty-of-science">Faculty of Science at Utrecht University</a>. The list is relevant to the six disciplines at the Faculty of Science: Biology, Chemistry, Mathematics, Information and Computing Sciences, Physics, and Pharmaceutical Sciences. For this purpose, a &quot;diamond journal&quot; is defined as a journal indexed in the <a href="https://www.doaj.org/">Directory of Open Access Journals (DOAJ)</a> that does not charge an article processing charge (APC).</p> <p>&nbsp;</p> <p><strong>Contents and Results</strong></p> <p>The Excel file titled &ldquo;Diamond_journals_faculty_of_science_UU&rdquo; contains the list of selected diamond journals based on the following criteria: they allow submissions in English, have a plagiarism screening policy, possess an electronic ISSN number, and accept submissions in Biology, Chemistry, Mathematics, Information and Computing Sciences, Physics, and Pharmaceutical Sciences. In this shortlist, 355 journals meet the criteria. Out of these 355 journals, only 29 have received a DOAJ seal, 150 journals are indexed in <a href="https://www.scopus.com/">Scopus</a>, and 94 journals are indexed in <a href="https://mjl.clarivate.com/home">Web of Science</a>.</p> <p>A detailed description of the methods employed to obtain this shortlist can be found in the Word file titled &quot;Methods_and_Results&quot;.</p> <p>The raw CSV data has been included under the name &quot;Raw_DOAJ_journal_metadata_2023_07_25&quot;.</p> <p>&nbsp;</p> <p><strong>Limitations</strong></p> <p>The compilers of this shortlist are aware that some current diamond journals could change their status to non-diamond by charging article processing fees at a later stage. Since the journal record is not always updated by the publishers, we strongly recommend the users double-check the latest open access status directly on the journal&#39;s homepage (journal URLs are provided in the Excel file). The same applies for Scopus and WOS indexations.</p>

opencc-by-4.0Aug 2023View details →
zenodo44/100

NWO and ZonMw Open Access Monitor 2022 - dataset

<p>This is the dataset underlying the report "NWO and ZonMw Open Access Monitor 2022".&nbsp;</p><p>Openly accessible metadata was used to research if and how publications from research funded by NWO and ZonMw were open access in 2022. 93% of the articles that were detected have been made available open access using one of the available routes (NWO: 93,1%; ZonMw: 92,5%). This is a slight increase compared to 2021 (90%) and 2020 (85%). At least 64% of the included NWO and ZonMw publications of 2022 (3.688 out of 5.763) has been published via Gold, Hybrid, under a transformative agreement, or Green OA, with a CC-BY license and is therefore fully Plan S compliant.&nbsp;</p>

opencc-zeroOct 2023View details →
zenodo40/100

Data for Research Assessment in the Transition to Open Science. 2019 EUA Open Science and Access Survey Results

<p>This database refers to the data collected by the European University Association (EUA) for its Open Science and Access Survey 2019, which gathered responses from universities and higher education institutions across Europe. The full report published by the association is available at <a href="https://eua.eu/resources/publications/888:research-assessment-in-the-transition-to-open-science.html">https://eua.eu/resources/publications/888:research-assessment-in-the-transition-to-open-science.html</a>.</p> <p>The data included in this database refers only to those universities and higher education institutions that accepted their data to be available in open access (n=174). All information that could lead to the identification of individual universities and higher education institutions was removed from the database (cf. cells highlighted in red). The following files are available:</p> <ul> <li>2019 EUA Open Science and Access Survey</li> <li>Database in the following formats: .xlsx (Microsoft Excel)</li> <li>Survey Codebook: includes information on all the variables and their coding.</li> </ul>

opencc-by-4.0Jan 2020View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record