Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
16
datasets available to search
ShareScore release 0.9.0
Dataset results
16 results for “Scholarly journals”
Scholarly journals publishing articles by family and community physicians in Brazil, up to December 2018
<p>This is the dataset of manuscript titled "In which journals do family and community physicians in Brazil publish? The <em>Trajetórias MFC</em> project". There are two spreadsheets: the dataset proper and the data dictionary. See the manuscript for background.</p> <p>All spreadsheets are in the CSV (comma-separated values) format, delimited with semicolons and encoded in UTF-8 with the byte-order mark (BOM). The spreadsheets can be opened with desktop or Web application software (LibreOffice Calc, Microsoft Excel, Google Sheets) or with statistical software such as R.</p> <p>A <a href="https://zenodo.org/record/3905255">previous version</a> of this dataset was used in a <a href="https://doi.org/10.1101/19005744">preprint</a>. This version should be cited by an upcoming article.</p> <p>See also the <a href="https://doi.org/10.1136/fmch-2020-000321">article</a>, <a href="https://doi.org/10.5281/zenodo.3376310">dataset</a> and <a href="https://doi.org/10.5281/zenodo.3381576">supplementary table</a> for an earlier milestone, about the postgraduate education of family and community physicians in Brazil.</p>
An open, collaborative, and scholarly digital edition of Anṭūn al-Jumayyil's monthly journal "al-Zuhūr" (Cairo, 1910--1913)
<p>This release has been necessitated by the need for documenting changes and improvements that took place over the last two years:</p> <ol> <li>New sets of facsimiles have been added using IIIF</li> <li>The TEI Boilerplate has been updated to the latest version</li> <li>The mark-up of entities and their links to our authority files have been much improved</li> </ol>
Founding, Running, and Improving Scholarly Journals_Questionnaire Raw Data
<p>This is the raw data set of a questionnaire on founding, running, and improving scholarly journals and the availability of educational material for these purposes. The questionnaire was prepared on the SoSci Survey platform and the raw data downloaded from it. The questionnaire had 93 respondents who were nearly all Editors or Scholarly Publishing Professionals working with academic journals. The questionnaire was prepared for the purposes of presenting a poster at the Society for Scholarly Publishing's (SSP) Annual Conference in Portland, Oregon, USA, in May 2023.</p> <p>The Findings of the project are available here: https://zenodo.org/record/7924220#.ZFy_SaXP3cs</p> <p>A Resource Compendium is available here: https://zenodo.org/record/7924268#.ZFy_DaXP3cs</p> <p> </p>
Factors affecting altmetrics attention to scholarly publication in peer-reviewed journals published in Iran and Turkey
<p>The goal of this study was to trace the altmetric measures of peer-reviewed journals in two non-English speaking countries na,ely Iran and Turkey, in order to understand their correlation with some website structure and design determinants, as well as the subject and the full-text language of the journals.</p> <p> </p>
Codes and Datasets for: 'The Largest Academic Publishers of Scholarly Journals: A Webscraping Approach'
<p>This set comprises:</p> <ul> <li>"<strong>count_publishers*</strong>": four R codes ("count_publishers*") that draw from <em>DOAJ</em>, <em>Publons</em>, <em>Scopus </em>and <em>SherpaRomeo</em> to extract scholarly publishers and the journal counts assigned to each publisher;</li> <li>"<strong>data*</strong>": two underlying data samples (from <em>DOAJ </em>and <em>Scopus</em>) - the two other samples are accessed via webscraping;</li> <li>"<strong>harmonize*</strong>": one text-file and one R-code for harmonizing publisher names;</li> <li>"<strong>alljournals.xlsx</strong>": the resulting list of scholarly publishers ordered by the highest number of journal counts assigned to them.</li> </ul>
SCOPING REVIEW OF EMPIRICAL LITERATURE ON EVALUATION OF EDITORIAL POLICIES SUPPORTING OPEN SCIENCE PRACTICES IN SCHOLARLY JOURNALS
<p>Excel file containing the description of methods & materials used in the scoping review, the whole dataset of included studies, and respective descriptive metadata.</p>
Latin American and Caribbean journals indexed in Google Scholar Metrics
<p>Dataset from a study aiming to analyze the coverage of Latin American and Caribbean journals in Google Scholar Metrics (GSM). Data from 8,205 journals from 24 countries of the region were downloaded from Latindex database. A Python script was used for automated title search and data extraction (titles, h5-index, h5-median, URLs) in GSM. For the journals not found, a manual search was carried out, with attempts by variations of the title. It was found 3,070 journals indexed in GSM, which corresponds to 37.42% of the Latindex list. The search was performed on the 2021 edition of GSM, which considers articles published between 2016 and 2020 and citations registered until July 2021. The number of all types of documents published (productivity) in the h5-index period (2016-2020) in Scopus, Journal Citation Reports, and SciELO of 1,314 journals was also identified. </p> <p>The present dataset is the result of this study, which is under peer-review in a scientific journal. </p> <p>The dataset comprises titles, h5-index; h5-median, URLs of 3,070 publications from Latin America and the Caribbean identified in Google Scholar Metrics, and the respective editorial information of the publications was extracted from Latindex</p> <p>The original language of the content was kept, mainly Spanish in the case of editorial data from Latindex. The columns descriptors are also shown in English.</p> <p>The productivity data refer to the number of all types of documents published by the journals in the period 2016-2020. Data were extracted from the InCities Journal Citation Reports, Scopus, and SciELO Citation Index (Web of Science database).</p> <p>In this version 2, only the productivity data were changed, covering a larger number of journals (1,314) and including all types of documents. Other data are the same as in the first version (https://doi.org/10.5281/zenodo.5572873).</p> <p> </p> <p> </p> <p> </p> <pre> </pre> <p> </p>
An open, collaborative, and scholarly digital edition of Anastās Mārī al-Karmalī's monthly journal "Lughat al-ʿArab" (Baghdad, 1911--14)
<p>This repository has seen a lot of edits since the last release:</p> <p>changes</p> <ul> <li>moved to central TEI boilerplate</li> <li>switched interface to Arabic</li> </ul> <p>edits</p> <ul> <li>added mastheads to vol.s 1 and 2</li> <li>some validation of automated structural mark-up</li> <li>manual validation of mark-up against the facsimile: vol. 1</li> <li>removed faulty automated mark-up of dates and periodical titles and re-added them through much improved process</li> <li>mark-up of named entities</li> <li>linked entity names to authority files</li> </ul>
An open, collaborative, and scholarly digital edition of ʿAbd al-Qādir al-Iskandarānī's monthly journal "al-Ḥaqāʾiq" (Damascus, 1910--12)
<p>The last release, v0.9, wasn't caught by Zenodo for some unknown reason. Nothing has been changed since. The following release message has been copied from v0.9:</p> <p>As the mark-up of al-Ḥaqāʾiq has been complete for some time now and we are waiting for our scans to be processed at Halle for some three years now, we have decided to release a practically complete version. There have been a couple of maintenance edits over the last years to unify the set-up of editions across the entire project. Thus, the folder with TEI files was renamed tei/. Entity linked as also considerably improved, particularly for periodical titles.</p>
Acknowledged scholars extracted from open access journals
<p><strong>Published paper</strong><br> For details of the data development: </p> <p>Kusumegi, K., Sano, Y. Dataset of identified scholars mentioned in acknowledgement statements. Sci Data 9, 461 (2022). https://doi.org/10.1038/s41597-022-01585-y</p> <p>To use the data, please cite the above mentioned paper .</p> <p> </p> <p><strong>Data on scholars mentioned in acknowledgments extracted from open access journals.</strong></p> <p>This data firstly updated the repository at University of Tsukuba (Division of Policy and Planning Sciences Commons).</p> <p>https://commons.sk.tsukuba.ac.jp/data_en</p> <pre> <strong>Each csv correspond to:</strong> biology.csv as PLOS Biology compbiology.csv as PLOS Computational Biology genetics.csv as PLOS Genetics medicine.csv as PLOS Medicine ntds.csv as PLOS Neglected Tropical Diseases pathogenes.csv as PLOS Pathogens plosone.csv as PLOS ONE srep.csv as Scientific Reports Data details: Doi: DOI of the paper PaperId: Paper ID in MAG AcknowledgedId: Acknowledged scholar's ID in MAG CollaborationApproach: Acknowledged scholar is identified by collaboration relationship CitationApproach: Acknowledged scholar is identified by citation relationship </pre> <p> </p>
Journal data for European scholarly journals
<p>This relates to the following study: https://doi.org/10.5281/zenodo.5909512</p> <p>The methodology is described in the linked manuscript.</p> <p> </p>
Network data for the paper: Intellectual and social similarity among scholarly journals.
<p>Network data used for the analysis contained in Baccini A, Barabesi L, Gingras Y, Kalfaoui M (2019) Intellectual and social similarity among scholarly journals. An exploratory comparison of the networks of editors, authors and co-citations.</p> <p>Data are in .net format for Pajek software</p> <p>CC indicates co-citation network.</p> <p>IA indicated Interlocking authorship network.</p> <p>IE indicates interlocking editorship network.</p> <p>Stat is for statistics; Econ is for economics; ILS is for information and library science.</p> <p> </p> <p> </p>
Mapping the Swiss Landscape of Diamond Open Access Journals. The PLATO Study on Scholar-Led Publishing. Dataset
<p><strong>Context</strong><br> From March to September 2022, the <a href="https://www.openscience.uzh.ch/en/openaccess/plato.html">«Platinum Open Access Funding» Project (PLATO)</a>, in collaboration with the Institute for Applied Data Science & Finance at the Bern University of Applied Sciences, undertook a bibliometric and empirical study of the Platinum/Diamond open access journal landscape in Switzerland. The PLATO project is an initiative of six Swiss universities – the University of Zurich, the University of Bern, the University of Geneva, the University of Neuchâtel, the Zurich University of the Arts and ETH Zurich –, dedicated to furthering community-led scholarly publishing in Switzerland. Diamond open access stands for a concept of equitable open access to and participation in scholarly publishing that is free for both authors and readers.</p> <p><strong>Presentation</strong><br> The main objective of the PLATO Study was to gain insight into the Platinum/Diamond open access publishing ecosystem in Switzerland through a mixed-method approach. The study consisted of three parts: First, bibliometric data were combined with inputs from Swiss open access publishers, institutional open access experts as well as information on journal websites to identify Swiss Diamond OA journals and their main characteristics. Second, seven semi-structured interviews with editors of select Diamond OA journals were conducted to generate a thorough understanding of their workflows, infrastructures, business models, challenges and opportunities. Third, based on the inputs from the interviews, three surveys were designed and sent to authors/reviewers, editors, and representatives of hosting and funding institutions of Swiss Diamond OA journals.</p> <p>The results of the study are published in the form of the following outputs:</p> <ul> <li>DOI Report: 10.5281/zenodo.7461728</li> <li>DOI Bibliometric List: <a href="https://zenodo.org/record/6992615#.YzK3ElJBw-Q">10.34914/olos:l2tys6tie5f63h35lqpqzzlx24</a></li> <li>DOI Data Set: 10.5281/zenodo.7461754</li> </ul> <p>The Data Set comprises the following files:</p> <ul> <li>Survey questionnaires</li> </ul> <p> SurveyQuestionnaire_Author.pdf<br> SurveyQuestionnaire_Editor.pdf<br> SurveyQuestionnaire_Publisher.pdf</p> <ul> <li>Data collected in three survey studies addressing journal editors, authors/reviewers as well as representatives of hosting and funding institutions:</li> </ul> <p> StudyData_Author.csv<br> StudyData_Editor.csv<br> StudyData_Publisher.csv</p> <ul> <li>Codebooks explaining the coding of the survey data files:</li> </ul> <p> StudyCodebook_Author.pdf<br> StudyCodebook_Editor.pdf<br> StudyCodebook_Publisher.pdf</p>
A dataset of scholarly journals in wikidata : (selected) external identifiers
<p> </p> <p><strong>For an updated list , see </strong></p> <p><a href="https://github.com/almugabo/openalex_qa/blob/main/coverage/OpenAlex_venues.md">Matching OpenAlex venues to Wikidata identifiers</a></p> <p><strong>Motivation : the selective/Inclusive approach in bibliometric databases </strong></p> <p>An important difference between bibliometric databases is their “<em>inclusion policy”</em>. </p> <p>Some databases like Web Of Science and Scopus select the sources they index, while others like Dimensions and OpenAlex are more inclusive (they index for example all data from a given source such as Crossref). </p> <p><a href="https://doi.org/10.1162/qss_a_00018">WOS</a></p> <p>“<em>selectivity remained a hallmark of coverage because Garfield had decided early on to focus on internationally influential journals.” (...)</em>.” </p> <p><a href="https://doi.org/10.1162/qss_a_00019">SCOPUS </a></p> <p>“<em>Serial content (i.e., journals, conference proceedings, and book series) submitted for possible inclusion in Scopus by editors and publishers is reviewed and selected, based on criteria of scientific quality and rigor. This selection process is carried out by an external Content Selection and Advisory Board (CSAB) of editorially independent scientists, each of which are subject matter experts in their respective fields. This ensures that only high-quality curated content is indexed in the database and affirms the trustworthiness of Scopus</em>”</p> <p> </p> <p><a href="https://doi.org/10.1162/qss_a_00020">Dimensions </a></p> <p><em>We have decided to take an “inclusive” approach to the publications we index in Dimensions. We believe that Dimensions should be a comprehensive data source, not a judgment call, and so we index as broad a swath of content as possible and have developed a number of features (e.g., the Dimensions API, journal list filters that limit search results to journals that appear in sources such as Pubmed or the 2015 Australian ERA</em><em>6</em><em> journal list) that allow users to filter and select the data that is most relevant to their specific needs.</em></p> <p><br> <br> <strong>Using wikidata to enable the filtering of “ venues subsets” in OpenAlex </strong></p> <p> </p> <p>We are interested in creating subsets of venues in OpenAlex (for example for comparative analysis with inclusive databases or other use cases). This would require matching identifiers of OpenAlex venues to other identifiers. </p> <p>Thanks to WikiCite, a project to record and link scholarly data, Wikidata has a large collection of metadata related to Scholarly journals. This repository provides a subset of the scholarly journals in Wikidata, focusing mainly on external identifiers.</p> <p>The dataset will be used to explore the extent to which wikidata journal external identifiers can be used to select the content in OpenAlex. </p> <p>(see here an <a href="https://github.com/almugabo/open_metadata/wiki/Journals">list of openly available lists of journals </a>)</p> <p><strong>Dataset creation & Documentation </strong></p> <ul> <li> <p>Wikidata dump from 2022-02-21</p> </li> <li> <p>Extract entities with following properties: </p> <ul> <li> <p>https://www.wikidata.org/wiki/Q5633421 # scientific journal (Q5633421)</p> </li> <li> <p>https://www.wikidata.org/wiki/Q737498 # academic journal (Q737498)</p> </li> </ul> </li> <li> <p>Extract the properties related to (selected) external identifiers </p> </li> </ul> <p> </p> <p>Some numbers : </p> <p>Number of journals in wikidata : 113,797 ; With issn_l 95,888 , With OpenAlex_venue id : 29,150</p> <p><strong>external identifiers</strong></p> <p>https://www.wikidata.org/wiki/Property:P236 # ext_id_issn</p> <p>https://www.wikidata.org/wiki/Property:P7363 # ext_id_issn_l</p> <p>https://www.wikidata.org/wiki/Property:P8375 # ext_id_crossref_journal_id</p> <p>https://www.wikidata.org/wiki/Property:P1055 # ext_id_nlm_unique_id</p> <p>https://www.wikidata.org/wiki/Property:P1058 # ext_id_era_journal_id</p> <p>https://www.wikidata.org/wiki/Property:P1250 # ext_id_danish_bif_id</p> <p>https://www.wikidata.org/wiki/Property:P10283 #ext_id_openalex_id</p> <p>https://www.wikidata.org/wiki/Property:P1156 # ext_id_scopus_source_id</p> <p><br> <strong>Indexing services</strong></p> <p>https://www.wikidata.org/wiki/Property:P8875</p> <p>https://www.wikidata.org/wiki/Q371467 # Scopus</p> <p>https://www.wikidata.org/wiki/Q104047209 # Science Citation Index Expanded</p> <p>https://www.wikidata.org/wiki/Q22908122 # Emerging Sources Citation Index</p> <p>https://www.wikidata.org/wiki/Q1090953 # Social Sciences Citation Index</p> <p>https://www.wikidata.org/wiki/Q713927 # Arts and Humanities Citation index</p> <p> </p> <p> </p> <p> </p> <p> </p>
An open dataset of scholarly publications highlighted by journal editors
<p><strong>Abstract: </strong></p> <p>We present a dataset of scholarly publications featured as outstanding by journal editors.<br> This first version, which covers the last 10 years, includes papers referenced in (a) Breakthroughs of the year by Science Magazine (b) Les 10 découvertes de l'année by La Recherche Magazine (c) research highlights by Nature journal and (c) Editors' choice by Science magazine.</p> <p>The rationale and process of its creation are described in details in the following pre-print:</p> <p>Mugabushaka, A.M. , Sadat, J and Dantas Faria, J.C. (2020). In Search of Outstanding Research Advances: prototyping the creation of an open dataset of "editorial highlights" . <a href="https://arxiv.org/abs/2011.07910">Arxiv 2011.07910</a></p> <p><br> <strong>Description of the dataset</strong></p> <p>In this version, the datasets are released in 4 Microsoft Excel Files.</p> <p>1. Science - Breakthroughs of the year</p> <p>2. La Recherche - les dix découvertes de l'année</p> <p>3. Nature Magazine - Research highlights</p> <p>4. Science magazine - Editors' choices</p> <p>The entries for each year are recorded in a separate sheet.</p> <p>In each sheet, the highlighting article and its metadata are recorded as well as the referenced papers together with their identifiers (doi)</p> <p> </p> <p> </p>
Core competencies of a scholarly journal editor. Database and citation search report and PRISMA_flow_diagram
<p><strong>Report on Database and Citation search and PRISMA_flow_diagram</strong></p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.