Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
18
datasets available to search
ShareScore release 0.7.1
Dataset results
18 results for “Open Infrastructures”
Data for "Breaking the Paywall: The role of Open Journal System as key Open Science infrastructure"
<h3><strong>Context</strong></h3> <p>This research was conducted within the NSF-SEEKCommons Project, a research initiative dedicated to supporting Open Science and Open Access in disciplinary research. The project has a special interest in understanding the role that critical infrastructure has in supporting open initiatives. The Open Journal System (OJS) serves as a long-standing fundamental piece for Open Access throughout the globe. Hence, it provides valuable information about experiences developing, deploying, and maintaining open technologies. </p> <h3><strong>Methods<br></strong></h3> <div> <div>We used mixed methods for our research, triangulating repository data, installation data, interviews, and documentary analysis. We collected repository data using a report generator (Kopp [2018] 2024) that uses repository metadata to present general statistics about a Git project. The resulting information was manually curated, disambiguated, and annotated to have a homogeneous set of developers with information about their institutional affiliation and country. </div> <div> </div> <div>Names are normalized based on the information in qualitative interviews and by browsing the full-extent commits in the GitHub repository. Other sources for this were the institutional materials (available in current and archived versions of the PKP website), meeting minutes, the user forum, and further project documentation available online. GitHub handles are homologated to their most comprehensive version. For institutional and country affiliation, we resorted to GitHub profiles, PKP documentation and forums, institutional domains available in emails, and researchers' ORCID IDs. </div> </div> <h3><strong>Available files</strong></h3> <ol> <li><strong>Information about the codebase</strong> (number of files, lines of code, and timestamp) organized by <strong>month, quarter, and semester. </strong><br>See file: OJS_GitStats_04-24.csv</li> <li>Information about the historical evolution of the codebase (number of files, lines of code, and timestamp), including <strong>a description of the top committers for each month</strong>. Commiters are described by including their institutional affiliation and country of origin. <br>See file: OJS_DevStats_Institution-Country_1.tsv</li> <li>Information about the <strong>historical evolution of the codebase </strong>focusing on <strong>top committers</strong>, along with their institution and country. This file is formatted to map the co-occurrence of developers and attributes by month between 2004-2024.<br>See file: OJS_DevStats_Institution-Country_2.tsv</li> <li>Selected fields to describe<strong> working and regularly maintained plugins for OJS as of October 2024.</strong> Includes name of the plugin, homepage, description, maintainer, and institutional affiliation. <br>See file: OJS_Plugins_2024_Processed.tsv</li> <li>Details of the aggregated <strong>information</strong> included in <strong>Table</strong> <strong>5</strong> of the article.<br>See file: OJS_Plugins_2024_Table5.tsv</li> <li><strong>Snapshot</strong> to XML information of the <strong>plugin gallery of OJS </strong>(October 21) retrieved from PKP website (Smecher 2024)<br>See file: OJS_Plugins_2024.csv</li> </ol> <h3>Funding</h3> <p><span>The SEEKCommons Project is funded by the U.S. National Science Foundation (NSF), grant #2226425</span></p>
Towards an open pipeline for the detection of Critical Infrastructure from satellite imagery – A case study on electrical substations in The Netherlands
<p><strong>Abstract.</strong> Critical infrastructure (CI) are at risk of failure due to the increased frequency and magnitude of climate extremes related to climate change. It is thus essential to include them in a risk management framework to identify risk hotspots, develop risk management policies and support adaptation strategies to enhance their resilience. However, the lack of information on the exposure of CI prevents their incorporation in large-scale risk assessment studies. This study sets out to improve the representation of CI for risk assessment studies by building a neural network model to detect CI assets from optical remote sensing imagery. We present a pipeline that extracts CI from OpenStreetMaps, processes the imagery and assets' masks, and trains a Mask R-CNN model that allows for instance segmentation of CI at the asset level. This study provides an overview of the pipeline and tests it with the detection of electrical substations assets in the Netherlands. Several experiments are presented for different under-sampling percentages of the majority class (25%, 50% and 100%) and hyperparameters settings (batch size and learning rate). The best metrics achieved are an Average Precision at an Intersection over Union of 50% of 30.93 and a tile F-score of 89.88%. This allows us to confirm the feasibility of the method and invite disaster risk researchers to use this pipeline for other infrastructure types. We conclude by exploring the different avenues to improve the pipeline by addressing the class imbalance, Transfer Learning and Explainable AI.</p>
Data for: The State of Open Infrastructure Grant Funding, 2024 State of Open Infrastructure Report
<p>The purpose of the analysis based on these data was to better understand the amount, distribution, impact, and limitations of grant funding to open infrastructures that support research and scholarship.</p> <p>The data were summarized and reported in the “2024 State of Open Infrastructure Report” section “The state of open infrastructure grant funding.” The full report is available at https://doi.org/10.5281/zenodo.10934089.</p> <p>A readme, data dictionary, and additional metadata definition file are provided with the dataset with additional detail.</p>
Data for: Open Infrastructure Governance: Current structures, nomenclature, composition, and service trends, 2024 State of Open Infrastructure Report
<p>The purpose of the analysis based on these data was to<span> record information about community governance groups for open infrastructures, focused primarily on the individuals and institutions that serve in these groups. The data were summarized and reported in the “2024 State of Open Infrastructure Report” section “Open infrastructure governance: Current structures, nomenclature, composition, and trends.” The full report is available at <a href="The%20data%20were%20summarized%20and%20reported%20in%20the%20&ldquo;2024%20State%20of%20Open%20Infrastructure%20Report&rdquo;%20section%20&ldquo;Open%20infrastructure%20governance:%20Current%20structures,%20nomenclature,%20composition,%20and%20trends,&rdquo;%20available%20at%20https:/doi.org/10.5281/zenodo.10934089.">https://doi.org/10.5281/zenodo.10934089</a>.</span></p> <p><span>A readme is provided with the dataset with additional detail.</span></p>
Recordings Q&A and matchmaking sessions for call for proposals 'Open Science Infrastructure'
<p>On Thursday, July 11, and Tuesday, July 16, 2024, Open Science NL organised two online Q&A sessions combined with a matchmaking opportunity for the Open Science NL call 'Open Science Infrastructure'.</p> <p>These are the two recordings of the two Q&A sessions. The Open Science NL team has drafted a Frequently Asked Questions document addressing all the questions that came up during the meetings. This is added as a seperate text-file (PDF). The slides presented during both meetings are shared as well as a PDF.</p> <p>For more information about the call and how to apply, please go to: <a href="https://www.openscience.nl/en/calls/open-science-infrastructure" target="_blank" rel="noopener">https://www.openscience.nl/en/calls/open-science-infrastructure</a></p>
Zambezi dataset to "WHAT-IF: an open-source decision support tool for water infrastructure investment planning within the Water-Energy-Food-Climate Nexus"
<p>This is the dataset used in the HESS publication "<a href="https://www.hydrol-earth-syst-sci-discuss.net/hess-2019-167/">WHAT-IF: an open-source decision support tool for water infrastructure investment planning within the Water-Energy-Food-Climate Nexus</a>"</p> <p>The dataset describes the water-energy-food nexus of the Zambezi River Basin used as input to the <a href="https://github.com/RaphaelPB/WHAT-IF">WHAT-IF model</a>.</p> <p>The file Data_Organization.pdf, summarizes the available data. For more info look at the <a href="https://www.hydrol-earth-syst-sci-discuss.net/hess-2019-167/">publication</a> and/or <a href="https://github.com/RaphaelPB/WHAT-IF">Github</a>.</p>
A lack of open data standards for large infrastructure projects hampers social-ecological research in the Brazilian Amazon
<p>List of papers used in literature review for "A lack of open data standards for large infrastructure projects hampers social-ecological research in the Brazilian Amazon"</p>
Data for: Characteristics of Selected Open Infrastructures, 2024 State of Open Infrastructure Report
<p>The State of Open Infrastructure report provides an annual snapshot of general characteristics for open infrastructures (OIs) listed in Invest in Open Infrastructure’s (IOI) open infrastructure selection tool, Infra Finder (https://infrafinder.investinopen.org/).</p> <p>The data were summarized and reported in the “2024 State of Open Infrastructure Report” section “Characteristics of selected open infrastructures.” The full report is available at https://doi.org/10.5281/zenodo.10934089.</p> <p>A readme, data dictionary, and additional metadata definition file are provided with the dataset with additional detail.</p>
Evaluating Institutional Commitments to Open Scholarly Infrastructure: A Review of Open Access Collection Development Policies
<p>Data prepared for the publication "Evaluating Institutional Commitments to Open Scholarly Infrastructure: A Review of Open Access Collection Development Policies."</p> <p><strong>oa-cd-policies.csv</strong></p> <p>Scope: This data represents collection development policies that contain substantial mention of open access.</p> <p>Data collection: The policies were sourced using an Advanced Google Search for "open access" AND "collection development policy" at ".edu" domains.</p> <p>Variables:</p> <ul> <li>institution: Free text, name of the institution.</li> <li>carnegie_class: One of <a href="https://carnegieclassifications.acenet.edu/carnegie-classification/classification-methodology/basic-classification/">these options</a>; the Carnegie classification of the institution.</li> <li>institution_type: One of public or private; the funding source of the institution.</li> <li>policy_name: Free text; the title of the policy.</li> <li>supplemental_policy: Link to a supplemental open access policy if linked in the collection development policy.</li> <li>cd_policy_link: Link to the policy.</li> <li>infrastructure: TRUE or FALSE; whether the policy includes a commitment to open access scholarly infrastructure development, including open source platforms, locally hosted platforms, consortia, or an institutional repository.</li> <li>excerpt: Free text; text from the policy that mentions infrastructure.</li> </ul> <p><strong>principles-policies.csv</strong></p> <p>Scope: This data represents those collection development policies from oa-cd-policies.csv that contain commitments in line with the <a href="https://openscholarlyinfrastructure.org/">Principles for Open Scholarly Infrastructure</a>.</p> <p>Variables:</p> <ul> <li>institution: Free text, name of the institution.</li> <li>carnegie_class: One of <a href="https://carnegieclassifications.acenet.edu/carnegie-classification/classification-methodology/basic-classification/">these options</a>; the Carnegie classification of the institution.</li> <li>institution_type: One of public or private; the funding source of the institution.</li> <li>policy_name: Free text; the title of the policy.</li> <li>supplemental_policy: Link to a supplemental open access policy if linked in the collection development policy.</li> <li>cd_policy_link: Link to the policy.</li> <li>principle: One of the three main <a href="https://openscholarlyinfrastructure.org/">Principles</a>.</li> <li>sub_principle: One of the <a href="https://openscholarlyinfrastructure.org/">Sub-Principles</a>.</li> <li>excerpt: Free text; text from the policy that illustrates the sub_principle.</li> </ul>
Open Data, Open Code, Open Infrastructure Schematic Diagram
<p>A schematic diagram of how social workflows, technical workflows, and project governance interact with the open data, open code, and open infrastructure (O3)</p>
Figures 2 and 3 in HighRes for Springer book chapter "Online Infrastructures For Open Educational Resources"
<p>Figure 2 and Figure 3 in high resolution for the book chapter:</p> <p>Marín, V. I., & Villar-Onrubia, D. (in press, 2022). Online Infrastructures For Open Educational Resources. In I. Jung & O. Zawacki-Richter (Eds.), <em>Handbook of Open, Distance and Digital Education</em> (Global Perspectives and Internationalization). Springer. <a href="https://doi.org/10.1007/978-981-19-0351-9_18-1">https://doi.org/10.1007/978-981-19-0351-9_18-1</a></p> <p>----</p> <p>Details for the figures:</p> <p>Fig. 2 Examples of national and regional digital infrastructures in Europe. (Note: The original figure of the Europe map was created by Commons user Alexrk2, CC BY-SA 3.0, shared via Wikimedia Commons)</p> <p>Figure 3. Examples of national and regional digital infrastructures in South America. (Note: The original figure of the South America map was created by TUBS, CC BY-SA 3.0, shared via Wikimedia Commons)</p>
Funding Covid-19 research: Insights from an exploratory analysis using open data infrastructures - Supplementary material
<p>This dataset contains supplementary material for the paper 'Funding Covid-19 research: Insights from an exploratory analysis using open data infrastructures' by Alexis-Michel Mugabushaka, Nees Jan van Eck, and Ludo Waltman.</p> <ul> <li>supplementary_material_1_dataset.ods: Dataset of Covid-19 publications.</li> <li>supplementary_material_2_sample.ods: Samples of publications used to assess the accuracy of funding data in the different databases.</li> <li>supplementary_material_3_tables_and_figures.ods: Statistics underlying the tables and figures presented in the paper.</li> </ul>
Reported funding data for open infrastructure
<p>Reported funding data for open infrastructure projects, focused primarily on the services in the <a href="https://investinopen.org/blog/funding-open-infrastructure-a-survey-of-available-data-sources/">pilot for Invest in Open Infrastructure's funding landscape research</a>, as well as other notable providers in the SCOMCat index.</p>
Mapping built infrastructure in semi-arid systems using data integration and open-source approaches for image classification
Open the record for dataset details and reuse information.
European power system infrastructure in the open energy system model PyPSA-Eur
<p>The image is created using the data and scripts in the European open energy system model <a href="https://github.com/PyPSA/pypsa-eur">PyPSA-Eur.</a></p>
Sustainable Open Infrastructure Cost Recovery Workshop Outputs
<p>These files reflect unattributed participant feedback on sustainablity and cost-recovery mechanisms generated through in-person workshops held in conjunction with the 2023 Charleston Library Conference, the 2024 Researcher to Reader Conference, and the 2024 OPERAS Conference. Information provided through asynchronous web-form consultations with governance representatives and virtual focus group participants are also included. Files include a spreadsheet and print-friendly pdf version of results from each workshop and votes from governance and focus group participants. A pdf of prelimary results collated after the Charleston Library Conference and Research to Reader workshops is also included. </p> <p>This work was made possible by the Mellon Foundation through the "Advancing to Launch by Developing IDS Governance Building Blocks" grant. Christina Drummond and Ursula Rabar designed and facilitated the community consultation processes that surfaced this information.</p>
Investing in Open Science infrastructure and repositories, shared repositories and Idea Challenge 2020-2021
<p>Recording of the OR2020 virtual session that includes:</p> <ul> <li>Investing in Open Science infrastructure and repositories: How IOI and SCOSS can help<strong> </strong>– Vanessa Proudman, SPARC Europe</li> <li>Shared repositories: building a multi-tenancy repository service at the British Library – Sara Gould, the British Library</li> <li>A year long Idea Challenge: the new beginning. Kicking-off Idea Challenge 2020-2021, for more information, please visit the <a href="https://or2020.sun.ac.za/a-year-long-idea-challenge/">Idea Challenge page</a>.</li> </ul>
Dataset: Scoping the Open Science Infrastructure Landscape in Europe
<p>"We see a diverse, interconnected, open, professional and viable, developing OS ecosystem in Europe on solid ground; one that is worth investing in. At the same time, this developing ecosystem faces a range of issues that challenge its path to a more open and sustainable future." This is a core conclusion of a new SPARC Europe report; the work is a result of a recent in-depth survey of infrastructure and/or services that are part of the European Open Science infrastructure (OSI) landscape. This dataset supports that work. </p> <p>The dataset is licensed CC0.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.