Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5,061
datasets available to search
ShareScore release 0.7.1
Dataset results
5,061 results for “access”
Monitoring and evaluation of UKRI's Open Access Policy: Exploring the use of open data sources to inform baseline values - Dataset
<p>This dataset accompanies the report <em>"Monitoring and evaluation of UKRI's Open Access Policy: Exploring the use of open data sources to inform baseline values"</em>, which is available via Zenodo.<br><br>It provides record-level data of UKRI-funded and UK-affiliated research output (limited to journal articles with Crossref DOIs) published between 2012 and 2022 - including bibliographic metadata as well as data on open access availability, publisher, national and international collaborations, citations, views and downloads, altmetrics and subjects (fields). All variables are documented in the data dictionary included in this Zenodo record.</p> <p>The code used to generate the dataset from open data sources is available on GitHub. </p> <p>The following data sources were used:</p> <ul> <li> <p>Gateway to Research (records downloaded between 2023-11-05 and 2023-11-13)</p> </li> <li> <p>Crossref (Metadata Plus snaphot 2023-10-31, Crossref member route API 2024-01-23)</p> </li> <li> <p>OpenAlex (data snapshot 2023-10-18)</p> </li> <li> <p>Unpaywall (data snapshot 2023-11-27)</p> </li> <li> <p>IRUS UK (2024-04-03)</p> </li> <li> <p>Crossref Event Data (2023-04-01)</p> </li> </ul> <p><strong></strong><br><br>The project made use of Curtin Open Knowledge Initiative (COKI) infrastructure, which is documented on GitHub: <a href="https://github.com/The-Academic-Observatory">https://github.com/The-Academic-Observatory</a>. </p>
Data accessibility in the chemical sciences: an analysis of recent practice in organic chemistry journals
<div> <p>Data is the analysis of the data outputs of 240 randomly selected research papers from 12 top-ranked journals published in early 2023. We investigate author compliance with recommended (but not compulsory) data policies, whether there is evidence to suggest that authors apply FAIR data guidance in their data publishing, and if the existence of specific recommendations for publishing NMR data by some journals encourages compliance. Files in the data package have been provided in both human and machine-readable forms. The main dataset is available in the Excel file Data worksheet.XLSX, the contents of which can also be found in Main_dataset.CSV, Data_types.CSV, and Article_selection.CSV with explanations of the variable coding used in the studies in Variable_names.CSV, Codes.CSV, and FAIR_variable_coding.CSV. The R code used for the article selection can be found in Article_selection.R. Data about article types from the journals that contain original research data is in Article_types.CSV. Data collected for analysis in our sister paper[4] can be found in Extended_Adherence.CSV, Extended_Crystallography.CSV, Extended_DAS.CSV, Extended_File_Types.CSV, and Extended_Submission_Process.CSV. A full list of files in the data package and a short description for each is given in README.TXT.</p> </div>
Interviews for New Business Models for Pharmaceutical Innovation and Access to Medicines - Case Study of the Oral Cholera Vaccine Development
<p>These supplementary materials represent the partial dataset in the form of semi-structured interviews, collected and analyzed in the research article "The 30-year evolution of oral cholera vaccines: A case study of a collaborative network alternative innovation model". This article is one of the outcomes of the "New Business Models for Pharmaceutical Innovation and Global Access to Medicines" research project, conducted at the Global Health Center, within the Geneva Graduate Institute. The dataset contains 8/16 interviews collected and used in this article, which are published with the informed consent of the interviewees.</p>
C2SMARTER Year 1 Project "Enhancing Transit Access and Safety Through Equitable Micromobility Solution"
<p>These 6 PDF files are the maps produced from Task 1 of the C2SMARTER Year 1 project Enhancing Transit Access and Safety Through Equitable Micromobility Solution.</p> <p>Site A and Site B are transit underserved areas (census tracts in El Paso, Texas) identified in Task 1 of this project.</p> <p>The first 2 maps shows the underserved areas overlaid with bus stops (taken from the General Transit Feed Specification or GTFS database).</p> <p>The next 2 maps shows the underserved areas overlaid with locations of crashes involving pedestrians and bicycles from 1/1/2024 to 7/30/2024..</p> <p>The last 2 maps color coded the streets in Site A and Site B with bicycle level of traffic stress (LTS).</p>
Environmental and AIS data collected during the EUMarineRobots Trans-National Access activities experiments using the NATO STO-CMRE Littoral Ocean Observatory Network testbed
<p>Environmental and AIS data collected during the H2020 project EUMarineRobots Trans-National Access activities experiments using the NATO STO-CMRE Littoral Ocean Observatory Network (LOON) testbed. Environmental data consists of temperature measured across the water column; sound velocity measured close to the surface and close to the sea bottom; meteorological data at the surface (i.e., pressure, temperature, wind speed and direction, humidity and rain). The environmental dataset is complemented with Automatic Identification System (AIS) data for the ships transiting close to the LOON area (Gulf of La Spezia, Italy)</p> <p>Temperature measured across the water column in the LOON area (Gulf of La Spezia, Italy). The dataset includes measurements for:<br> i) Nov 12, 19-20, 23-24 - 2020<br> ii) Dec 1-4, 14-20 - 2020<br> iii) Jan 12-13, 15, 18-24, 27-28 - 2021</p> <p><br> Meteorological data at the surface (i.e., pressure, temperature, wind speed and direction, humidity and rain) in the LOON area (Gulf of La Spezia, Italy). The dataset includes measurements for:<br> i) Nov 12, 19-20, 23-24 - 2020<br> ii) Dec 1-4, 14-20 - 2020<br> iii) Jan 12-13, 15, 18-24, 27-28 - 2021</p> <p><br> Sound velocity measured close to the surface (SVP1) and close to the sea bottom (SVP2) in the LOON area (Gulf of La Spezia, Italy). The dataset includes measurements for:<br> i) Nov 12, 19-20, 23-24 - 2020<br> ii) Dec 1-4, 14-20 - 2020<br> iii) Jan 12-13, 15, 18-24, 27-28 - 2021</p> <p>SVP2 data missing for Dec 14-20 (2020) and Jan 24, 27-28 (2021).</p> <p>Automatic Identification System (AIS) data for the ships transiting close to the LOON area (Gulf of La Spezia, Italy). The dataset includes AIS data for:<br> i) Nov 12, 19-20, 23-24 - 2020<br> ii) Dec 1-4, 14-20 - 2020<br> iii) Jan 12-13, 15, 18-24, 27-28 - 2021<br> </p> <p>For reference, see: "Environmental data collected on the CMRE LOON tested during the EUMR project: dataset description", Petroccia, Roberto; Zappa, Giovanni; Cimino, Giampaolo; Grati, Alberto; Alves, João. CMRE-DA-2021-001. July 2021, available at https://www.cmre.nato.int/research/publications/latest-techreports/1638-cmre-da-2021-001</p>
Open-Access Data for "Received SignalStrength Measurements with BLE Signals for Contact Tracing and Proximity Detection"
<p>This archive contains three folders which are supplementary material for the paper accepted for publishing in IEEE Sensors Journal.</p> <p><strong>Contents:</strong></p> <ul> <li> The folder `open-access-data/upb/` contains the measurements acquired at UPB. The subfolders are named as `upb_ble_*`, where an asterisk masks the directory number. Whenever UPB is specified, use the data sets from the corresponding directory.</li> <li>The folder `open-access-data/tau/` contains the measurements acquired at TAU. The subfolders are named as `tau_ble_*`, where an asterisk masks the directory number. Whenever TAU is specified, use the data sets from the corresponding directory.</li> <li>The folder `open-access-data/wifi-on-off/` contains a sample code to read the files and plot the data from Fig. 14 in `open-access-data/wifi-on-off/wifi_on_off_read_plot.py` and Fig. 15 in `open-access-data/wifi-on-off/wifi_on_off_read_plot.ipynb`.</li> </ul> <p><strong>Results based on the data have been presented in the paper:</strong><br> Flueratoru, L., Shubina, V., Niculescu, D., Lohan, E.S. (2021). On the High Fluctuations of Received Signal Strength Measurements with BLE Signals for Contact Tracing and Proximity Detection, IEEE Sensors, Special Issue on Advanced Sensors and Sensing Technologies for Indoor Positioning and Navigation</p> <p><strong>To cite these data sets please use the following:</strong><br> Laura Flueratoru, Viktoriia Shubina, Dragoș Niculescu, & Elena Simona Lohan. (2021). Open Access Data for "Received SignalStrength Measurements with BLE Signals for Contact Tracing and Proximity Detection" [Data set]. Zenodo. http://doi.org/10.5281/zenodo.4643668</p>
yangclaraliu/armslist_scraping: provide access to data
<p>This is the data used in the research letter:</p> <p>Drake, Coleman, Ashley M. Hernandez, Yang Liu, Adam H. Schwartz, and Maria E. Sundaram. "Evidence of Background Checks in an Online Firearms Marketplace." <em>American journal of preventive medicine</em> 57, no. 5 (2019): 718.</p> <p>https://www.ajpmonline.org/article/S0749-3797(19)30273-9/fulltext#%20</p>
Improving the data access control using blockchain for healthcare domain
<p>This research reviews the importance of blockchains in healthcare as they provide infinite possibilities to individuals, companies, and governments.</p>
Discovering dataset download link, or access via service, from DOI metadata
<p>Diagram showing how it can be possible to access a digital resource that a DOI identifies, either by direct download or via a web service, from the DOI's DataCite metadata. </p>
The Unofficial Guide on applying NCN Open Access rules to GitHub repositories.
<p><b>The Unofficial Guide on applying NCN Open Access rules to GitHub repositories.</b> <i>Some</i> HTML code can be used here.</p>
Restriction of access to the central cavity is a major contributor to substrate selectivity in plant ABCG transporters
<p>The input and main output files used for the paper <em><strong>"Restriction of access to the central cavity is a major contributor to substrate selectivity in plant ABCG transporters"</strong></em> are separated in the different tar files depending the MD stage they belong to.</p> <p><strong>Content</strong></p> <p>00_AlphaFold2: The models predicted from AlphaFold2</p> <p>01_build_system: The parameters for ATP and the initial pdb file used to build each system</p> <p>02_minimization: Minimization input files for each variant</p> <p>03_equilibration: Equilibration input files for each variant</p> <p>04_long_equilibration: ATP restrained equilibration input files for each variant</p> <p>05_free_equilibration: Free equilibration input files for each variant</p> <p>06_production: Free production input files for each variant</p> <p>07_caver: Caver calculations for each variant and replica</p> <p>08_transport_tools: Analysis of tunnel networks using TransportTools software</p> <p>09_tunnel_selection: Selection of the tunnels with widest bottleneck radius to perform CaverDock experiments</p> <p>10_caverdock: CaverDock calculations for each variant and for each ligand tested</p> <p>11_membrane_patch: MD simulations for liquiritigenin and POPC membrane only</p> <p>12_MD_analysis: Calculations of RMSD, RMSF of each system. Calculation of helical parameters for trans-membrane helices 2, 5, 8 and 11 (not for APO). Calculation of X1 and X2 angles for residue N1331 in each variant (not for APO)</p> <p>13_US_closed_to_open: Umbrella Sampling simulations to obtain the inward facing (IF) open state of each variant. Not used for PMF analysis.</p> <p>14_US_opening_energy: Umbrella Sampling simulations to obtain the Potential of Mean Force for the transition from IF-closed to IF-open conformations.</p> <p>15_US_equilibration: Equilibration input files for WT and F562L variants in IF-open states.</p> <p>16_US_production: Production input files for WT and F562L variants in IF-open states.</p> <p>17_US_caver: Caver calculations for WT and F562L variants in IF-open states.</p> <p>18_US_transport_tools: Analysis of tunnel networks using TransportTools software of WT and F562L variants in IF-open states.</p> <p>19_US_tunnel_selection: Selection of the tunnels with widest bottleneck radius to perform CaverDock experiments for WT and F562L variants in IF-open states.</p> <p>20_US_caverdock: CaverDock calculations for WT and F562L variants in IF-open states.</p> <p>21_US_MD_analysis: Calculations of RMSD, RMSF of each system. Calculation of helical parameters for trans-membrane helices 2, 5, 8 and 11.</p> <p>ABC_Sequences.fasta: Sequences from 1KP analysis</p>
Data from the OPERAS business models survey on open access books
<p>OPERAS (the European Research Infrastructure for the development of open scholarly communication in the social sciences and humanities) has conducted a survey of publishing organisations throughout Europe to identify and better understand existing and potential business models to support the Open Access publication of research monographs. The results of the survey are used to inform the formulation of recommendations about how to create a sustainable open access book publishing ecosystem within Europe.</p> <p>The survey was designed to serve two core aims: <br> 1. To further, better or improve our understanding of the scholarly publishing landscape and of the challenges that publishers face in the context of publishing OA monographs;<br> 2. To identify main trends (including opportunities and challenges) and the knowledge of collaborative funding and infrastructure models in OA publishing in SSH. </p> <p>The survey was open between 16 February and 14 April 2021.</p> <p>The results are presented in two versions of the white paper of the Open Access Business Models Special Interest Group: Stone, Graham, Błaszczyńska, Marta, Lebon, Chloé, Morka, Agata, Mosterd, Tom, Mounier, Pierre, Proudman, Vanessa, Speicher, Lara, & Melinščak Zlodi, Iva. (2021). Collaborative models for OA book publishers (1.0). Zenodo. https://doi.org/10.5281/zenodo.5494731 and the second version to be published in Spring 2023.</p>
GEOLAB - Transnational Access project QC-CEM - Mapping quick clay with geophysical methods
<p>Quick clay is characterised by complete collapse and liquid-like mobility when overloaded. Quick clay is found primarily in Norway and Sweden, but also exists in Finland, Russia, Canada and Alaska. Quick clay landslides, with their retrogression characteristics and extreme mobility, pose significant risk to human lives, infrastructure, property and surrounding ecosystems. Hence, the proper characterization of quick clay sites is essential for ensuring the safety and resilience of infrastructure in Norway and elsewhere in Europe.<br> The current practice for mapping quick clay in Norway relies heavily on borehole data with either rotary sounding or total sounding and core samples tested in the laboratory. The only method for identifying quick clay with certainty is physical testing in the laboratory, but it is time-consuming, expensive and gives limited information, i.e., only at the depths and locations where the samples are taken. In Norway, rotary sounding and total soundings are frequently used in mapping of quick clay. There is increasing interest in using geophysical methods such as Electrical Resistivity Tomography (ERT) to supplement the results from soundings, particularly in early stage of ground investigation for mapping of quick clay. ERT is a near surface geophysical method that uses direct current to measure the earth's electrical resistivity. The current is injected into the subsurface through steel electrodes installed 10-20 cm into the ground, and the apparent resistivity distribution along a profile or area is measured. Using data processing and inverse modelling a 2D or 3D resistivity model of the subsurface can be derived.<br> Geophysical methods such as ERT show capability to identify not quick clay such as sand, silt, dry crust, moraine and bed rock reasonably accurate, but the identification of quick clay is still generally limited. The detection of leached clay (thus potentially quick clay) is however possible.<br> Transnational Access project QC-CEM is funded through the 1st call for proposal for the GEOLAB project. This project aims at testing various geophysical methods for their capability for soil characterisation, particularly for detecting quick clay.</p> <p>The objectives of the QC-CEM project are:<br> (i) to test different configurations of Electrical Resistivity Tomography survey for detection of quick clay<br> (ii) to test innovative and efficient electromagnetic based methods for mapping of quick clay. Results from this investigation is not available to share at this stage.<br> (iii) to investigation the effectiveness of cross-interpretation using different geophysical methods for soil characterisation. The results from this activity will be published in open publication after they are processed.</p>
efantnu/drep-2021-collab-MINESParisTech-NTNU: Pre-print version IEEE Access
<p>This repository contains the data sets of the paper "Allocation of spinning reserves in autonomous grids considering frequency stability constraints and short-term solar power variations" authored by Erick F. Alves, Louis Polleux, Gilles Guerassimoff, Magnus Korpås, Elisabetta Tedeschi. With these files, it is possible to reproduce most simulations and results obtained in the paper.</p> <p>Folder organization</p> <ul> <li>Results: final results and values of intermediate steps of the optimization model implemented in Gurobi 9.1 and described in section III-A and III-B of the paper.</li> <li>Validation: test system implemented in OpenModelica for validation of the results and described in section III-C of paper.</li> <li>Solar convex hulls: hourly convex hulls obtained using the procedure detailed in Appendix A.</li> <li>Solar irradiance profiles: high-resolution irradiance timeseries from NREL used to identify the worst-case solar PV ramp scenarios.</li> </ul>
Project "Public services management system to improve the quality and accessibility of services" (01.2.2-LMT-K-718-03-0019) interviews
<p>The dataset of depersonalized qualitative semi-structured interviews with the representatives of public sector organisations providing public services in Lithuania. The data were collected in December-November, 2022 as a part of the project "Public services management system to improve the quality and accessibility of services" ("Viešųjų paslaugų vadybos sistema paslaugų kokybei ir prieinamumui gerinti"), grant no. 01.2.2-LMT-K-718-03-0019, funded by the Lithuanian research council. The interviews are in the Lithuanian language.</p>
Genome assembly and annotation files for Corylus americana accessions 'Rush' and 'Winkler'
<p>The native shrub American hazelnut (<em>Corylus americana</em>) is currently used in breeding programs that are aiming to develop commercially viable hazelnut varieties for the U.S. Upper Midwestern U.S. This species provides significant ecological benefits as it is a perennial crop and well-adapted to this region. Breeding cycles for perennial species are long, and may benefit from the use of predictive methods such as genomic selection to reduce cycle time and increase the efficiency of field trials.</p> <p>High-quality reference genome assemblies are very useful for the implementation marker-assisted selection and genomic prediction, and we therefore developed the first chromosome-scale reference assemblies for <em>C. americana</em>, using the accessions 'Rush' and 'Winkler'. Initial draft assemblies were created using HiFi PacBio reads and Arima Hi-C sequencing to assemble genomes into 11 pseudomolecules. We then utilized Oxford Nanopore reads and a high-density genetic map in order to perform error correction. N50 scores were calculated to be 31.9 Mb and 35.3 Mb for 'Rush' and 'Winkler', respectively, while 97.1% (for 'Winkler') and 90.2% (for 'Rush') of the total genome was assembled into the 11 pseudomolecules. Gene prediction was performed using both RNAseq libraries as well as protein homology data. 'Rush' had a BUSCO score of 99.0 for its assembly and 99.0 for its annotation, while 'Winkler' had corresponding scores of 96.9 and 96.5, indicating extremely high-quality assemblies.</p> <p>These two independent, de novo assemblies enable unbiased assessment of structural variation across the genome, as well as patterns of syntenic relationships within C. americana and the <em>Corylus</em> genus. These assemblies are also an important first step in providing a resource for using next-generation sequencing data in the improvement of <em>C. americana</em>. We demonstrate this utility through the generation of high-density SNP marker sets from genotyping-by-sequencing data for 1,343 <em>C. americana</em>, <em>C. avellana</em>, and <em>C. americana</em> x <em>C. avellana</em> hybrids, in order to assess population structure in natural and breeding populations. Finally, the transcriptomes of these assemblies, as well as several other recently published <em>Corylus</em> genomes, were utilized to perform phylogenetic analysis of sporophytic self-incompatibility (SSI) in hazelnut, providing further evidence of unique molecular pathways governing self-incompatibility in Corylus not exhibited in other well-studied SSI systems. We hope these assemblies will aide in the application of modern breeding methods to the development of commercially viable hazelnut varieties for the U.S. Upper Midwest.</p>
Dataset from VR Streaming Server (Emulated) and Radio Access Network for Streaming Traffic
<p>The dataset contains an experiment in a site where UEs attach to a gNodeB that provides access to a streaming server that is stressed with high demanding transcoding workloads to emulate VR/AR processes. The UEs are realized through the Remote UE mode enabled by Amarisoft Simbox emulator, and the gNodeB is realized through the Amarisoft Callbox, which also provides the user plane function. The emulated VR streaming server is deployed as a Nginx pod in a Kubernetes cluster.</p> <p>We rely on MonB5G sampling functions that feed monitoring data (CPU and RAN parameters) to the monitoring system. A streaming video server has been deployed with the help of a NGINX server. It provides video-on-demand and video streaming, which can be accessed by any user (or UE) for real-time reproduction. This VR video streaming emulation aids to assess the performance of the network and therefore the benefits that each solution has brought. The video “Big Buck Bunny” with h.264 encoding and a resolution of 1920x1080p has been used for the experiments.The description of dataset features are:<br> 1-) Index Number,<br> 2-) Time: Time of the experiment,<br> 3-) N: number of VR streaming clients,<br> 4-) C: Average CPU of VR streaming server [mc]<br> 5-) O: Outbound traffic at the server average outbound traffic (O) flowing from the data interface of the video server. <br> 6-) R: Instantaneous downlink bit rate [Mbps],</p> <p>The original video file information:</p> <table> <tbody> <tr> <td> <p>Video codec </p> </td> <td> <p>Advanced Video Codec (AVC) </p> </td> </tr> <tr> <td> <p>Width </p> </td> <td> <p>1920 pixels </p> </td> </tr> <tr> <td> <p>Height </p> </td> <td> <p>1080 pixels </p> </td> </tr> <tr> <td> <p>Display aspect radio </p> </td> <td> <p>16:9 </p> </td> </tr> <tr> <td> <p>Duration </p> </td> <td> <p>10 min 34 s </p> </td> </tr> <tr> <td> <p>Max Bitrate </p> </td> <td> <p>16.7 Mb/s </p> </td> </tr> <tr> <td> <p>Frame rate </p> </td> <td> <p>30 FPS </p> </td> </tr> </tbody> </table>
AgrImOnIA: Open Access dataset correlating livestock and air quality in the Lombardy region, Italy
<p>The AgrImOnIA dataset is a comprehensive dataset relating air quality and livestock (expressed as the density of bovines and swine bred) along with weather and other variables. The AgrImOnIA Dataset represents the first step of the <a href="http://www.agrimonia.net">AgrImOnIA project</a>. The purpose of this dataset is to give the opportunity to assess the impact of agriculture on air quality in Lombardy through statistical techniques capable of highlighting the relationship between the livestock sector and air pollutants concentrations.</p> <p>The building process of the dataset is detailed in the <strong>companion paper:</strong></p> <p>A. Fassò, J. Rodeschini, A. Fusta Moro, Q. Shaboviq, P. Maranzano, M. Cameletti, F. Finazzi, N. Golini, R. Ignaccolo, and P. Otto (2023). Agrimonia: a dataset on livestock, meteorology and air quality in the Lombardy region, Italy. <em>SCIENTIFIC DATA</em>, 1-19.</p> <p>available <a href="https://rdcu.be/c7T9H">here</a>.</p> <p>This dataset is a collection of estimated daily values for a range of measurements of different dimensions as: air quality, meteorology, emissions, livestock animals and land use. Data are related to Lombardy and the surrounding area for 2016-2021, inclusive. The surrounding area is obtained by applying a 0.3° buffer on Lombardy borders.</p> <p>The data uses several aggregation and interpolation methods to estimate the measurement for all days.</p> <p>The files in the record, renamed according to their version (es. .._v_3_0_0), are:</p> <ul> <li> <p>Agrimonia_Dataset.csv(.mat and .Rdata) which is built by joining the daily time series related to the AQ, WE, EM, LI and LA variables. In order to simplify access to variables in the Agrimonia dataset, the variable name starts with the dimension of the variable, i.e., the name of the variables related to the AQ dimension start with 'AQ_'. This file is archived also in the format for MATLAB and R software. </p> </li> <li> <p>Metadata_Agrimonia.csv which provides further information about the Agrimonia variables: e.g. sources used, original names of the variables imported, transformations applied.</p> </li> <li> <p>Metadata_AQ_imputation_uncertainty.csv which contains the daily uncertainty estimate of the imputed observation for the AQ to mitigate missing data in the hourly time series. </p> </li> <li> <p>Metadata_LA_CORINE_labels.csv which contains the label and the description associated with the CLC class. </p> </li> <li> <p>Metadata_monitoring_network_registry.csv which contains all details about the AQ monitoring station used to build the dataset. Information about air quality monitoring stations include: station type, municipality code, environment type, altitude, pollutants sampled and other. Each row represents a single sensor.</p> </li> <li> <p>Metadata_LA_SIARL_labels.csv which contains the label and the description associated with the SIARL class.</p> </li> <li> <p>AGC_Dataset.csv(.mat and .Rdata) that includes daily data of almost all variables available in the Agrimonia Dataset (excluding AQ variables) on an equidistant grid covering the Lombardy region and its surrounding area. </p> </li> </ul> <p>The Agrimonia dataset can be reproduced using the code available at the GitHub page: <a href="https://github.com/AgrImOnIA-project/AgrImOnIA_Data">https://github.com/AgrImOnIA-project/AgrImOnIA_Data</a></p> <p><strong>UPDATE 31/05/2023</strong> <strong>- NEW RELEASE - V 3.0.0</strong></p> <p>A new version of the dataset is released: Agrimonia_Dataset_v_3_0_0.csv (.Rdata and .mat), where variable <em>WE_rh_min, WE_rh_mean and WE_rh_max </em>have been recomputed due to some bugs<em>.</em></p> <p>In addition, two new columns are added, they are <em>LI_pigs_v2 and LI_bovine_v2 </em>and represents the density of the pigs and bovine (expressed as animals per kilometer squared) of a square of size ~ 10 x 10 km centered at the station localisation.</p> <p>A new dataset is released: the Agrimonia Grid Covariates (AGC) that includes daily information for the period from 2016 to 2020 of almost all variables within the Agrimonia Dataset on a equidistant grid containing the Lombardy region and its surrounding area. The AGC does not include AQ variables as they come from the monitoring stations that are irregularly spread over the area considered.</p> <p><strong>UPDATE 11/03/2023</strong> <strong>- NEW RELEASE - V 2.0.2</strong></p> <p>A new version of the dataset is released: Agrimonia_Dataset_v_2_0_2.csv (.Rdata), where variable <em>WE_tot_precipitation </em>have been recomputed due to some bugs<em>.</em></p> <p>A new version of the metadata is available: Metadata_Agrimonia_v_2_0_2.csv where the spatial resolution of the variable <em>WE_precipitation_t </em>is corrected.</p> <ul> </ul> <p><strong>UPDATE 24/01/2023</strong> <strong>- NEW RELEASE - V 2.0.1</strong></p> <p>minor bug fixed</p> <p><strong>UPDATE 16/01/2023</strong> <strong>- NEW RELEASE - V 2.0.0</strong></p> <p>A new version of the dataset is released, Agrimonia_Dataset_v_2_0_0.csv (.Rdata) and Metadata_monitoring_network_registry_v_2_0_0.csv. Some minor points have been addressed:</p> <ul> <li>Added values for <em>LA_land_use</em> variable for Switzerland stations (in Agrimonia Dataset_v_2_0_0.csv)</li> <li>Deleted incorrect values for <em>LA_soil_use</em> variable for stations outside Lombardy region during 2018 (in Agrimonia Dataset_v_2_0_0.csv)</li> <li>Fixed duplicate sensors corresponding to the same pollutant within the same station<em> </em>(in Metadata_monitoring_network_registry_v_2_0_0.csv)</li> </ul>
Project "Public services management system to improve the quality and accessibility of services" (01.2.2-LMT-K-718-03-0019) literature review screening results
<p>The results of the keyword query in Scopus search with the abstracts were screened using <i>abstractr </i>platform at <a href="http://abstrackr.cebm.brown.edu">http://abstrackr.cebm.brown.edu</a>. Four reviewers reviewed intersecting subsets of the overall list of publications in separate reviews, therefore duplicate records in the file are possible. The results from four reviews were combined into one file using functionality of <i>abstractr </i>platform. The reviews were finalized in February, 2022. Majority of the publications from the Scopus query results were automatically assigned low relevance scores thanks to the active learning algorithm used by <i>abstractr</i> and therefore were not reviewed manually.</p><p>Notes on the columns of the dataset:</p><ul><li>(internal) id - internal id added by <i>abstractr.</i></li><li>(source) id - Scopus document id followed by underscore and '1' (if publication has DOI) or 'n' (if publication has no DOI).</li><li>keywords - authors' keywords and Scopus keywords concatenated from the Scopus query results.</li><li>abstract - an abstract of a publication from the Scopus query results.</li><li>title - title of a publication from the Scopus query results.</li><li>journal - journal of a publication from the Scopus query results.</li><li>authors - authors of a publlication from the Scopus query results.</li><li>consensus - for publications reviewed by multiple reviewers the consensus decision is signified by '1', no consensus - by 'x' and unable to asses consensus by 'o'. These codes are generated by <i>abstrackr.</i></li><li>eg - the first reviewer id. Code '1' means the publication was selected based on title and abstract, '0' - unsure, '-1' rejected.</li><li>dj - the second reviewer id. Code '1' means the publication was selected based on title and abstract, '0' - unsure, '-1' rejected.</li><li>rp - the third reviewer id. Code '1' means the publication was selected based on title and abstract, '0' - unsure, '-1' rejected.</li><li>mp - the fourth reviewer id. Code '1' means the publication was selected based on title and abstract, '0' - unsure, '-1' rejected.</li><li>count (+1) - integer, the number of reviewers selecting the publication for fulltext reading. Calculated from 'eg ','dj ','rp' and 'mp' columns.</li><li>count (0) - integer, the number of reviewers not sure of selecting the publication for fulltext reading. Calculated from 'eg ','dj ','rp' and 'mp' columns.</li><li>count (-1) - integer, the number of reviewers not selecting the publication for fulltext reading. Calculated from 'eg ','dj ','rp' and 'mp' columns.</li><li>at leat once selected - binary integer, representing the final decision rule to select publications for fulltext reading and further analysis.</li></ul><p>The data were collected as a part of the project "Public services management system to improve the quality and accessibility of services" ("Viešųjų paslaugų vadybos sistema paslaugų kokybei ir prieinamumui gerinti"), grant no. 01.2.2-LMT-K-718-03-0019, funded by the Lithuanian research council.</p>
zbMATH Open Access Subset
<p>The dataset contains two tables as csv files.</p> <p>1) documents_in_oa_series</p> <p>is a list of zbMath Open documents in serials where the description contains the word "Open Access". Note that some documents did not appear yet or might have been retracted. Thus when fetching information from the oai-pmh API or the website, be prepared to handle non-existing documents.</p> <p>2) oa_links</p> <p>Lists all links from zbMATH Open documents to fulltext matched by either unpaywall or arxiv.</p> <p>The meaning of the fields is:</p> <ul> <li><strong>zbmath_id</strong> Unique identifier from zbMATH Open. Prefix with <code>https://zbmath.org/</code> to visit additional information on the article. For example, <code>5635019</code> is associated with <a href="https://zbmath.org/5635019">https://zbmath.org/5635019</a></li> <li><strong>link</strong> link to the fulltext. Note that those links are not fully reliable. We estimate a success rate of 90%</li> </ul> <p>Note that there is an overlap between 1 and 2. So some documents are published in OA serials and have arxiv or unpaywall links at the same time.</p> <p> </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.