Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
3,688
datasets available to search
ShareScore release 0.7.1
Dataset results
3,688 results for “Computer”
Dataset for the paper: Computational framework for radionuclide migration assessment in clay rocks
<p>This dataset contains OpenGeoSys-6 simulations results and scripts in Jupyter-notebook format for reproducing and visualizing the results of the paper:</p> <p>Garibay-Rodriguez J, Chen C, Shao H, Bilke L, Kolditz O, Montoya V and Lu R (2022) Computational Framework for Radionuclide Migration Assessment in Clay Rocks. Front. Nucl. Eng. 1:919541. doi: 10.3389/fnuen.2022.919541</p> <p>Details about the usage of the scripts can be found in the OpenGeoSys-6 user guide: https://www.opengeosys.org/</p>
Demonstration of computer assisted scholarly collation
<p>Demonstration of computer assisted scholarly collation, using the collation of General Prologue 488 in the Textual Communities interface. The software here used is a customization of the Collation Editor, written by Cat Smith of the Institute for Textual Scholarship and Electronic Editing at the University of Birmingham, with assistance from Troy Griffitts, itself an extension of the CollateX program, written by Ronald Dekkers and others.</p>
DFT Calculated xyz Files in Support of "Bidentate Rh(I)-Phosphine Complexes for the C-H Activation of Alkanes: Computational Modelling and Mechanistic Insight"
<p>Theoretically calculated xyz files for propane, carbon monoxide, butyraldehyde and multiple Rh-phosphine complexes as well as transition states relevant for the C-H activation and subsequent carbonylation of alkanes.</p> <p>All quantum chemical simulations were performed using the Gaussian 16 software package. Closed-shell equilibrium structures, i.e., minima and transition states (TSs) as well as electronic properties of educts, intermediates, and products involved in the C-H activation of propane (methyl group activation) and subsequent steps mediated by the Rh complexes were obtained at the DFT level of theory. The range-separated B97XD functional was employed. The def2-SVP basis set and the respective effective core potential (ECP) were utilized for all atoms. TSs were fully optimized at the same level of theory using the rational function optimization (RFO) approach as well as nudged elastic band (NEB) method as implemented in the pysisyphus<sup> </sup>software suite. Subsequently, a vibrational analysis was carried out for each stationary point to verify that a minimum or first-order saddle point was obtained on the 3<em>N</em>-6-dimensional potential energy (hyper)surface (PES).</p>
A computational study of the structure and function of human Zrt and Irt-like proteins metal transporters: An elevator-type transport mechanism predicted by AlphaFold2
<p>Data produced and analyzed in the manuscript "A computational study of the structure and function of human Zrt and Irt-like proteins metal transporters: An elevator-type transport mechanism predicted by AlphaFold2" by Pasquadibisceglie et al.</p> <p><br> If you include these data in your manuscript, please cite: Pasquadibisceglie A, Leccese A and Polticelli F (2022) A computational study of the structure and function of human Zrt and Irt-like proteins metal transporters: An elevator-type transport mechanism predicted by AlphaFold2. <em>Front. Chem.</em> 10:1004815. doi: 10.3389/fchem.2022.1004815</p>
Dataset for publication 'Automated computed tomography based parasitoid detection in mason bee rearings'
<p>Dataset corresponding to the publication <strong>Automated computed tomography based parasitoid detection in mason bee rearings </strong>published at PLOS One. This repository contains the training, validation and test data.</p>
A Greek Parliament Proceedings Dataset for Computational Linguistics and Political Analysis
<p>The dataset is a new version of the previous upload and includes the following files:</p> <p>1. <strong>dataset_versions/tell_all.csv: </strong>The initial dataset of 1,280,927 extracted speeches, before preprocessing and cleaning. The speeches extend chronologically from July 1989 up to July 2020 and were exported from 5,355 parliamentary sitting record files. The file has a total volume of 2.5 GB and includes the following columns:</p> <ul> <li>member_name: the name of the individual who spoke during a sitting.</li> <li>sitting_date: the date the sitting took place.</li> <li>parliamentary_period: the name and/or number of the parliamentary period that the speech took place in. A parliamentary period is defined as the time span between one general election and the next. A parliamentary period includes multiple parliamentary sessions.</li> <li>parliamentary_session: the name and/or number of the parliamentary session that the speech took place in. A session is defined as a time span of usually 10 months within a parliamentary period during which the parliament can convene and function as stipulated by the constitution. A session can fall into the following categories: regular, extraordinary or special. In the intervals between the sessions the parliament is in recess. A parliamentary session includes multiple parliamentary sittings.</li> <li>parliamentary_sitting: the name and/or number of the parliamentary sitting that the speech took place in. A sitting is defined as a meeting of parliament members.</li> <li>political_party: the political party of the speaker.</li> <li>government: the government in force when the speech took place.</li> <li>member_region: the electoral district the speaker belonged to.</li> <li>roles: information about the parliamentary roles and/or government position of the speaker.</li> <li>member_gender: the gender of the speaker</li> <li>speech: the speech that the individual gave during the parliamentary sitting.</li> </ul> <p>2. <strong>dataset_versions/tell_all_FILLED.csv: </strong>This file is an intermediate version of the dataset that includes improvements in the consistency and completeness of the dataset, with a total volume of 2.5 GB. Specifically, this file is produced by filling the missing names of chairmen of various parliamentary sittings of the "tell_all.csv". It includes the same columns as the "tell_all.csv" file.</p> <p>3.<strong> dataset_versions/tell_all_cleaned.csv: </strong>This version of the dataset is the result of further cleaning and preprocessing and is used for our word usage change study. It consists of 1,280,918 speech fragments of Greek parliament members in the order of the conversation that took place, with a total volume of 2.12 GB. It includes the same columns as the aforementioned versions. The preprocessing includes the replacement of all references to political parties with the symbol "@" followed by an abbreviation of the party name, using regular expressions that capture different grammatical cases and variations. It also includes the removal of accents, strings with length less than 2 characters, all punctuation except full stops, and the replacement of stopwords with "@sw".</p> <p>4. <strong>wiki_data</strong>: A folder of modern Greek female and male names and surnames and their available grammatical cases crawled from the entries of the Wiktionary Greek names category (https://en.wiktionary.org/wiki/Category:Greek_names). We produced the grammatical cases of the missing grammatical entries according to the rules of the Greek grammar and saved the files in the same folder by adding to their filenames the string "_populated.json".</p> <p>5. <strong>parl_members_activity_1989onwards_with_gender.csv</strong>: The Greek Parliament website provides a<br> <a href="https://www.hellenicparliament.gr/Vouleftes/Diatelesantes-Vouleftes-Apo-Ti-Metapolitefsi-Os-Simera/">list</a> of all the elected members of parliament since the fall of the military junta in Greece, in 1974. We collected and cleaned the data, added the gender and kept the elected members from 1989 onwards, matching the available parliament proceeding records. This dataset includes the full names of the members, the date range of their service, the political party they served, the electoral district they belonged to and their gender.</p> <p>6. <strong>formatted_roles_gov_members_data.csv</strong>: As government members we refer to individuals in ministerial or other government posts, regardless of whether they were elected in the parliament. This information is available in the website of the <a href="https://gslegal.gov.gr/?page_id=776&sort=time">Secretariat General for Legal and Parliamentary Affairs</a>. The government members dataset includes the full names of the official individuals, the name of the role they were given, the date range of their service at each specific role and their gender.</p> <p>7. <strong>governments_1989onwards.csv</strong>: A dataset of government information including the names of governments since 1989, their start and end dates, and a URL that points to the respective official government web page of each past government. The data is crawled from the website of the <a href="https://gslegal.gov.gr/?page_id=776&sort=time">Secretariat General for Legal and Parliamentary Affairs</a>.</p> <p>8. <strong>extra_roles_manually_collected.csv</strong>: A dataset with manually collected information from Wikipedia about additional government or parliament posts such as Chairman of the Parliament, party leaders, opposition leaders and other information.</p> <p>9. <strong>all_members_activity.csv</strong>: A dataset of all the information of the aforementioned files 3,4,5,6 merged. Each row of the file includes the full name of the individual, the start and end date of their term of office, the political party and electoral district they belonged to, their gender, the parliamentary and/or government positions that they held along with start and end dates, and the name of the government that was in power during their term of office. An individual can change political parties or become an independent member of the parliament during a parliamentary period, thus having more than one entries/rows in the file.</p> <p>10. <strong>freqs_for_semantic_shift_cleaned_data_decade1990.csv & freqs_for_semantic_shift_cleaned_data_decade2010.csv</strong>: Files of frequencies of words in the corpora of the decades 1990-1999 and 2010-2019.</p> <p>11. <strong>compass_top100.csv:</strong> Top 100 most changed words between the decades 1990-1999 and 2010-2019, as computed with the use of the Compass tool by V. D. Carlo et. al. [1].</p> <p>12. <strong>compass_fc_top100.csv</strong>: Top 100 most changed words between the decades 1990-1999 and 2010-2019, as computed with the use of the Compass tool [1] in combination with the frequency cut-offs of the Gonen et. al. approach [3]. For the frequency cut-offs, the files in bullet 8 are used.</p> <p>13. <strong> procrustes_top100.csv</strong>: Top 100 most changed words between the decades 1990-1999 and 2010-2019, as computed with the use of the Orthogonal Procrustes approach of Hamilton et. al. [2].</p> <p>14. <strong>nn_top100.csv</strong>: Top 100 most changed words between the decades 1990-1999 and 2010-2019, as computed with the use of the Gonen et. al. approach [3].</p> <p>15. <strong>second_order_top100.csv</strong>: Top 100 most changed words between the decades 1990-1999 and 2010-2019, as computed with the use of the Second-Order Similarity approach by Hamilton et. al. [4].</p> <p>16. <strong>top100_minfreq50.xls</strong>: An .xls file for convinient viewing of the top 100 most changed words per approach with minimum frequency of 50 occurrences, produced by merging the aforementioned files 11, 12, 13, 14, 15 and 16.</p> <p>17. <strong>freqs_for_semantic_shift_cleaned_data_period1997_2007.csv & freqs_for_semantic_shift_cleaned_data_period2008_2018.csv</strong>: Files of frequencies of words in the corpora of the decades before (1997_2007) and during (2008_2018) the Greek economic crisis.</p> <p>18. <strong>semantic_shifts_dichotomy_crisis_compass_1997_2007_2008_2018_atleast50.csv</strong>: A file with the top 100 most changed words between between the decades before (1997-2007) and during (2008-2018) the Greek economic crisis. The computations are implemented with the use of the Compass tool.</p> <p>19. <strong>selected_topics_shift_per_period_compass.csv</strong>: The usage change of selected topics/words of generic political interest between pairs of consecutive parliamentary periods. The computations are implemented with the use of the Compass tool.</p> <p>20. <strong>semantic_shifts_party_embeddings_per_period_merged_compass.csv</strong>: The usage change of selected political party names that have played an important role in recent political history, namely New Democracy (ND), the Panhellenic Socialist Movement (PASOK), the Coalition of the Radical Left - Progressive Alliance (SYRIZA), the Communist Party of Greece (KKE), the Coalition of the Left, of Movements and Ecology (SYN) and Golden Dawn (GD).</p> <p>-------------</p> <p><strong><em>Citations:</em></strong></p> <p>[1] Valerio Di Carlo, Federico Bianchi, and Matteo Palmonari. Training Temporal Word Em- beddings with a Compass. In <em>Proceedings of the Thirty–Third AAAI Conference on Artificial Intelligence</em>, AAAI’19, pages 6326–6334, 2019. doi: 10.1609/aaai.v33i01.33016326.</p> <p>[2] William L. Hamilton, Jure Leskovec, and Dan Jurafsky. Diachronic Word Embeddings Reveal Statistical Laws of Semantic Change. In <em>Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)</em>, ACL 2016, pages 1489– 1501, Berlin, Germany, August 2016. Association for Computational Linguistics. doi: 10. 18653/v1/P16-1141. URL https://www.aclweb.org/anthology/P16-1141.</p> <p>[3] Hila Gonen, Ganesh Jawahar, Djamé Seddah, and Yoav Goldberg. Simple, Interpretable and Stable Method for Detecting Words with Usage Change across Corpora. In <em>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics</em>, ACL 2020, pages 538– 555, Online, July 2020. Association for Computational Linguistics. doi: 10.18653/v1/2020.acl- main.51. URL https://aclanthology.org/2020.acl-main.51.</p> <p>[4] William L. Hamilton, Jure Leskovec, and Dan Jurafsky. Cultural Shift or Linguistic Drift? Comparing Two Computational Measures of Semantic Change. In <em>Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing</em>, EMNLP 2016, pages 2116–2121, Austin, Texas, November 2016. Association for Computational Linguistics. doi: 10.18653/v1/D16-1229. URL https://www.aclweb.org/anthology/D16-1229.</p> <p>-------------</p> <p><strong><em>Acknowledgments:</em></strong></p> <p>This work was supported by the European Union’s Horizon 2020 research and innovation program ``FASTEN'' under grant agreement No 825328 and the non profit data journalism organization iMEdD.org.</p>
Demonstration of Android Jank when Fibonacci computation is done on main thread
<p>This video demonstrates Android Ui jank which occurs when heavy fibonacci computation is done on main thread</p>
Video demonstrating results of applying show view updates on Android Activity without Fibonacci Computation on the main thread
<p>Video demonstrating results of applying show view updates on Android Activity without Fibonacci Computation on the main thread</p>
Video demonstrating results of applying show view updates on Android Activity with Fibonacci Computation on the main thread
<p>Video demonstrating results of applying show view updates on Android Activity with Fibonacci Computation on the main thread</p>
Target-focused library design by pocket-applied computer vision and fragment deep generative linking
<p>Data inputs and outputs used in</p> <pre>Target-focused library design by pocket-applied computer vision and fragment deep generative linking</pre> <p>Code: https://github.com/kimeguida/POEM</p> <p> </p>
Computational Prediction and Experimental Realisation of Earth Abundant Transparent Conducting Oxide Ga-doped ZnSb2O6 - computational datasets
<p>Additional computational data for our publication on Ga-doped ZnSb2O6. This repository contains two files related to the defect calculations ("new-defect-corrections.json" and "aide-plotter"), which contain data concerning the charge correction schemes and total energies of all defect species, and the charge transition levels (which can be used to repeat the SCFL analysis), respectively. The other two files are the AMSET output files, which can be used to plot all the calculated charge transport data used in this study. Documentation for plotting using these files is provided at https://github.com/hackingmaterials/amset.</p> <p>Contact:</p> <p>joe.willis.15@ucl.ac.uk; d.scanlon@ucl.ac.uk</p>
A computational method for predicting the most likely evolutionary trajectories in the stepwise accumulation of resistance mutations
<p>Supporting information dataset for <em>A computational method for predicting the most likely evolutionary trajectories in the stepwise accumulation of resistance mutations, </em>including Flex ddG binding free energy predictions, epistasis calculations, pathway probabilities, Rosetta files and structural files. </p>
Large-scale grid computing for content-based image retrieval
<p>The author presents an approach in which a large distributed processing Grid has been used to apply a range of content-based image retrieval methods to a substantial number of images. By massively distributing the required computational task across thousands of Grid nodes, we have achieved very high throughput at relatively low overheads.</p> <p> </p>
Cloud computing is one of the most popular and sophisticated technologies adopted by organizations worldwide. Some world-leading organizations enhance their efficiency and effectiveness by using cloud computing technology. Working from home (WFH) has been a popular trend among organizations during the coronavirus (COVID-19) pandemic. The COVID-19 saw a breakthrough in work cultures and environments where working from home was a remarkable success in remote working environments, despite being a rare phenomenon in Sri Lanka. Yet, it is argued that the deployment of work from home has not been effective among Sri Lankan business organizations due to a lack of IT infrastructure, facilities, and knowledge. The purpose of the study is to investigate the impact of cloud computing, embracing the service models (Infrastructure as a Service, Platform as a Service, and Software as a Service) as theoretical lenses and testing the COVID-19 as the moderator. The study has been conducted based on a deductive approach and adopted a stratified random sampling method. The sample consisted of 384 IT employees among those who had experienced working from home. The study utilized multiple regression and found that cloud computing service models significantly impact work from home with the moderating effect of COVID-19.
<p>Cloud computing is one of the most popular and sophisticated technologies adopted by organizations worldwide. Some world-leading organizations enhance their efficiency and effectiveness by using cloud computing technology. Working from home (WFH) has been a popular trend among organizations during the coronavirus (COVID-19) pandemic. The COVID-19 saw a breakthrough in work cultures and environments where working from home was a remarkable success in remote working environments, despite being a rare phenomenon in Sri Lanka. Yet, it is argued that the deployment of work from home has not been effective among Sri Lankan business organizations due to a lack of IT infrastructure, facilities, and knowledge. The purpose of the study is to investigate the impact of cloud computing, embracing the service models (Infrastructure as a Service, Platform as a Service, and Software as a Service) as theoretical lenses and testing the COVID-19 as the moderator. The study has been conducted based on a deductive approach and adopted a stratified random sampling method. The sample consisted of 384 IT employees among those who had experienced working from home. The study utilized multiple regression and found that cloud computing service models significantly impact work from home with the moderating effect of COVID-19.</p>
Code and data for N Le et. al "Scalable and robust quantum computing on qubit arrays with fixed coupling"
<p>Simulation code and data used in N Le et. al "Scalable and robust quantum computing on qubit arrays with fixed coupling."</p>
Nanopore sequencing data analysis using Microsoft Azure cloud computing service
<p>Genetic information provides insights into the exome, genome, epigenetics and structural organisation of the organism. Given the enormous amount of genetic information, scientists are able to perform mammoth tasks to improve the standard of health care such as determining genetic influences on outcome of allogeneic transplantation. Cloud-based computing has increasingly become a key choice for many scientists, engineers and institutions as it offers on-demand network access and users can conveniently rent rather than buy all required computing resources. With the positive advancements of cloud computing and nanopore sequencing data output, we were motivated to develop an automated and scalable analysis pipeline utilizing cloud infrastructure in Microsoft Azure to accelerate HLA genotyping service and improve the efficiency of the workflow at lower cost. In this study, we describe (i) the selection process for suitable virtual machine sizes for computing resources to balance between the best performance versus cost-effectiveness; (ii) the building of Docker containers to include all tools in the cloud computational environment; (iii) the comparison of HLA genotype concordance between the in-house manual method and the automated cloud-based pipeline to assess data accuracy. In conclusion, the Microsoft Azure cloud-based data analysis pipeline was shown to meet all the key imperatives for performance, cost, usability, simplicity and accuracy. Importantly, the pipeline allows for the ongoing maintenance and testing of version changes before implementation. This pipeline is suitable for data analysis from MinION sequencing platforms and could be adopted for other data analysis application processes.</p>
Supplemental data for "Estimating the Jones polynomial for Ising anyons on noisy quantum computers"
<p>This data supports "Estimating the Jones polynomial for Ising anyons on noisy quantum computers" by Chris N. Self, Sofyan Iblisdir, Gavin K. Brennen, and Konstantinos Meichanetzidis https://arxiv.org/abs/2210.11127</p> <p>Related code can be found in the GitHub repository: (https://github.com/chris-n-self/Ising-anyons-Jones-polynomials-for-NISQ). The 'analysis' folder here can be dropped inside the code repository in order to view the data using the 'view...' notebooks. The experimental data folders 'ibmq_...' contain the individual sets of results and can be used to generate new zero-noise-extrapolation fits.</p>
Geometries for "Linear Response Properties of Solvated Systems: A Computational Study"
<p>Geometry files (.xyz format) used in the paper "Linear Response Properties of Solvated Systems: A Computational Study"</p>
An experimental investigation on the water retention behaviour of a silty soil for the computation of the lateral earth thrust on a retaining wall.
<p>The dataset is linked with the conference paper available open access with the same title. In the excel file, each spreadsheet reports the data of the figures available in the manuscript.</p>
Old Skool Computer
Old skool computer 3d modelling, an old project. High Poly model. Suitable for interior design. Now belongs to everyone. Source: Objaverse 1.0 / Sketchfab
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.