Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
155
datasets available to search
ShareScore release 0.7.1
Dataset results
155 results for “Popular”
Keyword frequencies in popular tech media (01.2016-02.2020)
<p><strong>Sources with weights</strong></p> <pre> Arstechnica: 1/8, Euractiv: 1/8, Fastcompany: 1/8, The Register: 1/8, Techcrunch: 1/8, The Guardian: 1/8, Venturebeat: 1/8, The Verge: 1/8 </pre> <p><strong>Methodology</strong></p> <ul> <li>Frequency of appearances for all unigrams and bigrams in the texts</li> <li>Frequency: number of appearances of every term divided by the number of published articles (for every month and source)</li> <li>This measure reveals how many times an expression has been mentioned on average per article</li> <li>Several media sources: a representative index is calculated with weighted average (weights as above)</li> <li>Average monthly change in the analised term's frequency is calculated by OLS regressions</li> <li>The dependent variable of the estimation is the frequency index, while the number of months since the beginning of the analysed period (January 2016) is the independent variable</li> <li>The regression coefficient (referred to as coef) shows by how much on average the analysed expression’s frequency changed with every observed month (marginal change of the frequency), revealing which keywords had the biggest monthly growth</li> </ul> <p><strong>Columns</strong></p> <p>freq_months (e.g. freq_2019-04): the average frequency of the term</p> <p>coef: the regression coefficient</p> <p>coef_norm: the regression coefficient divided by the mean frequency of the keyword</p> <p>coef_norm_max: the regression coefficient divided by the maximum frequency of the keyword</p>
Most popular songs on 11/11/2023 from the SoundCloud platform: Top 50
<p>This dataset contains the top 50 songs of all genres from the SoundCloud website and stores them in a CSV. This project corresponds to a practice assignment for 'Tipologia i Cicle de Vida de les Dades' (UOC)</p>
Supporting data for "Space weather in the popular media, and the opportunities the upcoming solar maximum brings"
<p>This contains the supporting data for the Space Weather Editorial "Space weather in the popular media, and the opportunities the upcoming solar maximum brings". It contains the Google Trends data (multiTimeline.csv) and the F10.7 and Kp data (omniweb.txt). </p>
A Stimuli Set of Forty Popular Music Drum Patterns with Perceived Complexity Estimates (Data set)
<p>Stimuli and datasets for the study "A Stimuli Set of Forty Popular Music Drum Patterns with Perceived Complexity Estimates"</p>
Reddit's NetSec forum GitHub projects, typology, maturity and popularity indicators
<p>Dataset including information of GitHub projects shared on Reddit's forum /r/netsec. It includes information about the posts itselves, like the author, title, image, tags, number of upvotes and number of comments, as well as repository information, such as number of stars, forks, issues, collaborators and the repository URL and last modification date.</p>
Películas Populares de IMDb en abril 2024
<p>El dataset proporciona las 100 películas más populares de la web IMDb.com extraído en Abril 2024</p> <p><strong>Variables Categóricas:</strong></p> <ol> <li><strong>Título Original de la Película (original_title):</strong> El título original de la película en su idioma original.</li> <li><strong>Título en Español (title):</strong> El título de la película traducido al español, si está disponible.</li> <li><strong>Géneros (genre1, genre2, genre3):</strong> Los géneros a los que pertenece la película, divididos en hasta tres variables distintas.</li> <li><strong>Director (director):</strong> El nombre del director de la película.</li> <li><strong>Clasificación de Edad (classification):</strong> La clasificación de edad recomendada para la película, que también podría ser una variable numérica discreta.</li> </ol> <p><strong>Variables Numéricas:</strong></p> <ol> <li><strong>Orden de Popularidad (ranking):</strong> El orden de las películas según su popularidad.</li> <li><strong>Rating (rating):</strong> La calificación o puntuación asignada a la película.</li> <li><strong>Año de Estreno (year):</strong> El año en que la película fue estrenada.</li> <li><strong>Duración (duration):</strong> La duración de la película en minutos.</li> </ol>
Associated data files and scripts for 'Ancestry-inclusive dog genomics challenges popular breed stereotypes'
<p>Behavioral genetics in dogs has focused on modern breeds, isolated subgroups <200 years old with distinctive physical, and, purportedly, behavioral characteristics. We interrogated breed stereotypes by surveying owners of 18,371 purebred and mixed-breed dogs, and densely genotyping (~45 million markers) a subset of 2,155 dogs. Most behavioral traits are heritable (h<sup>2</sup>>25%), and admixture patterns in mixed-breed dogs can reveal breed propensities. However, breed poorly predicts an individual purebred dog's behavioral phenotype, explaining just 9% of variation. Using genome-wide association, we identify 11 loci significantly associated with howling and other behaviors, and show characteristic breed behaviors are genetically complex. Behavior-associated loci are not unusually differentiated in modern breeds, but breed propensities do align, albeit weakly, with ancestral function. We propose behaviors now perceived as characteristic of modern breeds likely derive from thousands of years of polygenic adaptation predating breed formation, with modern breeds distinguished primarily by aesthetic, not behavioral, traits.</p>
Base de Datos - Canciones Populares Infantiles Españolas con Eementos Trágicos
<p>Base de datos en Excel de canciones populares infantiles españolas con elementos textuales trágicos, producto de la Tesis doctoral del autor.</p>
Genomic divergence, local adaptation, and complex demographic history may inform management of a popular sportfish species complex
<p>We investigated genomic divergence, interspecific and intraspecific diversity and population structure, local directional selection, and complex demographic history in a popular freshwater species complex in the Central Interior Highlands, North America. Specifically, we assessed differentiation between the Neosho Bass (<em>Micropterus velox</em>) and Smallmouth Bass (<em>M. dolomieu</em>) where their ranges are parapatric in this ecoregion. We scanned the genome for signatures of local direction and mapped divergence at selected SNPs at the population level. Additionally, we identified stream populations with extensive admixture, and we used a model-testing framework to investigate complex demographic histories between the two species.</p> <p>Raw ddRADseq .fastq sequence files, along with intermediate processing files, finalized VCF files, a data summary report, and other data information, are given in the "smb_ddRAD_rawdata.tar file. Metadata, including sample_id, species designation, and stream population, are given in the "metadata.xlsx" file.</p>
Cloud computing is one of the most popular and sophisticated technologies adopted by organizations worldwide. Some world-leading organizations enhance their efficiency and effectiveness by using cloud computing technology. Working from home (WFH) has been a popular trend among organizations during the coronavirus (COVID-19) pandemic. The COVID-19 saw a breakthrough in work cultures and environments where working from home was a remarkable success in remote working environments, despite being a rare phenomenon in Sri Lanka. Yet, it is argued that the deployment of work from home has not been effective among Sri Lankan business organizations due to a lack of IT infrastructure, facilities, and knowledge. The purpose of the study is to investigate the impact of cloud computing, embracing the service models (Infrastructure as a Service, Platform as a Service, and Software as a Service) as theoretical lenses and testing the COVID-19 as the moderator. The study has been conducted based on a deductive approach and adopted a stratified random sampling method. The sample consisted of 384 IT employees among those who had experienced working from home. The study utilized multiple regression and found that cloud computing service models significantly impact work from home with the moderating effect of COVID-19.
<p>Cloud computing is one of the most popular and sophisticated technologies adopted by organizations worldwide. Some world-leading organizations enhance their efficiency and effectiveness by using cloud computing technology. Working from home (WFH) has been a popular trend among organizations during the coronavirus (COVID-19) pandemic. The COVID-19 saw a breakthrough in work cultures and environments where working from home was a remarkable success in remote working environments, despite being a rare phenomenon in Sri Lanka. Yet, it is argued that the deployment of work from home has not been effective among Sri Lankan business organizations due to a lack of IT infrastructure, facilities, and knowledge. The purpose of the study is to investigate the impact of cloud computing, embracing the service models (Infrastructure as a Service, Platform as a Service, and Software as a Service) as theoretical lenses and testing the COVID-19 as the moderator. The study has been conducted based on a deductive approach and adopted a stratified random sampling method. The sample consisted of 384 IT employees among those who had experienced working from home. The study utilized multiple regression and found that cloud computing service models significantly impact work from home with the moderating effect of COVID-19.</p>
Facade - Popular architecture
Fachada de construcción tradicional combinando canto rodado de formato medio o mampuesto, ladrillo y adobe Source: Objaverse 1.0 / Sketchfab
The futures of 193 Popular Financial Reporting of Awards Program 2017 and social economic elements
<p>Dataset The futures of 193 Popular Financial Reporting of Awards Program 2017 and social economic elements</p>
Eulerian modelling of the three-dimensional distribution of seven popular microplastic types in the global ocean dataset
<p>Dataset for the paper "Eulerian modelling of the three-dimensional distribution of seven popular microplastic types in the global ocean" by A. S. Mountford and M. A. Morales Maqueda.</p> <p>ORCA2_5d_00010101_00011231_ptrc_T_con.nc<a href="https://zenodo.org/api/files/c6c9cbac-d5db-452a-aade-fd852db07351/ORCA2_5d_00010101_00011231_ptrc_T_con.nc"> </a> - control experiment (year 50)</p> <p>ORCA2_5d_00010101_00011231_ptrc_T_30m.nc - 30 m year<sup>-1</sup> piston velocity sensitivity experiment (year 50)</p> <p>ORCA2_5d_00010101_00011231_ptrc_T_90m.nc - 90 m year<sup>-1</sup> piston velocity sensitivity experiment (year 50)</p> <p>ORCA2_5d_00010101_00011231_ptrc_T_50.nc - neutrally buoyant sensitivity simulation (year 50)</p> <p>plastic_1_ts.nc - positively buoyant time series</p> <p>plastic_2_ts.nc - neutrally buoyant time series</p> <p>plastic_3_ts.nc - negatively buoyant time series</p> <p>plastic_input_ORCA2.nc - plastic input data file</p> <p>ORCA2_5d_00010101_00011231_grid_U.nc & ORCA2_5d_00010101_00011231_grid_V.nc - ocean velocity files</p>
Web tracking data for 500 websites popular among Finnish web users
<p>This dataset includes observations of trackers present on the top 500 pages popular among Finnish web users as per Alexa. The data collection was conducted using TrackerTracker in five separate requests for five subsets of 100 sites each between 19.8.2017 and 20.8.2017. The tool used a tracker database from March 24, 2017. More methodology details are described in the associated journal article <a href="https://doi.org/10.23978/inf.87841">https://doi.org/10.23978/inf.87841</a></p>
Novel Datasets for Evaluating Song Popularity Prediction Tasks
<p>We present two novel datasets for hit song prediction (HSP): HSP-S and HSP-L. They are substantially larger than currently available dataset with respect to the number of features contained.</p> <p>Both datasets provide high- and low-level audio features stemming from <a href="https://acousticbrainz.org/data">AcousticBrainz</a> and short representative MP3 samples. Further, we include listener- and play-counts gathered from <a href="https://www.last.fm">last.fm</a> for both datasets. The larger dataset, HSP-L, contains 73,482 songs with audio features, listener- and play-counts, making it substantially larger than previous datasets. In addition, we provide the release year information for 65,575 songs as provided by the <a href="https://millionsongdataset.com">Million Song Dataset</a>.<br> The smaller dataset, HSP-S, contains 7,736 songs, listener- and play-counts as well as <a href="https://www.officialcharts.com/charts/billboard-hot-100-chart/">Billboard Hot 100</a> data for 50% of the songs, and release year information is available for 7,449 songs.</p>
Wild kangaroos become more social when caring for young and may maintain long-term affiliations with popular individuals
<p>Kangaroos are an iconic group of Australian fauna. Despite considerable research on kangaroo behaviour, key gaps remain in our understanding of their social organization in the wild. In particular, it remains largely unknown whether kangaroos form long-term social bonds and what factors might prompt individuals to associate or dissociate from one another. Over 6 years, we monitored the social affiliations of individually identified eastern grey kangaroos, Macropus giganteus, in a large wild population. We investigated the short-term and long-term relationships of kangaroos and the extent those relationships varied with age, sex and reproductive state. We found evidence that long-term relationships among eastern grey kangaroos are possible, especially between adult females. Those individuals that were more sociable within years were also more likely to establish affiliations across years. Contrary to previous studies, we observed females actively associating with other mothers in the years in which they had young. These data suggest that the fission-fusion dynamics of eastern grey kangaroo social behaviour allow females to modulate their social position with conspecifics according to their current reproductive state. We highlight the adaptive implications of the formation of long-term bonds and the changes in social behaviour observed in females.</p>
Proof of Concept database with inputs and outputs of the Master thesis: Analyzing Software Delivery Performance behavior in popular Open Source Software Projects on a Release timeline basis through delivery metrics
<p>The software has become one of the main assets to deliver services today. Thus, software delivery has been dealing with a competitive and dynamic environment where the demand for faster and more assertive deliverables, called here Releases, only increases. Agile development methods emerged helping to accelerate software delivery, embracing industry and open source community. Since then, the software delivery frequency has expanded and improved bringing more adopters of rapid release cycles to reduce their time-to-market. However, using only rapid releases can not be enough as measuring software delivery can answer essential questions, like how software delivery is happening and how it should be. Some approaches for measuring software delivery appeared such as Software Delivery Performance (SDP) where software delivery is measured as a consequence of capabilities evolution. Popularity in Open Source Software Projects (OSSP) means that a project is mature enough in the community to fit the software demand and, therefore, is likely to be ready to be measured through a software delivery approach like SDP. In light of it, this work offers means to analyze SDP behavior in popular Open Source Software Projects on a Release timeline basis through delivery metrics. The results demonstrated that popularity is efficient filtering, as it improves the OSSP delivery, supporting the work's reliability and accuracy. The source code and methodology are published as a replication package to encourage reproducibility and future research.</p>
Wild kangaroos become more social when caring for young and may maintain long-term affiliations with popular individuals
Open the record for dataset details and reuse information.
Associated data files and scripts for 'Ancestry-inclusive dog genomics challenges popular breed stereotypes'
Open the record for dataset details and reuse information.
Co-occurrences of trending keywords in popular tech media (01.2016-12.2019)
<p>Sources with weights</p> <ul> <li>Euractiv 5%</li> <li>The Conversation 5%</li> <li>Politico Europe 5 %</li> <li>IEEE Spectrum 5 %</li> <li>Techforge 5%</li> <li>Fastcompany 5%</li> <li>The Guardian (Tech) 12%</li> <li>Arstechnica 5%</li> <li>Reuters 5%</li> <li>Gizmodo 9%</li> <li>ZDNet 9%</li> <li>The Register 12%</li> <li>The Verge 9%</li> <li>TechCrunch 9%</li> </ul> <p>Methodology</p> <ul> <li>Exploring the relationship between topics</li> <li>Pairs of terms which are mentioned together in media articles</li> <li>Most trending social issues and technologies have been selected (e.g. 'gdpr', '5G')</li> <li>The co-occurrence analysis is calculated for pairs consisting of emerging social issues and trending uni/bigrams</li> <li>The number of times the terms appear in articles together with a social issue is divided by the number of times the social issue is mentioned across all articles</li> <li>A single index is constructed for all word pairs by weighted average (taking into account the prevalence of the given source)</li> </ul>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.