Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
2
datasets available to search
ShareScore release 0.9.0
Dataset results
2 results for “Spanish manuscript”
Data underlying the manuscript: "Analysis of Research Data Sharing in Scientific Articles on Climate Change in the Covid-19 Year. The Spanish case 2020".
<p>This is the research data for the manuscript "Analysis of Research Data Sharing in Scientific Articles on Climate Change in the Covid-19 Year. The Spanish case 2020".<br>The following is the original abstract: Introduction: Sharing research data on climate change would facilitate the development of solutions to curb its impact, for this, data needs to be shared in an optimal way. General objective: To identify how many Spanish scientific articles on climate change published during 2020 share their research data in some way. Specific objectives: a) Identify the attributes of shared research data b) Describe the characteristics of the case studies found on how research data are shared. Methodology: Qualitative and descriptive study analyzing nine attributes: availability (1), accessibility (2), format (3), license (4), linkage (5), funding (6), editorial policy (7), content (8), statistics (9). Results: We analyzed 2212 articles were analyzed, 1867 (84%) articles had no associated research data. The remaining 16% have associated research data: 152 (7%) articles deposited their data in repositories, 42 (2%) submitted their data as supplementary material, 136 (6%) will share their data upon request to the author and 15 (1%) do not have publication permissions. Conclusions: Researchers are willing to share their research data, but under different conditions. Researchers who reused research data did not share the new data they generated. There is a lack of training among researchers on how to manage their research data. There is information on the web on this topic, but it is not just a matter of publishing manuals, but also of creating training spaces within universities, institutes and research centers to build a community of researchers committed to Open Science.</p>
Data for manuscript: "The Prevalence of Prejudice Denoting Terms in Spanish Newspapers"
<p>This data set contains frequency counts of target words in 5 million news and opinion articles from 3 popular newspapers in Spain: El País, El Mundo and ABC. The target words are listed in the associated manuscript and are mostly words that denote some type of prejudice. A few additional words not denoting prejudice are also available since they are used in the manuscript for illustration purposes.</p> <p>The textual content of news and opinion articles from the outlets listed in Figure 1 of the main manuscript is available in the outlet's online domains and/or public cache repositories such as Google cache (https://webcache.googleusercontent.com), The Internet Wayback Machine (https://archive.org/web/web.php), and Common Crawl (https://commoncrawl.org). We used derived word frequency counts from original sources. Textual content included in our analysis is circumscribed to articles headlines and main body of text of the articles and does not include other article elements such as figure captions.</p> <p>Targeted textual content was located in HTML raw data using outlet specific xpath expressions. Tokens were lowercased prior to estimating frequency counts. To prevent outlets with sparse text content for a year from distorting aggregate frequency counts, we only include outlet frequency counts from years for which there is at least 1 million words of article content from an outlet. </p> <p>Yearly frequency usage of a target word in an outlet in any given year was estimated by dividing the total number of occurrences of the target word in all articles of a given year by the number of all words in all articles of that year. This method of estimating frequency accounts for variable volume of total article output over time.</p> <p>The list of compressed files in this data set is listed next:</p> <p>-analysisScripts.rar contains the analysis scripts used in the main manuscript </p> <p>-targetWordsInArticlesCounts.rar contains counts of target words in outlets articles as well as total counts of words in articles</p> <p>Usage Notes</p> <p>In a small percentage of articles, outlet specific XPath expressions can fail to properly capture the content of the article due to the heterogeneity of HTML elements and CSS styling combinations with which articles text content is arranged in outlets online domains. As a result, the total and target word counts metrics for a small subset of articles might not be precise. </p> <p>To conclude, in a data analysis of millions of news articles, we cannot manually check the correctness of frequency counts for every single article and hundred percent accuracy at capturing articles’ content is elusive due to the small number of difficult to detect boundary cases such as incorrect HTML markup syntax in online domains. Overall however, we are confident that our frequency metrics are representative of word prevalence in print news media content (see Figure 2 of main manuscript for supporting evidence).</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.