Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
165
datasets available to search
ShareScore release 0.9.0
Dataset results
165 results for “political”
Participant Data and Code for "Human Detection of Political Speech Deepfakes across Transcripts, Audio, and Video"
<p>All code produced to analyze the participant response data and the participant response data itself are included in this repository. </p>
Number of sources at individual stages of the continuum between the apolitical and political nature of the police in the Polish People's Republic and the Third Polish Republic
<p>Bäcker Roman, <span>Number of sources at individual stages of the continuum between the apolitical and political nature of the police in the Polish People’s Republic and the Third Polish Republic</span></p> <p> </p> <p><span>This dataset was elaborated for the research project <em>Civil Disorder in Pandemic-ridden European Union</em>. The latter was financially supported by the National Science Centre, Poland [grant number 2021/43/B/HS5/00290].</span></p>
Replication Package for "Identity Politics"
<p><span><span>This replication package contains the codes and data needed to reproduce all Figures and Tables in the text and in the online appendix of the paper “Identity Politics”, by Nicola Gennaioli and Guido Tabellini</span></span></p>
Data for manuscript: "Using Word Embeddings to Probe Sentiment Associations of Politically Loaded Terms in News and Opinion Articles from News Media Outlets"
<p>This data set contains material for the purpose of scientific reproducibility of the accompanying manuscript "Using Word Embeddings to Probe Sentiment Associations of Politically Loaded Terms in News and Opinion Articles from News Media Outlets".</p> <p>Note that this data set is distributed with an Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0) License. NonCommercial means you may not use the material for commercial purposes. NoDerivatives means if you remix, transform, or build upon the material, you may not distribute the modified material. Attribution means you must give appropriate credit, provide a link to the license, and indicate if changes were made. You may do so in any reasonable manner, but not in any way that suggests the licensor endorses you or your use. See attached license terms for details.</p> <p>The work "Using Word Embeddings to Probe Sentiment Associations of Politically Loaded Terms in News and Opinion Articles from News Media Outlets" describes an analysis of political associations in 27 million diachronic (1975-2019) news and opinion articles from 47 news media outlets popular in the United States. We use embedding models trained on individual outlets content to quantify outlet-specific latent associations between positive/negative sentiment words and terms loaded with political connotations such as those describing political orientation, party affiliation, names of influential politicians and ideologically aligned public figures. </p> <p>News and opinion articles from the outlets listed in Figure 3 are available in the outlet's online domains and/or public cache repositories such as Google cache, The Internet Wayback Machine [31] and Common Crawl [32]. This work has not analyzed video or audio content of news media organizations, except when the outlet explicitly provides a transcript of such content in article form.<br> The temporal coverage of articles from different news outlets is not uniform. For most media organizations, news articles availability in their online domains or Internet cache backups becomes sparse as a function of articles’ age. This is not the case for some news outlets, where availability of news articles goes back to the 1970s. The Supplementary Material (SM) illustrates the time ranges of article data analyzed based on news outlets articles online availability.</p> <p>Textual content included in our analysis is circumscribed to the articles’ headlines and main text and does not include other article elements such as figure captions. Targeted textual content was located in HTML raw data using outlet specific XPath expressions. Tokens were lowercased prior to estimating embedding models. Markup language tags, URLs, nonalphanumeric characters, punctuation, digits, 330 common stop words and multiple spaces were removed prior to estimating word embeddings models.<br> All the analysis scripts and the diachronic word embedding models built from each of the 47 news media outlets analyzed in this work are available in this repository.</p> <p>For the purpose of reproducibility, we also provide in the above repository the articles’ text used to train the news outlets embedding models with the caveat that outlets articles not accessible without a subscription have been excluded. Also, for the included articles, stop words have been removed and the remaining words have been randomly scrambled within a sliding window of size 10 to render the articles incomprehensible to a human reader. These steps have been taken to not infringe articles copyright. These preprocessing steps have only minor impact on Continuous Bag of Words (CBOW) word2vec and the results reported in this work are similar when using the scrambled articles text to train outlet-specific embedding models.</p> <p>We derived outlet-specific word embedding models at every five-year time intervals within the 1975-2019 time range. The gensim [33] implementation of word2vec was used to train the embedding models. The continuous bag of words (CBOW) architecture performed slightly better than the Skip-Gram architecture in commonly used validation metrics so it was used for all subsequent analysis. </p> <p>For training the word embedding models, the following parameters were used: vector dimensions=300, window size=10, negative sampling=10, down sampling frequent words = 0.0001, minimum frequency count of 5 (only terms that appear more than 5 times in the corpus were included into the word embedding model vocabulary), number of training iterations (epochs) through the corpus=5. The exponent used to shape the negative sampling distribution was the default 0.75. </p> <p>Outlet-specific embedding models performance across a range of commonly used semantic, syntactic and analogy tasks was similar to popular pre-trained embedding models trained on corpora such as Twitter or Google books on similarity, association and word analogy tasks, see Supplemeentary Material of the manuscript for detailed validation tests results.</p> <p> </p>
Replication package for: Forbidden Fruits: The Political Economy of Science, Religion, and Growth
<p>The zipped folder contains all the files need to replicate the results reported in the paper “Forbidden Fruits: The Political Economy of Science, Religion, and Growth” by Roland Bénabou, Davide Ticchi, and Andrea Vindigni, published in the Review of Economic Studies.</p>
Appendix B in R. Turunen (2021). Shades of Red: Evolution of the Political Language of Finnish Socialism from the Nineteenth Century until the Civil War of 1918. The Finnish Society for Labour History.
<p>This is Appendix B in R. Turunen (2021). <em>Shades of Red: Evolution of the Political Language of Finnish Socialism from the Nineteenth Century until the Civil War of 1918</em>. The Finnish Society for Labour History. The XLSX file contains words that appear more frequently than statistically expected in the socialist newspapers compared to the non-socialist newspapers. Keyness value was calculated using the log-likelihood (4-term) test. The dataset of socialist newspapers consists of the following lemmatized raw text files:<strong> </strong><em>Työmies</em> 1895–1917, <em>Kansan Lehti</em> 1898–1917, <em>Vapaa Sana</em> 1906–1917 and <em>Savon Työmies</em> 1906–1910. The dataset of non-socialist newspapers consists of the following lemmatized raw text files: <em>Uusi Suometar</em> 1895–1917, <em>Päivälehti </em>/ <em>Helsingin Sanomat</em> 1895–1917, <em>Aamulehti</em> 1895–1917, <em>Tampereen Sanomat</em> 1904–1917, <em>Pohjan Poika</em> 1906–1909, <em>Pohjanmaa</em> 1909–1912, <em>Pohjalainen</em> 1911–1915, <em>Savotar </em>1906–1910 and <em>Otava </em>1904–1909. Available from https://digi.kansalliskirjasto.fi/opendata. The analysis is described more carefully in R. Turunen (2021). <em>Shades of Red: Evolution of the Political Language of Finnish Socialism from the Nineteenth Century until the Civil War of 1918. </em>The Finnish Society for Labour History.</p>
Handwritten newspapers used in Turunen, R. (2021). Shades of Red: Evolution of the Political Language of Finnish Socialism from the Nineteenth Century until the Civil War of 1918
<p>This dataset contains XLSX files of five different handwritten newspapers: <em>Kuritus </em>(1909–1911), <em>Palveliatar</em> (1907–1913, 1917), <em>Yritys</em> (1915–1917), <em>Tehtaalainen</em> (1908–1914, 1917) and <em>Nuija </em>(1899–1903, 1907–1909, 1912, 1914–1915).</p> <p>The original sources and the coding scheme for the XLSX files are described in:</p> <p>Turunen, R. (2021). <em>Shades of Red: Evolution of the Political Language of Finnish Socialism from the Nineteenth Century until the Civil War of 1918</em>. The Finnish Society for Labour History.</p>
Research Data Set for PhD thesis "Political Expression in Web Defacements"
<p>The software used for collection and processing is available at <a href="https://github.com/mkrzmr/Political-Expression-in-Web-Defacements-Crawler">https://github.com/mkrzmr/Political-Expression-in-Web-Defacements-Crawler</a></p>
ChatGPT - Questions and Answers to the Political Coordinates Test in English, French, Italian and German
<p>This dataset contains the questions and answers resulting from the administration of the policy coordination test in English, French, Italian and German to ChatGPT. The version used was GPT-4.</p>
Political Participation and Open Data Maturity
<p>To discover a correlation between open data maturity and political participation of European Union countries data about the Open data Maturity, parliamentary voter turnout (%) and individuals using the internet for taking part in online consultations or voting (%) were used.</p>
Political Participation and Digital Maturity
<p>To discover a correlation between information society and political participation of European Union countries data about the DESI index (Digital Economy and Society Index), parliamentary voter turnout (%) and individuals using the internet for taking part in online consultations or voting (%) were used.</p>
The Correlation Between Psychological Capital, Intention to Leave, Work Job and Environment Satisfaction, and Political Skills Among Different Generations Nurses
ClinicalTrials.gov study NCT06495060. IPD Sharing: NO. Countries: 1. Publications: 24.
Scaling laws of political regime dynamics: Stability of democracies and autocracies in the 20th-century
Open the record for dataset details and reuse information.
Data from: Finding politically feasible conservation policies: the case of wildlife trafficking
Open the record for dataset details and reuse information.
The political and institutional challenges facing the crisis of water resources and water and sanitation services in Chile (in Spanish)
<p>Talk (in Spanish) with Alex Caldera, University of Guanajuato, Mexico, and Jose Esteban Castro, National Council of Scientific and Technical Research (CONICET), Argentina and Newcastle University, UK. In this issue we feature Hugo Maturana Aguilar, President of the Workers Union of Valparaiso’s Company of Sanitary Works (ESVAL S.A.), and President of the Executive Directorship, National Federation of Workers of Sanitary Works (FENATRAOS), Chile, who talks about the challenges and opportunities facing water resources and water and sanitation services in Chile. The talk also focuses on the crisis triggered in the city of Osorno in July 2019, when the city was left without water supply for over two weeks by failures of the privatized Los Lagos Company of Sanitary Services (ESSAL), owned by the French multinational Suez.</p> <p>This podcast is part of the Voices for Water Series: http://waterlat.org/public-engagement/podcasts/.</p> <p>The podcast is part of the activities taking place within the Cooperation Agreement between the WATERLAT-GOBACIT Network and the Public Services International (PSI) (http://waterlat.org/projects/cooperation-agreement-with-the-psi/ ), and the collaboration with the Confederation of Workers of Water, Sanitation, and Environment Services of the Americas (CONTAGUAS) (http://www.contaguas.org/).</p> <p> </p> <p> </p>
Data from: Processing political misinformation: comprehending the Trump phenomenon
This study investigated the cognitive processing of true and false political information. Specifically, it examined the impact of source credibility on the assessment of veracity when information comes from a polarizing source (Experiment 1), and effectiveness of explanations when they come from one's own political party or an opposition party (Experiment 2). These experiments were conducted prior to the 2016 Presidential election. Participants rated their belief in factual and incorrect statements that President Trump made on the campaign trail; facts were subsequently affirmed and misinformation retracted. Participants then re-rated their belief immediately or after a delay. Experiment 1 found that (i) if information was attributed to Trump, Republican supporters of Trump believed it more than if it was presented without attribution, whereas the opposite was true for Democrats and (ii) although Trump supporters reduced their belief in misinformation items following a correction, they did not change their voting preferences. Experiment 2 revealed that the explanation's source had relatively little impact, and belief updating was more influenced by perceived credibility of the individual initially purporting the information. These findings suggest that people use political figures as a heuristic to guide evaluation of what is true or false, yet do not necessarily insist on veracity as a prerequisite for supporting political candidates.
Data from: Association between the dopamine D4 receptor gene exon III variable number of tandem repeats and political attitudes in female Han Chinese
Twin and family studies suggest that political attitudes are partially determined by an individual's genotype. The dopamine D4 receptor gene (DRD4) exon III repeat region that has been extensively studied in connection with human behaviour, is a plausible candidate to contribute to individual differences in political attitudes. A first United States study provisionally identified this gene with political attitude along a liberal–conservative axis albeit contingent upon number of friends. In a large sample of 1771 Han Chinese university students in Singapore, we observed a significant main effect of association between the DRD4 exon III variable number of tandem repeats and political attitude. Subjects with two copies of the 4-repeat allele (4R/4R) were significantly more conservative. Our results provided evidence for a role of the DRD4 gene variants in contributing to individual differences in political attitude particularly in females and more generally suggested that associations between individual genes, and neurochemical pathways, contributing to traits relevant to the social sciences can be provisionally identified.
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases
Open the record for dataset details and reuse information.
Multimodal Fallacy Classification in Political Debates
<p>MM-USED-fallacy supplementary files</p>
POLITICAL RELATIONS AND STRUGGLES BETWEEN IRAN AND TRANSOXIANA DURING THE REIGN OF SHAH TAHMASP
Open the record for dataset details and reuse information.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.