Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
3,983
datasets available to search
ShareScore release 0.7.1
Dataset results
3,983 results for “Persons”
Processed metabolomic data from the EXPOsOMICS Personal Exposure Monitoring study
<p>Metabolomic data from the 'Variability of the Human Serum Metabolome over 3 Months in the EXPOsOMICS Personal Exposure Monitoring Study' paper <a href="https://doi.org/10.1021/acs.est.3c03233">DOI: 10.1021/acs.est.3c03233</a> . </p> <p>The data was originally collected and generated by the multicenter EXPOsOMICS Personal Exposure Monitoring study. Details on data collection and processing are described in the aforementioned paper. The statistical analysis from that paper is available at <a href="https://github.com/moosterwegel/variability-metabolites-paper">https://github.com/moosterwegel/variability-metabolites-paper</a> and may contain useful information/code to work with this data.</p> <p>`processed_covariate_data.csv`:<br> ```<br> Rows: 298<br> Columns: 7<br> $ subjectid: hashed identifier subject<br> $ sample_code: indicates if it's the first (A) or second (B) blood sample<br> $ centre: indicates in which centre the data was collected<br> $ age_cat: indicates age category at the time of a PEM session<br> $ sq_sex: indicates the sex of the participant (male, female) as filled in during the screening questionaire<br> $ traf: indicates the exposure to traffic (PM2.5 and UFP) as measured during the PEM sessions. <br> $ bmi_cat: indicates BMI category at the time of a PEM session<br> ```</p> <p>`processed_lcms_data data.csv` contains the processed LCMS data:<br> ```<br> Rows: 298<br> Columns: 4297<br> $ subjectid: hashed identifier subject<br> $ sample_code: indicates if it's the first (A) or second (B) blood sample<br> $ centre: indicates in which centre the data was collected<br> $ compounds: measured features (compounds) are prefixed by the letter X. The name contains information on the measured monoisotopicmass_retentiontime.<br> Non-detects (below limit of detection (LOD) are coded as 1 for the compounds.<br> ....<br> ```<br> In the datasets each row indicates a measurement on a day (`sample_code`) and person (`subjectid`). The datasets can be joined on these variables.</p> <p>The other data files (`annotations.xslx`, `ancestors_annotations.xlsx`, `annotations_plus_kegg_pathways.csv`) contain the annotations, ancestors of the annotations (to assign a class to a compound based on ChEBI ontology, see our paper for details), annotations plus KEGG pathways respectively. </p>
Agreeableness personality trait and social information encoding
Open the record for dataset details and reuse information.
GenBank accession numbers of the four marker genes and associated voucher specimens/tissues that were used in this study. For more details see Guo et al. (2014). Sequences of species in bold are unpublished and were provided by P. Guo as personal communication in Rediscovery of Andrea's keelback, Hebius andreae (Ziegler & Le, 2006): First country record for Laos and phylogenetic placement
GenBank accession numbers of the four marker genes and associated voucher specimens/tissues that were used in this study. For more details see Guo et al. (2014). Sequences of species in bold are unpublished and were provided by P. Guo as personal communication
Brazilian missing persons
<p>10.499 registers of Brazilian missing persons collected in 2018.</p>
In silico prediction of ARB resistance: A first step in creating personalized ARB therapy
<p><strong>AT1R Model preparation</strong><br> The crystal structure of human AT1R bound to olmesartan (PDB: 4ZUD) was downloaded from the RCSB Protein Data Bank. 4ZUD contains apocytochrome b562RIL fused to the amino terminus, and many of the flexible regions, as well as helix 8, are not resolved. In order to generate an appropriate starting structure, olmesartan and the apocytochrome b562RIL fusion were removed from 4ZUD, and the missing regions were added to the protein with MOE software (Chemical Computing Group ULC, Montreal, Canada). Specifically, the N-Terminus (residues 1 to 25), intracellular loop 2 (residues 134 to 140), extracellular loop 2 (residues 186 to 188), intracellular loop 3 (residues 223 to 234), and helix 8 (residues 305 to 316) were added to the AT1R in accordance to the human AT1R sequence and PDB:4YAY. The remaining carboxyl-tail of the AT1R (residues 317 to 359) was not modeled. The AT1R model then underwent an energy minimization within MOE using the Amber10:Extended Huckel Theory (EHT) force field.</p> <p><strong>Molecular dynamic (MD) simulations and analysis</strong><br> The MOE minimized AT1R was loaded into CHARMM-GUI. An 80 Å by 80 Å lipid bi-layer composed of 13% cholesterol and 87% Phosphatidylcholine (POPC) was generated around the receptor. Water was packed 17.5 Å above and below the lipid bi-layer, and 150 mM Na+ and Cl- ions were added to the system via Monte-Carlo ion placing. The all-atom CHARMM C36 force field for proteins and ions, and the CHARMM TIP3P force field for water were selected. A hard non-bonded cutoff of 8.0 angstroms was utilized. All molecular dynamics simulations were performed using the PMEMD module of the AMBER16 package with support for MPI multi-process control and GPU acceleration code. Orthorhombic periodic boundary conditions with a constant pressure of 1 atm was set via the NPT ensemble and temperature was set to 310.15°K (37°C) using Langevin dynamics. The SHAKE algorithm was used to constrain bonds containing hydrogens. The dynamics were propagated using Langevin dynamics with Langevin damping coefficient of 1 ps-1 and a time step of 2 fs. Before the production run, the AT1R model was minimized for 5000 steps using the steepest descent method and then equilibrated for 600 ps. The protein coordinates were saved in 10 ps intervals. The production run lasted 150 ns, at which point all three replicas were stable for at least the last 20 ns.</p>
Datasets from the RecSys 2020 article "Carousel Personalization in Music Streaming Apps with Contextual Bandits"
<p>We publicly release the anonymized <em>user_features.csv</em> and <em>playlist_features.csv</em> datasets, from the music streaming platform Deezer, as described in the article "<em>Carousel Personalization in Music Streaming Apps with Contextual Bandits"</em> published in the proceedings of the 14th ACM Conference on Recommender Systems (<em>RecSys 2020</em>). The paper is available <a href="https://arxiv.org/abs/2009.06546">here</a>.</p> <p>These datasets are used in the GitHub repository <a href="https://github.com/deezer/carousel_bandits">deezer/carousel_bandits</a> to reproduce experiments from the article.</p> <p>Please cite our paper if you use our code or data in your work.</p>
Data for Are Changes in Alcohol Use and Personality Traits associated? A Cohort Study among Young Swiss Men
<p>These are the data and metadata for the article </p> <p><strong>Are Changes in Alcohol Use and Personality Traits associated? A Cohort Study among Young Swiss Men</strong></p> <p>by </p> <p><strong>Gerhard Gmel, Simon Marmet, Joseph Studer, and Matthias Wicki</strong></p> <p><strong>to be published in Frontiers of Psychiatry</strong></p>
Refined personal name data from the census book of Vodskaja pjatina
<p>The data contains approximately 36,000 personal names derived from medieval Russian documentation. More preciously, names are collected from an edited version of the census book of Vodskaja pjatina, which was one of the five administrative areas in the late 15<sup>th</sup> century Novgorod.</p> <p>Editions were compiled in parts and the first two, which cover the northernmost region, are called <em>Переписная окладная книга по новугороду вотской пятины</em> (1851, 1852)(POKV I‒II). The third part of the book series <em>Новгородские пистсовые книги</em> (1868)(NPK III) covers the southern and western parts of the study area.</p> <p>The process of obtaining the personal from the inscription has been following: First, editions of the census book were obtained as scanned PDF files. These were transformed as editable copies by using OCR (=Optical Character Recognition) software Abbyy. The program read the original mid-19<sup>th</sup> century Russian text adequately with its old Russian alphabet package.</p> <p>After the initial corrections, a Python script was written to harvest the personal names. This was based on exploiting the systematic formalities in how most of the names were presented in the census book. The script looked for abbreviations “дв.” and “д.” and extracted all following capitalized words until section end markers “.”, “;” or “:”. As an output, a name to pogost matrix was produced, which held the raw frequencies of each word in each pogost.</p> <p>The process of cleaning the name data, in turn, has been done mostly by data wrangling program OpenRefine in following manner: For starters, all name forms shorter than four characters were removed as there were no personal names consisting of three or less letters. Furthermore, nouns that were not names were removed. This meant discarding expressions that described person’s special feature or profession, like such as being a widow (“вдова”) or working as a deacon (“діакъ”). For some reason, editors followed inconsistent conventions in capitalizing these non-name nouns.</p> <p>In addition, some orthographical and morphological harmonization was done on the data. The letter <em>ы </em>was cut from the end of bynames, where it denotes plurality. Similarity of so called soft and hard signs, <em>ь </em>and <em>ъ</em> caused some problems. As the latter one is not used in contemporary Russian and was not used in the original documents either (Неволин 1853 : 4 (in Appendix 1)) it was removed. The soft sign <em>ь </em>was also removed because it was absent in the original documents and it had been used inconsistently by the editors. The letter <em>ѣ</em> (yat) is rarely used in personal names but nevertheless, it was changed to <em>е </em>(like as it is in contemporary Russian) as since it was often confused with soft and hard signs (<em>ь </em>and <em>ъ</em>). Furthermore, the letter <em>ѳ </em>(fita) was often erroneously recognized as <em>о </em>or <em>е. </em>As it is only found in NPK III and only in the beginning of certain names, which all are also written with “Ф” (e.g. “Ѳедко” vs. “Федко”), it was replaced with <em>Ф</em>.</p> <p>In the second phase most of the erroneous orthographies were corrected. We do not detail herescribe all the OCR-errors here that were found, but in the following a short description is given of the most significant corrections. There were, for example, many letters whose similarity caused problems for the OCR-program (e.g. <em>и </em>/ <em>й </em>and <em>б </em>/ <em>в</em>). In these cases, the correct orthography was sought in the census book editions and accordingly, Openrefine was used to change erroneous forms to right correct ones.</p> <p>After the corrections were made, the number of name types (= name variants) was reduced from 4942 to 2748. The Overall overall number of name tokens was dropped as well: from 36,405 to 35,726. Of the name types, more than half (1484) have only one occurrence.</p> <p>The refined and harmonized data is published as pogost-by-name frequency tabulations (<em>pogost,</em> equivalent of English <em>parish</em>). The file is in tab-delimited file (.tsv) format.</p> <p>References:</p> <p>Неволин, К. А. 1853, О пятинах и погостах новгородских в XVI веке, с приложением карты, Санкт-Петербург (Из Записок Императорского русского географического общества, Кн. VIII).</p> <p>NPK III = Новгородские писцовые книги, Т. 3 : Переписная оброчная книга Вотской пятины, 1500 года, 1868, 1868, Санкт Петербург.</p> <p>POKV I, II = Переписная окладная книга по Новугороду Вотьской пятины, 1851, 1852, Имп. Моск. о-во истории и древностей рос., Москва.</p> <p> </p>
Can a Wi-Fi WLAN Support a First Person Shooter?
<p>Jose Saldana, Juan Luis de la Cruz, Luis Sequeira, Julian Fernandez-Navajas, Jose Ruiz-Mas, "Can a Wi-Fi WLAN Support a First Person Shooter?," NetGames 2015, The 14th International Workshop on Network and Systems Support for Games Zagreb, Croatia, December 3-4, 2015.<br /> ISBN: 978-1-5090-0067-8</p> <p>This work has been partially ?nanced by the EU H2020 Wi-5 project (Grant Agreement no: 644262), and European Social Fund in collaboration with the Government of Aragon.</p> <p><br /> - The file "netgames_2015_in_proc.pdf" contains the paper published in the proceedings of the conference.</p> <p><br /> - Support files and results:</p> <p><br /> "captures" directory contains the ".pcap" captures made for both tests using Wireshark.</p> <p><br /> "files" directory contains the filtered data from the captures or from the D-ITG raw output. The ".log" files are binary. They have been obtained with D-ITG (Distributed Internet Traffic Generator, http://traffic.comics.unina.it/software/ITG/)</p> <p>A. Botta, A. Dainotti, A. Pescapè, "A tool for the generation of realistic network workload for emerging networking scenarios", Computer Networks (Elsevier), 2012, Volume 56, Issue 15, pp 3531-3547. </p> <p> </p> <p>"src" directory contains the ".m" matlab script and functions used for processing the raw data obtained. MATLAB R2013 compatible. It also includes the ".fig" files, to be opened with the same version of MATLAB.</p> <p>NOTE: It is MANDATORY adding the directories to the MATLAB path.</p> <p>Juan Luis de la Cruz, September 2015</p> <p> </p> <p><em>Abstract</em>—In corporate and commercial environments, the deployment of a set of coordinated Wi-Fi APs is becoming a common solution to provide Internet coverage to moving users. In these scenarios, real-time services as online games can also be present. This paper presents a set of experiments developed in a test scenario where an end device moves between different APs while generating game traffic. A WLAN solution based on virtual APs is used, in order to make the handoffs transparent for Layer 3. The results show that it is possible to maintain an acceptable level of subjective quality during the handoff. At the same time, it is set clear that the fact of having a gamer in an AP could be taken into account by radio resource management algorithms, in order to provide a better quality.</p>
Assessing the typology of person portmanteaus (supplementary material)
<p>Supplementary material for 'Assessing the typology of person portmanteaus' (doi:10.1007/s11525-017-9305-z)</p> <p> </p>
Persons and Names of the Middle Kingdom
<p>The database "Persons and Names of the Middle Kingdom and early New Kingdom" (PNM) is developed as part of the projects "Umformung und Variabilität im Korpus altägyptischer Personennamen 2055–1550 v. Chr." and “Altägyptische Titel in amtlichen und familiären Kontexten, 2055-1352 v. Chr.”. The database includes data on Egyptian Middle Kingdom and early New Kingdom personal names, people, written sources, titles, and dossiers of persons attested in various sources.</p> <p>The online version of the database is published at <a href="https://pnm.uni-mainz.de/info">https://pnm.uni-mainz.de/info</a>. The source code of the web-interface is available at <a href="https://doi.org/10.5281/zenodo.1418714">https://doi.org/10.5281/zenodo.1418714.</a></p> <p>Version 6 covers the timeframe of the Middle Kingdom and the early New Kingdom (from Mentuhotep II to Amenhotep III). It comparison to version 5, it adds 249 persons to the dataset. </p> <p>The release includes the dump of the MySQL database (tested with MariaDB 10.5.29), the description of the database structure, the graphic representations of spellings, created with JSesh, and RDF representations of the dataset in the Turtle and RDF/XML notations. Additionally it includes the ontology developed for this dataset (<a href="https://pnm.uni-mainz.de/ontology/">https://pnm.uni-mainz.de/ontology/</a>) and R2RML mappings for transforming the MySQL database to RDF. A live SPARQL interface to the RDF version of the dataset can be found at <a href="https://pnm.uni-mainz.de/sparql">https://pnm.uni-mainz.de/sparql</a>.</p>
Variation in personality shaped by evolutionary history, genotype, and developmental plasticity in response to feeding modalities in the Arctic charr
<p>Animal personality has been shown to be influenced by both genetic and environmental factors and shaped by natural selection. Currently, little is known about mechanisms influencing the development of personality traits. This study examines the extent to which personality development is genetically influenced and/or environmentally responsive (plastic). We also investigated the role of evolutionary history, assessing whether personality traits could be canalized along a genetic and ecological divergence gradient. We tested the plastic potential of boldness in juveniles of five Icelandic Arctic charr morphs (<em>Salvelinus</em> <em>alpinus</em>), including two pairs of sympatric morphs, displaying various degrees of genetic and ecological divergence from the ancestral anadromous charr, split between treatments mimicking benthic vs. pelagic feeding modalities. We show that differences in mean boldness are mostly affected by genetics. While the benthic treatment led to bolder individuals overall, the environmental effect was rather weak, suggesting that boldness lies under strong genetic influence with reduced plastic potential. Finally, we found hints of differences by morphs in boldness canalization through reduced variance and plasticity, and higher consistency in boldness within morphs. These findings provide new insights into how behavioural development may impact adaptive diversification.</p>
Handling of Personal Data by Smart Home Equipment: an Exploratory Analysis in the Context of LGPD
<p>This dataset provides data about an exploratory research that analyzed the Privacy and Security Policies and the Instruction Manuals of 59 home automation equipment for Smart Home in order to verify which personal data was handled and how these documents were providing information about processes performed in personal data. The analysis was conducted with a quantitative approach followed by a qualitative analysis, using content analysis.</p>
Life cycle inventory database for consumption in Quebec - Personal hygiene
<p>These inventory datasets are essential for calculating the environmental impacts of an individual’s consumption in Quebec.</p> <p>Led by the CIRAIG, in collaboration with ESG-UQAM, this project aims to develop an inventory database of the life cycle of consumption in Quebec. These inventory datasets are essential for calculating the carbon footprint of an individual’s consumption in Quebec. The inventory is developed with a life cycle approach. Ultimately, it allows for evaluating carbon footprints at every step of the consumption life cycle (extraction of primary sources, transformation, transport, use of goods and services, end of life). The inventory is developed in a modular fashion for the different areas of individual consumption as Food; Transport; Housing; Clothing; Travel; Communications; Entertainment and Culture; Financial and Administrative Management; Health, Hygiene, and Beauty. These areas are developed and detailed as a priority, as they contribute most to an individual’s carbon footprint in Quebec. Other non-priority areas are roughly modelled in order to provide a complete (but more uncertain) portrait of individual consumption. The project is underway and the deliverables will be made available online as things progress. It is not, however, an objective of the project to create a carbon footprint calculation tool at the moment.</p> <p>https://ciraig.org/index.php/project/life-cycle-inventory-database-for-consumption-in-quebec/ </p>
What a "minor" event looks like to a marginalised person vs a privileged person
<p>alt-text:</p> <p> </p> <p>Left graph shows a graph that shows how trauma accumulates with time like steps that keep going up and is very high. Text underneath says: The left figure shows how I react to a “minor” event that triggers deep emotional reactions based on previous trauma that accumulates over time.</p> <p>Right graph no accumulation of trauma and thre are no steps at all and only shows the small impact (a single step) an outsider sees that is very low. Text underneath says: The right figure is how a person with privilege might view the same “minor” event and judge my reaction as an “overreaction”.</p>
Person Detection on Construction Sites
<p>All information can be found here:</p> <p><a href="https://github.com/aidresden/person_detection_construction_sites" target="_blank" rel="noopener">https://github.com/aidresden/person_detection_construction_sites</a></p> <p>The development of this data set was funded by the Federal Ministry of Labour and Social Affairs (BMAS) and the Federal Institute for Occupational Safety and Health (BAuA) under the administrative agreement ‘Artificial Intelligence in a Safe and Healthy Working Environment’.</p>
A personalized value-based justification in food swaps to stimulate healthy online food choices
<p>This study examined the effect of a personalized value-based justification in explaining the rationale behind healthy food swaps. Additionally, consumers' willingness to share their personal information with retailers to personalize swap recommendations is explored.</p>
Replication Package of: "From Anecdote to Evidence: The Relationship Between Personality and Need for Cognition"
<p>Several anecdotes suggests that software engineers enjoy engaging in solving puzzles and other cognitive efforts. This tendency to engage in and enjoy effortful thinking is referred to as a person's 'need for cognition.' An open question is, however, whether developers differ from the general population in their scores of need for cognition. To address this question, we conducted a large-scale sample study of 483 software engineers. Personality plays a significant role in people's behavior and is stable over time, and is therefore considered a defining characteristic of individuals. We analyzed the data using multiple Bayesian linear regression analyses. The results indicate that ca. 33% of variation in developers' need for cognition can be explained by personality traits. Given the importance of human factors for software developer performance in general, and problem solving skills in particular, answering this question has substantial implications, such as recruitment & retention, working behaviour, and teaming.</p>
Group composition of individual personalities alters social network structure in experimental populations of forked fungus beetles
<p><span>Social network structure is a critical group character that mediates the flow of information, pathogens, and resources among individuals in a population, yet little is known about what shapes social structures. In this study, we experimentally tested whether social network structure depends on the personalities of group members. Replicate groups of forked fungus beetles (<i>Bolitotherus cornutus</i>) were engineered to include only members previously assessed as either more social or less social. We found that individuals behaved consistently across social contexts, exhibiting repeatable numbers of interactions and numbers of partners. At the group level, networks composed of more social individuals had higher interaction rates, higher tie density, higher global clustering, and shorter average shortest paths than those composed of less social individuals. We highlight group composition of personalities as a source of variance in group traits and a potential mechanism by which networks could evolve.</span></p>
Dargwa: Person
<p>This lecture is part of the lecture series: Glottothèque: Languages of the Anatolia, Caucasus, Iran, Mesopotamia; grammatical snippets online (electronic resource). Bamberg, Cambridge, Göttingen, Moscow, Nicosia, Paris: LACIM network, at https://spw.unigoettingen.de/projects/lacim/, edited by Christiane Bulut, Anaïd Donabédian-Demopoulos, Geoffrey Haig, Geoffrey Khan, Pollet Samvelian, Stavros Skopeteas, Nina Sumbatova.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.