Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
676
datasets available to search
ShareScore release 0.9.0
Dataset results
676 results for “Manis”
Raw data to "Series expansions in closed and open quantum many-body systems with multiple quasiparticle types"
<p>This collection of data is complementary to the publication "Series expansions in closed and open quantum many-body systems with multiple quasiparticle types", Lea Lenke, Andreas Schellenberger, Kai Phillip Schmidt, <a href="https://arxiv.org/abs/2302.01000">arXiv:2302.01000</a> (<a href="https://arxiv.org/abs/2302.01000">https://arxiv.org/abs/2302.01000</a>).</p> <p>It contains all data used for Figure 2 given in the file `Figure_2_complementary_data.yaml` and all needed data to recalculate the energies of the visualized modes in the files `Figure_2_coefficients_expectation_values.yaml` and `Figure_2_broad_signum_coefficients_expectation_values.yaml`.</p> <p>For the last two files, we used a program to calculate the coefficients. The source code for coefficient calculation is openly available under GitHub (<a href="https://github.com/FAU-kpslab/pcstpp_CoefficientGenerator">https://github.com/FAU-kpslab/pcstpp_CoefficientGenerator</a>) including configuration files to reproduce the coefficients given here.</p> <p>All files are self-consistent, for further information we recommend the comments directly in the files.</p> <p>For further details on the used method pcst<sup>++ </sup>and discussion of the results we refer to the linked publication.</p> <p>If any question may arise, you are highly welcome to contact us (see e.g. contact information on the publication).</p>
Data relating to Chiedozie et al. How many medications do doctors in primary care use? An observational study of the DU90% indicator in primary care in England.
<p>Data in .csv format relating to the paper Chiedozie et al. 2020 "How many medications do doctors in primary care use? An observational study of the DU90% indicator in primary care in England." Also contains eTables 5-8 in Excel format, and Stata do file for deriving the DU90% indicator.</p>
Random-Phase Approximation in Many-Body Noncovalent Systems: Methane in a Dodecahedral Water Cage
<p>Supplementary information and raw data for Random-Phase Approximation in Many-Body Noncovalent Systems: Methane in a Dodecahedral Water Cage. </p>
Data supporting the publication "Many-body quantum sign structures as non-glassy Ising models"
<p>This repository contains all raw data that were used to draw conclusions and generate figures for the paper:</p> <p><strong>"Many-body quantum sign structures as non-glassy Ising models"</strong><br> by Westerhout, T., Katsnelson, M. I., & Bagrov, A. A.</p> <p><em>Abstract:</em> The non-trivial phase structure of the eigenstates of many-body quantum systems severely limits the applicability of quantum Monte Carlo, variational, and machine learning methods. Here, we study real-valued signful ground-state wave functions of frustrated quantum spin systems and, assuming that the tasks of finding wave function amplitudes and signs can be separated, show that the signs can be easily bootstrapped from the amplitudes. We map the problem of finding the sign structure to an auxiliary classical Ising model defined on a subset of the Hilbert space basis. We show that the Ising model does not exhibit significant frustrations even for highly frustrated parental quantum systems, and is solvable with a fully deterministic O(K log K)-time combinatorial algorithm (where K is the Ising model size). Given the ground state amplitudes, we reconstruct the signs of the ground states of several frustrated quantum models, thereby revealing the hidden simplicity of many-body sign structures.</p>
Data package for modeling the journey of Colonel William Leake in the southern Mani Peninsula, Greece, using least-cost analysis
<p>Data used to model Colonel William Leake's journey in the southern Mani Peninsula, Greece, in the year 1805. Leake's journey is described in the book, <em>Travels in the Morea: Volume I </em>(Leake 1830, pp. 233-321). The data may be used to calculate least-cost paths between the places where Leake stopped, taking into consideration the contemporary path network and calculating cost in time based on Tobler's hiking function and the Modified Tobler function. A paper interpreting these data, 'Reconstructing Historical Journeys with Least-Cost Analysis: Colonel William Leake in the Mani Peninsula, Greece,' is published in <em>Journal of Archaeological Science: Reports</em> and can be accessed here: <a href="http://doi.org/10.1016/j.jasrep.2019.01.014">https://doi.org/10.1016/j.jasrep.2019.01.014</a>. The article pre-print can be accessed here: <a href="https://works.bepress.com/rebecca-seifried/11/">https://works.bepress.com/rebecca-seifried/11/</a>.</p> <p>Dr. Rebecca M. Seifried mapped the pre-modern paths as part of a PhD dissertation completed in 2016 through the Department of Anthropology at the University of Illinois at Chicago, entitled 'Community Organization and Imperial Expansion in a Rural Landscape: The Mani Peninsula, Greece (AD 1000-1821)' (<a href="http://hdl.handle.net/10027/21274">https://hdl.handle.net/10027/21274</a>). Fieldwork was conducted in 2014 and 2016 under the auspices of the 5th Ephorate of Byzantine Antiquities in Sparta and in collaboration with the Diros Project, an archaeological survey and excavation co-directed by Dr. Giorgos Papathanassopoulos and Dr. Anastasia Papathanasiou through the Ephorate of Palaeoanthropology & Speleology of Southern Greece. The remaining datasets were created in collaboration with Dr. Chelsea A.M. Gardner as part of the 'CART-ography Project: Cataloguing Ancient Routes and Travels in the Mani Peninsula,' whose goal is to catalogue the historic accounts of travelers to Mani and to model their routes throughout the peninsula.</p> <p>This research was funded by the National Science Foundation (BCS-1346694), Marie Sklodowska-Curie Actions (H2020-MSCA-IF-2016 750843), the DigitalGlobe Foundation, the National Cadastre and Mapping Agency, SA (Ktimatologio), ArchaeoLandscapes Europe, the University of Illinois at Chicago, the Society of Women Geographers, the Archaeological Institute of America, and Mount Allison University.</p>
Image-based Many-language Programming Language Identification - Replication Package
<p>This dataset contains the data, software, and instructions needed to replicate the findings of the paper:</p> <p>Francesca Del Bonifro, Maurizio Gabbrielli, Antonio Lategano, and Stefano Zacchiroli. Image-based Many-language<br> Programming Language Identification. <a href="https://peerj.com/computer-science/"><em>PeerJ Computer Science</em></a>, 2021 (to appear). DOI: <a href="https://dx.doi.org/10.7717/peerj-cs.631">10.7717/peerj-cs.631</a></p> <p>After retrieving the full dataset, extract the replication-package.zip archive and follow the instructions described in the README.md file.</p>
Scalably learning quantum many-body Hamiltonians from dynamical data
<p>Our <a href="https://arxiv.org/abs/2209.14328">paper</a> on Hamiltonian learning for large quantum systems contains several numerical results. The results were produced with the <a href="https://github.com/frederikwilde/differentiable-tebd">differentiable-tebd package</a> which we developed for this study. The scripts and raw output data, as well as Jupyter notebooks for generating the plots shown in the paper are contained in this repository. For more information please refer to the <a href="https://github.com/frederikwilde/scalable-dynamical-hamiltonian-learning/">guiding repository</a>.</p> <p>For funding information please refer to the acknowledgement section of the <a href="https://arxiv.org/abs/2209.14328">paper</a>.</p>
A "short blanket" dilemma for a state-of-the-art neural network potential for water: Reproducing experimental properties or the underlying many-body physics?
<p>Deep neural network (DNN) potentials have recently gained popularity in computer simulations of a wide range of molecular systems, from liquids to materials.<br> In this study, we explore the possibility of combining the computational efficiency of the DeePMD framework and the demonstrated accuracy of the MB-pol data-driven many-body potential to train a DNN potential for large-scale simulations of water across its phase diagram.<br> We find that the DNN potential is able to reliably reproduce the MB-pol results for liquid water but provides a less accurate description of the vapor-liquid equilibrium properties.<br> This shortcoming is traced back to the inability of the DNN potential to correctly represent many-body interactions.<br> An attempt to explicitly include information about many-body effects results in a new DNN potential that exhibits the opposite performance, being able to correctly reproduce the MB-pol vapor-liquid equilibrium properties but losing accuracy in the description of the liquid properties.<br> These results suggest that DeePMD-based DNN potentials are not able to correctly "learn" and, consequently, represent many-body interactions, which implies that DNN potentials may have limited ability to predict properties for state points that are not explicitly included in the training process.<br> The computational efficiency of the DeePMD framework can still be exploited to train DNN potentials on data-driven many-body potentials, which can thus enable large-scale, "chemically accurate" simulations of various molecular systems, with the caveat that the target state points must have been adequately sampled by the reference data-driven many-body potential in order to guarantee a faithful representation of the associated properties.</p>
Data for GECCO2023 Paper "Many-objective (Combinatorial) Optimization is Easy"
<p><strong>Data for Paper "Many-objective (Combinatorial) Optimization is Easy"</strong></p> <ul> <li><strong>instances.tar.xz</strong> contains 𝜌mnk-landscape instances</li> <li><strong>metrics.csv</strong> contains the metric-values based on full enumeration</li> <li><strong>performance.csv</strong> contains the Pareto resolution and the hypervolume of the different algorithms one each instance</li> <li><strong>performance_neval.csv</strong> contains the number of evaluations performed by PLS</li> <li><strong>script.R</strong> is the R script for producing the figures</li> </ul> <p><strong>Reference</strong></p> <p>Arnaud Liefooghe and Manuel López-Ibáñez. 2023. Many-objective (Combinatorial) Optimization is Easy. In Genetic and Evolutionary Computation Conference (GECCO '23), July 15–19, 2023, Lisbon, Portugal. <a href="https://doi.org/10.1145/3583131.3590475">https://doi.org/10.1145/3583131.3590475</a></p> <p><strong>Abstract</strong></p> <p>It is a common held assumption that problems with many objectives are harder to optimize than problems with two or three objectives. In this paper, we challenge this assumption and provide empirical evidence that increasing the number of objectives tends to reduce the difficulty of the landscape being optimized. Of course, increasing the number of objectives brings about other challenges, such as an increase in the computational effort of many operations, or the memory requirements for storing non-dominated solutions. More precisely, we consider a broad range of multi- and many-objective combinatorial benchmark problems, and we measure how the number of objectives impacts the dominance relation among solutions, the connectedness of the Pareto set, and the landscape multimodality in terms of local optimal solutions and sets. Our analysis shows the limit behavior of various landscape features when adding more objectives to a problem. Our conclusions do not contradict previous observations about the inability of Pareto-optimality to drive search, but we explain these observations from a different perspective. Our findings have important implications for the design and analysis of many-objective optimization algorithms.</p>
Virtual memory on a many-core NoC: experimental data
<p>Experimental data that accompanies the thesis "Virtual Memory on a Many-Core NoC" (http://etheses.whiterose.ac.uk/25675/). The data is textual and compressed.</p>
How many words is a picture worth? Attention allocation on thumbnails versus title text regions: Dataset
<p>Dataset for the following publication: https://jainlab.cise.ufl.edu/eyetrack-onlineux.html</p> <p>How many words is a picture worth? Attention allocation on thumbnails versus title text regions, Yandandul, Chaitra and Paryani, Sachin and Le, Madison and Jain, Eakta, ACM Symposium on Eye Tracking Research & Applications. (ETRA)</p> <p> </p>
Data from: Properties of Markov chain Monte Carlo performance across many empirical alignments -- part I
<p>Nearly all current Bayesian phylogenetic applications rely on Markov chain Monte Carlo (MCMC) methods to approximate the posterior distribution for trees and other parameters of the model. These approximations are only reliable if Markov chains adequately converge and sample from the joint posterior distribution. While several studies of phylogenetic MCMC convergence exist, these have focused on simulated datasets or select empirical examples. Therefore, much that is considered common knowledge about MCMC in empirical systems derives from a relatively small family of analyses under ideal conditions. To address this, we present an overview of commonly applied phylogenetic MCMC diagnostics and an assessment of patterns of these diagnostics across more than 18,000 empirical analyses. Many analyses appeared to perform well and failures in convergence were most likely to be detected using the average standard deviation of split frequencies, a diagnostic that compares topologies among independent chains. Different diagnostics yielded different information about failed convergence, demonstrating that multiple diagnostics must be employed to reliably detect problems. The number of taxa and average branch lengths in analyses have clear impacts on MCMC performance, with more taxa and shorter branches leading to more difficult convergence. We show that the usage of models that include both Γ-distributed among-site rate variation and a proportion of invariable sites are not broadly problematic for MCMC convergence but are also unnecessary. Changes to heating and the usage of model-averaged substitution models can both offer improved convergence in some cases, but neither are a panacea.</p>
Data from: Properties of Markov chain Monte Carlo performance across many empirical alignments --part II
<p>Nearly all current Bayesian phylogenetic applications rely on Markov chain Monte Carlo (MCMC) methods to approximate the posterior distribution for trees and other parameters of the model. These approximations are only reliable if Markov chains adequately converge and sample from the joint posterior distribution. While several studies of phylogenetic MCMC convergence exist, these have focused on simulated datasets or select empirical examples. Therefore, much that is considered common knowledge about MCMC in empirical systems derives from a relatively small family of analyses under ideal conditions. To address this, we present an overview of commonly applied phylogenetic MCMC diagnostics and an assessment of patterns of these diagnostics across more than 18,000 empirical analyses. Many analyses appeared to perform well and failures in convergence were most likely to be detected using the average standard deviation of split frequencies, a diagnostic that compares topologies among independent chains. Different diagnostics yielded different information about failed convergence, demonstrating that multiple diagnostics must be employed to reliably detect problems. The number of taxa and average branch lengths in analyses have clear impacts on MCMC performance, with more taxa and shorter branches leading to more difficult convergence. We show that the usage of models that include both Γ-distributed among-site rate variation and a proportion of invariable sites are not broadly problematic for MCMC convergence but are also unnecessary. Changes to heating and the usage of model-averaged substitution models can both offer improved convergence in some cases, but neither are a panacea.</p>
Measurements of 406 medieval and post-medieval houses in the southern Mani Peninsula, Greece
<p>Data from field recording of domestic house architecture in the southern region of the Mani Peninsula, Greece. Measurements were collected for 406 houses dating to the Middle Byzantine to Ottoman periods (8th to 17th centuries AD), allowing for a better understanding of typical characteristics of the so-called "palaiomaniatika" (or "Old Maniat") architecture that dates to these periods. Additionally, total house counts were obtained for 32 settlements that were able to be mapped in their entirety, providing insight into the typical size of these settlements. A paper interpreting these data, 'The Stone-Built Palaiomaniatika of the Mani Peninsula, Greece,' will be published in the volume <em>Deserted Villages: Perspectives from the Eastern Mediterranean</em>, edited by Rebecca M. Seifried and Deborah Brown Stewart, Digital Press at the University of North Dakota, Grand Forks (forthcoming, 2021). The paper will be linked here after publication.</p> <p>Dr. Seifried collected the data as part of a PhD dissertation completed in 2016 through the Department of Anthropology at the University of Illinois at Chicago, entitled 'Community Organization and Imperial Expansion in a Rural Landscape: The Mani Peninsula, Greece (AD 1000-1821)' (<a href="http://hdl.handle.net/10027/21274">https://hdl.handle.net/10027/21274</a>). Fieldwork was conducted in 2014 and 2016 under the auspices of the 5th Ephorate of Byzantine Antiquities in Sparta and in collaboration with the Diros Project, an archaeological survey and excavation co-directed by Dr. Giorgos Papathanassopoulos and Dr. Anastasia Papathanasiou through the Ephorate of Palaeoanthropology & Speleology of Southern Greece. </p> <p>This research was funded by the National Science Foundation (BCS-1346694), the Marie Sklodowska-Curie Actions (H2020-MSCA-IF-2016 750843), the DigitalGlobe Foundation, the National Cadastre and Mapping Agency, SA (Ktimatologio), ArchaeoLandscapes Europe, the University of Illinois at Chicago.</p>
Matrix multiplication software and results bundle for paper "Tuning and optimization for a variety of many-core architectures without changing a single line of implementation code using the Alpaka library" for P^3MA submission
<p>This is the archive containing the matrix multiplication software and the results of the publication "<em>Tuning and optimization for a variety of many-core architectures without changing a single line of implementation code using the Alpaka library</em>" submitted to the P^3MA workshop 2017.</p> <p><strong>The archive has the following content:</strong></p> <ul> <li>Source code for the (tiled) matrix multiplication in "src": <ul> <li>regular version in "src/matmul": <ul> <li>Remote: https://github.com/theZiz/matmul.git (copy will be removed)</li> <li>Branch: topic-compatible-alpaka-0-1-0</li> <li>Commit: a63ba4810d6bfcca62c68dd57408af15028e78a3</li> </ul> </li> <li>forked version for XL in "src/matmul": <ul> <li>Remote: https://github.com/theZiz/matmul.git (copy will be removed)</li> <li>Branch: topic-xl-workaround</li> <li>Commit: 1fee028eccb8cf7b677e8071233e08aa9f81846a</li> </ul> </li> </ul> </li> <li>The compiled binaries and the results of the tuning and scaling runs are in "runs" in sub folders for each type of run and architectures.</li> </ul>
Figure 6 in Integrated morphological, CO1 and distributional analysis confirms many species in the Iridomyrmex anceps (Roger) complex of ants
Figure 6. Variation in gastric pubescence among selected species of the Iridomyrmex anceps complex. A. Sp. B (specimen RHYIR 070), dorsal view. B. Sp. K (specimen RHYIR 076), dorsal view. C. Sp. M (specimen IRIDO 249), lateral view. D. Sp. N (specimen RHYIR 095), lateral view. E. Sp. O (specimen IRIDO 251-16), dorsal view. F. Sp. P (specimen IRIDO 250-16), dorsal view. Scale bars = 0.1 mm.
Figure 5 in Integrated morphological, CO1 and distributional analysis confirms many species in the Iridomyrmex anceps (Roger) complex of ants
Figure 5. Distribution records of sequenced specimens from the Iridomyrmex anceps complex in Australia. Monsoonal tropics region shaded in grey. A. Species A, G, H and I. B. Species B, J, K, L, M, N, O and R. Abbreviations: WA—Western Australia, NT— Northern Territory, Qld—Queensland.
Figure 3 in Integrated morphological, CO1 and distributional analysis confirms many species in the Iridomyrmex anceps (Roger) complex of ants
Figure 3. Scape length in relation to head length for selected species. A. Species A, G, H, I and J. B. Species B, K, L, M and N.
Figure 2 in Integrated morphological, CO1 and distributional analysis confirms many species in the Iridomyrmex anceps (Roger) complex of ants
Figure 2. Summary CO1 tree of the 82 sequenced specimens of Iridomyrmex anceps. Maximum-likelihood phylogeny inferred using IQ-TREE. Black and red circles indicate bootstrap support values ≥90 and ≥70, respectively. The full CO1 tree is shown in Supplementary Figure 1. Abbreviations: WA—Western Australia, NT—Northern Territory, Qld—Queensland, PNG—Papua New Guinea.
Figure 1 in Integrated morphological, CO1 and distributional analysis confirms many species in the Iridomyrmex anceps (Roger) complex of ants
Figure 1. Head (A) and lateral (B) views of a typical member of the Iridomyrmex anceps complex (sp. A; specimen IRIDO217-16). Scale bars = 1 mm.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.