Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
123
datasets available to search
ShareScore release 0.9.0
Dataset results
123 results for “SNP genotyping”
Data from: Finding the right coverage: The impact of coverage and sequence quality on SNP genotyping error rates
Open the record for dataset details and reuse information.
Spatiotemporal monitoring of the rare Northern dragonhead, Dracocephalum ruyschiana (Lamiaceae): SNP genotyping and environmental niche modelling herbarium specimens
Open the record for dataset details and reuse information.
Data from: Origins of cattle on Chirikof Island, Alaska, elucidated from genome-wide SNP genotypes
Open the record for dataset details and reuse information.
Data from: Phylogeography and adaptation genetics of stickleback from the Haida Gwaii archipelago revealed using genome-wide SNP genotyping
Open the record for dataset details and reuse information.
Impatiens glandulifera SNP and SilicoDArT genotyping data
Open the record for dataset details and reuse information.
SNP genotyping of North Head and northern Sydney Long-nosed bandicoots (Perameles nasuta)
Open the record for dataset details and reuse information.
SNP genotypes for healthy and CIM-affected GSDs
Open the record for dataset details and reuse information.
Gila Trout neutral and outlier SNP genotype matrices
Open the record for dataset details and reuse information.
Data from: Characterisation of microsatellite and SNP markers from Miseq and genotyping-by-sequencing data among parapatric Urophora cardui (Tephritidae) populations
Open the record for dataset details and reuse information.
Construction of genetic linkage map based on SNP markers, QTL mapping and detection of candidate genes of growth-related traits in Pacific abalone using genotyping-by-sequencing
Open the record for dataset details and reuse information.
Data from: A new multiplex SNP genotyping assay for detecting hybridization and introgression between the M and S molecular forms of Anopheles gambiae
Open the record for dataset details and reuse information.
Data from: Use of genotyping-by-sequencing data to develop a high-throughput and multi-functional SNP panel for conservation applications in Pacific lamprey
Open the record for dataset details and reuse information.
Data from: More affordable and effective noninvasive SNP genotyping using high-throughput amplicon sequencing
<p>Non-invasive genotyping methods have become key elements of wildlife research over the last two decades, but their widespread adoption is limited by high costs, low success rates, and high error rates. <span>The information lost when genotyping success is low may lead to decreased precision in animal population densities, which could misguide conservation and management actions.</span> <span>Single nucleotide polymorphisms (SNPs) provide a promising alternative to traditionally used microsatellites as SNPs allow amplification of shorter DNA fragments, are less prone to genotyping errors, and produce results that are easily shared among laboratories.</span> Here, we outline a detailed protocol for cost-effective and accurate noninvasive SNP genotyping using multiplexed amplicon sequencing optimized for degraded DNA. <span>We validated this method for individual identification by genotyping 216 scats, 18 hairs and 15 tissues from coyotes (<i>Canis latrans</i>) using 26 SNPs. </span><a name="_Hlk33181599">Our genotyping success rate for scat samples was 93%, and 100% for hair and tissue, representing a substantial increase compared to previous microsatellite-based studies while remaining at a low cost of under $5 per PCR replicate (excluding labor). </a>The accuracy of the genotypes was further corroborated in that genotypes from scats matching known, GPS-collared coyotes were always located within the territory of the known individual. We also show that different levels of multiplexing produced similar results, but that PCR product cleanup strategies can have substantial effects on genotyping success. By making noninvasive genotyping more affordable, accurate, and efficient, this research may allow for a substantial increase in the use of noninvasive methods to monitor and conserve free-ranging wildlife populations.</p>
Data from: Estimations of linkage disequilibrium, effective population size and ROH-based inbreeding coefficients in Spanish Churra sheep using imputed high-density SNP genotypes
In this study, the availability of the Ovine HD SNP BeadChip (HD-chip) and the development of an imputation strategy provided an opportunity to further investigate the extent of linkage disequilibrium (LD) at short distances in the genome of the Spanish Churra dairy sheep breed. A population of 1686 animals, including 16 rams and their half-sib daughters, previously genotyped for the 50K-chip, was imputed to the HD-chip density based on a reference population of 335 individuals. After assessing the imputation accuracy for beagle v4.0 (0.922) and fimpute v2.2 (0.921) using a cross-validation approach, the imputed HD-chip genotypes obtained with beagle were used to update the estimates of LD and effective population size for the studied population. The imputed genotypes were also used to assess the degree of homozygosity by calculating runs of homozygosity and to obtain genomic-based inbreeding coefficients. The updated LD estimations provided evidence that the extent of LD in Churra sheep is even shorter than that reported based on the 50K-chip and is one of the shortest extents compared with other sheep breeds. Through different comparisons we have also assessed the impact of imputation on LD and effective population size estimates. The inbreeding coefficient, considering the total length of the run of homozygosity, showed an average estimate (0.0404) lower than the critical level. Overall, the improved accuracy of the updated LD estimates suggests that the HD-chip, combined with an imputation strategy, offers a powerful tool that will increase the opportunities to identify genuine marker-phenotype associations and to successfully implement genomic selection in Churra sheep.
Data from: Vitis phylogenomics: hybridization intensities from a SNP array outperform genotype calls
Understanding relationships among species is a fundamental goal of evolutionary biology. Single nucleotide polymorphisms (SNPs) identified through next generation sequencing and related technologies enable phylogeny reconstruction by providing unprecedented numbers of characters for analysis. One approach to SNP-based phylogeny reconstruction is to identify SNPs in a subset of individuals, and then to compile SNPs on an array that can be used to genotype additional samples at hundreds or thousands of sites simultaneously. Although powerful and efficient, this method is subject to ascertainment bias because applying variation discovered in a representative subset to a larger sample favors identification of SNPs with high minor allele frequencies and introduces bias against rare alleles. Here, we demonstrate that the use of hybridization intensity data, rather than genotype calls, reduces the effects of ascertainment bias. Whereas traditional SNP calls assess known variants based on diversity housed in the discovery panel, hybridization intensity data survey variation in the broader sample pool, regardless of whether those variants are present in the initial SNP discovery process. We apply SNP genotype and hybridization intensity data derived from the Vitis9kSNP array developed for grape to show the effects of ascertainment bias and to reconstruct evolutionary relationships among Vitis species. We demonstrate that phylogenies constructed using hybridization intensities suffer less from the distorting effects of ascertainment bias, and are thus more accurate than phylogenies based on genotype calls. Moreover, we reconstruct the phylogeny of the genus Vitis using hybridization data, show that North American subgenus Vitis species are monophyletic, and resolve several previously poorly known relationships among North American species. This study builds on earlier work that applied the Vitis9kSNP array to evolutionary questions within Vitis vinifera and has general implications for addressing ascertainment bias in array-enabled phylogeny reconstruction.
Data from: Development of highly reliable in silico SNP resource and genotyping assay from exome capture and sequencing: an example from black spruce (Picea mariana)
Picea mariana is a widely distributed boreal conifer across Canada and the subject of advanced breeding programs for which population genomics and genomic selection approaches are being developed. Targeted sequencing was achieved after capturing P. mariana exome with probes designed from the sequenced transcriptome of Picea glauca, a distant relative. A high capture efficiency of 75.9% was reached although spruce has a complex and large genome including gene sequences interspersed by some long introns. The results confirmed the relevance of using probes from congeneric species to perform successfully interspecific exome capture in the genus Picea. A bioinformatics pipeline was developed including stringent criteria that helped detect a set of 97 075 highly reliable in silico SNPs. These SNPs were distributed across 14 909 genes. Part of an Infinium iSelect array was used to estimate the rate of true positives by validating 4267 of the predicted in silico SNPs by genotyping trees from P. mariana populations. The true positive rate was 96.2%, for in silico SNPs compared to a genotyping success rate of 96.7% for a set 1115 P. mariana control SNPs recycled from previous genotyping arrays. These results indicate the high success rate of the genotyping array and the relevance of the selection criteria used to delineate the new P. mariana in silico SNP resource. Furthermore, in silico SNPs were generally of medium to high frequency in natural populations, thus providing high informative value for future population genomics applications.
SNP data for Syringa vulgaris genotypes
<p><span>Common lilac (<em>Syringa</em> <em>vulgaris</em> L.) is a popular landscaping plant. In the present study, its genotypes were investigated using genotyping-by-sequencing (GBS) methodology. Our aim was to obtain a large set of SNP markers, to reveal the precise identities of the investigated <em>S. vulgaris</em> accessions, and to discover genetic relationships among them. The studied plant material included local Finnish, previously unidentified accessions, known reference cultivars, and so-called historical accessions i.e., old shrubs growing in historic cultural landscapes. We intended to verify cultivar names for some valuable local common lilac accessions and to provide insights into the history of common lilac cultivation in Finland. In the analyses, we used a set of 15,007 SNP markers.</span> <span>First, polymorphic information contents (PIC) were calculated (mean 0.190, range 0.012–0.500 per marker). Then, to investigate genetic relationships among genotypes, a phylogenetic tree was constructed, and a principal coordinate analysis (PCoA) was conducted. A Bayesian analysis of population structure was carried out to determine the number and distribution of genetic clusters among samples. Genetic marker data combined with existing historical and phenotypic knowledge revealed novel information on the unidentified cultivars and on the genetic relationships among studied accessions, and solved the arrival and early history of common lilac in Finland. Overall, such comprehensive genomic characterization and deep understanding of genetic relationships of <em>S. vulgaris</em> can be used when utilizing present cultivars and developing new ones in future breeding programs.</span></p>
Genotype by sequencing SNP dataset for Malus domestica x Malus sieversii F1 populations GMAL4591 and GMAL4592
<p>Fire blight, a bacterial disease caused by <em>Erwinia</em> <em>amylovora</em>, is the most devastating disease of apples and a major threat to apple production. Most commercial apple cultivars are susceptible to fire blight driving the need to develop fire blight resistant cultivars. Although several major fire blight resistance QTLs have been identified from wild species of <em>Malus</em>, the challenges of breeding apples due to long juvenile phase and heterozygosity greatly limit their use. <em>M. sieversii</em>, the primary progenitor of domesticated apples, is one of the wild <em>Malus</em> species that is sexually compatible with <em>M. domestica</em> and has some favorable fruit quality traits. In this study, we performed QTL analysis on two F1 apple populations of <em>M. domestica</em> cv. 'Royal Gala' × <em>M. sieversi</em>i (GMAL4591 and GMAL4592) to identify fire blight resistance QTL. Parental linkage maps were constructed for each family using marker sets of approximately 20K GBS-SNPs. Phenotype data was collected from parents and progeny through controlled fire blight inoculations in the greenhouse for two subsequent years. A significant (P < 0.0001) moderate-effect fire blight resistance QTL on linkage group 7 of GMAL4591 was identified from the paternal parent<em> M. sieversii</em> 'KAZ 95 17-14' (<em>Msv_FB7</em>). <em>Msv_FB7</em> explains about 48–53% of the phenotyping variance across multiple years and time points. Additionally, a significant (P < 0.001) minor effect QTL explaining 18% of the phenotypic variance was identified in population GMAL4592 on LG10 from 'Royal Gala'. We developed diagnostic SSR markers flanking the Msv_FB7 QTL to use in apple breeding. These findings have the potential to accelerate the development of fire-blight-resistant cultivars.</p>
Data from: A 34K SNP genotyping array for Populus trichocarpa: Design, application to the study of natural populations and transferability to other Populus species
Open the record for dataset details and reuse information.
Genotype by sequencing SNP dataset for Malus domestica x Malus sieversii F1 populations GMAL4591 and GMAL4592
Open the record for dataset details and reuse information.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.