Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

3,481

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

3,481 results for “data set”

Learn how ShareScore rates datasets ↗
zenodo40/100

Dune Overwash and Breaching data set produced at the CIEM flume, Hydralab III

<p>The data set here presented helps to to improve the understanding of dune overwash and breaching processes during storm surges in nearly prototype scale. The experiments were done at the CIEM wave flume at UPC, Barcelona, as part of Hydralab III. The large scale movable-bed hydraulic experiments measure hydrodynamics and sediment processes involved in onshore and offshore sediment transport, dune breaching and overwash.</p> <p>Due to its size, the data set can not be placed on this repository and will be provided on demand. Please contact with the authors or with the data manager of the CIEM installation.</p> <p>More information can be found on the published papers:</p> <p>D'Alessandro, F.; Tomasicchio, R.; Alsina, J.; Caceres, I.; Fortes, C.J.E.M.; Ilic, S.; James, M.; Nagler, L.; Pinheiro, L.V.; Sanchez-Arcilla, A.; Sancho, F.; Shaw, E.; Schüttrumpf, H., 2010. Dune over wash and breaching, Coastlab 2010, Barcelona, Spain.</p> <p>&nbsp;</p>

opencc-by-4.0May 2018View details →
zenodo40/100

Swash zone response Under grouping Storm Conditions data set produced at the CIEM flume, Hydralab III

<p>The data set here presented reports the large-scale laboratory experiments on the influence of long waves, bichromatic wave groups and random waves on sediment transport in the surf and swash zones. The experiments were done at the CIEM wave flume at UPC, Barcelona, as part of the SUSCO (swash zone response under grouping storm conditions) experiment in the Hydralab III. Fourteen different wave conditions were used, encompassing monochromatic waves, bichromatic wave groups and random waves. The experiments were designed specifically to compare variations in beach profile evolution between monochromatic waves and unsteady waves with the same mean energy flux. Each test commenced with approximately the same initial profile</p> <p>Due to its size, the data set can not be placed on this repository and will be provided on demand. Please contact with the authors or with the data manager of the CIEM installation.</p> <p>More information can be found on the published papers:</p> <p>Baldock, T.E., Alsina, J.A., Caceres, I., Vicinanza, D., Contestabile, P., Power, H. and Sanchez-Arcilla, A., 2011. Large-scale experiments on beach profile evolution and surf and swash zone sediment transport induced by long waves, wave groups and random waves. Coastal Engineering, Vol. 58, pp. 214-227.</p> <p>Vicinanza, D., Baldock, T., Contestabile, P., Alsina, J., Cáceres, I., Brocchini, M., Conley, D., Andersen, T.L., Frigaard, P. and Ciavola, P., 2011. Swash zone response under various wave regimes. Journal of Hydraulic Research, Vol. 49, pp. 55-63.</p>

opencc-by-4.0May 2018View details →
zenodo40/100

Wave-induced steady current data set produced at the CIEM wave flume, Hydralab III

<p>The data set here presented reports the Wave-induced steady currents experiments done in the Barcelona CIEM flume. This experiment was part of the TA within Hydralab III. The aim of the experiments was to obtain new data of flow velocity in a large scale wave flume where the bottom boundary layer is in the turbulent regime. The measurements provide instantaneous velocity values along the vertical, offshore of the breaker line, in presence of an erodible bed and, in turn, in presence of small scale bedforms. The data is elaborated in order to obtain statistical quantities such as ensemble-averaged velocity profiles, steady velocity components, Reynolds stresses and eddy viscosity.</p> <p>Due to its size, the data set can not be placed on this repository and will be provided on demand. Please contact with the authors or with the data manager of the CIEM installation.</p> <p>More information can be found on the published papers:</p> <p>Scandura, P and Foti, E., 2011. Measurements of wave-induced steady currents outside the surf zone. Journal of Hydraulic Research, Vol. 49, 64-71</p>

opencc-by-4.0May 2018View details →
zenodo40/100

Coupled High Frequency Measurements of Swash Sediment Transport and Morphodynamic data set produced at the CIEM flume, Hydralab IV

<p>The data set here presented aims to increase the understanding of the nearshore sediment dynamics. The experiments aimed at obtaining high quality data of hydrodynamics, sediment concentration and beach-face evolution with an intra-wave time scale. The specific objectives of CoSSedM access project were i) to obtain information of the effect of the wave group periods on the beach morphological evolution; ii) To obtain detailed sediment transport information at the inner surf and swash zones with different bi-chromatic wave conditions and iii) To obtain intra-wave measurements of beach-face evolution.</p> <p>The present work was developed in the framework of the HYDRALAB IV Transnational Access projects. The experiments were carried out in the large scale wave flume CIEM at Universitat Politècnica de Catalunya (UPC), Barcelona. This is a wave flume 100 m long, 3 m wide, and 4.5 m deep. The working water depth was at around 2.5 m over the horizontal flume section and was varied slightly depending on the wave test. A beach was installed made of commercial well-sorted sand (d50 = 0.25 mm) with an overall mean beach gradient of approximately 1:15.</p> <p>Due to its size, the data set can not be placed on this repository and will be provided on demand. Please contact with the authors or with the data manager of the CIEM installation.</p> <p>More information can be found on the published papers:</p> <p>Alsina, J.M., Padilla, E.M. and Cáceres, I., 2016. Sediment transport and beach profile evolution induced by bi-chromatic wave groups with different group periods. Coastal Engineering, Vol. 114, 325-340.</p> <p>Van der Zanden, J.; Alsina, J.; Caceres, I.; Buijsrogge, R. H.; Ribberink, J. , 2015. Bed level motions and sheet flow processes in the swash zone : observations with a new conductivity-based concentration measuring technique (CCM+) . Coastal Engineering, Vol. 105, 47-65.</p>

opencc-by-4.0May 2018View details →
zenodo40/100

Slow motion in films and video clips: Music influences perceived duration and emotion, autonomic physiological activation and pupillary responses [data set]

<p>Data set for a study to be published by PLOS ONE.</p>

opencc-by-4.0Jun 2018View details →
zenodo40/100

Data Set for Article "Verification-Aided Debugging: An Interactive Web-Service for Exploring Error Witnesses", Proc. CAV'16

<p>This is the description of the supplementary archive of example interactive reports for the approach described in the article &quot;Verification-Aided Debugging: An Interactive Web-Service for Exploring Error Witnesses&quot;, Proc. CAV&#39;16.</p> <p>This archive contains a static snapshot of our system that allows the reader to<br> a) experience the features of our web-service without relying on its online availability and<br> b) reproduce the bug reports displayed in this static snapshot by validating the provided witnesses against the source code and the corresponding specifications using CPAchecker.</p> <p>The witness database is available at:<br> &nbsp; static/index.html<br> The supplied verification tasks can be found at:<br> &nbsp; static/programs/<br> The supplied error witnesses are grouped by their corresponding verification tasks and can be found at:<br> &nbsp; static/witnesses/<br> The software verifier CPAchecker is placed at:<br> &nbsp; CPAchecker/</p> <p>To browse the witness database and explore the supplied error reports, we recommend using the Firefox web browser,<br> because not all features of our bug reports are guaranteed to be available in other browsers.</p> <p>Like the supplementary archive originally provided to the reviewers, this witness database contains only a small selection of the witnesses harvested from the &quot;Competition on Software Verification 2016&quot;, because we do not want to burden the reader with an enormous amount of data that likely is not relevant for understanding the concepts. Also, error witnesses produced by some competition candidates that were not even syntactically correct were removed, because they do not add any value to the evaluation. However, the full data is still available online via our web service, for example, the list of witnesses for a verification task can be requested by computing the SHA-1 hash of the verification task&#39;s source code and submitting the following query:<br> &nbsp; http://vcloud.sosy-lab.org/webclient/master/witness?inputFile=&lt;program-hash&gt;<br> The resulting JSON data contains all hashes of witnesses stored for the given program.<br> A witness stored in the database can be requested via its SHA-1 hash by submitting the following query:<br> &nbsp; https://vcloud.sosy-lab.org/webclient/files/&lt;hash&gt;<br> All verification tasks are available at the SV-COMP repository:<br> &nbsp; https://github.com/dbeyer/sv-benchmarks<br> If you use verification tasks from the repository and are interested in validating witnesses produced for SV-COMP &#39;16,<br> please use the &#39;svcomp16&#39; tag, because the tasks and their hashes might have changed since then.</p> <p>You can use CPAchecker to validate a witness for a verification task and generate an error report.<br> First, navigate to the CPAchecker directory:</p> <p>&nbsp; cd CPAchecker/</p> <p>Now, perform the validation by providing the verification task (consisting of specification and program source code) and a witness:</p> <p>&nbsp; scripts/cpa.sh -generateReport -witness-validation \<br> &nbsp;&nbsp;&nbsp; -spec &lt;specification&gt; \<br> &nbsp;&nbsp;&nbsp; &lt;source-code&gt; \<br> &nbsp;&nbsp;&nbsp; -spec &lt;witness&gt;</p> <p>For example:</p> <p>&nbsp;scripts/cpa.sh -generateReport -witness-validation \<br> &nbsp;&nbsp;&nbsp; -spec ../static/programs/loop-acceleration/ALL.prp \<br> &nbsp;&nbsp;&nbsp; ../static/programs/loop-acceleration/array_false-unreach-call3.i \<br> &nbsp;&nbsp;&nbsp; -spec ../static/witnesses/loop-acceleration/array_false-unreach-call3.i/a4572a0c1b505b1d1170b7347e48a2a93cb3f4c1</p> <p>The report will be generated in the subdirectory<br> &nbsp; output/report/</p> <p>&nbsp;</p>

opencc-by-sa-4.0Jul 2016View details →
zenodo40/100

Data sets for the Simulated AMPI (SAMPI) load balancing simulation workflow and Ondes3D performance analysis (Companion to CCPE paper)

<p>This package contains data sets and scripts (in&nbsp;an Org-mode file) related to our submission to the&nbsp; journal &quot;Concurrency and Computation: Practice and Experience&quot;, under the title&nbsp;<em>&quot;Performance Modeling of a Geophysics Application to Accelerate the Tuning of Over-decomposition Parameters through Simulation&quot;</em>.</p>

opencc-by-sa-4.0Jun 2018View details →
zenodo40/100

e-ROSA initiatives' data set

<p>This dataset holds the metadata for the initiatives&nbsp;uploaded/mapped on&nbsp;http://map.aginfra.eu/</p> <p>&nbsp;</p>

opencc-by-nc-4.0Jul 2018View details →
zenodo40/100

Sample ACC Data Set for LOFAR Station IE613 LBA whole sky observation

<p>This is a large sample of data from Station IE613 suitable for use with <a href="https://github.com/creaneroDIAS/beamModelTester">beamModelTester</a> and <a href="https://github.com/2baOrNot2ba/iLiSA">iLiSA</a>.&nbsp; To use with these systems, transfer it to a directory of the form :</p> <p><em>{STN_ID}_YYYYMMDD_HHMMSS_rcu{RCU_MODE}<em>dur{DURATION}</em>{SOURCE}_acc</em><br> e.g. IE613_20180406_091321_rcu3_dur85635_CasA_acc</p>

opencc-by-4.0Aug 2018View details →
zenodo40/100

Data set: Control design, implementation and evaluation for an in-field 500 kW wind turbine with a fixed-displacement hydraulic drivetrain

<p>Data set of Wind Energy Science (WES) paper:&nbsp;Control design, implementation and evaluation for an in-field 500 kW wind turbine with a fixed-displacement hydraulic drivetrain</p>

opencc-by-4.0May 2018View details →
zenodo40/100

A ferrofluid-based sensor to measure bottom shear stresses under currents and waves. Data set: VelocityProfilies_2018_Musumarra

<p>The experimental campaign was devoted to study the velocity profile inside the small scale flume for several bottom configurations. In particular, the following configurations were considered: thin sand (D<sub>50</sub>=0.25 mm); coarse sand (D<sub>50</sub>=0.56 mm); mixed sand: 10% coarse sand and 90% thin sand; mixed sand: 20% coarse sand and 80% thin sand; mixed sand: 30% coarse sand and 70% thin sand; mixed sand: 40% coarse sand and 60% thin sand; small gravel (diameter between 3 and 5 mm); gravel (diameter between 9 and 14 mm); small gravel and thin sand; &nbsp;gravel and thin sand.</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

Distinguishing between pan assay interference compounds (PAINS) that are promiscuous or represent dark chemical matter - data set and prediction models

<p>Data sets of promiscuous PAINS (PROM_PAINS) and dark chemical matter PAINS (DCM_PAINS) are provided and support vector machine models built on the basis of original and balanced training data (see readme.txt).<br> &nbsp;</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

NCBI::Taxonomy data set

<p>This is a data set for NCBI::Taxonomy Module.</p> <p>The sizes of the decompressed files are:</p> <pre><code class="language-bash">all_taxonomy.bin 25 GBytes nodes.bin 185 MBytes nucl_est_taxonomy.bin 5.1 GBytes nucl_gb_taxonomy.bin 6.1 GBytes nucl_gss_taxonomy.bin 4.5 GBytes nucl_wgs_taxonomy.bin 14 GBytes pdb_taxonomy.bin 4.2 GBytes prot_taxonomy.bin 13 GBytes ranks.bin 325 Bytes </code></pre> <p>&nbsp;</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

RACMO2.3p2 South Greenland data set, 2007

<p>This data set contains RACMO2.3p2 output for South Greenland for 2007 on a 20 km grid. Netcdf files are provided every 3 hours. The most relevant variables are declared in RACMO2.3p2_output_variable_declaration.pdf. This data set is used in the manuscript &quot;A module to convert spectral to narrowband snow albedo for use in climate models: SNOWBAL v1.0&quot;.</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

Raw data sets from Jones et al. 2018 QSR publication: A multi-proxy approach to understanding complex responses of saltlake catchments to climate variability and human pressure: A Late Quaternary case study from south-eastern, Spain

<p>Attached are the raw data sets containing the pollen data, DXR, Grain size and C14 ages from the recent publication:&nbsp;Jones et al. 2018 QSR publication: A multi-proxy approach to understanding complex responses of saltlake catchments to climate variability and human pressure: A Late Quaternary case study from south-eastern, Spain.</p> <p>Note that these data sets do contain hiatuses and a major age-reversal due to erosian which have&nbsp;likely been caused by increased seasonal wetness at the onset of the Holocene. A full explanation is provided in our 2018 publication. If you do wish to use the data, it is essential that you read&nbsp;the publication inorder to interpret the results correctly. We also require that when using this data that you correctly cite it&nbsp;(Bibliographic reference and the doi number of the data set). There were some problems uploading the XRF (geochemical)&nbsp;data sets, so I haven&#39;t included these yet, but hopefully will do eventually.&nbsp;</p> <p>Below I have also included the abstract from our publication, which provides an overview of the purpose of our work and a brief summary of the main findings.</p> <p>Abstract of Jones et al. 2018:</p> <p>The article focuses on a former salt lake in the upper Vinalopo Valley in south-eastern Spain. The study spans the Late Pleistocene through to the Late Holocene, although with particular focus on the period between 11 ka cal BP and 3000 ka cal BP (which spans the Mesolithic and part of the Bronze Age). High resolution multi-proxy analysis (including pollen, non pollen palynomorphs, grain size, X-ray fluorescence,&nbsp;and X-ray diffraction) was undertaken on the lake sediments. The results show strong sensitivity to<br> both long term and small changes in the evaporation/precipitation ratio, affecting the surrounding vegetation composition, lake-biota and sediment geochemistry. To summarise the key findings the main general trends identified include: 1) Hyper-saline conditions<br> and low lake levels at the end of the Late Glacial 2) Increasing wetness and temperatures which witnessed an expansion of mesophilic woodland taxa, lake infilling and the establishment of a more perennial lake system at the onset of the Holocene 3) An increase in solar insolation after 9 ka cal BP which saw the re-establishment of pine forests 4) A continued trend towards increasing dryness (climatic optimum) at 7 ka cal BP but with continued freshwater input 5) An increase in sclerophyllous open woody vegetation (anthropogenic?), and increasing wetness (climatic?) is represented in the lake record between 5.9 and 3 ka cal BP 6) The Holocene was also punctuated by several aridity pulses, the most prominent corresponding to the 8.2 ka cal BP event. These events, despite a paucity of well dated archaeological sites in the surrounding area, likely altered the carrying capacity of this area both regionally and locally, particularly during the Mesolithic-Neolithic transition, in terms of fresh water supply for human/animal consumption, wild plant food reserves and suitable land for crop growth.</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

Data set used in the Article "Evaluation of Monte Carlo tools for high-energy atmospheric physics II: relativistic runaway electron avalanches"

<p>Data used for the Relativistic Runaway Electron Avalanches (RREA) simulations of the Article : &quot;Evaluation of Monte Carlo tools for high energy atmospheric physics II: relativistic runaway electron avalanches&quot; by D. Sarria et al.</p> <p>Includes two set of results : The Probability of Generating RREAs, and the Characterizations of RREAs.</p> <p>Link to the article :&nbsp;<a href="https://www.geosci-model-dev.net/11/4515/2018/gmd-11-4515-2018.html">https://www.geosci-model-dev.net/11/4515/2018/gmd-11-4515-2018.html</a></p> <p>DOI of the article:&nbsp;10.5194/gmd-2018-119</p>

opencc-by-4.0Oct 2018View details →
zenodo40/100

Data set for "Columnar clusters in the human motion complex reflect consciously perceived motion axis"

<p>Accompanying data for manuscript &ldquo;Columnar clusters in the human motion complex reflect consciously perceived motion axis&rdquo; written by Marian Schneider, Valentin Kemper, Thomas Emmerling, Federico De Martino, Rainer Goebel, submitted, November 2018.</p> <p>Imaging files<br> -------------<br> * T1w and PDw images, only acquired in session 1<br> * 2 runs task-MotLoc, only acquired in session 2<br> * 5-6 runs task-ambiguous (called &quot;Experiment 1&quot; in accompanying manuscript, divided across 2 scanning sessions)<br> * 5-6 runs task-unambiguous (called &quot;Experiment 2&quot; in accompanying manuscript, divided across 2 scanning sessions)</p> <p><br> Acquisition details<br> -------------------<br> For visualization of the functional results, we acquired scans with structural information in the first scanning session. At high magnetic fields, MR images exhibit high signal intensity variations that result from heterogeneous RF coil profiles. We therefore acquired both T1w images and PDw images using a magnetization-prepared 3D rapid gradient-echo (3D MPRAGE) sequence (TR: 3100 ms (T1w) or 1440 ms (PDw), voxel size = 0.6 mm isotropic, FOV = 230 x 230 mm2, matrix = 384 x 384, slices = 256, TE = 2.52 ms, FA = 5&deg;). Acquisition time was reduced by using 3&times; GRAPPA parallel imaging and 6/8 Partial Fourier in phase encoding direction (acquisition time (TA): 8 min 49 s (T1w) and 4 min 6 s (PDw)).</p> <p>To determine our region of interest, we acquired two hMT+ localiser runs. We used a 2D gradient echo (GE) echo planar imaging (EPI) sequence (1.6 mm isotropic nominal resolution; TE/TR = 18/2000 ms; in-plane field of view (FoV) 150&times;150 mm; matrix size 94 x 94; 28 slices; nominal flip angle (FA) = 69&deg;; echo spacing = 0.71 ms; GRAPPA factor = 2, partial Fourier = 7/8; phase encoding direction head - foot; 240 volumes). We ensured that the area of acquisition had bilateral coverage of the posterior inferior temporal sulci, where we expected the hMT+ areas. Before acquisition of the first functional run, we collected 10 volumes for distortion correction - 5 volumes with the settings specified here and 5 more volumes with identical settings but opposite phase encoding (foot - head), here called &quot;phase1&quot; and &quot;phase2&quot;.</p> <p>For the sub-millimetre measurements (Experiments 1: here called &quot;task-ambiguous&quot; and Experiments 2: here called &quot;task-unambiguous&quot;), we used a 2D GE EPI sequence (TE/TR = 25.6/2000 ms; in-plane FoV 148&times;148 mm; matrix size 186 x 186; slices = 28; nominal FA = 69&deg;; echo spacing = 1.05 ms; GRAPPA factor = 3, partial Fourier = 6/8; phase encoding direction head - foot; 300 volumes), yielding a nominal resolution of 0.8 mm isotropic. Placement of the small functional slab was guided by online analysis of the hMT+ localizer data recorded immediately at the beginning of the first session. This allowed us to ensure bilateral coverage of area hMT+ for every subject. In the second scanning session, the slab was placed using Siemens auto-align functionality and manual corrections. Before acquisition of the first functional run, we collected 10 volumes for distortion correction (5 volumes with opposite phase encoding: foot - head). During acquisition, runs for the ambiguous and unambiguous motion experiments were interleaved.</p>

opencc-by-4.0Nov 2018View details →
zenodo40/100

Multi-publication data set on experienced-based and description-based risky choice

<p>The data of 28 publications studying the description-experience gap, involving experience- and description-based risky choices, that served as the basis for the meta analysis&nbsp;reported in&nbsp;Wulff, D. U., Mergenthaler-Canseco, M., &amp; Hertwig, R. (2018). A meta-analytic review of two modes of learning and the description-experience gap.&nbsp;<em>Psychological Bulletin, 144</em>(2), 140-176.</p>

opencc-by-sa-4.0Nov 2018View details →
zenodo40/100

The Collaborative Organization of Knowledge: Data Set

<p>Wikipedia is an ongoing endeavor to create a free encyclopedia through an open computer-mediated collaborative effort. How does Wikipedia grow and maintain its coverage? This page contains supporing material relevant to a publication that examines this question.</p> <ul> <li>Diomidis Spinellis and Panagiotis Louridas. The collaborative organization of knowledge. Communications of the ACM, 51(8):68&ndash;73, August 2008. (<a href="http://dx.doi.org/10.1145/1378704.1378720">doi:10.1145/1378704.1378720</a>)</li> </ul> <p>In the above paper, a longitudinal study of Wikipedia&#39;s evolution shows that although Wikipedia&#39;s scope is increasing, its coverage is not deteriorating. This can be explained by the fact that referring to an non-existing entry typically leads to the establishment of an article for it. Wikipedia&#39;s evolution also demonstrates the creation of a large real world scale-free graph through a combination of incremental growth and preferential attachment.</p> <p>Though this data set you can download the processed results. The file starts with a header giving various attributes of the processed data set.</p> <pre>% Number of bins: 72 % Total revisions: 28247658 % Maximum revisions: 28273 (George W. Bush) % Maximum reverts: 9218 (George W. Bush) % Number of moves: 81380 % Total pages: 1898139 % Revisions from IP addresses: 8518913 % Total contributors: 230130 % Maximum different contributors: 2539 (George W. Bush) % Redirected pages: 631567 % Restricted pages: 2441 % Maximum number of contained references: 17577 (List of all three letter acrony ms) % Pages with at least one revert: 211704 % Total number of reverts across all pages: 1147151 % Total time between reverts: 54524346346 % Moved pages: 80332 </pre> <p>Next comes one line of data for each one of Wikipedia&#39;s entries. Here is an example.</p> <pre>A (musical note):1128386876:Mailer diablo:1130566991:MrD9:10:7:18:0:0:0:0:0:0:0: 0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0:0: 0:0:0:0:0:0:0:0:0:0:0:1:1:1:2:2:2:2:2:2:2:2:2:2:2:2:E </pre> <p>Each line contains the following fields.</p> <ul> <li>Entry name</li> <li>Time of first definition (in seconds since Unix epoch)</li> <li>Name of the contributor who first defined the entry</li> <li>Time of first reference (in seconds since Unix epoch)</li> <li>Name of the contributor who first referenced the entry</li> <li>Number of references</li> <li>Number of contributors</li> <li>Number of revisions</li> <li>Number of reverts</li> <li>For each one of the time period bins (72 in this file) the number of references to the entry</li> <li>The letter &quot;E&quot;</li> </ul> <p>The fields are colon-separated. Colons in the input data are converted to an underscore.</p> <p>Finally, come lines summarizing the data set&#39;s characteristics for each time period. Here is an example.</p> <pre>2001-07-01 4851 0 27106 15129 13458 531 </pre> <p>Each line contains the following fields.</p> <ul> <li>Start date of this period</li> <li>Number of entries</li> <li>Number of entries that are stubs</li> <li>Number of references</li> <li>Number of referenced articles</li> <li>Number of undefined entries</li> <li>Number of active contributors in this period</li> </ul>

opencc-by-4.0Jun 2008View details →
zenodo40/100

Data sets for modeling double strand break susceptibility and interrogating structural variation in cancer

<p>This is data used and produced for the study of &quot;Modeling double strand break susceptibility to interrogate structural variation in cancer&quot;.&nbsp;</p> <p><strong>Background: </strong>Structural variants (SVs) are known to play important roles in a variety of cancers, but their origins and functional consequences are still poorly understood. Many SVs are thought to emerge from errors in the repair processes following DNA double strand breaks (DSBs).</p> <p><strong>Results:</strong> We used experimentally quantified DSB frequencies in cell lines with matched chromatin and sequence features to derive the first quantitative genome-wide models of DSB susceptibility. These models are accurate and provide novel insights into the mutational mechanisms generating DSBs. Models trained in one cell type can be successfully applied to others, but a substantial proportion of DSBs appear to reflect cell type specific processes. Using model predictions as a proxy for susceptibility to DSBs in tumours, many SV-enriched regions appear to be poorly explained by selectively neutral mutational bias alone. A substantial number of these regions show unexpectedly high SV breakpoint frequencies given their predicted susceptibility to mutation and are therefore credible targets of positive selection in tumours. These putatively positively selected SV hotspots are enriched for genes previously shown to be oncogenic. In contrast, several hundred regions across the genome show unexpectedly low levels of SVs, given their relatively high susceptibility to mutation. These novel coldspot regions appear to be subject to purifying selection in tumours and are enriched for active promoters and enhancers.</p> <p><strong>Conclusions:</strong> We conclude that models of DSB susceptibility offer a rigorous approach to the inference of SVs putatively subject to selection in tumours.</p>

opencc-by-4.0Jan 2019View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record