Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
4
datasets available to search
ShareScore release 0.9.0
Dataset results
4 results for “Nephele”
NEPHELE Tunable Laser Performance
<p><em>Time domain and frequency domain measurements describing the switching performance of the NEPHELE POD switch subsystem. The dataset will describe the methodology for measuring switching times, extinction ratio and crosstalk in the POD switch that combines the NEPHELE WSS and tunable laser subsystems with passive optical components such as optical couplers and AWGs. The scope of the research context is on dynamically reconfigurable optical networks for application in datacenters.</em></p>
NEPHELE big data experiment
<p>The big data experiments use HiBench for benchmarking, which generates data using a number of data generators. These generators generate text data, web-data that follows a Zipfian distribution, numeric data for k-means clustering following a uniform or Gaussian distribution. Each of those benchmark instances consists of ten micro benchmarks. Some of those micro benchmarks provide multiple implementations. These differ in language of implementation and framework. The frameworks includes Hadoop and Spark. The Spark-based implementations feature code in Java, Python, and Scala. Below is a list of evaluated workloads and data:</p> <p>1. Sort – This micro benchmark sorts text data, which is generated by RandomTextWriter.</p> <p>2. WordCount – This micro-benchmark counts the occurrences of words in the input data, which is generated by RandomTextWriter.</p> <p>3. TeraSort – This micro benchmark sorts a large data set generated by TeraGen. This micro benchmark has been used to push the boundaries of sorting very large data sets.</p> <p>4. Sleep – This micro benchmark sleeps for a set amount of time to test the scheduler.</p> <p>5. SQL (Scan, Join, Aggregate) – This benchmark tests typical SQL commands using Hive queries. The benchmark operates on generated hyperlinks.</p> <p>6. PageRank – This benchmark computes the PageRank on generated hyperlinks.</p> <p>7. Nutch indexing – Nutch is an Apache search engine and implements an indexing algorithm. The benchmark operates on generated web sites with words and hyperlinks.</p> <p>8. Bayesian classification – This benchmarks uses Spark MLLib and Mahout to perform Naive Bayesian classification. The data consists of generated text documents.</p> <p>9. K-means clustering – In this benchmark the k-means clustering algorithm in Spark MLLib/Mahout is used to cluster generated input data. </p> <p>10. Enhanced DFSIO – This benchmark tests the HDFS throughput by generating a large number of tasks that simultaneously write and read data from HDFS.</p>
NEPHELE POD switch performance
<p><em>Time domain and frequency domain measurements describing the switching performance of the NEPHELE POD switch subsystem. The dataset will describe the methodology for measuring switching times, extinction ratio and crosstalk in the POD switch that combines the NEPHELE WSS and tunable laser subsystems with passive optical components such as optical couplers and AWGs. The scope of the research context is on dynamically reconfigurable optical networks for application in datacenters.</em></p>
NEPHELE WSS switch performance
<p>Time domain and frequency domain measurements describing the switching performance of the NEPHELE WSS subsystem. The dataset will describe the methodology for measuring switching times and crosstalk between WDM channels in the WSS subsystem. The scope of the research context is on dynamically reconfigurable optical networks for application in datacenters or metropolitan networks.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.