Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5
datasets available to search
ShareScore release 0.9.0
Dataset results
5 results for “document processing”
Collection of videos documenting the silk weaving process (Silk pilot, Mingei)
<p>Documentation videos of the Weaving process from the Silk pilot of the Mingei project.</p>
Research data, sources and documents for thesis on Exploring Complexity Metrics for Artifact-Centric Business Process Models
<p>Research data, sources and documents for thesis on Exploring Complexity Metrics for Artifact-Centric Business Process Models This repository contains the supplemental material for the <a href="https://pqdtopen.proquest.com/pubnum/10759956.html">thesis "Exploring Complexity Metrics for Artifact-Centric Business Process Models" by Marin, Mike A., Ph.D., University of South Africa (South Africa), 2017.</a></p>
Dataset for Paper: A System for Processing and Recognition of Greek Byzantine and Post-Byzantine Documents
<p>Dataset for the paper: "A System for Processing and Recognition of Greek Byzantine and Post-Byzantine Documents", P. Kaddas, K. Palaiologos, B. Gatos, V. Katsouros, K. Christopoulou, 17th International Conference on Document Analysis and Recognition (ICDAR), San Jose, California, USA</p> <p>The dataset consists of 57 pages from the third edition of the Greek New Testament published by Robert Estienne (1503–1559), who was appointed “Royal Typographer” by the King of France François I (1494–1547). Robert Estienne produced this edition in 1550 using the grecs du roi typeface, produced by Claude Garamont on the basis of the Greek minuscule style of the calligrapher Angelos Vergikios (1505–1569) from Crete, who active copying Greek manuscripts in Venice and France. The dataset consists of 2045 cropped text line images in .png format with their corresponding OCR in .txt format, where 1431 used for training, 204 for validation and 410 for test. Initial images acquired from: https://bibles-online.net/flippingbook/1550/</p>
Assisted Data Annotation for Business Process Information Extraction from Textual Documents
Open the record for dataset details and reuse information.
Usability Evaluation Process for Domain-Specific Languages - Documents and survey data
<p>Usability Evaluation Process for Domain-Specific Languages - Documents and survey data</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.