Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

396

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

396 results for “Books”

Learn how ShareScore rates datasets ↗
zenodo48/100

2020 Citizen Science Event Book: #LongCovid

<p>The Chinese Version "2020 年公民科學事件簿:#長新冠(#Long Covid) ": <a href="https://pansci.asia/archives/370282">https://pansci.asia/archives/370282</a></p><p>The English full text : <a href="https://details-or-fragments.blogspot.com/2023/12/LCEventBook.html">https://details-or-fragments.blogspot.com/2023/12/LCEventBook.html</a></p><p>Long Covid comes from a "patient-created term" in the spring of 2020. On October 6, 2021, the WHO announced its official definition. Although it used "post-COVID-19 condition", the Long Covid is still the most common term. This bottom-up grassroots movement of public participation in scientific concepts in online communities has reached the social conscience of the public and driven scientific development, and finally led to the establishment of relevant policies and scientific progress. This is what sociologists called "citizen science".&nbsp;</p><ul><li>How did it all begin?</li><li>Patient symptom stories: COVID-19 affects more than just the lungs</li><li>Long COVID Citizen Campaign: Responses from health services</li><li>The openness of online social media</li></ul><p>The positive actions of these online community and the collective consensus reached are enough to convincingly prove to medical institutions, including the WHO, that Long Covid is a real disease despite the lack of traditional evidence-based medicine. A group of online citizens collectively wrote the first textbook on Long Covid in 2020. At this moment, we are witnessing the mass power of the online community, which not only promotes real changes in the real world, ensures recognition of medical care supply, but also stimulate a new scientific <i>research</i> stage.</p>

opencc-by-4.0Dec 2023View details →
zenodo48/100

MOBO: The MOvie and BOok reviews dataset

<p>The <strong>MOBO</strong>&nbsp;dataset.</p> <p>The <em><strong>MOvie and BOok reviews dataset</strong></em> is a collection made up of movie and book reviews, paired with their related plots.<br> The reviews come from different publicly available datasets: the Stanford&#39;s IMDB movie reviews [1], the GoodReads [2] and the Amazon reviews dataset [3]. With the help of 15 annotators, we further labeled more than 18,000 reviews&#39; sentences (~6000 per corpus), marking the sentence polarity (<em>Positive</em>,&nbsp;<em>Negative</em>), or whether a sentence describes its corresponding movie/book&nbsp;<em>Plot</em>, or none of the above (<em>None</em>). In the&nbsp;<code>dataset</code>&nbsp;folder, we have shared an excerpt of the annotated sentences for each dataset.</p>

opencc-by-4.0Mar 2022View details →
zenodo48/100

Polifonia Corpus - Books Module Metadata - French Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Polifonia Corpus - Books Module Metadata - Dutch Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Polifonia Corpus - Books Module Metadata - German Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Polifonia Corpus - Books Module Metadata - Spanish Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Polifonia Corpus - Books Module Metadata - Italian Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →
zenodo48/100

Censored Books during the Portuguese Estado Novo: Transcription Dataset of the Censorship Commission's Card Files (1934-74)

<p>This spreadsheet contributes to a new bibliography of censored books under the Portuguese Estado Novo dictatorial regime.</p> <p>It contains the transcription of the data fields of 1,015 card files of censored books, which are indexed by author surname in letters A and B. These files are available at the Arquivo Nacional da Torre do Tombo, in Lisbon, Portugal (PT/TT/SNI-DSC/7, &quot;Fichas de Autores de Obras Proibidas e Autorizadas&quot;, <a href="https://digitarq.arquivos.pt/details?id=4326912">https://digitarq.arquivos.pt/details?id=4326912</a>).</p> <p>The card files document data about the books censored by the Estado Novo Censorship Commission (1934-74). Data fields include file number, book report number, decision, date, author, title, origin, destination, observations, notices, and author or book process number.</p> <p>All card files have been photographed from very poor-quality photocopies and manually transcribed by &Aacute;lvaro Sei&ccedil;a during 2020/21. Letters C-Z are ongoing work and will be added to this dataset.</p> <p>This project received funding from the European Union&rsquo;s Horizon 2020 research and innovation programme under the Marie Skłodowska-Curie grant agreement no. 793147, ARTDEL.</p> <p>More info at https://artdel.net</p>

opencc-by-4.0Oct 2021View details →
zenodo48/100

Books of hours as codified compilations of compilations: dataset and code

<p><strong>Dataset, code </strong>to produce the figures, additional figures and visualizations for the following contribution:</p> <p><strong>Stutzmann</strong>, Dominique, and <strong>Chevalier</strong>, Louis, &quot;Books of hours as codified compilations of compilations Textual networks and hybrid liturgical uses&quot;, <em>Journal of Historical Network Research</em> 9 (2023) 146 &ndash; 199, doi:10.25517/jhnr.v9i1.139</p> <pre><code>@article{stutzmann_books_2023, title = {Books of hours as codified compilations of compilations {Textual} networks and hybrid liturgical uses}, volume = {9}, doi = {10.25517/jhnr.v9i1.139}, language = {en}, journal = {Journal of Historical Network Research}, author = {Stutzmann, Dominique and Chevalier, Louis}, year = {2023}, pages = {146--199}, } </code></pre> <p>Books of hours were the medieval best-seller. The referenced article aims to change how we study the textual content of books of hours by tackling the most common texts at a large scale. Intended for lay people and imitating the model of liturgical books, books of hours contain a core of votive offices and appear to have a very standardized content. The choice and the order of the chants, readings and orations may vary within the offices according to not only the liturgical destination, but also the place of production, the target export market, and the choices of the client. Variations are therefore difficult to characterize and analyze. Here, we focus on an analysis of the Hours of the Virgin and the Office of the Dead as both compilations and networks of compilations. At the level of texts and liturgical uses, we highlight and study textual commonalities based on geography or other historical links (e.g.<br> Germany, the Dominican order and Southern France for the Hours of the Virgin, Flanders and Scandinavia, Poitiers and Bordeaux, Auxerre and Bayeux for the Office of the Dead). Since patrons or copyists could also modify the expected contents, our last part analyses the uses by Utrecht and Bruges, and how uses specific to one institution may either be faithfully reproduced or give way to hybridization. This phenomenon is characterized, for example, by changes between pieces from another use into a well-identified set.</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

100 Bestselller books during COVID-19 in Spain

<p>Description of&nbsp;100 Bestselller books in Amazon&nbsp;during COVID-19 in Spain.&nbsp;</p> <p>Data for the books:&nbsp;</p> <ul> <li>Name</li> <li>Author</li> <li>Value</li> <li>Num reviews</li> <li>Category</li> <li>Price</li> <li>Pages</li> <li>Editor</li> <li>Language</li> <li>ISBN-10</li> <li>ISBN-13</li> <li>Colection</li> <li>Edad recomendada</li> </ul> <p>Data for review:&nbsp;</p> <ul> <li>Name book</li> <li>User</li> <li>Punctuation</li> <li>Title</li> <li>Date</li> <li>Text</li> <li>Vots</li> </ul>

opencc-by-4.0May 2020View details →
zenodo44/100

Historical uncertainty in Gregory of Tours's History of the Franks (book 7)

<p>Our goal was to create a research dataset based on geographical and chronological uncertainties in the work of Gregory of Tours&#39;s *History of the Franks* (book 7). We used and modified a topology of geographical and chronological uncertainty based on a rudimentary schema that would be universal when analysing an historical source :</p> <p>Chronological :<br> * uncertain dating<br> * uncertain method of dating<br> * lack of dating<br> * precise dating</p> <p>Geographical:&nbsp;<br> * uncertain location<br> * general location (region, country)<br> * lack of location&nbsp;<br> * precise location</p> <p>After working on book 7 for a while, that schema was reworked as those 9 types of uncertainty :&nbsp;</p> <p>Chronological :<br> * uncertain_dating<br> * uncertain_method_dating<br> * event_dating_null<br> * precise_dating</p> <p>Geographical:&nbsp;<br> * uncertain_location<br> * general_location&nbsp;<br> * event_location_null<br> * uncertain_method_location<br> * precise_location</p> <p><br> The geographical and chronological focus makes it possible to identify where and when, in a source, the historical uncertainty is higher.&nbsp;</p> <p>Using python, that dataset was then automatically cleaned and enhanced with bounding box based on geo-mapping information for the entries of geographical uncertainty. Those were classified as either precise_location or general_location.&nbsp;</p> <p>For example, anything relating to a city general area (like &nbsp;&quot;in the Rouen area&quot;) creates a general_location bounding box encompassing the *current* geographical space occupied by the municipality of Rouen (in the format&nbsp;&#39;LongMin&#39;, &#39;LongMax&#39;, &#39;LatMin&#39;, &#39;LatMax&#39;&nbsp;in a single column &quot;bbox&quot;). Anything described as a unique point in space (like &quot;in Paris&quot;) creates a precise_location and its corresponding lat/long system of coordinates.&nbsp;</p> <p>This is an arbitrary way to translate slightly undefined geographical concepts&nbsp;of uncertainty into formal data, but at least it can be fully explained explicitly.<br> &nbsp;&nbsp;</p> <p>Translation used:&nbsp;Tours G. <em>et alii</em>, <em>The history of the franks</em>, Penguin Books Limited, 1974, <a href="https://books.google.ch/books?id=4Lx-M2RHGgoC">https://books.google.ch/books?id=4Lx-M2RHGgoC</a>.</p>

opencc-by-4.0Jan 2021View details →
zenodo44/100

Explore maps as you read a comic book

<p>This repository contains three documents accompanying the article 'Explore maps as you read a comic book.</p> <p>The first document compiles several types of breaks encountered in various pan-scalar maps. Each type is also assigned a progression note after an analysis of their effects on the user.</p> <p>The next table illustrates an example of good generalisation for each of the hydrographic patterns observed in the article. This generalisation is based on three key concepts of progressivity, promoting a continuity of meaning, coherence, and rhythm.</p> <p>The final document represents a scale master inspired by the methodology of (Brewer and Buttenfield, 2007). In the article, we demonstrate how to adapt it to pan-scalar maps to better visualize sequences, as well as two types of rhythms: map cadence and abstraction cadence.</p>

opencc-by-4.0Dec 2023View details →
zenodo44/100

Curated Estonian National Bibliography - books

<p>This curated dataset is derived from the books subset of the Estonian National Bibliography (ENB), a comprehensive catalog of publications written in Estonian, published in Estonia, or focusing on Estonian culture and people. Designed for computational analysis, this dataset adapts the original catalog for research and cultural exploration.</p> <p>Through a systematic process of filtering, cleaning, and harmonizing, the ENB dataset is presented in a streamlined tabular format that retains rich metadata while improving accessibility. Fields selected for inclusion are harmonized and, where possible, linked to external sources, offering an optimized and reproducible resource for historical, cultural, and bibliographic research.</p>

opencc-by-4.0Nov 2024View details →
zenodo44/100

Percentage of Population Who Read Books In European Regions

<p>The indicator is created from the Eurobarometer 79.2 survey&rsquo;s <a href="https://search.gesis.org/research_data/ZA5688">GESIS datafile</a> using regional subsamples. The regional subsamples were recoded to the NUTS 2016 regional boundary definitions with the <a href="https://regions.dataobservatory.eu/">regions</a> R package. In the larger countries, where only NUTS1 level information was present (for example, in Germany and the United Kingdom), we imputed the NUTS1 territorial average values to the constituent NUTS2 regions.</p> <p>A &lsquo;dirty averaging&rsquo; was used to create regional averages, with scale national post-stratification weights to an expected value of 1. Al respondents who read at least one book in the previous 12 months were coded to have read a book.</p> <p>This indicator was used in the<br> Bal&aacute;zs Bod&oacute;, D&aacute;niel Antal, Zolt&aacute;n Puha: <a href="https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0242509">Can scholarly pirate libraries bridge the knowledge access gap?</a> An empirical study on the structural conditions of book piracy in global and European academia, in Plos ONE (Published: December 3, 2020.)</p>

opencc-by-4.0Nov 2021View details →
zenodo44/100

Gamma-ray picture book.

<p>A set of plots related to gamma-ray and hadron initated air showers.</p> <p>Note that this is a rather random selection of plots and prepared a long time ago (in 2005).</p> <p><a href="https://github.com/GernotMaier/gamma-ray-picturebook/blob/main/gamma_picturebook.pdf">gamma_picturebook.pdf</a> gives an overview of typical distributions important for ground-based gamma-ray astronomy. Additional distributions can be found in the folder <a href="https://github.com/GernotMaier/gamma-ray-picturebook/blob/main/shower-distributions">shower-distributions</a>.</p>

opencc-by-4.0Feb 2022View details →
zenodo44/100

Treebank of Artemidorus Onir. book V

<p>Collection of Ancient Greek annotated trees&nbsp;of Artemidorus&#39; Oneirocritica Book 5.&nbsp; Part of the Open Projects in Digital Classics at the College of Letters and Sciences of the State University of S&atilde;o Paulo in Araraquara, S&atilde;o Paulo, Brazil.<br> <br> The trees were annotated manually on Perseids Platform using the Arethusa tool. The treebank tagset and guidelines used were those from&nbsp;<a href="https://github.com/PerseusDL/treebank_data/blob/master/AGDT2/guidelines/Greek_guidelines.md">The Ancient Greek Dependency Treebank&nbsp;</a>with the morphological and syntactic layer, which was based on Bamman&#39;s and Crane&#39;s 2008&nbsp;<em>Guidelines for the Syntactic Annotation of the Ancient Greek Dependency Treebank (1.1).&nbsp;</em>We translated it into Portuguese with a few additions from some specifications provided by a forum maintained in 2013 by Alpheios.net,&nbsp;<a href="https://web.archive.org/web/20160401072609/http://treebank.alpheios.net/book">which is no longer online</a>. The trees are visible&nbsp;in Perseids Collection as UNESP-trees at&nbsp;<a href="https://perseids-publications.github.io/unesp-trees">https://perseids-publications.github.io/unesp-trees</a>.</p>

opencc-by-4.0Feb 2022View details →
zenodo44/100

Aligned translation of Artemidorus Onir. book V

<p>Aligned translation of Artemidorus' Oneirocritica Book 5 in 95 chapters and a prologue divided into four sections. Part of the Open Projects in Digital Classics at the College of Letters and Sciences of the State University of S&atilde;o Paulo in Araraquara, S&atilde;o Paulo, Brazil. That is a second version of the translation. It was aligned on the Ugarit Platform and is visible at <a href="http://ugarit.ialigner.com/userProfile.php?userid=15&amp;tgid=9056">https://ugarit.ialigner.com/userProfile.php?userid=15&amp;tgid=14643</a>. The Greek text source was the digitized Pack's 1963 edition from CTS Perseids: urn:cts:greekLit:tlg0553.tlg001.1st1K-grc1:5. The Portuguese text is the revised translation (urn:cts.greekLit:tlg:0553.tlg001.ferreira2:5) of a previous&nbsp;<a href="https://www.culturaacademica.com.br/catalogo/oneirokritika-de-artemidoro-de-daldis-seculo-ii-d-c/">ebook</a>&nbsp;published by Cultura Acad&ecirc;mica in 2014 (digitized and available at https://furman-editions-in-progress.github.io/UNESP_FU/ as urn:cts:greekLit:tlg0553.tlg001.ferreira1:5).</p>

opencc-by-4.0Feb 2022View details →
zenodo44/100

Video: Introducing the Open Book Collective (OBC)

<p><em>This video is a Deliverable (D2.11) of the COPIM project.</em></p> <p>Open Book Collective is a non-profit that connects academics, librarians, publishers, and service providers to collectively sustain the infrastructures, relationships, and organizations vital for the success of open access book publishing.</p> <p>Through the Open Book Collective platform, publishers and publishing service providers can offer research institutions the option to financially support their work through library membership programs. Librarians can access the platform to explore and assess different initiatives, support Open Access collections, and manage their subscriptions.</p> <p>OBC&rsquo;s legal and governance structure ensures it can&rsquo;t be co-opted for profit, and that stakeholders have a meaningful say in its future. We provide financial support to new open access initiatives and connect book publishers to sustainable revenue streams.</p> <p>Open Book Collective is helping build a world where open access books are produced and distributed collaboratively and anti-competitively, without technical or economic barriers.</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

Raw data for the book chapter "Review of Haptic and Computerized (Simulation) Games on Climate Change"

<p>Raw data used for the book chapter &quot;Gerber, A., Ulrich, M., W&auml;ger, P. (2021). Review of Haptic and Computerized (Simulation) Games on Climate Change. In: Wardaszko, M., Meijer, S., Lukosch, H., Kanegae, H., Kriz, W.C., Grzybowska-Brzezińska, M. (eds) Simulation Gaming Through Times and Disciplines. ISAGA 2019. Lecture Notes in Computer Science(), vol 11988. Springer, Cham. <a href="https://doi.org/10.1007/978-3-030-72132-9_24">https://doi.org/10.1007/978-3-030-72132-9_24</a>&quot;</p> <p>The documents include the raw data (both as .csv and .xlsx files with the same content), as well as the publication (.pdf file). The data collection process and the data itself are described in the publication. The data is published as &quot;supplementary material&quot; on the publisher&#39;s homepage.</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

Polifonia Corpus - Books Module Metadata - English Language (Full)

<p>We release the Metadata of the Books module of the Polifonia Textual Corpus. According to the availability from the source origin, the Metadata may include the URL from which a text of the Books corpus is accessible, along with the title, the author, the year of publication, and the publisher. Metadata allows for a complete reconstruction of the corpus as we cannot make the actual texts available because they are subject to heterogeneous licensing.</p> <p>Full description at <a href="http://github.com/polifonia-project/Polifonia-Corpus">https://github.com/polifonia-project/Polifonia-Corpus</a></p>

opencc-by-4.0Jun 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record