Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5
datasets available to search
ShareScore release 0.9.0
Dataset results
5 results for “French Novel”
Collection de romans français du dix-huitième siècle (1750-1800) / Collection of Eighteenth-Century French Novels (1750-1800)
<p><strong>Key information</strong>: This collection of Eighteenth-Century French Novels contains digital texts of novels created or first published between 1751 and 1800. The collection is created in the context of Mining and Modeling Text, a project at the Trier Center for Digital Humanities (TCDH) at Trier University, Germany (2019-2023). The current release contains 200 novels.</p><p><strong>Further information</strong>: <a href="https://github.com/MiMoText/roman18">https://github.com/MiMoText/roman18</a></p><p><strong>Citation suggestion: </strong><i>Collection de romans français du dix-huitième siècle (1751-1800) / Eighteenth-Century French Novels (1751-1800)</i>, edited by Julia Röttgermann, with contributions from Julia Dudar, Henning Gebhard, Anne Klee, Johanna Konstanciak, Damir Padieu, Amelie Probst, Sarah Rebecca Ondraszek and Christof Schöch. Release v.1.2.0. Trier: TCDH, 2023. URL: <a href="https://github.com/mimotext/roman18">https://github.com/mimotext/roman18</a>; DOI: <a href="https://doi.org/10.5281/zenodo.10349902">10.5281/zenodo.10349902</a>. </p><p> </p><p> </p>
Collection de romans français du dix-huitième siècle (1751-1800) / Collection of Eighteenth Century French Novels 1751-1800
<p>This collection of Eighteenth-Century French Novels contains 200 digital texts of novels created or first published between 1751 and 1800. The collection is created in the context of <a href="https://www.mimotext.uni-trier.de/en">Mining and Modeling Text</a> (2019-2023), a project which is located at the Trier Center for Digital Humanities (<a href="https://tcdh.uni-trier.de/en">TCDH</a>) at Trier University.</p> <h2>Metadata</h2> <p>There is a short and an extensive metadata description in TSV for all TEI/XML files:</p> <ul> <li>Metadata, short version: <a href="https://github.com/MiMoText/roman18/blob/master/XML-TEI/xml-tei_metadata.tsv">https://github.com/MiMoText/roman18/blob/master/XML-TEI/xml-tei_metadata.tsv</a></li> <li>Metadata, long version: <a href="https://github.com/MiMoText/roman18/blob/master/XML-TEI/xml-tei_full_metadata.tsv">https://github.com/MiMoText/roman18/blob/master/XML-TEI/xml-tei_full_metadata.tsv</a></li> </ul> <p>Please find further information on our <a href="https://github.com/MiMoText/roman18/tree/v1.2.0">corpus balancing</a> .</p> <h2>Licence</h2> <p>All texts and scripts are in the public domain and can be reused without restrictions. We don't claim any copyright or other rights on the transcription, markup or metadata. If you use our texts, for example in research or teaching, please reference this collection using the citation suggestion below.</p> <h2>Citation suggestion</h2> <p><em>Collection de romans français du dix-huitième siècle (1751-1800) / Eighteenth-Century French Novels (1751-1800)</em>, edited by Julia Röttgermann, with contributions from Julia Dudar, Henning Gebhard, Anne Klee, Johanna Konstanciak, Damir Padieu, Amelie Probst, Sarah Rebecca Ondraszek and Christof Schöch. Release v 1.2.1. Trier: TCDH, 2023. URL: https://github.com/mimotext/roman18. DOI: https://doi.org/10.5281/zenodo.4061903.</p> <h2>Funding</h2> <p>Forschungsinitiative des Landes Rheinland-Pfalz 2019-2023</p>
Word2Vec Models built from a Collection of French 20th-Century Novels
<p>The models were trained using the Gensim library for Python, developed by Radim Rehurek, in 2017. All models are based on the same collection of 20th century French novels that covers the period from 1900 to 2010, with a large range of authors and genres respresented. The collection contains approximately 1,200 novels and about 60 million tokens.</p> <p>The models were created using the SGNS (Skip-Gram with Negative Sampling) architecture, the context window was always of size of 6 + 6 around the target word, and the texts were lemmatised and POS-tagged beforehand. POS-Tags remain attached to each token (as in "souris_nom"). Other parameters vary by model: some have 200, some have 300 dimensional vectors; the minimum frequency of the words in the model varies with values of 50, 100 and 200, something which influences the size of the vocabulary and the size of the model. </p>
French Novel Collection (ELTeC-fra) for TXM
<p>TXM annotated binary corpus of the French Novel Collection (ELTeC-fra) in the "European Literary Text Collection" (ELTeC) produced by the COST Action "Distant Reading for European Literary History" (CA16204). The current version is based on v1.0.0 of the French Novel Collection (ELTeC-fra).</p> <p>Authors represented in this collection (see metadata table for details): Amédée Achard; Juliette Adam; Gustave Aimard; Alphonse Allais; Marguerite Audoux; Honoré de Balzac; Henri Barbusse; Maurice Barrès; René Bazin; Stella Blandy; Adèle Blondel; Fortuné du Boisgobey; Paul Bourget; Zulma Carraud; Claide de Chandeneux; Michel Corday; Georges Darien; Comtesse Dash; Alphonse Daudet; Lucie de Delarue-Mardrus; Xavier de Montépin; Roger Dombre; Alexandre (père) Dumas; Georges Eekhoud; Émile Erckmann; Octave Feuillet; Paul Henri Corentin Féval; Gustave Flaubert; Zénaïde Fleuriot; Anatole France; Émile Gaboriau; Arnould Galopin; Judith Gautier; Théophile Gautier; Marie Françoise Sophie Gay; Jehan Gilbert; Delphine de Girardin; Julie Gouraud; Henry Gréville; Gyp (Sibylle Riquetti de Mirabeau); Paul de Kock; Maurice Leblanc; Jules Lermina; Gustave Le Rouge; Gaston Leroux; Daniel Lesueur; Pierre Loti; Hector Malot; Auguste Maquet; Guy de Maupassant; Catulle Mendès; Pierre Mille; Octave Mirbeau; Édouard Montagne; Emile Moselly; Anna de Noailles; Georges Ohnet; Joséphin Péladan; Pierre Alexis Ponson du Terrail; Marcel Proust; Marie Rattazzi; Hugues Rebell; Fanny Reybaud; Romain Rolland; Jules Sandeau; George Sand; Daniel Stern; Madame de Stolz; Mario Uchard; Victor Vaillant; Cécile de Valgand; Jules Verne; Horace de Viel-Castel; Claude Vignon.</p>
French contemporary novel subgenre samples
<p>Set of 30.000 short samples from 600 French novels first published between 1970 and 1999, and categorized by author, year of publication and subgenre. </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.