Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
545
datasets available to search
ShareScore release 0.9.0
Dataset results
545 results for “Latin”
Raw data for Infrastructure and Awareness Landscape Analysis in Latin America
<p>Persistent Identifiers (PIDs), such as Digital Object Identifiers (DOIs), are foundational to connecting and enhancing the visibility of Latin American research within a global framework. Although the region is rich in diverse and impactful research, many repositories remain only partially integrated into international registries and aggregators, limiting their discoverability and reach. The adoption of PIDs across repositories in Latin America varies widely, underscoring the need for increased awareness about the role of open PIDs in advancing research accessibility and visibility.</p> <p>This dataset offers a comprehensive overview of the current landscape of repositories, publishing systems, and Open Science policies across Latin America, shedding light on the institutional and national efforts that support an open and inclusive research infrastructure. It highlights the importance of collaboration among researchers, institutions, funders, librarians, and government agencies in fostering Open Science practices and encouraging strategic PID adoption. By expanding these open practices and strengthening PID adoption, Latin American research can achieve greater integration and impact within the global research ecosystem.</p> <p>You can read the full report titled "Infrastructure and Awareness Landscape Analysis in Latin America" at <a href="https://doi.org/10.5281/zenodo.14010858" target="_blank" rel="noopener">https://doi.org/10.5281/zenodo.14010858</a> </p>
An analysis of meiofauna knowledge generated by Latin American researchers
<p>Bibliographic databases used to analyse the document production of benthic meiofauna in Latin American countries. To be opened on R, bibliometrix package.</p> <p> </p>
Late Latin Charter Treebank 1 (LLCT1), version 1.2
<p>Version 1.2 of the Late Latin Charter Treebank 1 (LLCT1). Contains a number of minor corrections, replaces the version 1.0 published at Zenodo in 2018. Early Medieval Latin documentary texts from Italy between AD 714-869 with morphological and syntactic annotation. Latin Dependency Treebank (LDT) compatible linguistic annotation, Prague style treebank format (PML). For a detailed description of the Late Latin Charter Treebanks, see the pre-print of the paper 'Late Latin Charter Treebank: contents and annotation', to be published in Corpora, 16:2 (2021), at the <a href="https://researchportal.helsinki.fi/fi/publications/late-latin-charter-treebank-contents-and-annotation">institutional repository of the University of Helsinki</a>. See also Korkiakangas, T. and Lassila, M. (2013), <a href="https://www.academia.edu/5491302/Korkiakangas_Timo_and_Lassila_Matti_Abbreviations_fragmentary_words_formulaic_language_treebanking_mediaeval_charter_material_"><em>Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material</em></a>, in Mambrini, F., Passarotti, M. and Sporleder, C., <em>Proceedings of the third workshop on annotation of corpora for research in the humanities</em>, pp. 61–72, and Korkiakangas, T. and Passarotti, M. (2011), <a href="https://pdfs.semanticscholar.org/6825/a8ad70fe6e2a77540d9ff2774b8f34804fd0.pdf?_ga=2.234021022.1247020789.1580545874-419947753.1580545874"><em>Challenges in Annotating Medieval Latin Charters</em></a>, in «Journal of Language Technology and Computational Linguistics», 26, pp. 103–114.</p>
Test Data from a Study on Latin Vocabulary Acquisition (Cicero)
<p>The dataset contains test results from an intervention study with intermediate learners in two high schools in Berlin. In total, 58 students participated in three groups (= classes). The intervention materials and tests are published as well.</p> <p>The study was the first to collect empirical data on what German students actually know about Latin vocabulary and how they handle their vocabulary knowledge. One of the main goals of the research project is to establish a broad understanding of vocabulary knowledge in Latin lessons in Germany, which aims at a versatile education of (cross-linguistically helpful) vocabulary competence.</p>
Test Data from a Study on Latin Vocabulary Acquisition (Ovid)
<p>The dataset contains test results from an intervention study with intermediate learners in two high schools in Berlin (2018-2019). In total, 60 students participated in three groups (= classes). The intervention materials and tests are published as well.</p> <p>A key question of the still ongoing research project is: How can vocabulary competence in a historical language such as Latin be acquired and deepened by using corpus-based, i.e. context-based, methods? This question is based on a broad understanding of vocabulary that refers back to theories of the mental lexicon.</p>
Latin embeddings
<p>Lemma embeddings for Latin can be downloaded <a href="https://embeddings.lila-erc.eu/samples/download/">here</a>.</p> <p>Embeddings have been evaluated on a novel benchmark for the synonym selection task available in the file <em>syn-selection-benchmark-Latin.tsv</em>.</p> <p>You can visually explore the embeddings online: <a href="https://embeddings.lila-erc.eu/">https://embeddings.lila-erc.eu/</a>.</p> <p>These resources are licensed under a <a href="http://creativecommons.org/licenses/by-nc-sa/4.0/">Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International License</a>.</p>
List of Links to Digital Resources for Latin and Ancient Greek
<p>List of Links to Digital Resources for Latin and Ancient Greek</p> <p>The list was produced as an appendix to the German publication "Wie die Digitalisierung unseren Umgang mit den Alten Sprachen verändert hat" (How Digitization Changed the Way We Deal with Latin and Ancient Greek) in the journal "Forum Classicum", scheduled for release at the end of the year 2020.</p> <p>It contains references to various resources, such as text editions, databases, teaching materials, newspaper articles, tools for natural language processing and more. Most of them are available in English, some only in German. The list is sorted by the appearance of links in the article.</p> <p>Changelog:</p> <p>Version 2.0: Added headings from the paper to indicate topics for each part of the link list. English translations for the German headings are given in brackets.</p> <p>The list:</p> <p>Wie die Digitalisierung unseren Umgang mit den Alten Sprachen verändert hat / Linkliste (How Digitization Changed the Way We Deal with Latin and Ancient Greek / Link List)<br>A. Umgang mit der Literatur und anderen Wissensbeständen (Dealing with Literature and Other Data Collections)<br>1. Digitale Textsammlungen sind schnell verfügbar und unterstützen Lehre und Forschung. (Digital text collections are quickly accessible and support teaching as well as research.)<br>https://www.degruyter.com/view/db/btltll <br>http://stephanus.tlg.uci.edu/ <br>https://cil.bbaw.de/ <br>https://latin.packhum.org/ <br>http://cite-architecture.org/cts/ <br>https://referenceworks.brillonline.com/entries/brill-s-new-pauly/ancient-authors-and-titles-of-works-Ancient_Authors_and_Titles_of_Works <br>http://www.perseus.tufts.edu/hopper/collection?collection=Perseus:collection:Greco-Roman <br>https://tesserae.caset.buffalo.edu/<br>2. Digitale Datenbanken ermöglichen schnelle systematische Suchanfragen in großen Text- oder Informationsbeständen, auch über disziplinäre Grenzen hinweg. (Digital databases enable quick systematic queries for large collections of texts and other information, even beyond disciplinary boundaries.)<br>https://about.brepolis.net/lannee-philologique-aph/ <br>https://www.gbd.digital/metaopac/start.do?View=gnomon <br>https://referenceworks.brillonline.com/browse/brill-s-new-pauly <br>https://www.navigium.de/ <br>https://www.navigium.de/latein-unterrichten.html <br>http://lehrerportal.ccbuchner.de/Textanalyse/Default.aspx <br>https://open-educational-resources.de/ <br>https://github.com/sommerschield/ancient-text-restoration <br>3. Digitale Datenbestände werden vernetzt und für neue Anwendungszwecke kombiniert. (Digital data collections can be interconnected and combined for new use cases.)<br>https://www.w3.org/standards/semanticweb/data <br>https://lila-erc.eu/ <br>https://peripleo.pelagios.org/ <br>https://medium.com/pelagios/linked-open-data-to-navigate-the-past-using-peripleo-in-class-4286b3089bf3 <br>https://topostext.org/ <br>4. Die maschinelle sprachliche Vorverarbeitung antiker Texte erleichtert den Zugang für Lernende und Forschende. (Natural language processing of ancient texts facilitates access for both teachers and researchers.)<br>http://www.lemlat3.eu/ <br>https://d.iogen.es/ <br>https://alpheios.net/</p> <p>B. Umgang mit dem Spracherwerb (Dealing with Language Acquisition)<br>5. Die Digitalisierung fördert einen multimodalen und inklusiven Spracherwerb. (Digitization supports multimodal and inclusive language acquisition.)<br>https://www.hearinglink.org/living/loops-equipment/hearing-loops/what-is-a-hearing-loop/<br>http://www.cross-plus-a.com/balabolka.htm<br>https://propylaeum.de/e-learning/historische-aussprache-des-lateinischen-und-altgriechischen<br>https://www.youtube.com/watch?v=R5vdg_2i_pU<br>https://www.lesediagnostik.de/eye-tracking/<br>https://www.youtube.com/watch?v=8QocWsWd7fc<br>https://www.speechtexter.com/<br>https://etherpad.org/<br>https://moodle.org<br>6. Der Spracherwerb kann flexibel und personalisiert gestaltet werden. (Language acquisition can be designed in a flexible and personalized manner.)</p> <p>C. Umgang mit der Öffentlichkeit (Dealing with the Public)<br>https://www.che.de/third-mission/<br>7. Social Media ermöglichen eine schnelle Interessens- und Wissensvernetzung innerhalb und vor allem außerhalb einer definierten Gemeinschaft. (Social Media enable us to quickly connect interests and knowledge inside and especially outside of a specific community.)<br>https://la.wikipedia.org/wiki/Vicipaedia_Latina<br>http://forum.latein24.de/<br>https://twitter.com/RomAthen<br>https://www.projekte.hu-berlin.de/de/callidus/blog-2017-2018<br>https://www.superprof.de/blog/lateinische-begriffe-im-deutschen/<br>https://www.facebook.com/klassphil/?__tn__=%2Cd%2CP-R&eid=ARDXqBAnvPxAePqFMxWrKxnFG2nfqqzKDWdoHdSg1CBNwBmcZbHwF5f8IWuQZXEODH6VKzqzWvUvUzfU<br>https://www.instagram.com/fs_klassphil_tuebingen/<br>https://hu-berlin.academia.edu/MarkusAsper<br>https://www.researchgate.net/profile/Monica_Berti<br>https://www.br.de/alphalernen/faecher/latein/latein-einfach-erklaert-100.html<br>https://www.pinterest.de/pin/5418462037462026/<br>https://www.youtube.com/channel/UChB8TYnAEtSIL1mY7FuBoqA<br>https://learnattack.de/latein/saetze-uebersetzen?utm_campaign=Learnattack_Kanal&utm_source=youtube.com&utm_medium=social&utm_content=saetze-uebersetzen-latein&kanal=youtube#video-wie-du-einen-lateinischen-satz-%C3%BCbersetzt<br>https://vimeo.com/276706092<br>8. Der digitale weltweite Zugang zu und Austausch von Wissen fördert das informelle Lernen und die Open-Science-Bewegung. (The worldwide digital access to and exchange of knowledge supports informal learning and the Open Science movement.)<br>https://www.udemy.com/course/an-introduction-to-classical-latin/<br>https://www.coursera.org/learn/roman-architecture<br>https://www.coursera.org/learn/plato<br>https://www.youtube.com/channel/UCNW1n7ctSkW3cgYFCzKPK3A/videos<br>https://scholar.google.de/<br>https://www.kim.uni-konstanz.de/openscience/onlinekurs-open-science-von-daten-zu-publikationen/<br>https://www.go-fair.org/fair-principles/<br>https://zenodo.org/record/3601182<br>https://zenodo.org/record/3816709<br>https://scm.cms.hu-berlin.de/callidus<br>https://www.ianus-fdz.de/<br>https://opr.degruyter.com/<br>http://ahropenreview.com/<br>https://arxiv.org/help/trackback<br>https://www.propylaeum.de/<br>https://journals.ub.uni-heidelberg.de/index.php/dco/index<br>http://www.pegasus-onlinezeitschrift.de/<br>https://www.schule-bw.de/faecher-und-schularten/sprachen-und-literatur/latein<br>https://www.schule-bw.de/faecher-und-schularten/sprachen-und-literatur/griechisch<br>https://www.bmbf.de/de/citizen-science-wissenschaft-erreicht-die-mitte-der-gesellschaft-225.html<br>https://pleiades.stoa.org/home</p> <p>Fazit (Conclusion)<br>http://pom.bbaw.de/cmg/</p>
Accompanying dataset; 'Agroforestry enhances biological activity, diversity and soil-based ecosystem functions in mountain agroecosystems of Latin America: A meta-analysis.'
<p>The database created as part of the meta-analysis is designed to facilitate the comparison of biological activity, diversity (BIAD), and ecosystem functions (EFs) between agroforestry systems (AFS) and other land-use types. It incorporates data extracted from selected studies, each record comprising a mean value, sample size, and a variance measure to compute standard deviation. The database also categorizes data according to 22 explanatory variables, including geographical coordinates, climate classification, soil type, AFS classification, and more, to characterize the sites and management systems involved. This detailed classification enables a nuanced analysis of how different factors might influence the BIAD and EFs in the context of AFS. The database supports the meta-analysis by allowing for the estimation of effect sizes using response ratios, which compare the relative difference in BIAD and EFs between AFS and other land uses. Data extraction from primary studies was meticulous, employing both direct and indirect methods such as graph digitizing software, and missing data were supplemented using reliable sources or direct communication with the original study authors. The comprehensive nature of this database ensures that the analysis can account for a wide range of variables that may affect the outcomes of interest in the meta-analysis. </p><p>For an in-depth exploration of the study's findings and methodology, refer to the comprehensive meta-analysis available in Global Change Biology (2024), entitled "<i>Agroforestry Enhances Biological Activity, Diversity, and Soil-Based Ecosystem Functions in Mountain Agroecosystems of Latin America: A Meta-Analysis</i>."</p>
Wikipedia: wikipedia-la (Latin)
Wikipedia is a multilingual, web-based, free-content encyclopedia project supported by the Wikimedia Foundation and based on a model of openly editable content. EOL harvests articles from wikipedia that are indexed as species or higher taxa.<p></p><p></p>https://la.wikipedia.org/
Revised database of the Soil Information System of Latin America and the Caribbean, SISLAC
<p>This dataset contains the revised version of the SISLAC database in three formats: comma-separated values (.csv), microsoft access (.mdb) and PostGIS database (.backup). This database was reviewed and the inconsistencies found in the profiles and in the description of their horizons were corrected. Consists of two tables, one for the description of the profiles and the other with the description of the horizons and their properties. The key field between both tables is the profile identifier, column <strong><em>profile_id</em></strong>.</p>
Dataset for "Authorship concentration in health sciences journals from Latin America and the Caribbean"
<p>Authorship concentration indexes and other data for journals in the LILACS (Latin American and the Caribbean Literature on Health Sciences) bibliographic database, from 2015 to 2019 (FONTENELLE, 2022). These data are read and created by <a href="https://doi.org/10.5281/zenodo.6127497">analytic code in Zenodo</a>.</p> <ul> <li><em>authorship_concentration.csv</em> - dataset derived in Fontenelle (2022) from raw data exported from LILACS. This is the main file, and it's CC-BY because other researchers might have curated the raw data differently and thus derived different data. CSV file encoded with ASCII.</li> <li><em>authorship_concentration_datadictionary.csv</em> - data dictionary for the previous file. This file is actually CC0. CSV file encoded with ASCII.</li> <li><em>journals.csv</em> - dataset about the journals indexed in LILACS between 2015 and 2019. As a result of simply converting and filtering the original TITLE database, this is actually CC0 by the Pan American Health Organization (PAHO). CSV file encoded with UTF-8.</li> <li><em>journal_subjects.csv</em> - DeCS descriptors for the journals identified by the ISSN. CC0 by the Pan American Health Organization (PAHO), as above. CS file encoded with ASCII.</li> </ul>
European Investment Bank Projects in ACP, OCT, Africa, Asia, and Latin America (1957-2024)
<p>This dataset offers a comprehensive analysis of European Investment Bank (EIB) projects in Africa, the Caribbean, and the Pacific (ACP) regions, Overseas Countries and Territories (OCT), Asia, and Latin America, spanning from 1975 to 2023. The dataset includes information on 2,558 projects; each entry in the dataset includes key project details such as the project’s sector, date of signature, and financial commitments. All numbers are in 2015 euros.</p>
Revised database of the Soil Information System of Latin America and the Caribbean, SISLAC version 1.2
<p>The SISLAC_database version 1.2 contains the revised version of the SISLAC database in comma-separated values (csv) format. This database was reviewed and the inconsistencies found in the profiles and in the description of their horizons were corrected. The key field between both tables is the profile identifier, column <strong><em>profile_id</em></strong>.</p>
Translation Alignment: Ancient Greek to Latin. Annotation Style Guide and Gold Standard
<p>This dataset contains guidelines and a gold standard for the alignment of Ancient Greek texts with Latin scholarly translations. </p> <p>The gold standard consists of 100 fragments randomly selected from the <em>Digital Fragmenta Historicorum Graecorum </em>(DFHG) (https://www.dfhg-project.org/), which were aligned manually by Chiara Palladino and David J. Wright using Ugarit (https://ugarit.ialigner.com/). The Annotation Style Guide was developed for this project. The resulting Inter-Annotator-Agreement (IAA) is 90.5%. </p> <p>The materials available in this repository can be used to perform and evaluate alignments of various texts in Ancient Greek, to create gold standards, and to train automated translation alignment models. </p> <p>The Guidelines can be further adapted to address similar language pairs including inflected languages, or can provide a structure for the alignment of other historical texts against modern translations. However, the guidelines are not project-specific: they were specifically intended for the scenario of machine translation. Different research questions, such as translation history or pedagogy, may need further tweaking of these guidelines. </p> <p>For further information on Ugarit and translation alignment of historical languages, see http://ugarit.aligner.com/bib.php and follow us on Twitter (@ugarit_ty). <br> </p> <p> </p>
Fusarium associated with Banana - DArT-seq Cuban and Latin-American samples
<p>Using genotyping-by-sequencing and whole genome comparisons, we investigated the genetic diversity across this suite of isolates and compared it with the genetic diversity in a global <em>Fusarium</em> panel.</p>
ERA5-Land selected indicators daily aggregates for the Latin America region, 1951
<p>This deposit contains NetCDF files with daily aggregates from Copernicus Era5-Land eight selected indicators, covering the Latin America region, for 1951.</p><p>Each file represents one indicator aggregation for one month of the year. Inside each NetCDF file, the layers contain the daily aggregates.</p><p>For 2m dewpoint pressure, 10m u component of wind, 10m v component of wind, surface pressure, the mean function was used for aggregation. For total precipitation, the sum function was used for aggregation. For 2m temperature, the functions maximum, mean and minimum were used for aggregation.</p><p>Those files were created using the <a href="https://github.com/ErikKusch/KrigR">KrigR</a> package.</p>
ERA5-Land selected indicators daily aggregates for the Latin America region, 1950
<p>This deposit contains NetCDF files with daily aggregates from Copernicus Era5-Land eight selected indicators, covering the Latin America region, for 1950.</p><p>Each file represents one indicator aggregation for one month of the year. Inside each NetCDF file, the layers contain the daily aggregates.</p><p>For 2m dewpoint pressure, 10m u component of wind, 10m v component of wind, surface pressure, the mean function was used for aggregation. For total precipitation, the sum function was used for aggregation. For 2m temperature, the functions maximun, mean and minimum were used for aggregation.</p><p>Those files were created using the <a href="https://github.com/ErikKusch/KrigR">KrigR</a> package.</p>
ERA5-Land selected indicators daily aggregates for the Latin America region, 1959
<p>This deposit contains NetCDF files with daily aggregates from Copernicus Era5-Land eight selected indicators, covering the Latin America region, for 1959.</p><p>Each file represents one indicator aggregation for one month of the year. Inside each NetCDF file, the layers contain the daily aggregates.</p><p>For 2m dewpoint pressure, 10m u component of wind, 10m v component of wind, surface pressure, the mean function was used for aggregation. For total precipitation, the sum function was used for aggregation. For 2m temperature, the functions maximum, mean and minimum were used for aggregation.</p><p>Those files were created using the <a href="https://github.com/ErikKusch/KrigR">KrigR</a> package.</p>
ERA5-Land selected indicators daily aggregates for the Latin America region, 1956
<p>This deposit contains NetCDF files with daily aggregates from Copernicus Era5-Land eight selected indicators, covering the Latin America region, for 1956.</p><p>Each file represents one indicator aggregation for one month of the year. Inside each NetCDF file, the layers contain the daily aggregates.</p><p>For 2m dewpoint pressure, 10m u component of wind, 10m v component of wind, surface pressure, the mean function was used for aggregation. For total precipitation, the sum function was used for aggregation. For 2m temperature, the functions maximum, mean and minimum were used for aggregation.</p><p>Those files were created using the <a href="https://github.com/ErikKusch/KrigR">KrigR</a> package.</p>
ERA5-Land selected indicators daily aggregates for the Latin America region, 1957
<p>This deposit contains NetCDF files with daily aggregates from Copernicus Era5-Land eight selected indicators, covering the Latin America region, for 1957.</p><p>Each file represents one indicator aggregation for one month of the year. Inside each NetCDF file, the layers contain the daily aggregates.</p><p>For 2m dewpoint pressure, 10m u component of wind, 10m v component of wind, surface pressure, the mean function was used for aggregation. For total precipitation, the sum function was used for aggregation. For 2m temperature, the functions maximum, mean and minimum were used for aggregation.</p><p>Those files were created using the <a href="https://github.com/ErikKusch/KrigR">KrigR</a> package.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.