Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

192

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

192 results for “Websites”

Learn how ShareScore rates datasets ↗
zenodo48/100

The website 'La lengua maya de Yucatán'

<p>Presentation/demo of the website 'La lengua maya de Yucatán' (https://www.christianlehmann.eu/ling/sprachen/maya/index.php): Yucatec Maya description online, onomasiological grammar, semasiological grammar, comments on the description, lexical database, texts.</p>

opencc-by-4.0Dec 2023View details →
zenodo48/100

Curlie Enhanced with LLM Annotations: Two Datasets for Advancing Homepage2Vec's Multilingual Website Classification

<h3>Advancing Homepage2Vec with LLM-Generated Datasets for Multilingual Website Classification</h3> <p>This dataset contains two subsets of labeled website data, specifically created to enhance the performance of Homepage2Vec, a multi-label model for website classification. The datasets were generated using Large Language Models (LLMs) to provide more accurate and diverse topic annotations for websites, addressing a limitation of existing Homepage2Vec training data.</p> <p><strong>Key Features:</strong></p> <ul> <li><strong>LLM-generated annotations:</strong>&nbsp;Both datasets feature website topic labels generated using LLMs,&nbsp;a novel approach to creating high-quality training data for website classification models.</li> <li><strong>Improved multi-label classification:</strong> Fine-tuning Homepage2Vec with these datasets has been shown to improve its macro F1 score from 38% to 43% evaluated on a human-labeled dataset, demonstrating their effectiveness in capturing a broader range of website topics.</li> <li><strong>Multilingual applicability:</strong> The datasets facilitate classification of websites in multiple languages, reflecting the inherent multilingual nature of Homepage2Vec.</li> </ul> <p><strong>Dataset Composition:</strong></p> <ul> <li><strong>curlie-gpt3.5-10k:</strong> 10,000 websites labeled using GPT-3.5, context 2 and 1-shot</li> <li><strong>curlie-gpt4-10k:</strong> 10,000 websites labeled using GPT-4, context 2 and zero-shot</li> </ul> <p><strong>Intended Use:</strong></p> <ul> <li>Fine-tuning and advancing Homepage2Vec or similar website classification models</li> <li>Research on LLM-generated datasets for text classification tasks</li> <li>Exploration of multilingual website classification</li> </ul> <p><strong>Additional Information:</strong></p> <ul> <li><strong>Project and report repository:</strong> https://github.com/CS-433/ml-project-2-mlp</li> </ul> <p><strong>Acknowledgments:</strong></p> <p>This dataset was created as part of a project at EPFL's Data Science Lab (DLab) in collaboration with <a href="https://people.epfl.ch/robert.west">Prof. Robert West</a> and <a href="https://tizianopiccardi.github.io/" rel="nofollow">Tiziano Piccardi.</a></p>

opencc-by-4.0Dec 2023View details →
zenodo48/100

UK Parliament Petition Website: Hourly count data

<p>This dataset contains hourly counts of the number of signatures on each petition posted to the UK Parliament petition website from&nbsp;2015-07-20 to&nbsp;2016-09-12. The file is gzip compressed text data in tab-separated values format with four columns:</p> <ul> <li>db_id: An internal id number</li> <li>pet_id: The id of the petition on the Parliament website (e.g., 131215 corresponds to the petition https://petition.parliament.uk/archived/petitions/131215 )</li> <li>sigs: The number of signatures observed at datetime.</li> <li>datetime: The date and time of the observation&nbsp;in YYYY-MM-DD HH:mm:SS format (e.g., 2015-07-20 18:24:35)</li> </ul> <p>This data was collected via a Python scrapping script and initially stored in a MySQL database.</p>

opencc-by-nc-sa-4.0Jun 2019View details →
zenodo48/100

Dataset for Website Personality Detection

<p>This dataset supports research on identifying the personality of websites. It contains data from 3,000 websites, covering five distinct website categories, and provides quantitative elements extracted from these sites. Additionally, the dataset includes information about the selected website categories, as well as details on website personality "Facets" and "Items." The dataset is accompanied by survey results related to the research.</p>

opencc-by-4.0Nov 2024View details →
zenodo44/100

Indexed scientific articles for the seven journals listed on the Design Society website and all papers indexed for the DESIGN and ICED conferences

<p>Includes all articles indexed by Scopus® for the seven journals listed on the Design Society website and all papers indexed for DESIGN and ICED (accessed on 05/Nov/2016).</p> <p>The full search query used to extract the articles is:</p> <p>"( SRCTITLE ( "Intelligence for Engineering Design, Analysis and Manufacturing: AIEDAM"  OR  "Journal of Engineering Design"  OR  "Design Studies"  OR  "Research in Engineering Design"  OR  "CoDesign"  OR  "Journal of Design Research"  OR  "Design Science Journal"  OR  "The International Journal of Design Creativity and Innovation"  OR  "International Conference on Engineering Design"  OR  "INTERNATIONAL DESIGN CONFERENCE" ) )  AND  ( "data collection"  OR  "data acquisition"  OR  "data source"  OR  database  OR  "empirical data"  OR  "empirical grounding"  OR  interview  OR  documents  OR  "data logs"  OR  "case study"  OR  observation  OR  experiment*  OR  "empirical finding*"  OR  "empirical result*" )  AND  ( EXCLUDE ( EXACTSRCTITLE ,  "Hardware Software Codesign Proceedings Of The International Workshop" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Journal Of Engineering Design And Technology" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Research In Engineering Design Theory Applications And Concurrent Engineering" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Chinese Journal Of Engineering Design" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Codes Isss 2005 International Conference On Hardware Software Codesign And System Synthesis" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Codes Isss 12 Proceedings Of The 10th ACM International Conference On Hardware Software Codesign And System Synthesis Co Located With Esweek" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Codes Isss 2006 Proceedings Of The 4th International Conference On Hardware Software Codesign And System Synthesis" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Codes Isss 2007 International Conference On Hardware Software Codesign And System Synthesis" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Embedded Systems Week 2008 Proceedings Of The 6th IEEE ACM IFIP International Conference On Hardware Software Codesign And System Synthesis Codes Isss 2008" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Second IEEE ACM IFIP International Conference On Hardware Software Codesign And Systems Synthesis Codes Isss 2004" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "Embedded Systems Week 2011 Esweek 2011 Proceedings Of The 9th IEEE ACM IFIP International Conference On Hardware Software Codesign And System Synthesis Codes Isss 11" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "2010 IEEE ACM IFIP International Conference On Hardware Software Codesign And System Synthesis Codes Isss 2010" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "2013 International Conference On Hardware Software Codesign And System Synthesis Codes Isss 2013" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "2014 International Conference On Hardware Software Codesign And System Synthesis Codes Isss 2014" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "2015 ACM IEEE International Conference On Formal Methods And Models For Codesign Memocode 2015" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "8th ACM IEEE International Conference On Formal Methods And Models For Codesign Memocode 2010" ) )  AND  ( EXCLUDE ( EXACTSRCTITLE ,  "2015 International Conference On Hardware Software Codesign And System Synthesis Codes Isss 2015" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "9th ACM IEEE International Conference On Formal Methods And Models For Codesign Memocode 2011" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "11th ACM IEEE International Conference On Formal Methods And Models For Codesign Memocode 2013" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "10th ACM IEEE International Conference On Formal Methods And Models For Codesign Memocode 2012" )  OR  EXCLUDE ( EXACTSRCTITLE ,  "A Practical Introduction To Hardware Software Codesign" ) )"</p> <p>The keywords used in the article "DATA-DRIVEN ENGINEERING DESIGN RESEARCH: OPPORTUNITIES USING OPEN DATA" are:</p> <p>"data collection" OR "data acquisition" OR "data source" OR database OR "empirical data" OR "empirical grounding" OR interview OR documents OR "data logs" OR "case study" OR observation OR experiment* OR "empirical finding*" OR "empirical result*"</p>

opencc-by-4.0Nov 2016View details →
zenodo44/100

Web requests analysis of Italy websites which use Google Analytics

<p>List of 504,038 domains of Italy found to contain Google Analytics.</p> <p>The front page for Italy-related domain names has been accessed through HTTPS or HTTP and analysed with webbkoll and jq to gather data about third-party requests, cookies and other privacy-invasive features. Together with the actual URL visited, the user/property ID is provided for 495,663 domains (extracted either from the cookies deposited or the URL of requests to Google Analytics). MX and TXT records for the domains are also provided.</p> <p>The most common ID found was 23LNSPS7Q6, with over 35k domains calling it (seemingly associated with italiaonline.it). The most common responding IP addresses were 3 AWS IPv4 addresses (over 40k domains) and 2 CloudFlare IPv6 addresses (over 12k domains).</p>

opencc-zeroJul 2022View details →
zenodo44/100

Perceptions on the utility of community question and answer websites like Stack Overflow to software developers (Replication package)

<p>Interview Questions on the perception of the utility of CQAs like Stack Overflow to software developers. In this study, we focused on the questions highlighted in yellow.</p>

opencc-by-4.0Mar 2022View details →
zenodo44/100

115th U.S. Congress Member Website (Full JavaScript-enabled Scrape) Collection

<p>This data set represents a point-in-time full JavaScript-enabled scrape of all available 115th U.S. Congress member web sites. The data collection originated and completed on 2018-04-13 and the results are in ndjson/jsonlines/streaming JSON format. File format information is in the enclosed README.md file.</p> <p>The data was used to evaluate the privacy profiles of each U.S. Congress members&#39; official (.gov hosted) websites for the discussion in &lt;https://rud.is/b/2018/04/13/does-congress-really-care-about-your-privacy/&gt;.</p> <p>ScrapingHub&#39;s &quot;Splash&quot; platform (&lt;https://github.com/scrapinghub/splash&gt;) was used along with the &quot;splashr&quot; R package (&lt;https://github.com/hrbrmstr/splashr&gt;) to retrieve the content.</p>

opencc-by-4.0Apr 2018View details →
zenodo44/100

Recognising innovative companies by using a diversified stacked generalisation method for website classification – the raw results

<p><strong>Introduction</strong></p> <p>The classification models were trained out by using the Classification and Regression Training package (caret) [1]. The models&#39; parameters were fine-tuned by the 10-fold cross-validation procedure [2].</p> <p><strong>Cluster parameters</strong></p> <p>Most computations were carried out on a cluster having the following parameters:</p> <ul> <li>GPU: NVIDIA Tesla P100;</li> <li>CPU: 2.0 GHz Intel&reg; Xeon&reg; Platinum 8167M;</li> <li>The number of GPUs: 2;</li> <li>The number of CPU cores: 28;</li> <li>The number of CPU threads: 56;</li> <li>RAM: 192 GB;</li> <li>Storage: 3 TB.</li> </ul> <p>Only one model (k-nn) was calculated on a cluster having the following parameters:</p> <ul> <li>Processor: Intel(R) Core(TM) i7-4770 CPU @ 3.40GHz 3.40 GHz;</li> <li>RAM: 16 GB;</li> <li>Windows 64 bit.</li> </ul> <p><strong>Performance statistics</strong></p> <p>All performance statistics are stored in cvs files. Each file corresponds to a particular machine learning method such as a file, &quot;methodName-stat.csv&quot; contains all data regarding a method, &quot;methodName.&quot; All files cover the following columns:</p> <ul> <li><em>dataSetName &ndash; </em>a name of a data set on which evaluation was carried out; there are three possible values: (i) <em>firstPages</em> refers to the first data set (<em>L<sub>D</sub></em>) that contains textual description of a company; (ii) &nbsp;<em>firstPageLabels</em> refers to the second data set (<em>L<sub>L</sub></em>) that involves link labels that were extracted from an index page; (iii) <em>aggregateDocument</em> refers to the third data set (<em>L<sub>B</sub></em>) that consists of a so-called big document;</li> <li><em>fmeasure</em> - the number of features that were taken into account during &nbsp;evaluation;</li> <li><em>method</em>&nbsp; - the name of function in the caret package;</li> <li><em>parameters</em> - the values of parameters received from a tuning phase of a given classification method;</li> <li><em>precision </em>&ndash; the value of method&rsquo;s precision;</li> <li><em>recall </em>&ndash; the value of method&rsquo;s recall;</li> <li><em>fmeasure</em>&nbsp; - the value of method&rsquo;s F-measure;&nbsp;</li> <li><em>error</em> - the value of method&rsquo;s error;</li> <li><em>acc </em>&ndash; the value of method&rsquo;s.</li> </ul> <p><strong>Time processing statistics</strong></p> <p>All time processing statistics, like the performance statistics, are stored in cvs files. Each file corresponds to a particular machine learning method such as a file, &quot;methodName-time.csv&quot;. All files cover the following columns:</p> <ul> <li><em>dataSetName &ndash; </em>a name of a data set on which evaluation was carried out; there are three possible values: (i) <em>firstPages</em> refers to the first data set (<em>L<sub>D</sub></em>) that contains textual description of a company; (ii) &nbsp;<em>firstPageLabels</em> refers to the second data set (<em>L<sub>L</sub></em>) that involves link labels that were extracted from an index page; (iii) <em>aggregateDocument</em> refers to the third data set (<em>L<sub>B</sub></em>) that consists of a so-called big document;</li> <li><em>featureNo</em> - the number of features that were taken into account during &nbsp;evaluation;</li> <li><em>method</em>&nbsp; - the name of function in the caret package;</li> <li><em>user</em> - user time elapsed for executing a <em>method</em> as an R process;</li> <li><em>system</em>&nbsp; - system time elapsed for executing a <em>method</em> as an R process;</li> <li><em>elapsed</em> - total time elapsed for executing a <em>method</em> as an R process.</li> </ul> <p>For more information about user, system and total elapsed time, please see documentation [3].</p> <p><strong>References</strong></p> <p>[1] https://cran.r-project.org/web/packages/caret/</p> <p>[2] https://topepo.github.io/caret/model-training-and-tuning.html</p> <p>[3] https://stat.ethz.ch/R-manual/R-devel/library/base/html/proc.time.htm</p>

opencc-by-4.0Jan 2019View details →
zenodo44/100

South African Disinformation [Fake News] Website Data - 2020

<p>See publication:&nbsp;<strong>Is it Fake? News Disinformation Detection on South African News Websites</strong></p> <p>We used, as sources, investigations by the news websites MyBroadband (<a href="https://mybroadband.co.za/forum/threads/list-of-known-fake-news-sites-in-south-africa-and-beyond.879854/">https://mybroadband.co.za/forum/threads/list-of-known-fake-news-sites-in-south-africa-and-beyond.879854/</a>) and News24 (<a href="https://exposed.news24.com/the-website-blacklist/">https://exposed.news24.com/the-website-blacklist/</a>). These articles covered investigations into disinformation websites in South Africa in 2018. They compiled lists of websites that were suspected to be disinformation. During the period from those articles to present, a number of the websites have become inaccessible or offline. We attempted to use the internet archives <a href="https://archive.org/web/">WayBack Machine</a>&nbsp;we could only get partial snapshots and error messages.</p> <p>A web-scraper only worked for one of the sources although manual editing was still required to clean the text from Javascript code and some paragraph duplicates. On most of the other websites, a web-scraper did not work well as there were too many advertisements and broken parts of pages. Because of all these problems, most of the articles were manually copied and pasted and cleaned in flat files. In some cases, the text of articles could not be copied and was not made part of the South African disinformation corpus.</p> <p><strong>Citing the dataset</strong></p> <blockquote> <p>@inproceedings{de2021fake, title={Is it Fake? News Disinformation Detection on South African News Websites}, author={de Wet, Harm and Marivate, Vukosi}, booktitle={2021 IEEE AFRICON}, pages={1--6}, year={2021}, organization={IEEE} }</p> </blockquote>

opencc-by-sa-4.0Jul 2021View details →
zenodo44/100

Analysis of the content of the H2020 project websites related to LEAs and IA.

<p>This is the dataset used in the article entitled &quot;The disconnect between the goals of trustworthy AI for law enforcement and the EU research agenda&quot;. You can find more information about the results obtained, as well as the methodology used in the paper.</p>

opencc-by-4.0Dec 2022View details →
zenodo44/100

Official-Websites-of-the-Portuguese-Municipalities: v1.0.0

<p>This repository contains a list of all 308 Portuguese municipalities, with their respective district and website. The information presented here was based on the update of January 19, 2023, and comes from the website of the General Directorate of Local Authorities (DGAL), which is the central service of the direct administration of the State integrated into the Ministry of Territorial Cohesion of Portugal [1]. Data collection took place on January 26, 2023. We also carried out a manual review of all the website addresses on the list, where we identified and corrected some incorrect email addresses through a Google search of the corresponding municipality.</p>

openother-openJan 2023View details →
zenodo44/100

Phishing Website Dataset

<p>This dataset contains a collection of legitimate and phishing websites, along with information on the target brands (brands.csv) being impersonated in the phishing attacks. The dataset includes a total of 10,395 websites, 5,244 of which are legitimate and 5,151 of which are phishing websites. These websites impersonate a total of 86 different target brands.</p> <p>For phishing datasets, the files can be downloaded in a zip file with a &quot;phishing&quot; prefix, while for legitimate websites, the files can be downloaded in a zip file with a &quot;not-phishing&quot; prefix.<br> <br> In addition, the dataset includes features such as screenshots, text, CSS, and HTML structure for each website, as well as domain information (WHOIS data), IP information, and SSL information. Each website is labeled as either legitimate or phishing and includes additional metadata such as the date it was discovered, the target brand being impersonated, and any other relevant information.</p> <p>The dataset has been curated for research purposes and can be used to analyze the effectiveness of phishing attacks, develop and evaluate anti-phishing solutions, and identify trends and patterns in phishing attacks. It is hoped that this dataset will contribute to the advancement of research in the field of cybersecurity and help improve our understanding of phishing attacks.</p>

opencc-by-4.0Jul 2023View details →
zenodo40/100

Artifact for the EASE 2020 Paper: How Can I Contribute? A Qualitative Analysis of Community Websites of 25 Unix-Like Distributions

<p>Artifact for the EASE 2020 Paper: How Can I Contribute? A Qualitative Analysis of Community Websites of 25 Unix-Like Distributions</p> <p>Jacob Kr&uuml;ger, Sebastian Nielebock, Robert Heum&uuml;ller</p> <p>&nbsp;</p> <p>Please refer to the readme for more details</p>

opencc-by-4.0Feb 2020View details →
zenodo40/100

Applications on the F-Droid Website (May-June 2019)

<p>The applications&nbsp;on the f-droid website (https://f-droid.org/en/). These applications were scraped around May-June 2019</p>

opencc-by-4.0Mar 2020View details →
zenodo40/100

Agenda Cultural Alcobendas - Open data website

<p>For this dataset, the data set of Agenda Cultural has been chosen, accessible through the following link: https://datos.alcobendas.org/dataset/2458d605-1fd0-4d16-a1ef-795bc243e152/resource/90836555-54b5-45af-ae67-e7545937f591/download/recurso.json</p> <p>We have chosen Alcobendas open data website: https://datos.alcobendas.org/ as it provides the exposed data sets are offered under open property licenses.</p> <p>For this dataset some of the fields provided by the json web have been selected: Tem&aacute;ticas, Barrio, Nombre del evento, Subtemas, FechaInicio, HoraInicio, FechaFin, HoraFin, Tipos, Perfiles y URL_Ficha.</p>

openodc-byApr 2020View details →
zenodo40/100

Colophons and Scholars website source tables

<p>These excel tables are the source files for the projec site&nbsp;<a href="https://colophons-and-scholars.com">https://colophons-and-scholars.com</a>.</p>

opencc-by-4.0Jan 2021View details →
zenodo40/100

Checkbot API raw results from Libraries, Archives and Museums websites for evaluating a data-driven Search Engine Optimization methodology

<p>Results from Checkbot API to measure and collect 341 websites compatibility on multiple SEO variables (34 variables). Checkbot API indexes the website&#39;s code to find features capable of impacting SEO performance. Each website has been tested with&nbsp;the maximum number of links allowed to be crawled equally to 10.000 per test. In this way, we retrieved data about the overall websites performance including their sub-pages, and not only the main domain names. &nbsp;A scale from 0 (lowest rate) to 100 (highest rate) was adopted for each examined variable. This constitutes a useful managerial indicator of dealing with the quantification of websites performance while avoiding complex measurement systems that are difficult to be adopted by administrators. Websites tested were also categorized by the CMS type used. More information about the variables and the meaning of the results can be found at&nbsp;https://www.checkbot.io/&nbsp;</p>

opencc-by-4.0Jun 2021View details →
zenodo40/100

Output of webXray analysis of German library websites

<p>[1] was analyzed with [2]. The deposited files are the result.</p> <p>&nbsp;</p> <p>[1] Steeg, Fabian et al.. (2016). URLs von Webseiten mit Typ Bibliothek aus Lobid.org. Zenodo. 10.5281/zenodo.50969</p> <p>[2] Tim Libert et al.. (2016). webXray: First Release. Zenodo. 10.5281/zenodo.57272</p> <p>&nbsp;</p>

opencc-zeroAug 2016View details →
zenodo40/100

Animal Communicator International Scan of Books, Websites, Animal Communicator Directory

<p>Four datasets are included focusing on animal communicators&nbsp;(AC): practitioners of intuitive interspecies communication (IIC).</p> <p><strong>Dataset #1:</strong> International English language websites data set (n = 400). To be included and coded, websites had to meet the following 3 criteria: (1) have an English language version of the website, (2) currently offer private AC consultations, and (3) be identified before website analysis was determined comprehensive enough to represent the international scope of AC (Oct. 30, 2020). CITE AS:<strong>&nbsp;</strong>Barrett, M. J., Zmud, L., Mathur, A., &amp; &nbsp;Hoessler, C. (2024). Animal communicator website scan [Data set]. Zenodo. DOI 10.5281/zenodo.15131964</p> <p>&nbsp;<strong>Dataset #2:</strong> International English language published books. Total books (n = 191): Books were identified through internet searches, including searches on Amazon and used bookstores such as Abe Books, from practicing AC websites, and from the directory: Book Authority Website for &ldquo;53 Best Animal Communication Books of All Time,&rdquo; https://bookauthority.org/books/best-animal-communication-books. Dataset includes books&rsquo; titles, descriptions, front and back covers and tables of contents where available. Where it was not clear whether the author was an animal communicator who consulted with clients, we did additional online searches to make this determination. All books had an English language printed copy of the book. We excluded books that were available in electronic copy only. We expanded our initial content inclusion criteria used in the website report beyond individuals currently offering consultations as professional animal communicators to include: (1) professional animal communicators who have since retired; (2) individuals who work intensively with animals in other capacities such as healing, but also report instances of IIC; (3) individuals who may not have worked as professional animal communicators, but write about their own lived experience of the phenomena; and (4) books written by individuals who are not ACs but have interviewed ACs. Where a book was written by two authors, or in some cases, an animal communicator with an additional author, we included both authors in our formal citation. Books were written by ACs who offered professional services (182), or authors who were not ACs (9). CITE AS:<strong> </strong>Barrett, M. J., Mathur, A., &amp;&nbsp; Ghoreishi, Z., Hoessler, C., Kuppenbender, S. (2024). Animal communicator Book Scan [Data set]. Zenodo. DOI 10.5281/zenodo.15131964</p> <p><strong>Dataset #3:</strong> Directory of practicing ACs (1990-2011; n = 80&nbsp;issues). To provide a snapshot of growth in numbers of practitioners over time, we compiled listings of ACs published in the Animal Communicator Directory from <em>Species Link: The Journal of Interspecies Telepathic Communication</em>. From 1990-2011, the publication included a directory of practicing ACs; after 2011, the directory went fully online, and data from each year is not available. CITE AS:<strong> </strong>Barrett, M.J. &amp; Hoessler, C. (2022).&nbsp; Animal Communicator Directory. [Data set]. Zenodo. DOI 10.5281/zenodo.15131964</p> <p><strong>Dataset #4:&nbsp;</strong>Reference List of 15 Analyzed Animal Communicator Books with formal or informal &ldquo;How-To&rdquo; Sections. Compiled to identify and summarize the ways in which ACs were describing essential processes for conducting a successful intuitive communication session with animals, and thus provide an overview of what happens in an IIC session. Selection criteria: We were seeking succinct summaries from ACs. The intent was not to dig into and analyze processes in detail, but rather to summarize and synthesize the essential processes and steps as the ACs were reporting them.&nbsp; As such, it was beyond the scope of this study to analyze reported example communications or analyze processes in books where the entirety of the book was describing communications with animals. Publication date range: 1998-2019. Close to 300 pages were analyzed. The actual "how-to" excerpts are not included as they are subject to copyright. CITE AS:<strong> </strong>Barrett, M.J. &amp; Kuppenbender, S. (2022).&nbsp; Reference List of 15 Analyzed Animal Communicator Books [Data set]. Zenodo. DOI 10.5281/zenodo.15131964</p> <p>For further information on data collection and analysis details contact M.J. Barrett, PhD.&nbsp;&nbsp;mj.barrett@usask.ca&nbsp;</p> <p>Funded by the Social Sciences and Humanities Research Council of Canada<em>&nbsp;</em>Insight Development Grant: <em>Deepening Connection in Pursuit of Environmental Sustainability: Assessing a Promising Lever for Shifting Assumptions of Separation </em>(Grant # 430-2019-01023).</p>

opencc-by-4.0Mar 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record