Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

12

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

12 results for “neo4j”

Learn how ShareScore rates datasets ↗
zenodo44/100

Goblin: Neo4J Maven Central dependency graph

<p>This repository contains a Neo4j dump of&nbsp; Maven Central dependency graph generated using <a href="https://github.com/Goblin-Ecosystem/goblinDependencyMiner">goblinDependencyMiner</a>.<br>To import this graph into neo4j, <strong>please use a version 4.x</strong>.</p> <p>Our dependency graph structure and metamodel are shown in images "goblin_dg_structure" and "metamodel".</p> <p>The latest available version dates from April 20, 2025, contains <span>16,939,391</span> nodes (712,509 libraries and 16,226,882 releases) and <span>152,434,085</span> edges (136,207,203 dependencies and 16,226,882 versioning edges).</p> <p>This repository contains two dump of the database:</p> <ul> <li><strong>goblin_maven_20_04_25.dump: </strong>This dataset contains the entire Maven Central dependency graph.</li> <li><strong>with_metrics_goblin_maven_20_04_25.dump</strong>: This dataset is the same as the previous one, but enriched with new &ldquo;AddedValue&rdquo; nodes (49,393,155 new nodes) representing the following metrics: CVE (dated may 13, 2025), freshness, popularity and speed. More information in this <a href="https://github.com/Goblin-Ecosystem/goblinTutorial">tutorial</a>.</li> </ul> <p>More details in the dedicated paper: <strong>Goblin: A Framework For Enriching And Querying the Maven Central Dependency Graph </strong>(https://doi.org/10.1145/3643991.3644879)<strong> </strong>- 21st International Conference on Mining Software Repositories (MSR'24).<br>If you use it, please <strong>cite</strong> this paper: <a href="https://dl.acm.org/doi/10.1145/3643991.3644879">https://dl.acm.org/doi/10.1145/3643991.3644879</a></p> <p>⚠️ This dataset is the subject of the <strong>Mining Challenge at the MSR 2025 conference</strong>, more information&nbsp;<a href="https://2025.msrconf.org/track/msr-2025-mining-challenge">here</a>.</p>

opencc-by-4.0Jan 2024View details →
zenodo44/100

A graph-based representation of the Hack Forums using Neo4j

<p>A graph-based representation of the Hack Forums using Neo4j.</p> <p>This dataset&nbsp;contains data to complement the paper &quot;A Graph-based Stratified Sampling Methodology for the Analysis of (Underground) Forums&quot; to appear in IEEE Transactions on Information Forensics and Security (TIFS).&nbsp;</p> <p>Due to ethical reasons, the data is anonymized and access to the actual content stored in the graph is subject to a formal data-sharing agreement with the Cambridge Cybercrime Centre. Please visit&nbsp;<a href="https://www.cambridgecybercrime.uk/process.html">this page</a>&nbsp;for more details on the process.&nbsp;</p>

opencc-by-4.0Aug 2023View details →
zenodo40/100

DODO neo4j dump

<p>Neo4j dump of the public instance of DODO that can be loaded into neo4j/docker instance.&nbsp;</p> <p>See <a href="https://github.com/Elysheba/DODO">Elysheba/DODO</a> for more info.</p>

opencc-by-4.0Oct 2024View details →
zenodo40/100

Neo4J Maven Central dependency graph

<p>⚠️ Updated archive here: <a href="../records/11104819">https://zenodo.org/records/11104819</a> ⚠️</p> <p>Neo4j dump of the dependency graph of Maven Central as of April 04, 2023.<br>&nbsp;</p>

opencc-by-4.0May 2023View details →
zenodo36/100

Product Recommendations through Neo4j by Analyzing Patterns in Customer Purchases

<p><span>Recommendation system grows more important each day as user interaction on the internet grows in size and complexity. To achieve better user experience and personalized choice of products for each user, it is important to create a recommendation system that takes all the interaction of a user on the internet and analyzes it thoroughly to get a better understanding of the user. Understanding the user will benefit the business more as each user will get a personalized experience based on how they act. This study focuses on utilizing the graph database to gain insight into the user behavior and develop a recommendation system based on how users act on the internet. The recommender system will use the Neo4j database as it provides much functionality to work with, such as the Graph Data Science library and the Jaccard Similarity method. Using all the graph technologies that exist today, this study will enable businesses to give a personalized experience to user by providing a detailed, accurate, effective, and efficient recommendation to user.</span></p>

opencc-by-4.0Dec 2023View details →
zenodo36/100

docker-compose for neo4j with paradise papers data loaded: Release v0.1-43

<p><code>docker-compose</code> for neo4j with paradise papers data loaded</p>

openother-openDec 2021View details →
zenodo32/100

Netflix Movies and TV Shows Recommendation with Neo4j

<p>Recommendation for movies can help discover new and enjoyable movies. This study uses the Neo4j Graph Database to create a recommendation system using the Netflix Movie Dataset. The objective of this research is the development of a movie recommendation algorithm using the k-NN similarity algorithm and FastRP node embedding machine learning. The results have provided recommendations based on similar attributes, such as actors, directors, country, type, and rating.</p>

opencc-by-4.0Dec 2023View details →
zenodo32/100

DirectedSmallMoleculesNetwork (DSMN) graph database for Neo4j (3.x)

<p>Release of the Neo4j metabolic interaction database for species human (Homo sapiens)&nbsp;(first unzip before using!). The data is licensed under the&nbsp;<a href="https://creativecommons.org/share-your-work/public-domain/cc0/">CCZero waiver</a>. This file contains data from the following pathway databases: Reactome, LIPID MAPS, WikiPathways.</p> <p>&nbsp;</p>

openother-openSep 2022View details →
zenodo32/100

datasets Neo4j for CIKM'18

<p>These are Jneo4j datasets used by CIKM&#39;18.</p> <p>Every zip file&nbsp;contains 2 datasets, original dataset and aggregate dataset.</p>

opencc-by-4.0May 2018View details →
zenodo28/100

Neo4j Database Dump of Corona-Warn-App Repository Provenance Graphs

<p>Provenance database (Neo4j 4.1) dumps of the following <a href="https://github.com/corona-warn-app/">Corona-Warn-App</a>&nbsp;repositories:</p> <ol> <li>cwa-app-android</li> <li>cwa-app-ios</li> <li>cwa-server</li> <li>cwa-documentation</li> </ol> <p>Username: covid</p> <p>Password: covid19</p>

opencc-by-4.0Jul 2020View details →
zenodo28/100

Reactome Data 86 - Neo4j & MySQL

Open the record for dataset details and reuse information.

opencc-by-4.0Oct 2023View details →
zenodo24/100

Analysis Titanic Survival through Graph-Based Node with Neo4j: Unraveling Social Dynamics (Dataset & Cypher)

Open the record for dataset details and reuse information.

openDec 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record