Skip to main content
zenodoopen

Deezer listening events dataset

<p><strong>What does this dataset contain?</strong></p> <p>This dataset contains over 700 million time-stamped listening events collected from 3.4M anonymised users on the music streaming service Deezer, occurred between March and August 2022. It includes 50k anonymised songs, among the most popular ones on the service as well as their pre-trained embedding vectors, calculated by our internal model. All files are in parquet format which could be read by using&nbsp; <code>pandas.read_parquet</code> function.</p> <p>&nbsp;</p> <p><strong>What could this dataset be used for?</strong></p> <p>This dataset could be used for collaborative filtering as well as sequential recommendation (including both next-item and next-session recommendations).</p> <p>&nbsp;</p> <p><strong>Citation</strong></p> <p>If you use this dataset, please cite following paper:</p> <pre><span>@inproceedings</span>{<span>tran-recsys2024</span>, <span>title</span>=<span><span>{</span>Transformers Meet ACT-R: Repeat-Aware and Sequential Listening Session Recommendation<span>}</span></span>, <span>author</span>=<span><span>{</span>Viet-Anh Tran, Guillaume Salha-Galvan, Bruno Sguerra and Romain Hennequin<span>}</span></span>, <span>booktitle</span> = <span><span>{</span>Proceedings of the 18th ACM Conference on Recommender Systems<span>}</span></span>, <span>year</span> = <span><span>{</span>2024<span>}</span></span> }</pre>

ShareScore

32/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
12
Reuse readiness
8
Engagement
0

Topics