Skip to content
JournalsWorldThe Global Research Discovery Platform
Featured Dataset

Song Describer Dataset

The Song Describer Dataset: a Corpus of Audio Captions for Music-and-Language EvaluationA retro-futurist drum machine groove drenched in bubbly synthetic sound effects and a hint of an acid bassline.The Song Describer Dataset (SDD) contains ~1.1k ca

👤
CreatorManco, Ilaria
📅
Published2023-11-16
🔗
DOI10.5281/zenodo.10072001
📊
Downloads10,761
⚖️
Licensecc-by-sa-4.0
File Size3.1 GB
Data TypeDataset
Published2023
Licensecc-by-sa-4.0
Total Views6,062
Total Downloads10,761

The Song Describer Dataset: a Corpus of Audio Captions for Music-and-Language Evaluation

A retro-futurist drum machine groove drenched in bubbly synthetic sound effects and a hint of an acid bassline.

The Song Describer Dataset (SDD) contains ~1.1k captions for 706 permissively licensed music recordings. It is designed for use in evaluation of models that address music-and-language (M&L) tasks such as music captioning, text-to-music generation and music-language retrieval. More information about the data, collection method and validation is provided in the paper describing the dataset.

If you use this dataset, please cite our paper:

The Song Describer Dataset: a Corpus of Audio Captions for Music-and-Language Evaluation, Manco, Ilaria and Weck, Benno and Doh, Seungheon and Won, Minz and Zhang, Yixiao and Bogdanov, Dmitry and Wu, Yusong and Chen, Ke and Tovstogan, Philip and Benetos, Emmanouil and Quinton, Elio and Fazekas, György and Nam, Juhan, Machine Learning for Audio Workshop at NeurIPS 2023, 2023

📤 Share this page

Found this useful? Share it with your network.

✓ Link copied! Paste it on ResearchGate / Academia.edu
📦
Song Describer Dataset (Full Dataset)3.1 GB
⬇
📄
ReadmeVia DOI record
↗

Files are hosted on the source repository. Click download to access the full dataset.

Manco, Ilaria (2023). Song Describer Dataset. https://doi.org/10.5281/zenodo.10072001