Skip to content
JournalsWorldThe Global Research Discovery Platform
Featured Dataset

MERGE Dataset

The MERGE dataset is a collection of audio, lyrics, and bimodal datasets for conducting research on Music Emotion Recognition. A complete version is provided for each modality. The audio datasets provide 30-second excerpts for each sample, while full lyrics are provided in the relevant datasets.

👤
CreatorLima Louro, Pedro
📅
Published2024-10-16
🔗
DOI10.5281/zenodo.13939205
📊
Downloads2,815
⚖️
Licensecc-by-nc-4.0
File Size3.4 GB
Data TypeDataset
Published2024
Licensecc-by-nc-4.0
Total Views3,212
Total Downloads2,815

The MERGE dataset is a collection of audio, lyrics, and bimodal datasets for conducting research on Music Emotion Recognition. A complete version is provided for each modality. The audio datasets provide 30-second excerpts for each sample, while full lyrics are provided in the relevant datasets. The amount of available samples in each dataset is the following:

  • MERGE Audio Complete: 3554
  • MERGE Audio Balanced: 3232
  • MERGE Lyrics Complete: 2568
  • MERGE Lyrics Balanced: 2400
  • MERGE Bimodal Complete: 2216
  • MERGE Bimodal Balanced: 2000

Additional Contents

Each dataset contains the following additional files:

  • av_values: File containing the arousal and valence values for each sample sorted by their identifier;
  • tvt_dataframes: Train, validate, and test splits for each dataset. Both a 70-15-15 and a 40-30-30 split are provided.

Metadata

A metadata spreadsheet is provided for each dataset with the following information for each sample, if available:

  • Song (Audio and Lyrics datasets) – Song identifiers. Identifiers starting with MT were extracted from the AllMusic platform, while those starting with A or L were collected from private collections;
  • Quadrant – Label corresponding to one of the four quadrants from Russell’s Circumplex Model;
  • AllMusic Id –  For samples starting with A or L, the matching AllMusic identifier is also provided. This was used to complement the available information for the samples originally obtained from the platform;
  • Artist – First performing artist or band;
  • Title – Song title;
  • Relevance – AllMusic metric representing the relevance of the song in relation to the query used;
  • Duration – Song length in seconds;
  • Moods – User-generated mood tags extracted from the AllMusic platform and available in Warriner’s affective dictionary;
  • MoodsAll – User-generated mood tags extracted from the AllMusic platform;
  • Genres – User-generated genre tags extracted from the AllMusic platform;
  • Themes – User-generated theme tags extracted from the AllMusic platform;
  • Styles – User-generated style tags extracted from the AllMusic platform;
  • AppearancesTrackIDs – All AllMusic identifiers related with a sample;
  • Sample – Availability of the sample in the AllMusic platform;
  • SampleURL – URL to the 30-second excerpt in AllMusic;
  • ActualYear – Year of song release.

Citation

If you use some part of the MERGE dataset in your research, please cite the following article:

Louro, P. L. and Redinho, H. and Santos, R. and Malheiro, R. and Panda, R. and Paiva, R. P. (2024). MERGE – A Bimodal Dataset For Static Music Emotion Recognition. arxiv. URL: https://arxiv.org/abs/2407.06060.

BibTeX:

, 
    &nb

📤 Share this page

Found this useful? Share it with your network.

✓ Link copied! Paste it on ResearchGate / Academia.edu
📦
MERGE Dataset (Full Dataset)3.4 GB
⬇
📄
ReadmeVia DOI record
↗

Files are hosted on the source repository. Click download to access the full dataset.

Lima Louro, Pedro (2024). MERGE Dataset. https://doi.org/10.5281/zenodo.13939205