Datasets › JamendoMaxCaps

JamendoMaxCaps

Introduced by Abhinaba Roy et al. in JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata11 Feb 2025 archive 2025-07-28

📊 Dataset Details

  • Name: JamendoMaxCaps
  • URL: https://huggingface.co/datasets/amaai-lab/JamendoMaxCaps
  • Content: 362,238 songs with captions generated by Qwen2-Audio
Metadata Fields
  • genre
  • speed
  • variable tags

🎯 Rationale

  1. Scale: Over a quarter-million tracks, offering a rich multimodal corpus.
  2. Quality: Captions generated by the state-of-the-art Qwen2-Audio model.
  3. Use cases: Supports audio captioning, retrieval, genre classification, tempo/speed analysis, and more.

Benchmarks archive 2025-07-28

No leaderboard in the archive resolves to this dataset.

Papers archive 2025-07-28

No paper in the archive has a leaderboard row on this dataset; the archive counts 1 paper for it but never published that list.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

Creative Commons Attribution Share Alike 3.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • JamendoMaxCaps

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections