Version 1.2.0

Dataset for ICASSP 2026 Cadenza Challenge: Predicting Lyric Intelligibility (CLIP)

Roa Dabike, Gerardo;Cox, Trevor;Barker, Jon P

Description

CadenzaThis is the training and validation data for the ICASSP 2026 Cadenza Challenge: Predicting Lyrics Intelligibility (CLIP1)Understanding the lyrics in music is key for music enjoyment [1]. People with hearing loss can have difficulties to clearly and effortlessly hearing lyrics [2], however. In speech technology, having metrics to automatically evaluate intelligibility has driven improvements in speech enhancement. We want to do the same for music with lyrics!A detailed description of the challenge can be found on our websiteLicenseThis dataset is shared under the CPC-4 License (Creative Production Commons 4.0).CitationIf you use this dataset in your work, please cite the papers, not the Zenodo record:Roa-Dabike, G., Cox, T. J., Barker, J. P., Fazenda, B. M., Graetzer, S., Vos, R. R., Akeroyd, M. A., Firth, J., Whitmer, W. M., Bannister, S., & Greasley, A. (submitted). The Cadenza Lyric Intelligibility Prediction (CLIP) Dataset. Data in Brief.Roa-Dabike, G., Barker, J. P., Cox, T. J., Akeroyd, M. A., Bannister, S., Fazenda, B., Firth, J., Graetzer, S., Greasley, A., Vos, R. R., & Whitmer, W. M. (2026). Overview of the ICASSP 2026 Cadenza Challenge: Predicting Lyric Intelligibility (Manuscript submitted for publication). IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP).Bibtext@article{roa2025clip,  title  = {The Cadenza Lyric Intelligibility Prediction (CLIP) Dataset},  author = {Roa-Dabike, Gerardo and Cox, Trevor J. and Barker, Jon P. and Fazenda, Bruno M. and Graetzer, Simone and Vos, Rebecca R. and Akeroyd, Michael A. and Firth, Jennifer and Whitmer, William M. and Bannister, Scott and Greasley, Alinka},  journal = {Data in Brief},  year  = {2025}}@inproceedings{roa2026cadenza_overview,  title        = {Overview of the ICASSP 2026 Cadenza Challenge: Predicting Lyric Intelligibility},  author       = {Roa-Dabike, Gerardo and Barker, Jon P. and Cox, Trevor J. and                   Akeroyd, Michael A. and Bannister, Scott and                   Fazenda, Bruno and Firth, Jennifer and Graetzer, Simone and                   Greasley, Alinka and Vos, Rebecca R. and Whitmer, William M.},  booktitle    = {Proc. IEEE ICASSP},  year         = {2026},  note         = {To appear}}

Citations (0)

Mentions (0)

Metrics

Dataset Index

0.5

FAIR Score

85%

Citations

0

Mentions

0

Metrics Over Time

Publication Details

DOI

Publisher

Cadenza Project

License

Creative Commons Attribution 4.0 International

Assigned Domain

Subfield

Signal Processing

Field

Computer Science

Domain

Physical Sciences

Confidence Score

39%

Source

Scholar Data Model

Keywords

Lyrics IntelligibilityMachine LearningSignal ProcessingMusicAcoustics

Normalization Factors

FT

57.69

CTw

1.00

MTw

1.00