Unsupervised Melody-to-Lyrics Generation

Yufei Tian; Anjali Narayan-Chen; Shereen Oraby; Alessandra Cervone; Gunnar Sigurdsson; Chenyang Tao; Wenbo Zhao; Yiwen Chen; Tagyoung Chung; Jing Huang; Nanyun Peng

doi:10.18653/v1/2023.acl-long.513

Unsupervised Melody-to-Lyrics Generation

Yufei Tian, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone, Gunnar Sigurdsson, Chenyang Tao, Wenbo Zhao, Yiwen Chen, Tagyoung Chung, Jing Huang, Nanyun Peng

Abstract

Automatic melody-to-lyric generation is a task in which song lyrics are generated to go with a given melody. It is of significant practical interest and more challenging than unconstrained lyric generation as the music imposes additional constraints onto the lyrics. The training data is limited as most songs are copyrighted, resulting in models that underfit the complicated cross-modal relationship between melody and lyrics. In this work, we propose a method for generating high-quality lyrics without training on any aligned melody-lyric data. Specifically, we design a hierarchical lyric generation framework that first generates a song outline and second the complete lyrics. The framework enables disentanglement of training (based purely on text) from inference (melody-guided text generation) to circumvent the shortage of parallel data. We leverage the segmentation and rhythm alignment between melody and lyrics to compile the given melody into decoding constraints as guidance during inference. The two-step hierarchical design also enables content control via the lyric outline, a much-desired feature for democratizing collaborative song creation. Experimental results show that our model can generate high-quality lyrics that are more on-topic, singable, intelligible, and coherent than strong baselines, for example SongMASS, a SOTA model trained on a parallel dataset, with a 24% relative overall quality improvement based on human ratings. Our code is available at https://github.com/amazon-science/unsupervised-melody-to-lyrics-generation.

Anthology ID:: 2023.acl-long.513
Original:: 2023.acl-long.513v1
Version 2:: 2023.acl-long.513v2
Volume:: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: July
Year:: 2023
Address:: Toronto, Canada
Editors:: Anna Rogers, Jordan Boyd-Graber, Naoaki Okazaki
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 9235–9254
Language:
URL:: https://aclanthology.org/2023.acl-long.513/
DOI:: 10.18653/v1/2023.acl-long.513
Bibkey:
Cite (ACL):: Yufei Tian, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone, Gunnar Sigurdsson, Chenyang Tao, Wenbo Zhao, Yiwen Chen, Tagyoung Chung, Jing Huang, and Nanyun Peng. 2023. Unsupervised Melody-to-Lyrics Generation. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 9235–9254, Toronto, Canada. Association for Computational Linguistics.
Cite (Informal):: Unsupervised Melody-to-Lyrics Generation (Tian et al., ACL 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.acl-long.513.pdf
Video:: https://aclanthology.org/2023.acl-long.513.mp4

PDF (v2) PDF (v1) Cite Search Video Fix data