UniLG: A Unified Structure-aware Framework for Lyrics Generation
Tao Qian, Fan Lou, Jiatong Shi, Yuning Wu, Shuai Guo, Xiang Yin, Qin Jin
Abstract
As a special task of natural language generation, conditional lyrics generation needs to consider the structure of generated lyrics 1 and the relationship between lyrics and music. Due to various forms of conditions, a lyrics generation system is expected to generate lyrics conditioned on different signals, such as music scores, music audio, or partially-finished lyrics, etc. However, most of the previous works have ignored the musical attributes hidden behind the lyrics and the structure of the lyrics. Additionally, most works only handle limited lyrics generation conditions, such as lyrics generation based on music score or partial lyrics, they can not be easily extended to other generation conditions with the same framework. In this paper, we propose a unified structure-aware lyrics generation framework named UniLG. Specifically, we design compound templates that incorporate textual and musical information to improve structure modeling and unify the different lyrics generation conditions. Extensive experiments demonstrate the effectiveness of our framework. Both objective and subjective evaluations show significant improvements in generating structural lyrics. * *Corresponding Author. 1 The structure of lyrics in our work means that lyrics have the chorus and verse parts among the sentences. 𝑏 ! , 𝑏 " , … , 𝑏 # Beat Music Score Lyric Polished Lyric Music Score with Lyric Audio Audio with Lyric 爱真的需要勇气 来面对流言蜚语 ………… Lyric UniLG (Masked) Lyric Bar Beat Segment Intro-position
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d395887b-a38c-4c3e-b636-24c0415e390eCited by top-tier papers2
- ToneCraft: Cantonese Lyrics Generation with Harmony of Tones and PitchesJunyu Cheng, Chang Pan, Shuangyin LiEMNLP 2025
- Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following AbilityYusuke Sakai, Hidetaka Kamigaito, Taro WatanabeACL 2025
Builds on6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Video Background Music Generation with Controllable Music TransformerShangzhe Di, Zeren Jiang, Si Liu, Zhaokai Wang et al.ACM MM 2021 · 87 citations
- SongMASS: Automatic Song Writing with Pre-training and Alignment ConstraintZhonghao Sheng, Kaitao Song, Xu Tan, Yi Ren et al.AAAI 2021 · 84 citations
- DeepSinger: Singing Voice Synthesis with Data Mined From the WebYi Ren, Xu Tan, Tao Qin, Jian Luan et al.KDD 2020 · 72 citations
- TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage MethodZeqian Ju, Peiling Lu, Xu Tan, Rui Wang et al.EMNLP 2022 · 19 citations
Related papers
- Unsupervised Melody-to-Lyrics GenerationYufei Tian, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone et al.ACL 2023 · 6 citations
- AI-Lyricist: Generating Music and Vocabulary Constrained LyricsXichu Ma, Ye Wang, Min-Yen Kan, Wee Sun LeeACM MM 2021 · 22 citations
- S²MILE: Semantic-and-Structure-Aware Music-Driven Lyric GenerationMu You, Fang Zhang, Shuai Zhang, Linli XuAAAI 2025
- SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-TrainingJiaxing Yu, Xinda Wu, Yunfei Xu, Tieyao Zhang et al.AAAI 2025 · 2 citations
- CSL-L2M: Controllable Song-Level Lyric-to-Melody Generation Based on Conditional Transformer with Fine-Grained Lyric and Musical ControlsLi Chai, Donglin WangAAAI 2025 · 1 citation
