UniLG: A Unified Structure-aware Framework for Lyrics Generation
Tao Qian, Fan Lou, Jiatong Shi, Yuning Wu, Shuai Guo, Xiang Yin, Qin Jin
摘要
As a special task of natural language generation, conditional lyrics generation needs to consider the structure of generated lyrics 1 and the relationship between lyrics and music. Due to various forms of conditions, a lyrics generation system is expected to generate lyrics conditioned on different signals, such as music scores, music audio, or partially-finished lyrics, etc. However, most of the previous works have ignored the musical attributes hidden behind the lyrics and the structure of the lyrics. Additionally, most works only handle limited lyrics generation conditions, such as lyrics generation based on music score or partial lyrics, they can not be easily extended to other generation conditions with the same framework. In this paper, we propose a unified structure-aware lyrics generation framework named UniLG. Specifically, we design compound templates that incorporate textual and musical information to improve structure modeling and unify the different lyrics generation conditions. Extensive experiments demonstrate the effectiveness of our framework. Both objective and subjective evaluations show significant improvements in generating structural lyrics. * *Corresponding Author. 1 The structure of lyrics in our work means that lyrics have the chorus and verse parts among the sentences. 𝑏 ! , 𝑏 " , … , 𝑏 # Beat Music Score Lyric Polished Lyric Music Score with Lyric Audio Audio with Lyric 爱真的需要勇气 来面对流言蜚语 ………… Lyric UniLG (Masked) Lyric Bar Beat Segment Intro-position
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- ToneCraft: Cantonese Lyrics Generation with Harmony of Tones and PitchesJunyu Cheng, Chang Pan, Shuangyin LiEMNLP 2025
- Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following AbilityYusuke Sakai, Hidetaka Kamigaito, Taro WatanabeACL 2025
它引用的顶会 Paper6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Video Background Music Generation with Controllable Music TransformerShangzhe Di, Zeren Jiang, Si Liu, Zhaokai Wang 等ACM MM 2021 · 被引用 87 次
- SongMASS: Automatic Song Writing with Pre-training and Alignment ConstraintZhonghao Sheng, Kaitao Song, Xu Tan, Yi Ren 等AAAI 2021 · 被引用 84 次
- DeepSinger: Singing Voice Synthesis with Data Mined From the WebYi Ren, Xu Tan, Tao Qin, Jian Luan 等KDD 2020 · 被引用 72 次
- TeleMelody: Lyric-to-Melody Generation with a Template-Based Two-Stage MethodZeqian Ju, Peiling Lu, Xu Tan, Rui Wang 等EMNLP 2022 · 被引用 19 次
相关 Paper
- Unsupervised Melody-to-Lyrics GenerationYufei Tian, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone 等ACL 2023 · 被引用 6 次
- AI-Lyricist: Generating Music and Vocabulary Constrained LyricsXichu Ma, Ye Wang, Min-Yen Kan, Wee Sun LeeACM MM 2021 · 被引用 22 次
- S²MILE: Semantic-and-Structure-Aware Music-Driven Lyric GenerationMu You, Fang Zhang, Shuai Zhang, Linli XuAAAI 2025
- SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-TrainingJiaxing Yu, Xinda Wu, Yunfei Xu, Tieyao Zhang 等AAAI 2025 · 被引用 2 次
- CSL-L2M: Controllable Song-Level Lyric-to-Melody Generation Based on Conditional Transformer with Fine-Grained Lyric and Musical ControlsLi Chai, Donglin WangAAAI 2025 · 被引用 1 次
