DeepRapper: Neural Rap Generation with Rhyme and Rhythm Modeling
Lanqing Xue, Kaitao Song, Duocai Wu, Xu Tan, Nevin L. Zhang, Tao Qin, Wei-Qiang Zhang, Tie-Yan Liu
Abstract
Rap generation, which aims to produce lyrics and corresponding singing beats, needs to model both rhymes and rhythms. Previous works for rap generation focused on rhyming lyrics but ignored rhythmic beats, which are important for rap performance. In this paper, we develop DeepRapper, a Transformer-based rap generation system that can model both rhymes and rhythms. Since there is no available rap dataset with rhythmic beats, we develop a data mining pipeline to collect a largescale rap dataset, which includes a large number of rap songs with aligned lyrics and rhythmic beats. Second, we design a Transformerbased autoregressive language model which carefully models rhymes and rhythms. Specifically, we generate lyrics in the reverse order with rhyme representation and constraint for rhyme enhancement and insert a beat symbol into lyrics for rhythm/beat modeling. To our knowledge, DeepRapper is the first system to generate rap with both rhymes and rhythms. Both objective and subjective evaluations demonstrate that DeepRapper generates creative and high-quality raps with rhymes and rhythms. Code will be released on GitHub 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55377cb2-55bc-41b8-918f-087b7b7c32bbCited by top-tier papers6
- ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language ModelsJonas Belouadi, Steffen EgerACL 2023 · 12 citations
- Unsupervised Melody-to-Lyrics GenerationYufei Tian, Anjali Narayan-Chen, Shereen Oraby, Alessandra Cervone et al.ACL 2023 · 6 citations
- Drop the Beat! Freestyler for Accompaniment Conditioned Rapping Voice GenerationZiqian Ning, Shuai Wang, Yuepeng Jiang, Jixun Yao et al.AAAI 2025 · 5 citations
- Multi-Modal Experience Inspired AI CreationQian Cao, Xu Chen, Ruihua Song, Hao Jiang et al.ACM MM 2022 · 3 citations
- ToneCraft: Cantonese Lyrics Generation with Harmony of Tones and PitchesJunyu Cheng, Chang Pan, Shuangyin LiEMNLP 2025
Builds on5
- Pop Music Transformer: Beat-based Modeling and Generation of Expressive Pop Piano CompositionsYu-Siang Huang, Yi-Hsuan YangACM MM 2020 · 265 citations
- PopMAG: Pop Music Accompaniment GenerationYi Ren, Jinzheng He, Xu Tan, Tao Qin et al.ACM MM 2020 · 91 citations
- SongMASS: Automatic Song Writing with Pre-training and Alignment ConstraintZhonghao Sheng, Kaitao Song, Xu Tan, Yi Ren et al.AAAI 2021 · 84 citations
- Automatic Poetry Generation from Prosaic TextTim Van de CruysACL 2020 · 54 citations
- Rigid Formats Controlled Text GenerationPiji Li, Haisong Zhang, Xiaojiang Liu, Shuming ShiACL 2020 · 2 citations
Related papers
- RapVerse: Coherent Vocals and Whole-Body Motion Generation from TextJiaben Chen, Xin Yan, Yihang Chen, Siyuan Cen et al.ICCV 2025 · 7 citations
- DeepSinger: Singing Voice Synthesis with Data Mined From the WebYi Ren, Xu Tan, Tao Qin, Jian Luan et al.KDD 2020 · 72 citations
- UniMuMo: Unified Text, Music, and Motion GenerationHan Yang, Kun Su, Yutong Zhang, Jiaben Chen et al.AAAI 2025 · 3 citations
- DanceFormer: Music Conditioned 3D Dance Generation with Parametric Motion TransformerBuyu Li, Yongchi Zhao, Zhelun Shi, Lu ShengAAAI 2022 · 182 citations
- YuE: Scaling Open Foundation Models for Long-Form Music GenerationRuibin Yuan, Hanfeng Lin, Shuyue Guo, Ge Zhang et al.ICLR 2026 · 112 citations
