Optimizing the Songwriting Process: Genre-Based Lyric Generation Using Deep Learning Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Cai, Tracy, Liang, Wilson, Townes, Donte |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SongComposer: A Large Language Model for Lyric and Melody Generation in Song Composition
por: Ding, Shuangrui, et al.
Publicado: (2024)
por: Ding, Shuangrui, et al.
Publicado: (2024)
Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion
por: Frohmann, Markus, et al.
Publicado: (2025)
por: Frohmann, Markus, et al.
Publicado: (2025)
Sing it, Narrate it: Quality Musical Lyrics Translation
por: Ye, Zhuorui, et al.
Publicado: (2024)
por: Ye, Zhuorui, et al.
Publicado: (2024)
Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints
por: Meng, Hao, et al.
Publicado: (2026)
por: Meng, Hao, et al.
Publicado: (2026)
Joint Learning of Wording and Formatting for Singable Melody-to-Lyric Generation
por: Ou, Longshen, et al.
Publicado: (2023)
por: Ou, Longshen, et al.
Publicado: (2023)
SongCreator: Lyrics-based Universal Song Generation
por: Lei, Shun, et al.
Publicado: (2024)
por: Lei, Shun, et al.
Publicado: (2024)
REFFLY: Melody-Constrained Lyrics Editing Model
por: Zhao, Songyan, et al.
Publicado: (2024)
por: Zhao, Songyan, et al.
Publicado: (2024)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
por: Zhuo, Le, et al.
Publicado: (2023)
por: Zhuo, Le, et al.
Publicado: (2023)
Music Genre Classification: A Comparative Analysis of Classical Machine Learning and Deep Learning Approaches
por: Prajuli, Sachin, et al.
Publicado: (2026)
por: Prajuli, Sachin, et al.
Publicado: (2026)
SimClass: A Classroom Speech Dataset Generated via Game Engine Simulation For Automatic Speech Recognition Research
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models
por: Hsiao, Chi-Yuan, et al.
Publicado: (2026)
por: Hsiao, Chi-Yuan, et al.
Publicado: (2026)
CSL-L2M: Controllable Song-Level Lyric-to-Melody Generation Based on Conditional Transformer with Fine-Grained Lyric and Musical Controls
por: Chai, Li, et al.
Publicado: (2024)
por: Chai, Li, et al.
Publicado: (2024)
VQ-CTAP: Cross-Modal Fine-Grained Sequence Representation Learning for Speech Processing
por: Qiang, Chunyu, et al.
Publicado: (2024)
por: Qiang, Chunyu, et al.
Publicado: (2024)
Towards Advanced Speech Signal Processing: A Statistical Perspective on Convolution-Based Architectures and its Applications
por: Kapu, Nirmal Joshua, et al.
Publicado: (2024)
por: Kapu, Nirmal Joshua, et al.
Publicado: (2024)
Amplifying Emotional Signals: Data-Efficient Deep Learning for Robust Speech Emotion Recognition
por: Vu, Tai
Publicado: (2025)
por: Vu, Tai
Publicado: (2025)
NTPP: Generative Speech Language Modeling for Dual-Channel Spoken Dialogue via Next-Token-Pair Prediction
por: Wang, Qichao, et al.
Publicado: (2025)
por: Wang, Qichao, et al.
Publicado: (2025)
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
por: Huang, Jiawen, et al.
Publicado: (2024)
por: Huang, Jiawen, et al.
Publicado: (2024)
Efficient Dialect-Aware Modeling and Conditioning for Low-Resource Taiwanese Hakka Speech Processing
por: Peng, An-Ci, et al.
Publicado: (2026)
por: Peng, An-Ci, et al.
Publicado: (2026)
Genre Controlled Music Generation via Activation Steering
por: Narashiman, Swathi, et al.
Publicado: (2025)
por: Narashiman, Swathi, et al.
Publicado: (2025)
Continuous Modeling of the Denoising Process for Speech Enhancement Based on Deep Learning
por: Guo, Zilu, et al.
Publicado: (2023)
por: Guo, Zilu, et al.
Publicado: (2023)
A Computational Analysis of Lyric Similarity Perception
por: Kim, Haven, et al.
Publicado: (2024)
por: Kim, Haven, et al.
Publicado: (2024)
Relationships between Keywords and Strong Beats in Lyrical Music
por: Liao, Callie C., et al.
Publicado: (2024)
por: Liao, Callie C., et al.
Publicado: (2024)
UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions
por: Qiang, Chunyu, et al.
Publicado: (2026)
por: Qiang, Chunyu, et al.
Publicado: (2026)
Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
por: Majumder, Navonil, et al.
Publicado: (2024)
por: Majumder, Navonil, et al.
Publicado: (2024)
Kling-Foley: Multimodal Diffusion Transformer for High-Quality Video-to-Audio Generation
por: Wang, Jun, et al.
Publicado: (2025)
por: Wang, Jun, et al.
Publicado: (2025)
On Barriers to Archival Audio Processing
por: Sullivan, Peter, et al.
Publicado: (2025)
por: Sullivan, Peter, et al.
Publicado: (2025)
TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization
por: Hung, Chia-Yu, et al.
Publicado: (2024)
por: Hung, Chia-Yu, et al.
Publicado: (2024)
SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-Training
por: Yu, Jiaxing, et al.
Publicado: (2024)
por: Yu, Jiaxing, et al.
Publicado: (2024)
Device-Directed Speech Detection for Follow-up Conversations Using Large Language Models
por: Ognjen, et al.
Publicado: (2024)
por: Ognjen, et al.
Publicado: (2024)
Query-by-Example Keyword Spotting Using Spectral-Temporal Graph Attentive Pooling and Multi-Task Learning
por: Wang, Zhenyu, et al.
Publicado: (2024)
por: Wang, Zhenyu, et al.
Publicado: (2024)
Audio Contrastive-based Fine-tuning: Decoupling Representation Learning and Classification
por: Wang, Yang, et al.
Publicado: (2023)
por: Wang, Yang, et al.
Publicado: (2023)
A Self-Refining Framework for Enhancing ASR Using TTS-Synthesized Data
por: Chou, Cheng-Kang, et al.
Publicado: (2025)
por: Chou, Cheng-Kang, et al.
Publicado: (2025)
Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
por: Attia, Ahmed Adel, et al.
Publicado: (2025)
A Neural Model for Contextual Biasing Score Learning and Filtering
por: Huang, Wanting, et al.
Publicado: (2025)
por: Huang, Wanting, et al.
Publicado: (2025)
Auffusion: Leveraging the Power of Diffusion and Large Language Models for Text-to-Audio Generation
por: Xue, Jinlong, et al.
Publicado: (2024)
por: Xue, Jinlong, et al.
Publicado: (2024)
BAT: Learning to Reason about Spatial Sounds with Large Language Models
por: Zheng, Zhisheng, et al.
Publicado: (2024)
por: Zheng, Zhisheng, et al.
Publicado: (2024)
Weak Supervision Techniques towards Enhanced ASR Models in Industry-level CRM Systems
por: Wang, Zhongsheng, et al.
Publicado: (2025)
por: Wang, Zhongsheng, et al.
Publicado: (2025)
From Words to Waves: Analyzing Concept Formation in Speech and Text-Based Foundation Models
por: Ersoy, Asım, et al.
Publicado: (2025)
por: Ersoy, Asım, et al.
Publicado: (2025)
FMSD-TTS: Few-shot Multi-Speaker Multi-Dialect Text-to-Speech Synthesis for Ü-Tsang, Amdo and Kham Speech Dataset Generation
por: Liu, Yutong, et al.
Publicado: (2025)
por: Liu, Yutong, et al.
Publicado: (2025)
Get Large Language Models Ready to Speak: A Late-fusion Approach for Speech Generation
por: Shen, Maohao, et al.
Publicado: (2024)
por: Shen, Maohao, et al.
Publicado: (2024)
Ejemplares similares
-
SongComposer: A Large Language Model for Lyric and Melody Generation in Song Composition
por: Ding, Shuangrui, et al.
Publicado: (2024) -
Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion
por: Frohmann, Markus, et al.
Publicado: (2025) -
Sing it, Narrate it: Quality Musical Lyrics Translation
por: Ye, Zhuorui, et al.
Publicado: (2024) -
Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints
por: Meng, Hao, et al.
Publicado: (2026) -
Joint Learning of Wording and Formatting for Singable Melody-to-Lyric Generation
por: Ou, Longshen, et al.
Publicado: (2023)