MusicRL: Aligning Music Generation to Human Preferences
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cideron, Geoffrey, Girgin, Sertan, Verzetti, Mauro, Vincent, Damien, Kastelic, Matej, Borsos, Zalán, McWilliams, Brian, Ungureanu, Victor, Bachem, Olivier, Pietquin, Olivier, Geist, Matthieu, Hussenot, Léonard, Zeghidour, Neil, Agostinelli, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diversity-Rewarded CFG Distillation
von: Cideron, Geoffrey, et al.
Veröffentlicht: (2024)
von: Cideron, Geoffrey, et al.
Veröffentlicht: (2024)
WARM: On the Benefits of Weight Averaged Reward Models
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
Learning in Mean Field Games: A Survey
von: Laurière, Mathieu, et al.
Veröffentlicht: (2022)
von: Laurière, Mathieu, et al.
Veröffentlicht: (2022)
BOND: Aligning LLMs with Best-of-N Distillation
von: Sessa, Pier Giuseppe, et al.
Veröffentlicht: (2024)
von: Sessa, Pier Giuseppe, et al.
Veröffentlicht: (2024)
WARP: On the Benefits of Weight Averaged Rewarded Policies
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
Live Music Models
von: Lyria Team, et al.
Veröffentlicht: (2025)
von: Lyria Team, et al.
Veröffentlicht: (2025)
Bench-MFG: A Benchmark Suite for Learning in Stationary Mean Field Games
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2026)
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2026)
Population-aware Online Mirror Descent for Mean-Field Games with Common Noise by Deep Reinforcement Learning
von: Wu, Zida, et al.
Veröffentlicht: (2025)
von: Wu, Zida, et al.
Veröffentlicht: (2025)
SpectroStream: A Versatile Neural Codec for General Audio
von: Li, Yunpeng, et al.
Veröffentlicht: (2025)
von: Li, Yunpeng, et al.
Veröffentlicht: (2025)
MAD Speech: Measures of Acoustic Diversity of Speech
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
von: Wu, Zida, et al.
Veröffentlicht: (2024)
von: Wu, Zida, et al.
Veröffentlicht: (2024)
On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
von: Agarwal, Rishabh, et al.
Veröffentlicht: (2023)
von: Agarwal, Rishabh, et al.
Veröffentlicht: (2023)
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
Audio-to-Image Bird Species Retrieval without Audio-Image Pairs via Text Distillation
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion
von: Flet-Berliac, Yannis, et al.
Veröffentlicht: (2024)
von: Flet-Berliac, Yannis, et al.
Veröffentlicht: (2024)
Self-Improving Robust Preference Optimization
von: Choi, Eugene, et al.
Veröffentlicht: (2024)
von: Choi, Eugene, et al.
Veröffentlicht: (2024)
NatureLM-audio: an Audio-Language Foundation Model for Bioacoustics
von: Robinson, David, et al.
Veröffentlicht: (2024)
von: Robinson, David, et al.
Veröffentlicht: (2024)
Nash Learning from Human Feedback
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
Brain2Music: Reconstructing Music from Human Brain Activity
von: Denk, Timo I., et al.
Veröffentlicht: (2023)
von: Denk, Timo I., et al.
Veröffentlicht: (2023)
Solving robust MDPs as a sequence of static RL problems
von: Zouitine, Adil, et al.
Veröffentlicht: (2024)
von: Zouitine, Adil, et al.
Veröffentlicht: (2024)
Compact Hypercube Embeddings for Fast Text-based Wildlife Observation Retrieval
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
Beyond the Baseband: Adaptive Multi-Band Encoding for Full-Spectrum Bioacoustics Classification
von: Sarkar, Eklavya, et al.
Veröffentlicht: (2026)
von: Sarkar, Eklavya, et al.
Veröffentlicht: (2026)
Conditional Language Policy: A General Framework for Steerable Multi-Objective Finetuning
von: Wang, Kaiwen, et al.
Veröffentlicht: (2024)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2024)
Rethinking Social Action through Music
von: Baker, Geoffrey
Veröffentlicht: (2021)
von: Baker, Geoffrey
Veröffentlicht: (2021)
Do Recommender Systems Promote Local Music? A Reproducibility Study Using Music Streaming Data
von: Matrosova, Kristina, et al.
Veröffentlicht: (2024)
von: Matrosova, Kristina, et al.
Veröffentlicht: (2024)
Aligned Music Notation and Lyrics Transcription
von: Fuentes-Martínez, Eliseo, et al.
Veröffentlicht: (2024)
von: Fuentes-Martínez, Eliseo, et al.
Veröffentlicht: (2024)
RL-AVIST: Reinforcement Learning for Autonomous Visual Inspection of Space Targets
von: El-Hariry, Matteo, et al.
Veröffentlicht: (2025)
von: El-Hariry, Matteo, et al.
Veröffentlicht: (2025)
Aligning Spoken Dialogue Models from User Interactions
von: Wu, Anne, et al.
Veröffentlicht: (2025)
von: Wu, Anne, et al.
Veröffentlicht: (2025)
Multi Agents Semantic Emotion Aligned Music to Image Generation with Music Derived Captions
von: Shi, Junchang, et al.
Veröffentlicht: (2025)
von: Shi, Junchang, et al.
Veröffentlicht: (2025)
V2Meow: Meowing to the Visual Beat via Video-to-Music Generation
von: Su, Kun, et al.
Veröffentlicht: (2023)
von: Su, Kun, et al.
Veröffentlicht: (2023)
MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection
von: Lu, Tongyu, et al.
Veröffentlicht: (2025)
von: Lu, Tongyu, et al.
Veröffentlicht: (2025)
Averaging log-likelihoods in direct alignment
von: Grinsztajn, Nathan, et al.
Veröffentlicht: (2024)
von: Grinsztajn, Nathan, et al.
Veröffentlicht: (2024)
Synthesis of Novel Schiff Bases with Piperidine Rings and Investigation of Their Antioxidant Capacities and Anticholinesterase and Carbonic Anhydrase Enzyme Inhibition Properties
von: Sertan Aytaç
Veröffentlicht: (2025)
von: Sertan Aytaç
Veröffentlicht: (2025)
Aligning Text-to-Music Evaluation with Human Preferences
von: Huang, Yichen, et al.
Veröffentlicht: (2025)
von: Huang, Yichen, et al.
Veröffentlicht: (2025)
SAME: A Semantically-Aligned Music Autoencoder
von: Parker, Julian D., et al.
Veröffentlicht: (2026)
von: Parker, Julian D., et al.
Veröffentlicht: (2026)
ShiQ: Bringing back Bellman to LLMs
von: Clavier, Pierre, et al.
Veröffentlicht: (2025)
von: Clavier, Pierre, et al.
Veröffentlicht: (2025)
Soundtracks of Our Lives: How Age Influences Musical Preferences
von: Golubovikj, Arsen Matej, et al.
Veröffentlicht: (2025)
von: Golubovikj, Arsen Matej, et al.
Veröffentlicht: (2025)
Adaptative Bilingual Aligning Using Multilingual Sentence Embedding
von: Kraif, Olivier
Veröffentlicht: (2024)
von: Kraif, Olivier
Veröffentlicht: (2024)
Varieties of modal algebras without the congruence extension property
von: Gyenis, Zalán, et al.
Veröffentlicht: (2024)
von: Gyenis, Zalán, et al.
Veröffentlicht: (2024)
Emotion-Aligned Contrastive Learning Between Images and Music
von: Stewart, Shanti, et al.
Veröffentlicht: (2023)
von: Stewart, Shanti, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Diversity-Rewarded CFG Distillation
von: Cideron, Geoffrey, et al.
Veröffentlicht: (2024) -
WARM: On the Benefits of Weight Averaged Reward Models
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024) -
Learning in Mean Field Games: A Survey
von: Laurière, Mathieu, et al.
Veröffentlicht: (2022) -
BOND: Aligning LLMs with Best-of-N Distillation
von: Sessa, Pier Giuseppe, et al.
Veröffentlicht: (2024) -
WARP: On the Benefits of Weight Averaged Rewarded Policies
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)