Reinforcement Learning for Generative AI: A Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Cao, Yuanjiang, Sheng, Quan Z., McAuley, Julian, Yao, Lina |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Hint for Reinforcement Learning
by: Xia, Yu, et al.
Published: (2026)
by: Xia, Yu, et al.
Published: (2026)
Causality-Aware Transformer Networks for Robotic Navigation
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
Bridging Conversational and Collaborative Signals for Conversational Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2024)
by: Rabiah, Ahmad Bin, et al.
Published: (2024)
DITTO: Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Steering Autoregressive Music Generation with Recursive Feature Machines
by: Zhao, Daniel, et al.
Published: (2025)
by: Zhao, Daniel, et al.
Published: (2025)
Skill-R1: Agent Skill Evolution via Reinforcement Learning
by: Vishe, Yash, et al.
Published: (2026)
by: Vishe, Yash, et al.
Published: (2026)
PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing
by: Long, Phillip, et al.
Published: (2024)
by: Long, Phillip, et al.
Published: (2024)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
by: Xu, Xin, et al.
Published: (2026)
by: Xu, Xin, et al.
Published: (2026)
Presto! Distilling Steps and Layers for Accelerating Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
A Survey on Data-Centric AI: Tabular Learning from Reinforcement Learning and Generative AI Perspective
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
Unified Generation, Reconstruction, and Representation: Generalized Diffusion with Adaptive Latent Encoding-Decoding
by: Liu, Guangyi, et al.
Published: (2024)
by: Liu, Guangyi, et al.
Published: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
by: Yang, Dayu, et al.
Published: (2025)
by: Yang, Dayu, et al.
Published: (2025)
Low-Resource Guidance for Controllable Latent Audio Diffusion
by: Novack, Zachary, et al.
Published: (2026)
by: Novack, Zachary, et al.
Published: (2026)
Symbolic Representation for Any-to-Any Generative Tasks
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
Futga: Towards Fine-grained Music Understanding through Temporally-enhanced Generative Augmentation
by: Wu, Junda, et al.
Published: (2024)
by: Wu, Junda, et al.
Published: (2024)
Composer Vector: Style-steering Symbolic Music Generation in a Latent Space
by: Jiang, Xunyi, et al.
Published: (2026)
by: Jiang, Xunyi, et al.
Published: (2026)
A Survey Analyzing Generalization in Deep Reinforcement Learning
by: Korkmaz, Ezgi
Published: (2024)
by: Korkmaz, Ezgi
Published: (2024)
FedCLF -- Towards Efficient Participant Selection for Federated Learning in Heterogeneous IoV Networks
by: Wijethilake, Kasun Eranda, et al.
Published: (2025)
by: Wijethilake, Kasun Eranda, et al.
Published: (2025)
Plasticity Loss in Deep Reinforcement Learning: A Survey
by: Klein, Timo, et al.
Published: (2024)
by: Klein, Timo, et al.
Published: (2024)
PCGRL+: Scaling, Control and Generalization in Reinforcement Learning Level Generators
by: Earle, Sam, et al.
Published: (2024)
by: Earle, Sam, et al.
Published: (2024)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
by: Bhat, Vishvesh, et al.
Published: (2025)
by: Bhat, Vishvesh, et al.
Published: (2025)
A Comprehensive Survey on Inverse Constrained Reinforcement Learning: Definitions, Progress and Challenges
by: Liu, Guiliang, et al.
Published: (2024)
by: Liu, Guiliang, et al.
Published: (2024)
Learning Local Constraints for Reinforcement-Learned Content Generators
by: Bhaumik, Debosmita, et al.
Published: (2026)
by: Bhaumik, Debosmita, et al.
Published: (2026)
GSPRec: Temporal-Aware Graph Spectral Filtering for Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
A Survey of Continual Reinforcement Learning
by: Pan, Chaofan, et al.
Published: (2025)
by: Pan, Chaofan, et al.
Published: (2025)
CSyMR: Benchmarking Compositional Music Information Retrieval in Symbolic Music Reasoning
by: Wang, Boyang, et al.
Published: (2025)
by: Wang, Boyang, et al.
Published: (2025)
Diffusion Models for Reinforcement Learning: A Survey
by: Zhu, Zhengbang, et al.
Published: (2023)
by: Zhu, Zhengbang, et al.
Published: (2023)
A Survey on Explainable Deep Reinforcement Learning
by: Cheng, Zelei, et al.
Published: (2025)
by: Cheng, Zelei, et al.
Published: (2025)
Independence Constrained Disentangled Representation Learning from Epistemological Perspective
by: Wang, Ruoyu, et al.
Published: (2024)
by: Wang, Ruoyu, et al.
Published: (2024)
Emerging Synergies in Causality and Deep Generative Models: A Survey
by: Zhou, Guanglin, et al.
Published: (2023)
by: Zhou, Guanglin, et al.
Published: (2023)
Automatic Pair Construction for Contrastive Post-training
by: Xu, Canwen, et al.
Published: (2023)
by: Xu, Canwen, et al.
Published: (2023)
How to Train Data-Efficient LLMs
by: Sachdeva, Noveen, et al.
Published: (2024)
by: Sachdeva, Noveen, et al.
Published: (2024)
Towards Data-Centric AI: A Comprehensive Survey of Traditional, Reinforcement, and Generative Approaches for Tabular Data Transformation
by: Wang, Dongjie, et al.
Published: (2025)
by: Wang, Dongjie, et al.
Published: (2025)
A Survey of State Representation Learning for Deep Reinforcement Learning
by: Echchahed, Ayoub, et al.
Published: (2025)
by: Echchahed, Ayoub, et al.
Published: (2025)
Calibration-Disentangled Learning and Relevance-Prioritized Reranking for Calibrated Sequential Recommendation
by: Jeon, Hyunsik, et al.
Published: (2024)
by: Jeon, Hyunsik, et al.
Published: (2024)
Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators
by: Novack, Zachary, et al.
Published: (2026)
by: Novack, Zachary, et al.
Published: (2026)
A Survey of Constraint Formulations in Safe Reinforcement Learning
by: Wachi, Akifumi, et al.
Published: (2024)
by: Wachi, Akifumi, et al.
Published: (2024)
Similar Items
-
Learning to Hint for Reinforcement Learning
by: Xia, Yu, et al.
Published: (2026) -
Causality-Aware Transformer Networks for Robotic Navigation
by: Wang, Ruoyu, et al.
Published: (2024) -
DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024) -
Bridging Conversational and Collaborative Signals for Conversational Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2024) -
DITTO: Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024)