Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Poddar, Sriyash, Wan, Yanming, Ivison, Hamish, Gupta, Abhishek, Jaques, Natasha |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Infer Human's Intentions Before Following Natural Language Instructions
por: Wan, Yanming, et al.
Publicado: (2024)
por: Wan, Yanming, et al.
Publicado: (2024)
MaxMin-RLHF: Alignment with Diverse Human Preferences
por: Chakraborty, Souradip, et al.
Publicado: (2024)
por: Chakraborty, Souradip, et al.
Publicado: (2024)
Learning Personalized Agents from Human Feedback
por: Liang, Kaiqu, et al.
Publicado: (2026)
por: Liang, Kaiqu, et al.
Publicado: (2026)
Mental Modeling of Reinforcement Learning Agents by Language Models
por: Lu, Wenhao, et al.
Publicado: (2024)
por: Lu, Wenhao, et al.
Publicado: (2024)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
por: Singh, Anukriti, et al.
Publicado: (2025)
por: Singh, Anukriti, et al.
Publicado: (2025)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
por: Li, Haozhan, et al.
Publicado: (2025)
por: Li, Haozhan, et al.
Publicado: (2025)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
por: Xie, Tianbao, et al.
Publicado: (2023)
por: Xie, Tianbao, et al.
Publicado: (2023)
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods
por: Cao, Yuji, et al.
Publicado: (2024)
por: Cao, Yuji, et al.
Publicado: (2024)
Parameter Efficient Reinforcement Learning from Human Feedback
por: Sidahmed, Hakim, et al.
Publicado: (2024)
por: Sidahmed, Hakim, et al.
Publicado: (2024)
Distributionally Robust Reinforcement Learning with Human Feedback
por: Mandal, Debmalya, et al.
Publicado: (2025)
por: Mandal, Debmalya, et al.
Publicado: (2025)
The RL/LLM Taxonomy Tree: Reviewing Synergies Between Reinforcement Learning and Large Language Models
por: Pternea, Moschoula, et al.
Publicado: (2024)
por: Pternea, Moschoula, et al.
Publicado: (2024)
Enhancing Personalized Multi-Turn Dialogue with Curiosity Reward
por: Wan, Yanming, et al.
Publicado: (2025)
por: Wan, Yanming, et al.
Publicado: (2025)
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
por: Lee, Harrison, et al.
Publicado: (2023)
por: Lee, Harrison, et al.
Publicado: (2023)
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
por: Nam, Hyunji, et al.
Publicado: (2026)
por: Nam, Hyunji, et al.
Publicado: (2026)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
por: Zheng, Qinqing, et al.
Publicado: (2024)
por: Zheng, Qinqing, et al.
Publicado: (2024)
Personalized Language Modeling from Personalized Human Feedback
por: Li, Xinyu, et al.
Publicado: (2024)
por: Li, Xinyu, et al.
Publicado: (2024)
Learning to Cooperate with Humans using Generative Agents
por: Liang, Yancheng, et al.
Publicado: (2024)
por: Liang, Yancheng, et al.
Publicado: (2024)
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
por: Huang, Zilin, et al.
Publicado: (2024)
por: Huang, Zilin, et al.
Publicado: (2024)
Off-Policy Corrected Reward Modeling for Reinforcement Learning from Human Feedback
por: Ackermann, Johannes, et al.
Publicado: (2025)
por: Ackermann, Johannes, et al.
Publicado: (2025)
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
por: Zhang, Shun, et al.
Publicado: (2024)
por: Zhang, Shun, et al.
Publicado: (2024)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
por: Yuan, Yifu, et al.
Publicado: (2024)
por: Yuan, Yifu, et al.
Publicado: (2024)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
por: Hwang, Minjune, et al.
Publicado: (2026)
por: Hwang, Minjune, et al.
Publicado: (2026)
Predictive Preference Learning from Human Interventions
por: Cai, Haoyuan, et al.
Publicado: (2025)
por: Cai, Haoyuan, et al.
Publicado: (2025)
Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning
por: Mo, Shentong
Publicado: (2026)
por: Mo, Shentong
Publicado: (2026)
RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences
por: Cheng, Jie, et al.
Publicado: (2024)
por: Cheng, Jie, et al.
Publicado: (2024)
Reinforcement Learning with Backtracking Feedback
por: Sel, Bilgehan, et al.
Publicado: (2026)
por: Sel, Bilgehan, et al.
Publicado: (2026)
Reinforcement Learning from Human Feedback with Active Queries
por: Ji, Kaixuan, et al.
Publicado: (2024)
por: Ji, Kaixuan, et al.
Publicado: (2024)
RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
por: Chaudhari, Shreyas, et al.
Publicado: (2024)
por: Chaudhari, Shreyas, et al.
Publicado: (2024)
ProVox: Personalization and Proactive Planning for Situated Human-Robot Collaboration
por: Grannen, Jennifer, et al.
Publicado: (2025)
por: Grannen, Jennifer, et al.
Publicado: (2025)
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning
por: Li, Jianxiong, et al.
Publicado: (2024)
por: Li, Jianxiong, et al.
Publicado: (2024)
LLM-based Multi-Agent Reinforcement Learning: Current and Future Directions
por: Sun, Chuanneng, et al.
Publicado: (2024)
por: Sun, Chuanneng, et al.
Publicado: (2024)
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
por: Chittepu, Yaswanth, et al.
Publicado: (2025)
por: Chittepu, Yaswanth, et al.
Publicado: (2025)
Swap-guided Preference Learning for Personalized Reinforcement Learning from Human Feedback
por: Kim, Gihoon, et al.
Publicado: (2026)
por: Kim, Gihoon, et al.
Publicado: (2026)
Adaptive Querying for Reward Learning from Human Feedback
por: Anand, Yashwanthi, et al.
Publicado: (2024)
por: Anand, Yashwanthi, et al.
Publicado: (2024)
Gradient Regularization Prevents Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards
por: Ackermann, Johannes, et al.
Publicado: (2026)
por: Ackermann, Johannes, et al.
Publicado: (2026)
Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases
por: Hahm, Dongyoon, et al.
Publicado: (2026)
por: Hahm, Dongyoon, et al.
Publicado: (2026)
Vocal Sandbox: Continual Learning and Adaptation for Situated Human-Robot Collaboration
por: Grannen, Jennifer, et al.
Publicado: (2024)
por: Grannen, Jennifer, et al.
Publicado: (2024)
Learning to summarize user information for personalized reinforcement learning from human feedback
por: Nam, Hyunji, et al.
Publicado: (2025)
por: Nam, Hyunji, et al.
Publicado: (2025)
Residual Reward Models for Preference-based Reinforcement Learning
por: Cao, Chenyang, et al.
Publicado: (2025)
por: Cao, Chenyang, et al.
Publicado: (2025)
Batch Active Learning of Reward Functions from Human Preferences
por: Bıyık, Erdem, et al.
Publicado: (2024)
por: Bıyık, Erdem, et al.
Publicado: (2024)
Ejemplares similares
-
Infer Human's Intentions Before Following Natural Language Instructions
por: Wan, Yanming, et al.
Publicado: (2024) -
MaxMin-RLHF: Alignment with Diverse Human Preferences
por: Chakraborty, Souradip, et al.
Publicado: (2024) -
Learning Personalized Agents from Human Feedback
por: Liang, Kaiqu, et al.
Publicado: (2026) -
Mental Modeling of Reinforcement Learning Agents by Language Models
por: Lu, Wenhao, et al.
Publicado: (2024) -
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
por: Singh, Anukriti, et al.
Publicado: (2025)