Principled Reinforcement Learning with Human Feedback from Pairwise or $K$-wise Comparisons
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Banghua, Jiao, Jiantao, Jordan, Michael I. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024)
by: Zhou, Xinglin, et al.
Published: (2024)
Towards Interactive Reinforcement Learning with Intrinsic Feedback
by: Poole, Benjamin, et al.
Published: (2021)
by: Poole, Benjamin, et al.
Published: (2021)
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
by: Huang, Zilin, et al.
Published: (2024)
by: Huang, Zilin, et al.
Published: (2024)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Iterative Data Smoothing: Mitigating Reward Overfitting and Overoptimization in RLHF
by: Zhu, Banghua, et al.
Published: (2024)
by: Zhu, Banghua, et al.
Published: (2024)
Understanding Impact of Human Feedback via Influence Functions
by: Min, Taywon, et al.
Published: (2025)
by: Min, Taywon, et al.
Published: (2025)
Vibrotactile Preference Learning: Uncertainty-Aware Preference Learning for Personalized Vibration Feedback
by: Zhang, Rongtao, et al.
Published: (2026)
by: Zhang, Rongtao, et al.
Published: (2026)
Quantifying the Effect of Feedback Frequency in Interactive Reinforcement Learning for Robotic Tasks
by: Harnack, Daniel, et al.
Published: (2022)
by: Harnack, Daniel, et al.
Published: (2022)
On the Exponential Convergence for Offline RLHF with Pairwise Comparisons
by: Chen, Zhirui, et al.
Published: (2024)
by: Chen, Zhirui, et al.
Published: (2024)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
Cognitive Exoskeleton: Augmenting Human Cognition with an AI-Mediated Intelligent Visual Feedback
by: Xu, Songlin, et al.
Published: (2025)
by: Xu, Songlin, et al.
Published: (2025)
Reinforcement Learning Driven Generalizable Feature Representation for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2025)
by: Ye, Xiaozhou, et al.
Published: (2025)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
by: Hiranaka, Ayano, et al.
Published: (2024)
by: Hiranaka, Ayano, et al.
Published: (2024)
Sentiment Analysis in Learning Management Systems Understanding Student Feedback at Scale
by: Almutairi, Mohammed
Published: (2025)
by: Almutairi, Mohammed
Published: (2025)
How many labelers do you have? A closer look at gold-standard labels
by: Cheng, Chen, et al.
Published: (2022)
by: Cheng, Chen, et al.
Published: (2022)
Abstracted Trajectory Visualization for Explainability in Reinforcement Learning
by: Takagi, Yoshiki, et al.
Published: (2024)
by: Takagi, Yoshiki, et al.
Published: (2024)
Dodgersort: Uncertainty-Aware VLM-Guided Human-in-the-Loop Pairwise Ranking
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
dUltra: Ultra-Fast Diffusion Language Models via Reinforcement Learning
by: Chen, Shirui, et al.
Published: (2025)
by: Chen, Shirui, et al.
Published: (2025)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Information Plane Analysis Visualization in Deep Learning via Transfer Entropy
by: Moldovan, Adrian, et al.
Published: (2024)
by: Moldovan, Adrian, et al.
Published: (2024)
Learning to Decide with AI Assistance under Human-Alignment
by: Benz, Nina Corvelo, et al.
Published: (2026)
by: Benz, Nina Corvelo, et al.
Published: (2026)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Domain-Adversarial Anatomical Graph Networks for Cross-User Human Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2025)
by: Ye, Xiaozhou, et al.
Published: (2025)
Reciprocal Learning of Intent Inferral with Augmented Visual Feedback for Stroke
by: Xu, Jingxi, et al.
Published: (2024)
by: Xu, Jingxi, et al.
Published: (2024)
Introducing User Feedback-based Counterfactual Explanations (UFCE)
by: Suffian, Muhammad, et al.
Published: (2024)
by: Suffian, Muhammad, et al.
Published: (2024)
Integrating Policy Summaries with Reward Decomposition for Explaining Reinforcement Learning Agents
by: Septon, Yael, et al.
Published: (2022)
by: Septon, Yael, et al.
Published: (2022)
Towards Human-Centered Construction Robotics: A Reinforcement Learning-Driven Companion Robot for Contextually Assisting Carpentry Workers
by: Wu, Yuning, et al.
Published: (2024)
by: Wu, Yuning, et al.
Published: (2024)
Making RL with Preference-based Feedback Efficient via Randomization
by: Wu, Runzhe, et al.
Published: (2023)
by: Wu, Runzhe, et al.
Published: (2023)
AURA: A Reinforcement Learning Framework for AI-Driven Adaptive Conversational Surveys
by: Tang, Jinwen, et al.
Published: (2025)
by: Tang, Jinwen, et al.
Published: (2025)
Conformalized Interactive Imitation Learning: Handling Expert Shift and Intermittent Feedback
by: Zhao, Michelle, et al.
Published: (2024)
by: Zhao, Michelle, et al.
Published: (2024)
Learning a Canonical Basis of Human Preferences from Binary Ratings
by: Vodrahalli, Kailas, et al.
Published: (2025)
by: Vodrahalli, Kailas, et al.
Published: (2025)
AI, Meet Human: Learning Paradigms for Hybrid Decision Making Systems
by: Punzi, Clara, et al.
Published: (2024)
by: Punzi, Clara, et al.
Published: (2024)
PL-DCP: A Pairwise Learning framework with Domain and Class Prototypes for EEG emotion recognition under unseen target conditions
by: Li, Guangli, et al.
Published: (2024)
by: Li, Guangli, et al.
Published: (2024)
Predictive AI Can Support Human Learning while Preserving Error Diversity
by: He, Vivianna Fang, et al.
Published: (2025)
by: He, Vivianna Fang, et al.
Published: (2025)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
by: Okumura, Ryota, et al.
Published: (2025)
by: Okumura, Ryota, et al.
Published: (2025)
Coprocessor Actor Critic: A Model-Based Reinforcement Learning Approach For Adaptive Brain Stimulation
by: Pan, Michelle, et al.
Published: (2024)
by: Pan, Michelle, et al.
Published: (2024)
HARP: Human-Assisted Regrouping with Permutation Invariant Critic for Multi-Agent Reinforcement Learning
by: Hu, Huawen, et al.
Published: (2024)
by: Hu, Huawen, et al.
Published: (2024)
Culturally-Attuned Moral Machines: Implicit Learning of Human Value Systems by AI through Inverse Reinforcement Learning
by: Oliveira, Nigini, et al.
Published: (2023)
by: Oliveira, Nigini, et al.
Published: (2023)
A Light-weight Deep Human Activity Recognition Algorithm Using Multi-knowledge Distillation
by: Chen, Runze, et al.
Published: (2021)
by: Chen, Runze, et al.
Published: (2021)
Human-Centric Aware UAV Trajectory Planning in Search and Rescue Missions Employing Multi-Objective Reinforcement Learning with AHP and Similarity-Based Experience Replay
by: Ramezani, Mahya, et al.
Published: (2024)
by: Ramezani, Mahya, et al.
Published: (2024)
Similar Items
-
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024) -
Towards Interactive Reinforcement Learning with Intrinsic Feedback
by: Poole, Benjamin, et al.
Published: (2021) -
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
by: Huang, Zilin, et al.
Published: (2024) -
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024) -
Iterative Data Smoothing: Mitigating Reward Overfitting and Overoptimization in RLHF
by: Zhu, Banghua, et al.
Published: (2024)