Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Yifu, Hao, Jianye, Ma, Yi, Dong, Zibin, Liang, Hebin, Liu, Jinyi, Feng, Zhixin, Zhao, Kai, Zheng, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024)
by: Zhou, Xinglin, et al.
Published: (2024)
RLHF May Not Reflect Genuine Preferences
by: Ghafouri, Bijean, et al.
Published: (2026)
by: Ghafouri, Bijean, et al.
Published: (2026)
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
Demonstrating HumanTHOR: A Simulation Platform and Benchmark for Human-Robot Collaboration in a Shared Workspace
by: Wang, Chenxu, et al.
Published: (2024)
by: Wang, Chenxu, et al.
Published: (2024)
Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework
by: Metz, Yannick, et al.
Published: (2024)
by: Metz, Yannick, et al.
Published: (2024)
Same Feedback, Different Source: How AI vs. Human Feedback Shapes Learner Engagement
by: Morris, Caitlin, et al.
Published: (2026)
by: Morris, Caitlin, et al.
Published: (2026)
Impact of Cognitive Load on Human Trust in Hybrid Human-Robot Collaboration
by: Guo, Hao, et al.
Published: (2024)
by: Guo, Hao, et al.
Published: (2024)
Facilitating Human Feedback for GenAI Prompt Optimization
by: Sherson, Jacob, et al.
Published: (2024)
by: Sherson, Jacob, et al.
Published: (2024)
MoHeat: A Modular Platform for High-Responsive Non-Contact Thermal Feedback Interactions
by: Xu, Jiayi, et al.
Published: (2024)
by: Xu, Jiayi, et al.
Published: (2024)
Integrating Human Feedback into a Reinforcement Learning-Based Framework for Adaptive User Interfaces
by: Gaspar-Figueiredo, Daniel, et al.
Published: (2025)
by: Gaspar-Figueiredo, Daniel, et al.
Published: (2025)
When Benchmarks Talk: Re-Evaluating Code LLMs with Interactive Feedback
by: Pan, Jane, et al.
Published: (2025)
by: Pan, Jane, et al.
Published: (2025)
JetUnit: Rendering Diverse Force Feedback in Virtual Reality Using Water Jets
by: Zhang, Zining, et al.
Published: (2024)
by: Zhang, Zining, et al.
Published: (2024)
UbiPhysio: Support Daily Functioning, Fitness, and Rehabilitation with Action Understanding and Feedback in Natural Language
by: Wang, Chongyang, et al.
Published: (2023)
by: Wang, Chongyang, et al.
Published: (2023)
PPTP: Performance-Guided Physiological Signal-Based Trust Prediction in Human-Robot Collaboration
by: Guo, Hao, et al.
Published: (2025)
by: Guo, Hao, et al.
Published: (2025)
Same Feedback, Different Source: How AI vs. Human Feedback Attribution and Credibility Shape Learner Behavior in Computing Education
by: Morris, Caitlin, et al.
Published: (2026)
by: Morris, Caitlin, et al.
Published: (2026)
The Future of Open Human Feedback
by: Don-Yehiya, Shachar, et al.
Published: (2024)
by: Don-Yehiya, Shachar, et al.
Published: (2024)
Fostering Human Learning in Sequential Decision-Making: Understanding the Role of Evaluative Feedback
by: Gupta, Piyush, et al.
Published: (2023)
by: Gupta, Piyush, et al.
Published: (2023)
A Comprehensive Survey of Electrical Stimulation Haptic Feedback in Human-Computer Interaction
by: Yang, Simin, et al.
Published: (2025)
by: Yang, Simin, et al.
Published: (2025)
Exploring Re-inforcement Learning via Human Feedback under User Heterogeneity
by: Shashidhar, Sarvesh, et al.
Published: (2026)
by: Shashidhar, Sarvesh, et al.
Published: (2026)
Evolving Agents: Interactive Simulation of Dynamic and Diverse Human Personalities
by: Li, Jiale, et al.
Published: (2024)
by: Li, Jiale, et al.
Published: (2024)
Use of Winsome Robots for Understanding Human Feedback (UWU)
by: Eggers, Jessica, et al.
Published: (2025)
by: Eggers, Jessica, et al.
Published: (2025)
IELTS Writing Revision Platform with Automated Essay Scoring and Adaptive Feedback
by: Ramancauskas, Titas, et al.
Published: (2025)
by: Ramancauskas, Titas, et al.
Published: (2025)
Does Positive Reinforcement Work?: A Quasi-Experimental Study of the Effects of Positive Feedback on Reddit
by: Lambert, Charlotte, et al.
Published: (2024)
by: Lambert, Charlotte, et al.
Published: (2024)
Feedstack: Layering Structured Representations over Unstructured Feedback to Scaffold Human AI Conversation
by: Nguyen, Hannah Vy, et al.
Published: (2025)
by: Nguyen, Hannah Vy, et al.
Published: (2025)
The Observability Gap: Why Output-Level Human Feedback Fails for LLM Coding Agents
by: Wang, Yinghao, et al.
Published: (2026)
by: Wang, Yinghao, et al.
Published: (2026)
PointAloud: An Interaction Suite for AI-Supported Pointer-Centric Think-Aloud Computing
by: Gmeiner, Frederic, et al.
Published: (2026)
by: Gmeiner, Frederic, et al.
Published: (2026)
manvr3d: A Platform for Human-in-the-loop Cell Tracking in Virtual Reality
by: Pantze, Samuel, et al.
Published: (2025)
by: Pantze, Samuel, et al.
Published: (2025)
ProVoice: Designing Proactive Functionality for In-Vehicle Conversational Assistants using Multi-Objective Bayesian Optimization to Enhance Driver Experience
by: Susak, Josh, et al.
Published: (2026)
by: Susak, Josh, et al.
Published: (2026)
State-Dependent Refusal and Learned Incapacity in RLHF-Aligned Language Models
by: Lee, TK
Published: (2025)
by: Lee, TK
Published: (2025)
Hybrid Team Tetris: A New Platform For Hybrid Multi-Agent, Multi-Human Teaming
by: Mcdowell, Kaleb, et al.
Published: (2025)
by: Mcdowell, Kaleb, et al.
Published: (2025)
A Benchmark to Assess Common Ground in Human-AI Collaboration
by: Poelitz, Christian, et al.
Published: (2026)
by: Poelitz, Christian, et al.
Published: (2026)
To Ask or Not to Ask: Learning to Require Human Feedback
by: Pugnana, Andrea, et al.
Published: (2025)
by: Pugnana, Andrea, et al.
Published: (2025)
Tool Compensation and User Strategy during Human-Robot Teleoperation are Impacted by System Dynamics and Kinesthetic Feedback
by: Carducci, Jacob D., et al.
Published: (2024)
by: Carducci, Jacob D., et al.
Published: (2024)
MEDebiaser: A Human-AI Feedback System for Mitigating Bias in Multi-label Medical Image Classification
by: Shi, Shaohan, et al.
Published: (2025)
by: Shi, Shaohan, et al.
Published: (2025)
Exploring Uni-manual Around Ear Off-Device Gestures for Earables
by: Shimon, Shaikh Shawon Arefin, et al.
Published: (2024)
by: Shimon, Shaikh Shawon Arefin, et al.
Published: (2024)
Not Too Long, Not Too Short: Goldilocks Principle of 'Optimal' Reflection Time on Online Deliberation Platforms
by: Yeo, ShunYi, et al.
Published: (2024)
by: Yeo, ShunYi, et al.
Published: (2024)
CollaClassroom: An AI-Augmented Collaborative Learning Platform with LLM Support in the Context of Bangladeshi University Students
by: Sayeed, Salman, et al.
Published: (2025)
by: Sayeed, Salman, et al.
Published: (2025)
Perceptual Analysis of Groups of Virtual Humans Animated using Interactive Platforms
by: Montanha, Rubens, et al.
Published: (2024)
by: Montanha, Rubens, et al.
Published: (2024)
Designing Wine Tasting Experiences for All: The role of Human Diversity and Personal food memory
by: Shan, Xinyang, et al.
Published: (2025)
by: Shan, Xinyang, et al.
Published: (2025)
AutoLegend: A User Feedback-Driven Adaptive Legend Generator for Visualizations
by: Liu, Can, et al.
Published: (2024)
by: Liu, Can, et al.
Published: (2024)
Similar Items
-
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024) -
RLHF May Not Reflect Genuine Preferences
by: Ghafouri, Bijean, et al.
Published: (2026) -
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025) -
Demonstrating HumanTHOR: A Simulation Platform and Benchmark for Human-Robot Collaboration in a Shared Workspace
by: Wang, Chenxu, et al.
Published: (2024) -
Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework
by: Metz, Yannick, et al.
Published: (2024)