Learning a Reward Function for User-Preferred Appliance Scheduling
Fuente:
arXiv
Saved in:
| Main Authors: | Čović, Nikolina, Cremer, Jochen L., Pandžić, Hrvoje |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transferable Graph Learning for Transmission Congestion Management via Busbar Splitting
by: Rajaei, Ali, et al.
Published: (2025)
by: Rajaei, Ali, et al.
Published: (2025)
Robot Operation of Home Appliances by Reading User Manuals
by: Zhang, Jian, et al.
Published: (2025)
by: Zhang, Jian, et al.
Published: (2025)
Towards Generalization of Graph Neural Networks for AC Optimal Power Flow
by: Arowolo, Olayiwola, et al.
Published: (2025)
by: Arowolo, Olayiwola, et al.
Published: (2025)
MTRec: Learning to Align with User Preferences via Mental Reward Models
by: Zhao, Mengchen, et al.
Published: (2025)
by: Zhao, Mengchen, et al.
Published: (2025)
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)
by: Bıyık, Erdem, et al.
Published: (2024)
A Generalized Acquisition Function for Preference-based Reward Learning
by: Ellis, Evan, et al.
Published: (2024)
by: Ellis, Evan, et al.
Published: (2024)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
by: Dennler, Nathaniel, et al.
Published: (2024)
by: Dennler, Nathaniel, et al.
Published: (2024)
Reward Learning From Preference With Ties
by: Liu, Jinsong, et al.
Published: (2024)
by: Liu, Jinsong, et al.
Published: (2024)
Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
by: Wang, Haoxiang, et al.
Published: (2024)
by: Wang, Haoxiang, et al.
Published: (2024)
RealAppliance: Let High-fidelity Appliance Assets Controllable and Workable as Aligned Real Manuals
by: Gao, Yuzheng, et al.
Published: (2025)
by: Gao, Yuzheng, et al.
Published: (2025)
Search-Based Robot Motion Planning With Distance-Based Adaptive Motion Primitives
by: Kraljusic, Benjamin, et al.
Published: (2025)
by: Kraljusic, Benjamin, et al.
Published: (2025)
Implementing Rational Choice Functions with LLMs and Measuring their Alignment with User Preferences
by: Karnysheva, Anna, et al.
Published: (2025)
by: Karnysheva, Anna, et al.
Published: (2025)
Rectifying Shortcut Behaviors in Preference-based Reward Learning
by: Ye, Wenqian, et al.
Published: (2025)
by: Ye, Wenqian, et al.
Published: (2025)
Exploring and Addressing Reward Confusion in Offline Preference Learning
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Multi-Objective Reinforcement Learning for Power Grid Topology Control
by: Lautenbacher, Thomas, et al.
Published: (2025)
by: Lautenbacher, Thomas, et al.
Published: (2025)
Preference as Reward, Maximum Preference Optimization with Importance Sampling
by: Jiang, Zaifan, et al.
Published: (2023)
by: Jiang, Zaifan, et al.
Published: (2023)
Understanding User Preferences in Explainable Artificial Intelligence: A Survey and a Mapping Function Proposal
by: Hashemi, Maryam, et al.
Published: (2023)
by: Hashemi, Maryam, et al.
Published: (2023)
Preference Poisoning Attacks on Reward Model Learning
by: Wu, Junlin, et al.
Published: (2024)
by: Wu, Junlin, et al.
Published: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
by: Cao, Chenyang, et al.
Published: (2025)
by: Cao, Chenyang, et al.
Published: (2025)
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
by: Choi, Heewoong, et al.
Published: (2024)
by: Choi, Heewoong, et al.
Published: (2024)
Selective Preference Optimization via Token-Level Reward Function Estimation
by: Yang, Kailai, et al.
Published: (2024)
by: Yang, Kailai, et al.
Published: (2024)
Towards Comprehensive Preference Data Collection for Reward Modeling
by: Hu, Yulan, et al.
Published: (2024)
by: Hu, Yulan, et al.
Published: (2024)
Comparing Post-Hoc Explainable AI Methods for Interpreting Black-Box EEG Models in Depression Detection
by: Šarčević, Antonia, et al.
Published: (2026)
by: Šarčević, Antonia, et al.
Published: (2026)
Learning Transferable Latent User Preferences for Human-Aligned Decision Making
by: Hyk, Alina, et al.
Published: (2026)
by: Hyk, Alina, et al.
Published: (2026)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025)
by: Huang, Changxin, et al.
Published: (2025)
Provable Reward-Agnostic Preference-Based Reinforcement Learning
by: Zhan, Wenhao, et al.
Published: (2023)
by: Zhan, Wenhao, et al.
Published: (2023)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)
by: Misra, Dipendra, et al.
Published: (2026)
Similarity as Reward Alignment: Robust and Versatile Preference-based Reinforcement Learning
by: Rajaram, Sara, et al.
Published: (2025)
by: Rajaram, Sara, et al.
Published: (2025)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024)
by: Metcalf, Katherine, et al.
Published: (2024)
Energy Disaggregation & Appliance Identification in a Smart Home: Transfer Learning enables Edge Computing
by: Shahab, M. Hashim, et al.
Published: (2023)
by: Shahab, M. Hashim, et al.
Published: (2023)
Automated Machine Learning: A Case Study on Non-Intrusive Appliance Load Monitoring
by: Moin, Armin, et al.
Published: (2022)
by: Moin, Armin, et al.
Published: (2022)
Aligning Crowd Feedback via Distributional Preference Reward Modeling
by: Li, Dexun, et al.
Published: (2024)
by: Li, Dexun, et al.
Published: (2024)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Overthinking Reduction with Decoupled Rewards and Curriculum Data Scheduling
by: Jiang, Shuyang, et al.
Published: (2025)
by: Jiang, Shuyang, et al.
Published: (2025)
In-Context Reward Adaptation for Robust Preference Modeling
by: Sun, Zhenyu, et al.
Published: (2026)
by: Sun, Zhenyu, et al.
Published: (2026)
Capturing Individual Human Preferences with Reward Features
by: Barreto, André, et al.
Published: (2025)
by: Barreto, André, et al.
Published: (2025)
APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
by: Li, Zhuo, et al.
Published: (2025)
by: Li, Zhuo, et al.
Published: (2025)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
by: Hwang, Minjune, et al.
Published: (2026)
by: Hwang, Minjune, et al.
Published: (2026)
Users as Annotators: LLM Preference Learning from Comparison Mode
by: Cai, Zhongze, et al.
Published: (2025)
by: Cai, Zhongze, et al.
Published: (2025)
Similar Items
-
Transferable Graph Learning for Transmission Congestion Management via Busbar Splitting
by: Rajaei, Ali, et al.
Published: (2025) -
Robot Operation of Home Appliances by Reading User Manuals
by: Zhang, Jian, et al.
Published: (2025) -
Towards Generalization of Graph Neural Networks for AC Optimal Power Flow
by: Arowolo, Olayiwola, et al.
Published: (2025) -
MTRec: Learning to Align with User Preferences via Mental Reward Models
by: Zhao, Mengchen, et al.
Published: (2025) -
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)