Reward Centering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Naik, Abhishek, Wan, Yi, Tomar, Manan, Sutton, Richard S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Swift-Sarsa: Fast and Robust Linear Control
von: Javed, Khurram, et al.
Veröffentlicht: (2025)
von: Javed, Khurram, et al.
Veröffentlicht: (2025)
Adaptive and Explainable AI Agents for Anomaly Detection in Critical IoT Infrastructure using LLM-Enhanced Contextual Reasoning
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
Epigraph-Guided Flow Matching for Safe and Performant Offline Reinforcement Learning
von: Tayal, Manan, et al.
Veröffentlicht: (2026)
von: Tayal, Manan, et al.
Veröffentlicht: (2026)
Small Language Models for Agentic Systems: A Survey of Architectures, Capabilities, and Deployment Trade offs
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
von: Sharma, Raghav, et al.
Veröffentlicht: (2025)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
Robust Offline Reinforcement learning with Heavy-Tailed Rewards
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
Transductive Reward Inference on Graph
von: Qu, Bohao, et al.
Veröffentlicht: (2024)
von: Qu, Bohao, et al.
Veröffentlicht: (2024)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
Mutual-Taught for Co-adapting Policy and Reward Models
von: Shi, Tianyuan, et al.
Veröffentlicht: (2025)
von: Shi, Tianyuan, et al.
Veröffentlicht: (2025)
Reward Models in Deep Reinforcement Learning: A Survey
von: Yu, Rui, et al.
Veröffentlicht: (2025)
von: Yu, Rui, et al.
Veröffentlicht: (2025)
Cross Domain Adaptation using Adversarial networks with Cyclic loss
von: Kaur, Manpreet, et al.
Veröffentlicht: (2024)
von: Kaur, Manpreet, et al.
Veröffentlicht: (2024)
APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
Reinforcement Learning-based Approach for Vehicle-to-Building Charging with Heterogeneous Agents and Long Term Rewards
von: Liu, Fangqi, et al.
Veröffentlicht: (2025)
von: Liu, Fangqi, et al.
Veröffentlicht: (2025)
Eliminating Inductive Bias in Reward Models with Information-Theoretic Guidance
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
Step-size Optimization for Continual Learning
von: Degris, Thomas, et al.
Veröffentlicht: (2024)
von: Degris, Thomas, et al.
Veröffentlicht: (2024)
Attention-Based Reward Shaping for Sparse and Delayed Rewards
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
Reward Hacking Mitigation using Verifiable Composite Rewards
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
von: Tarek, Mirza Farhan Bin, et al.
Veröffentlicht: (2025)
Repairing Reward Functions with Feedback to Mitigate Reward Hacking
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2025)
Intrinsic Reward Policy Optimization for Sparse-Reward Environments
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
von: Cho, Minjae, et al.
Veröffentlicht: (2026)
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
MetaOptimize: A Framework for Optimizing Step Sizes and Other Meta-parameters
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2024)
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2024)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models
von: Kim, Yeongmin, et al.
Veröffentlicht: (2026)
von: Kim, Yeongmin, et al.
Veröffentlicht: (2026)
Learning a Diffusion Model Policy from Rewards via Q-Score Matching
von: Psenka, Michael, et al.
Veröffentlicht: (2023)
von: Psenka, Michael, et al.
Veröffentlicht: (2023)
Probabilistic Consensus through Ensemble Validation: A Framework for LLM Reliability
von: Naik, Ninad
Veröffentlicht: (2024)
von: Naik, Ninad
Veröffentlicht: (2024)
SemiReward: A General Reward Model for Semi-supervised Learning
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
von: Li, Siyuan, et al.
Veröffentlicht: (2023)
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
von: Chen, Yang, et al.
Veröffentlicht: (2025)
von: Chen, Yang, et al.
Veröffentlicht: (2025)
Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
von: Wang, Chaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Chaoqi, et al.
Veröffentlicht: (2025)
Tiered Reward: Designing Rewards for Specification and Fast Learning of Desired Behavior
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
Neural Reward Machines
von: Umili, Elena, et al.
Veröffentlicht: (2024)
von: Umili, Elena, et al.
Veröffentlicht: (2024)
Numeric Reward Machines
von: Levina, Kristina, et al.
Veröffentlicht: (2024)
von: Levina, Kristina, et al.
Veröffentlicht: (2024)
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning?
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2025)
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2025)
Which Rewards Matter? Reward Selection for Reinforcement Learning under Limited Feedback
von: Chaudhari, Shreyas, et al.
Veröffentlicht: (2025)
von: Chaudhari, Shreyas, et al.
Veröffentlicht: (2025)
Diffusion Reinforcement Learning via Centered Reward Distillation
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2026)
AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models
von: Liu, Qi, et al.
Veröffentlicht: (2025)
von: Liu, Qi, et al.
Veröffentlicht: (2025)
Auxiliary task discovery through generate-and-test
von: Rafiee, Banafsheh, et al.
Veröffentlicht: (2022)
von: Rafiee, Banafsheh, et al.
Veröffentlicht: (2022)
Intentional Updates for Streaming Reinforcement Learning
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2026)
von: Sharifnassab, Arsalan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Swift-Sarsa: Fast and Robust Linear Control
von: Javed, Khurram, et al.
Veröffentlicht: (2025) -
Adaptive and Explainable AI Agents for Anomaly Detection in Critical IoT Infrastructure using LLM-Enhanced Contextual Reasoning
von: Sharma, Raghav, et al.
Veröffentlicht: (2025) -
Epigraph-Guided Flow Matching for Safe and Performant Offline Reinforcement Learning
von: Tayal, Manan, et al.
Veröffentlicht: (2026) -
Small Language Models for Agentic Systems: A Survey of Architectures, Capabilities, and Deployment Trade offs
von: Sharma, Raghav, et al.
Veröffentlicht: (2025) -
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)