Evaluation-Aware Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Deshmukh, Shripad Vilasrao, Schwarzer, Will, Niekum, Scott |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Descriptive and Normative Theory of Human Beliefs in RLHF
by: Dandekar, Sylee, et al.
Published: (2025)
by: Dandekar, Sylee, et al.
Published: (2025)
Explaining RL Decisions with Trajectories
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
by: Chittepu, Yaswanth, et al.
Published: (2025)
by: Chittepu, Yaswanth, et al.
Published: (2025)
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
by: Kim, Kihyun, et al.
Published: (2026)
by: Kim, Kihyun, et al.
Published: (2026)
Training ML Models with Predictable Failures
by: Schwarzer, Will, et al.
Published: (2026)
by: Schwarzer, Will, et al.
Published: (2026)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Null Counterfactual Factor Interactions for Goal-Conditioned Reinforcement Learning
by: Chuck, Caleb, et al.
Published: (2025)
by: Chuck, Caleb, et al.
Published: (2025)
COSAC: Counterfactual Credit Assignment in Sequential Cooperative Teams
by: Deshmukh, Shripad, et al.
Published: (2026)
by: Deshmukh, Shripad, et al.
Published: (2026)
Pareto-Optimal Learning from Preferences with Hidden Context
by: Bahlous-Boldi, Ryan, et al.
Published: (2024)
by: Bahlous-Boldi, Ryan, et al.
Published: (2024)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
by: Chittepu, Yaswanth, et al.
Published: (2026)
by: Chittepu, Yaswanth, et al.
Published: (2026)
Learning Action-based Representations Using Invariance
by: Rudolph, Max, et al.
Published: (2024)
by: Rudolph, Max, et al.
Published: (2024)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
Adaptive Margin RLHF via Preference over Preferences
by: Chittepu, Yaswanth, et al.
Published: (2025)
by: Chittepu, Yaswanth, et al.
Published: (2025)
Supervised Reward Inference
by: Schwarzer, Will, et al.
Published: (2025)
by: Schwarzer, Will, et al.
Published: (2025)
On Zero-Shot Reinforcement Learning
by: Jeen, Scott
Published: (2025)
by: Jeen, Scott
Published: (2025)
Automated Discovery of Functional Actual Causes in Complex Environments
by: Chuck, Caleb, et al.
Published: (2024)
by: Chuck, Caleb, et al.
Published: (2024)
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
by: Jajoo, Pranaya, et al.
Published: (2026)
by: Jajoo, Pranaya, et al.
Published: (2026)
A Competition Winning Deep Reinforcement Learning Agent in microRTS
by: Goodfriend, Scott
Published: (2024)
by: Goodfriend, Scott
Published: (2024)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
CALF: Communication-Aware Learning Framework for Distributed Reinforcement Learning
by: Purves, Carlos, et al.
Published: (2026)
by: Purves, Carlos, et al.
Published: (2026)
Causal-Aware Generative Adversarial Networks with Reinforcement Learning
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
Zero-Shot Reinforcement Learning Under Partial Observability
by: Jeen, Scott, et al.
Published: (2025)
by: Jeen, Scott, et al.
Published: (2025)
Zero-Shot Reinforcement Learning from Low Quality Data
by: Jeen, Scott, et al.
Published: (2023)
by: Jeen, Scott, et al.
Published: (2023)
Object-Centric World Models for Causality-Aware Reinforcement Learning
by: Nishimoto, Yosuke, et al.
Published: (2025)
by: Nishimoto, Yosuke, et al.
Published: (2025)
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem
by: Strauß, Niklas, et al.
Published: (2024)
by: Strauß, Niklas, et al.
Published: (2024)
REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge
by: Zhang, Yasi, et al.
Published: (2026)
by: Zhang, Yasi, et al.
Published: (2026)
Learning by Doing: An Online Causal Reinforcement Learning Framework with Causal-Aware Policy
by: Cai, Ruichu, et al.
Published: (2024)
by: Cai, Ruichu, et al.
Published: (2024)
Learning When to Act: Interval-Aware Reinforcement Learning with Predictive Temporal Structure
by: Di Gioia, Davide
Published: (2026)
by: Di Gioia, Davide
Published: (2026)
Fast Adaptation with Behavioral Foundation Models
by: Sikchi, Harshit, et al.
Published: (2025)
by: Sikchi, Harshit, et al.
Published: (2025)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
Inverse Reinforcement Learning from Non-Stationary Learning Agents
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization
by: Zhao, Yunyi, et al.
Published: (2025)
by: Zhao, Yunyi, et al.
Published: (2025)
Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
by: Talvitie, Erin J., et al.
Published: (2024)
by: Talvitie, Erin J., et al.
Published: (2024)
Time-Varying Constraint-Aware Reinforcement Learning for Energy Storage Control
by: Jeong, Jaeik, et al.
Published: (2024)
by: Jeong, Jaeik, et al.
Published: (2024)
Entropy-Aware Model Initialization for Effective Exploration in Deep Reinforcement Learning
by: Jang, Sooyoung, et al.
Published: (2021)
by: Jang, Sooyoung, et al.
Published: (2021)
From Observations to Events: Event-Aware World Model for Reinforcement Learning
by: Peng, Zhao-Han, et al.
Published: (2026)
by: Peng, Zhao-Han, et al.
Published: (2026)
Similar Items
-
A Descriptive and Normative Theory of Human Beliefs in RLHF
by: Dandekar, Sylee, et al.
Published: (2025) -
Explaining RL Decisions with Trajectories
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023) -
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
by: Chittepu, Yaswanth, et al.
Published: (2025) -
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
by: Kim, Kihyun, et al.
Published: (2026) -
Training ML Models with Predictable Failures
by: Schwarzer, Will, et al.
Published: (2026)