Supervised Reward Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Schwarzer, Will, Schneider, Jordan, Thomas, Philip S., Niekum, Scott |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Training ML Models with Predictable Failures
by: Schwarzer, Will, et al.
Published: (2026)
by: Schwarzer, Will, et al.
Published: (2026)
Evaluation-Aware Reinforcement Learning
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2025)
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2025)
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
by: Chittepu, Yaswanth, et al.
Published: (2025)
by: Chittepu, Yaswanth, et al.
Published: (2025)
Bayesian Robust Optimization for Imitation Learning
by: Brown, Daniel S., et al.
Published: (2020)
by: Brown, Daniel S., et al.
Published: (2020)
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation
by: Tripathi, Tuhina, et al.
Published: (2025)
by: Tripathi, Tuhina, et al.
Published: (2025)
On the Benefits of Inducing Local Lipschitzness for Robust Generative Adversarial Imitation Learning
by: Memarian, Farzan, et al.
Published: (2021)
by: Memarian, Farzan, et al.
Published: (2021)
Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
by: Rafailov, Rafael, et al.
Published: (2024)
by: Rafailov, Rafael, et al.
Published: (2024)
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
by: Chittepu, Yaswanth, et al.
Published: (2026)
by: Chittepu, Yaswanth, et al.
Published: (2026)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
Pareto-Optimal Learning from Preferences with Hidden Context
by: Bahlous-Boldi, Ryan, et al.
Published: (2024)
by: Bahlous-Boldi, Ryan, et al.
Published: (2024)
Adaptive Margin RLHF via Preference over Preferences
by: Chittepu, Yaswanth, et al.
Published: (2025)
by: Chittepu, Yaswanth, et al.
Published: (2025)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
A Descriptive and Normative Theory of Human Beliefs in RLHF
by: Dandekar, Sylee, et al.
Published: (2025)
by: Dandekar, Sylee, et al.
Published: (2025)
Are Deep Speech Denoising Models Robust to Adversarial Noise?
by: Schwarzer, Will, et al.
Published: (2025)
by: Schwarzer, Will, et al.
Published: (2025)
Learning Action-based Representations Using Invariance
by: Rudolph, Max, et al.
Published: (2024)
by: Rudolph, Max, et al.
Published: (2024)
Reward-RAG: Enhancing RAG with Reward Driven Supervision
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
Scaling Reward Modeling without Human Supervision
by: Fan, Jingxuan, et al.
Published: (2026)
by: Fan, Jingxuan, et al.
Published: (2026)
Features as Rewards: Scalable Supervision for Open-Ended Tasks via Interpretability
by: Prasad, Aaditya Vikram, et al.
Published: (2026)
by: Prasad, Aaditya Vikram, et al.
Published: (2026)
Bandit Simulation for Average Reward Inference
by: Praharaj, Samya, et al.
Published: (2026)
by: Praharaj, Samya, et al.
Published: (2026)
Automated Discovery of Functional Actual Causes in Complex Environments
by: Chuck, Caleb, et al.
Published: (2024)
by: Chuck, Caleb, et al.
Published: (2024)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
by: Jajoo, Pranaya, et al.
Published: (2026)
by: Jajoo, Pranaya, et al.
Published: (2026)
Transductive Reward Inference on Graph
by: Qu, Bohao, et al.
Published: (2024)
by: Qu, Bohao, et al.
Published: (2024)
Reward Machine Inference for Robotic Manipulation
by: Baert, Mattijs, et al.
Published: (2024)
by: Baert, Mattijs, et al.
Published: (2024)
Null Counterfactual Factor Interactions for Goal-Conditioned Reinforcement Learning
by: Chuck, Caleb, et al.
Published: (2025)
by: Chuck, Caleb, et al.
Published: (2025)
Which Rewards Matter? Reward Selection for Reinforcement Learning under Limited Feedback
by: Chaudhari, Shreyas, et al.
Published: (2025)
by: Chaudhari, Shreyas, et al.
Published: (2025)
TMS: Trajectory-Mixed Supervision for Reward-Free, On-Policy SFT
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2026)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2026)
Semi-Supervised Reward Modeling via Iterative Self-Training
by: He, Yifei, et al.
Published: (2024)
by: He, Yifei, et al.
Published: (2024)
Inference-Time Reward Hacking in Large Language Models
by: Khalaf, Hadi, et al.
Published: (2025)
by: Khalaf, Hadi, et al.
Published: (2025)
Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision
by: Jayalath, Dulhan, et al.
Published: (2025)
by: Jayalath, Dulhan, et al.
Published: (2025)
Self-Supervised Moving Object Segmentation of Sparse and Noisy Radar Point Clouds
by: Schwarzer, Leon, et al.
Published: (2025)
by: Schwarzer, Leon, et al.
Published: (2025)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
Reward-SQL: Boosting Text-to-SQL via Stepwise Reasoning and Process-Supervised Rewards
by: Zhang, Yuxin, et al.
Published: (2025)
by: Zhang, Yuxin, et al.
Published: (2025)
Label-free Monitoring of Self-Supervised Learning Progress
by: Xu, Isaac, et al.
Published: (2024)
by: Xu, Isaac, et al.
Published: (2024)
SkiLD: Unsupervised Skill Discovery Guided by Factor Interactions
by: Wang, Zizhao, et al.
Published: (2024)
by: Wang, Zizhao, et al.
Published: (2024)
Optimal Design for Reward Modeling in RLHF
by: Scheid, Antoine, et al.
Published: (2024)
by: Scheid, Antoine, et al.
Published: (2024)
Position: Benchmarking is Limited in Reinforcement Learning Research
by: Jordan, Scott M., et al.
Published: (2024)
by: Jordan, Scott M., et al.
Published: (2024)
On the Limits of Test-Time Compute: Sequential Reward Filtering for Better Inference
by: Yu, Yue, et al.
Published: (2025)
by: Yu, Yue, et al.
Published: (2025)
Similar Items
-
Training ML Models with Predictable Failures
by: Schwarzer, Will, et al.
Published: (2026) -
Evaluation-Aware Reinforcement Learning
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2025) -
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
by: Chittepu, Yaswanth, et al.
Published: (2025) -
Bayesian Robust Optimization for Imitation Learning
by: Brown, Daniel S., et al.
Published: (2020) -
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
by: Sikchi, Harshit, et al.
Published: (2024)