Aligning LLMs with Domain Invariant Reward Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, David, Choudhury, Sanjiban |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Process Reward Models for LLM Agents: Practical Framework and Directions
by: Choudhury, Sanjiban
Published: (2025)
by: Choudhury, Sanjiban
Published: (2025)
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024)
by: Choudhury, Sanjiban, et al.
Published: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
Imitation Learning from a Single Temporally Misaligned Video
by: Huey, William, et al.
Published: (2025)
by: Huey, William, et al.
Published: (2025)
Inverse Reinforcement Learning without Reinforcement Learning
by: Swamy, Gokul, et al.
Published: (2023)
by: Swamy, Gokul, et al.
Published: (2023)
Distilling Realizable Students from Unrealizable Teachers
by: Kim, Yujin, et al.
Published: (2025)
by: Kim, Yujin, et al.
Published: (2025)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
by: Kedia, Kushal, et al.
Published: (2023)
by: Kedia, Kushal, et al.
Published: (2023)
Multi-Turn Code Generation Through Single-Step Rewards
by: Jain, Arnav Kumar, et al.
Published: (2025)
by: Jain, Arnav Kumar, et al.
Published: (2025)
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
by: Swamy, Gokul, et al.
Published: (2025)
by: Swamy, Gokul, et al.
Published: (2025)
Efficient Imitation under Misspecification
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
Hybrid Inverse Reinforcement Learning
by: Ren, Juntao, et al.
Published: (2024)
by: Ren, Juntao, et al.
Published: (2024)
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
by: Ren, Juntao, et al.
Published: (2025)
by: Ren, Juntao, et al.
Published: (2025)
One-Shot Imitation under Mismatched Execution
by: Kedia, Kushal, et al.
Published: (2024)
by: Kedia, Kushal, et al.
Published: (2024)
Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization
by: Ma, Qiyao, et al.
Published: (2026)
by: Ma, Qiyao, et al.
Published: (2026)
Imitation Learning via Focused Satisficing
by: Shah, Rushit N., et al.
Published: (2025)
by: Shah, Rushit N., et al.
Published: (2025)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
by: Jain, Arnav Kumar, et al.
Published: (2024)
by: Jain, Arnav Kumar, et al.
Published: (2024)
A Smooth Sea Never Made a Skilled SAILOR: Robust Imitation via Learning to Search
by: Jain, Arnav Kumar, et al.
Published: (2025)
by: Jain, Arnav Kumar, et al.
Published: (2025)
Domain Agnostic Conditional Invariant Predictions for Domain Generalization
by: Wang, Zongbin, et al.
Published: (2024)
by: Wang, Zongbin, et al.
Published: (2024)
Self-Aligned Reward: Towards Effective and Efficient Reasoners
by: Han, Peixuan, et al.
Published: (2025)
by: Han, Peixuan, et al.
Published: (2025)
Continual Learning of Domain-Invariant Representations
by: Janetzky, Pascal, et al.
Published: (2026)
by: Janetzky, Pascal, et al.
Published: (2026)
Prominent Roles of Conditionally Invariant Components in Domain Adaptation: Theory and Algorithms
by: Wu, Keru, et al.
Published: (2023)
by: Wu, Keru, et al.
Published: (2023)
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
by: Lee, Hyunseok, et al.
Published: (2024)
by: Lee, Hyunseok, et al.
Published: (2024)
Reward Collapse in Aligning Large Language Models
by: Song, Ziang, et al.
Published: (2023)
by: Song, Ziang, et al.
Published: (2023)
ALaRM: Align Language Models via Hierarchical Rewards Modeling
by: Lai, Yuhang, et al.
Published: (2024)
by: Lai, Yuhang, et al.
Published: (2024)
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data
by: Zeng, Thomas, et al.
Published: (2025)
by: Zeng, Thomas, et al.
Published: (2025)
Tree Reward-Aligned Search for TReASURe in Masked Diffusion Language Models
by: Yu, Zichao, et al.
Published: (2025)
by: Yu, Zichao, et al.
Published: (2025)
Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
by: Wu, Shirley, et al.
Published: (2025)
by: Wu, Shirley, et al.
Published: (2025)
Learning Causally Invariant Reward Functions from Diverse Demonstrations
by: Ovinnikov, Ivan, et al.
Published: (2024)
by: Ovinnikov, Ivan, et al.
Published: (2024)
Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
by: Nakamura, Mason, et al.
Published: (2025)
by: Nakamura, Mason, et al.
Published: (2025)
Class-Invariant Test-Time Augmentation for Domain Generalization
by: Lin, Zhicheng, et al.
Published: (2025)
by: Lin, Zhicheng, et al.
Published: (2025)
Domain Generalization In Robust Invariant Representation
by: Gupta, Gauri, et al.
Published: (2023)
by: Gupta, Gauri, et al.
Published: (2023)
X-Sim: Cross-Embodiment Learning via Real-to-Sim-to-Real
by: Dan, Prithwish, et al.
Published: (2025)
by: Dan, Prithwish, et al.
Published: (2025)
Self-Rewarding PPO: Aligning Large Language Models with Demonstrations Only
by: Zhang, Qingru, et al.
Published: (2025)
by: Zhang, Qingru, et al.
Published: (2025)
Aligning Few-Step Diffusion Models with Dense Reward Difference Learning
by: Zhang, Ziyi, et al.
Published: (2024)
by: Zhang, Ziyi, et al.
Published: (2024)
Aligning Molecular Graph Explanations with Chemical Identity via InChIfied Invariants
by: Guidotti, Emanuele, et al.
Published: (2026)
by: Guidotti, Emanuele, et al.
Published: (2026)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
by: Prabhudesai, Mihir, et al.
Published: (2023)
by: Prabhudesai, Mihir, et al.
Published: (2023)
Diversity as a Reward: Fine-Tuning LLMs on a Mixture of Domain-Undetermined Data
by: Ling, Zhenqing, et al.
Published: (2025)
by: Ling, Zhenqing, et al.
Published: (2025)
GRIFDIR: Graph Resolution-Invariant FEM Diffusion Models in Function Spaces over Irregular Domains
by: Rowbottom, James, et al.
Published: (2026)
by: Rowbottom, James, et al.
Published: (2026)
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference
by: Qiu, Wenjie, et al.
Published: (2025)
by: Qiu, Wenjie, et al.
Published: (2025)
Similar Items
-
Process Reward Models for LLM Agents: Practical Framework and Directions
by: Choudhury, Sanjiban
Published: (2025) -
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
by: Wu, David, et al.
Published: (2024) -
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024) -
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024) -
Imitation Learning from a Single Temporally Misaligned Video
by: Huey, William, et al.
Published: (2025)