Robust Reward Modeling via Causal Rubrics
Fuente:
arXiv
Saved in:
| Main Authors: | Srivastava, Pragya, Singh, Harman, Madhavan, Rahul, Patil, Gandharv, Addepalli, Sravanti, Suggala, Arun, Aravamudhan, Rengarajan, Sharma, Soumya, Laha, Anirban, Raghuveer, Aravindan, Shanmugam, Karthikeyan, Precup, Doina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time-Reversal Provides Unsupervised Feedback to LLMs
by: Varun, Yerram, et al.
Published: (2024)
by: Varun, Yerram, et al.
Published: (2024)
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts?
by: Addepalli, Sravanti, et al.
Published: (2024)
by: Addepalli, Sravanti, et al.
Published: (2024)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
by: Patil, Gandharv, et al.
Published: (2022)
by: Patil, Gandharv, et al.
Published: (2022)
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
by: Vijayan, Sushant, et al.
Published: (2025)
by: Vijayan, Sushant, et al.
Published: (2025)
Fairness under Covariate Shift: Improving Fairness-Accuracy tradeoff with few Unlabeled Test Samples
by: Havaldar, Shreyas, et al.
Published: (2023)
by: Havaldar, Shreyas, et al.
Published: (2023)
Learning from Label Proportions: Bootstrapping Supervised Learners via Belief Propagation
by: Havaldar, Shreyas, et al.
Published: (2023)
by: Havaldar, Shreyas, et al.
Published: (2023)
ProFeAT: Projected Feature Adversarial Training for Self-Supervised Learning of Robust Representations
by: Addepalli, Sravanti, et al.
Published: (2024)
by: Addepalli, Sravanti, et al.
Published: (2024)
Diverse Reactivity of 9‐Fluorene Propargylic Alcohol and 1,3‐Dicarbonyls with Bronsted and Lewis Acid Catalysts: Synthesis of Spiro, Conjugated, and 9‐Ethynyl Fluorene Derivatives
by: Aravamudhan Subhashini, et al.
Published: (2025)
by: Aravamudhan Subhashini, et al.
Published: (2025)
Bayesian Collaborative Bandits with Thompson Sampling for Improved Outreach in Maternal Health Program
by: Dasgupta, Arpan, et al.
Published: (2024)
by: Dasgupta, Arpan, et al.
Published: (2024)
Online Bidding under RoS Constraints without Knowing the Value
by: Vijayan, Sushant, et al.
Published: (2025)
by: Vijayan, Sushant, et al.
Published: (2025)
Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks
by: Patil, Gandharv, et al.
Published: (2026)
by: Patil, Gandharv, et al.
Published: (2026)
Diversity-Enriched Option-Critic
by: Kamat, Anand, et al.
Published: (2020)
by: Kamat, Anand, et al.
Published: (2020)
Functional Acceleration for Policy Mirror Descent
by: Chelu, Veronica, et al.
Published: (2024)
by: Chelu, Veronica, et al.
Published: (2024)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
by: Alver, Safa, et al.
Published: (2022)
by: Alver, Safa, et al.
Published: (2022)
Leveraging Vision-Language Models for Improving Domain Generalization in Image Classification
by: Addepalli, Sravanti, et al.
Published: (2023)
by: Addepalli, Sravanti, et al.
Published: (2023)
Learning to Call: A Field Trial of a Collaborative Bandit Algorithm for Improved Message Delivery in Mobile Maternal Health
by: Dasgupta, Arpan, et al.
Published: (2025)
by: Dasgupta, Arpan, et al.
Published: (2025)
Algorithmic Guarantees for Distilling Supervised and Offline RL Datasets
by: Gupta, Aaryan, et al.
Published: (2025)
by: Gupta, Aaryan, et al.
Published: (2025)
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
by: Singh, Sagalpreet, et al.
Published: (2025)
by: Singh, Sagalpreet, et al.
Published: (2025)
The Swarbhanu Event Horizon: A Comparative Analysis of the Tamil 'Shadow Planet' Axis and the Primordial Black Hole Hypothesis
by: Shanmugam, Karthikeyan
Published: (2026)
by: Shanmugam, Karthikeyan
Published: (2026)
Bandits with Mean Bounds
by: Sharma, Nihal, et al.
Published: (2020)
by: Sharma, Nihal, et al.
Published: (2020)
Causal ATE Mitigates Unintended Bias in Controlled Text Generation
by: Madhavan, Rahul, et al.
Published: (2023)
by: Madhavan, Rahul, et al.
Published: (2023)
A Personalized Exercise Assistant using Reinforcement Learning (PEARL): Results from a four-arm Randomized-controlled Trial
by: Lee, Amy Armento, et al.
Published: (2025)
by: Lee, Amy Armento, et al.
Published: (2025)
Balancing Plasticity and Stability with Fast and Slow Successor Features
by: Chua, Raymond, et al.
Published: (2026)
by: Chua, Raymond, et al.
Published: (2026)
On the Privacy of Selection Mechanisms with Gaussian Noise
by: Lebensold, Jonathan, et al.
Published: (2024)
by: Lebensold, Jonathan, et al.
Published: (2024)
Offline-to-online hyperparameter transfer for stochastic bandits
by: Sharma, Dravyansh, et al.
Published: (2025)
by: Sharma, Dravyansh, et al.
Published: (2025)
Code as Reward: Empowering Reinforcement Learning with VLMs
by: Venuto, David, et al.
Published: (2024)
by: Venuto, David, et al.
Published: (2024)
General Identifiability and Achievability for Causal Representation Learning
by: Varıcı, Burak, et al.
Published: (2023)
by: Varıcı, Burak, et al.
Published: (2023)
FRACTAL: Fine-Grained Scoring from Aggregate Text Labels
by: Makhija, Yukti, et al.
Published: (2024)
by: Makhija, Yukti, et al.
Published: (2024)
LLP-Bench: A Large Scale Tabular Benchmark for Learning from Label Proportions
by: Brahmbhatt, Anand, et al.
Published: (2023)
by: Brahmbhatt, Anand, et al.
Published: (2023)
Aggregating Data for Optimal and Private Learning
by: Agarwal, Sushant, et al.
Published: (2024)
by: Agarwal, Sushant, et al.
Published: (2024)
Linear Causal Representation Learning from Unknown Multi-node Interventions
by: Varıcı, Burak, et al.
Published: (2024)
by: Varıcı, Burak, et al.
Published: (2024)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
by: Carr, Jonathan Colaço, et al.
Published: (2023)
by: Carr, Jonathan Colaço, et al.
Published: (2023)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
by: Jain, Arushi, et al.
Published: (2024)
by: Jain, Arushi, et al.
Published: (2024)
Fluid-Agent Reinforcement Learning
by: Sharma, Shishir, et al.
Published: (2026)
by: Sharma, Shishir, et al.
Published: (2026)
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
by: Alver, Safa, et al.
Published: (2024)
by: Alver, Safa, et al.
Published: (2024)
Bandits with Stochastic Experts: Constant Regret, Empirical Experts and Episodes
by: Sharma, Nihal, et al.
Published: (2021)
by: Sharma, Nihal, et al.
Published: (2021)
CDQuant: Greedy Coordinate Descent for Accurate LLM Quantization
by: Nair, Pranav Ajit, et al.
Published: (2024)
by: Nair, Pranav Ajit, et al.
Published: (2024)
Improving Generalization via Meta-Learning on Hard Samples
by: Jain, Nishant, et al.
Published: (2024)
by: Jain, Nishant, et al.
Published: (2024)
Additive Large Language Models for Semi-Structured Text
by: K, Karthikeyan, et al.
Published: (2025)
by: K, Karthikeyan, et al.
Published: (2025)
Similar Items
-
Time-Reversal Provides Unsupervised Feedback to LLMs
by: Varun, Yerram, et al.
Published: (2024) -
Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts?
by: Addepalli, Sravanti, et al.
Published: (2024) -
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
by: Patil, Gandharv, et al.
Published: (2022) -
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
by: Vijayan, Sushant, et al.
Published: (2025) -
Fairness under Covariate Shift: Improving Fairness-Accuracy tradeoff with few Unlabeled Test Samples
by: Havaldar, Shreyas, et al.
Published: (2023)