The Representation-Rationalizability Tradeoff in Reward Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Jing, Yu, Yaoliang, Pourpart, Pascal |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convergence to Nash Equilibrium and No-regret Guarantee in (Markov) Potential Games
by: Dong, Jing, et al.
Published: (2024)
by: Dong, Jing, et al.
Published: (2024)
Uncoupled and Convergent Learning in Monotone Games under Bandit Feedback
by: Dong, Jing, et al.
Published: (2024)
by: Dong, Jing, et al.
Published: (2024)
Last-iterate Convergence in Regularized Graphon Mean Field Game
by: Dong, Jing, et al.
Published: (2024)
by: Dong, Jing, et al.
Published: (2024)
Prudent Rationalizability and the Best Rationalization Principle
by: De Vito, Nicodemo
Published: (2025)
by: De Vito, Nicodemo
Published: (2025)
Decongestion by Representation: Learning to Improve Economic Welfare in Marketplaces
by: Nahum, Omer, et al.
Published: (2023)
by: Nahum, Omer, et al.
Published: (2023)
Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond
by: Bacchiocchi, Francesco, et al.
Published: (2025)
by: Bacchiocchi, Francesco, et al.
Published: (2025)
Rationalizable Screening and Disclosure under Unawareness
by: Francetich, Alejandro, et al.
Published: (2025)
by: Francetich, Alejandro, et al.
Published: (2025)
Provable Policy Gradient Methods for Average-Reward Markov Potential Games
by: Cheng, Min, et al.
Published: (2024)
by: Cheng, Min, et al.
Published: (2024)
Model Selection for Average Reward RL with Application to Utility Maximization in Repeated Games
by: Masoumian, Alireza, et al.
Published: (2024)
by: Masoumian, Alireza, et al.
Published: (2024)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
by: Ren, Hang, et al.
Published: (2025)
by: Ren, Hang, et al.
Published: (2025)
On the Decomposition of Differential Game
by: Zhou, Nanxiang, et al.
Published: (2024)
by: Zhou, Nanxiang, et al.
Published: (2024)
Blind Inverse Game Theory: Jointly Decoding Rewards and Rationality in Entropy-Regularized Competitive Games
by: Virk, Hamza, et al.
Published: (2025)
by: Virk, Hamza, et al.
Published: (2025)
Rationalizability, Iterated Dominance, and the Theorems of Radon and Carathéodory
by: Long, Roy
Published: (2024)
by: Long, Roy
Published: (2024)
Near-Optimal Regret-Queue Length Tradeoff in Online Learning for Two-Sided Markets
by: Yang, Zixian, et al.
Published: (2025)
by: Yang, Zixian, et al.
Published: (2025)
Intelligent Agents for Auction-based Federated Learning: A Survey
by: Tang, Xiaoli, et al.
Published: (2024)
by: Tang, Xiaoli, et al.
Published: (2024)
Proximal Regret and Proximal Correlated Equilibria: A New Tractable Solution Concept for Online Learning and Games
by: Cai, Yang, et al.
Published: (2025)
by: Cai, Yang, et al.
Published: (2025)
HiBid: A Cross-Channel Constrained Bidding System with Budget Allocation by Hierarchical Offline Deep Reinforcement Learning
by: Wang, Hao, et al.
Published: (2023)
by: Wang, Hao, et al.
Published: (2023)
Free-Rider and Conflict Aware Collaboration Formation for Cross-Silo Federated Learning
by: Chen, Mengmeng, et al.
Published: (2024)
by: Chen, Mengmeng, et al.
Published: (2024)
Learning not to Regret
by: Sychrovský, David, et al.
Published: (2023)
by: Sychrovský, David, et al.
Published: (2023)
PAC Learning with Improvements
by: Attias, Idan, et al.
Published: (2025)
by: Attias, Idan, et al.
Published: (2025)
Learning Local Stackelberg Equilibria from Repeated Interactions with a Learning Agent
by: Ananthakrishnan, Nivasini, et al.
Published: (2025)
by: Ananthakrishnan, Nivasini, et al.
Published: (2025)
Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models
by: Fang, Zeyu, et al.
Published: (2024)
by: Fang, Zeyu, et al.
Published: (2024)
Online Learning with Bounded Recall
by: Schneider, Jon, et al.
Published: (2022)
by: Schneider, Jon, et al.
Published: (2022)
Learning in Structured Stackelberg Games
by: Balcan, Maria-Florina, et al.
Published: (2025)
by: Balcan, Maria-Florina, et al.
Published: (2025)
Incentivized Collaboration in Active Learning
by: Cohen, Lee, et al.
Published: (2023)
by: Cohen, Lee, et al.
Published: (2023)
Learning to Steer Learners in Games
by: Zhang, Yizhou, et al.
Published: (2025)
by: Zhang, Yizhou, et al.
Published: (2025)
Learning Social Welfare Functions
by: Pardeshi, Kanad Shrikar, et al.
Published: (2024)
by: Pardeshi, Kanad Shrikar, et al.
Published: (2024)
Learning to Price Homogeneous Data
by: Chen, Keran, et al.
Published: (2024)
by: Chen, Keran, et al.
Published: (2024)
Corrupted Learning Dynamics in Games
by: Tsuchiya, Taira, et al.
Published: (2024)
by: Tsuchiya, Taira, et al.
Published: (2024)
Learning to Allocate Resources with Censored Feedback
by: Montanari, Giovanni, et al.
Published: (2026)
by: Montanari, Giovanni, et al.
Published: (2026)
Look-Ahead Reasoning on Learning Platforms
by: Zhu, Haiqing, et al.
Published: (2025)
by: Zhu, Haiqing, et al.
Published: (2025)
Learning in Budgeted Auctions with Spacing Objectives
by: Fikioris, Giannis, et al.
Published: (2024)
by: Fikioris, Giannis, et al.
Published: (2024)
Dynamic Pricing and Advertising with Demand Learning
by: Agrawal, Shipra, et al.
Published: (2023)
by: Agrawal, Shipra, et al.
Published: (2023)
Learning to Play Against Unknown Opponents
by: Arunachaleswaran, Eshwar Ram, et al.
Published: (2024)
by: Arunachaleswaran, Eshwar Ram, et al.
Published: (2024)
Algorithmic Collective Action in Machine Learning
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
Regulation Games for Trustworthy Machine Learning
by: Yaghini, Mohammad, et al.
Published: (2024)
by: Yaghini, Mohammad, et al.
Published: (2024)
Barriers to Welfare Maximization with No-Regret Learning
by: Anagnostides, Ioannis, et al.
Published: (2024)
by: Anagnostides, Ioannis, et al.
Published: (2024)
Bayesian Learning in Episodic Zero-Sum Games
by: Yueh, Chang-Wei, et al.
Published: (2026)
by: Yueh, Chang-Wei, et al.
Published: (2026)
Learning Payment-Free Resource Allocation Mechanisms
by: Zeng, Sihan, et al.
Published: (2023)
by: Zeng, Sihan, et al.
Published: (2023)
Learning and Computation of $Φ$-Equilibria at the Frontier of Tractability
by: Zhang, Brian Hu, et al.
Published: (2025)
by: Zhang, Brian Hu, et al.
Published: (2025)
Similar Items
-
Convergence to Nash Equilibrium and No-regret Guarantee in (Markov) Potential Games
by: Dong, Jing, et al.
Published: (2024) -
Uncoupled and Convergent Learning in Monotone Games under Bandit Feedback
by: Dong, Jing, et al.
Published: (2024) -
Last-iterate Convergence in Regularized Graphon Mean Field Game
by: Dong, Jing, et al.
Published: (2024) -
Prudent Rationalizability and the Best Rationalization Principle
by: De Vito, Nicodemo
Published: (2025) -
Decongestion by Representation: Learning to Improve Economic Welfare in Marketplaces
by: Nahum, Omer, et al.
Published: (2023)