Saved in:
| Main Authors: | Rank, Ben, Triantafyllou, Stelios, Mandal, Debmalya, Radanovic, Goran |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.09838 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performative Reinforcement Learning with Linear Markov Decision Process
by: Mandal, Debmalya, et al.
Published: (2024)
by: Mandal, Debmalya, et al.
Published: (2024)
On Corruption-Robustness in Performative Reinforcement Learning
by: Pollatos, Vasilis, et al.
Published: (2025)
by: Pollatos, Vasilis, et al.
Published: (2025)
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025)
by: Mandal, Debmalya, et al.
Published: (2025)
Agent-Specific Effects: A Causal Effect Propagation Analysis in Multi-Agent MDPs
by: Triantafyllou, Stelios, et al.
Published: (2023)
by: Triantafyllou, Stelios, et al.
Published: (2023)
Independent Learning in Performative Markov Potential Games
by: Sahitaj, Rilind, et al.
Published: (2025)
by: Sahitaj, Rilind, et al.
Published: (2025)
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback
by: Nika, Andi, et al.
Published: (2026)
by: Nika, Andi, et al.
Published: (2026)
Corruption Robust Offline Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2024)
by: Mandal, Debmalya, et al.
Published: (2024)
Sparse Offline Reinforcement Learning with Corruption Robustness
by: Tran, Nam Phuong, et al.
Published: (2025)
by: Tran, Nam Phuong, et al.
Published: (2025)
Stochastic Principal-Agent Problems: Efficient Computation and Learning
by: Gan, Jiarui, et al.
Published: (2023)
by: Gan, Jiarui, et al.
Published: (2023)
Corruption-Robust Offline Two-Player Zero-Sum Markov Games
by: Nika, Andi, et al.
Published: (2024)
by: Nika, Andi, et al.
Published: (2024)
Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences
by: Nika, Andi, et al.
Published: (2024)
by: Nika, Andi, et al.
Published: (2024)
Policy Teaching via Data Poisoning in Learning from Human Preferences
by: Nika, Andi, et al.
Published: (2025)
by: Nika, Andi, et al.
Published: (2025)
Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making
by: Triantafyllou, Stelios, et al.
Published: (2024)
by: Triantafyllou, Stelios, et al.
Published: (2024)
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks?
by: Sasnauskas, Paulius, et al.
Published: (2025)
by: Sasnauskas, Paulius, et al.
Published: (2025)
Reward Design for Justifiable Sequential Decision-Making
by: Sukovic, Aleksa, et al.
Published: (2024)
by: Sukovic, Aleksa, et al.
Published: (2024)
Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms
by: Nöther, Jonathan, et al.
Published: (2025)
by: Nöther, Jonathan, et al.
Published: (2025)
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints
by: Nöther, Jonathan, et al.
Published: (2025)
by: Nöther, Jonathan, et al.
Published: (2025)
Reinforcement Learning for Durable Algorithmic Recourse
by: Ceccon, Marina, et al.
Published: (2025)
by: Ceccon, Marina, et al.
Published: (2025)
Learning Embeddings for Sequential Tasks Using Population of Agents
by: Mahajan, Mridul, et al.
Published: (2023)
by: Mahajan, Mridul, et al.
Published: (2023)
Strategyproof Reinforcement Learning from Human Feedback
by: Buening, Thomas Kleine, et al.
Published: (2025)
by: Buening, Thomas Kleine, et al.
Published: (2025)
MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems
by: Nöther, Jonathan, et al.
Published: (2026)
by: Nöther, Jonathan, et al.
Published: (2026)
GRAPHLCP: Structure-Aware Localized Conformal Prediction on Graphs
by: Baghershahi, Peyman, et al.
Published: (2026)
by: Baghershahi, Peyman, et al.
Published: (2026)
Surprisingly Popular Voting for Concentric Rank-Order Models
by: Hosseini, Hadi, et al.
Published: (2024)
by: Hosseini, Hadi, et al.
Published: (2024)
Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
by: Kim, Hyunwoo, et al.
Published: (2025)
by: Kim, Hyunwoo, et al.
Published: (2025)
Symmetric Linear Bandits with Hidden Symmetry
by: Tran, Nam Phuong, et al.
Published: (2024)
by: Tran, Nam Phuong, et al.
Published: (2024)
Gradual Domain Adaptation for Graph Learning
by: Lei, Pui Ieng, et al.
Published: (2025)
by: Lei, Pui Ieng, et al.
Published: (2025)
Deep learning enhanced mixed integer optimization: Learning to reduce model dimensionality
by: Triantafyllou, Niki, et al.
Published: (2024)
by: Triantafyllou, Niki, et al.
Published: (2024)
Assessing the Impact of Distribution Shift on Reinforcement Learning Performance
by: Fujimoto, Ted, et al.
Published: (2024)
by: Fujimoto, Ted, et al.
Published: (2024)
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2025)
by: Omura, Motoki, et al.
Published: (2025)
Gradual Optimization Learning for Conformational Energy Minimization
by: Tsypin, Artem, et al.
Published: (2023)
by: Tsypin, Artem, et al.
Published: (2023)
Learning the Expected Core of Strictly Convex Stochastic Cooperative Games
by: Tran, Nam Phuong, et al.
Published: (2024)
by: Tran, Nam Phuong, et al.
Published: (2024)
On Double Descent in Reinforcement Learning with LSTD and Random Features
by: Brellmann, David, et al.
Published: (2023)
by: Brellmann, David, et al.
Published: (2023)
Collaborative Distributed Machine Learning
by: Jin, David, et al.
Published: (2023)
by: Jin, David, et al.
Published: (2023)
Gradual Domain Adaptation: Theory and Algorithms
by: He, Yifei, et al.
Published: (2023)
by: He, Yifei, et al.
Published: (2023)
Distance Matters For Improving Performance Estimation Under Covariate Shift
by: Roschewitz, Mélanie, et al.
Published: (2023)
by: Roschewitz, Mélanie, et al.
Published: (2023)
Estimating the Joint Probability of Scenario Parameters with Gaussian Mixture Copula Models
by: Reichenbächer, Christian, et al.
Published: (2025)
by: Reichenbächer, Christian, et al.
Published: (2025)
Intrusion Detection System Using Supervised Machine Learning Models on the KDD Cup 1999 Dataset
by: Hadjioannou, Christos Stelios
Published: (2025)
by: Hadjioannou, Christos Stelios
Published: (2025)
Gradual Domain Adaptation via Normalizing Flows
by: Sagawa, Shogo, et al.
Published: (2022)
by: Sagawa, Shogo, et al.
Published: (2022)
Gradual Fine-Tuning for Flow Matching Models
by: Thorkelsdottir, Gudrun, et al.
Published: (2026)
by: Thorkelsdottir, Gudrun, et al.
Published: (2026)
FedShift: Robust Federated Learning Aggregation Scheme in Resource Constrained Environment via Weight Shifting
by: Seo, Jungwon, et al.
Published: (2024)
by: Seo, Jungwon, et al.
Published: (2024)
Similar Items
-
Performative Reinforcement Learning with Linear Markov Decision Process
by: Mandal, Debmalya, et al.
Published: (2024) -
On Corruption-Robustness in Performative Reinforcement Learning
by: Pollatos, Vasilis, et al.
Published: (2025) -
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025) -
Agent-Specific Effects: A Causal Effect Propagation Analysis in Multi-Agent MDPs
by: Triantafyllou, Stelios, et al.
Published: (2023) -
Independent Learning in Performative Markov Potential Games
by: Sahitaj, Rilind, et al.
Published: (2025)