Safe Exploitative Play with Untrusted Type Beliefs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Tongxin, Handina, Tinashe, Ren, Shaolei, Wierman, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Understanding Model Selection For Learning In Strategic Environments
von: Handina, Tinashe, et al.
Veröffentlicht: (2024)
von: Handina, Tinashe, et al.
Veröffentlicht: (2024)
Online Budgeted Matching with General Bids
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
Last-Iterate Convergence of Payoff-Based Independent Learning in Zero-Sum Stochastic Games
von: Chen, Zaiwei, et al.
Veröffentlicht: (2024)
von: Chen, Zaiwei, et al.
Veröffentlicht: (2024)
Breaking the Curse of Multiagency in Robust Multi-Agent Reinforcement Learning
von: Shi, Laixi, et al.
Veröffentlicht: (2024)
von: Shi, Laixi, et al.
Veröffentlicht: (2024)
Competitive Algorithms for Multi-Agent Ski-Rental Problems
von: Wang, Xuchuang, et al.
Veröffentlicht: (2025)
von: Wang, Xuchuang, et al.
Veröffentlicht: (2025)
Learning to Play Against Unknown Opponents
von: Arunachaleswaran, Eshwar Ram, et al.
Veröffentlicht: (2024)
von: Arunachaleswaran, Eshwar Ram, et al.
Veröffentlicht: (2024)
Playing Markov Games Without Observing Payoffs
von: Ablin, Daniel, et al.
Veröffentlicht: (2025)
von: Ablin, Daniel, et al.
Veröffentlicht: (2025)
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
Learning Safely Without Knowing the World:COMPASS-Hedge
von: Hu, Ting, et al.
Veröffentlicht: (2026)
von: Hu, Ting, et al.
Veröffentlicht: (2026)
Fast and Furious Symmetric Learning in Zero-Sum Games: Gradient Descent as Fictitious Play
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems
von: Nöther, Jonathan, et al.
Veröffentlicht: (2026)
von: Nöther, Jonathan, et al.
Veröffentlicht: (2026)
Pure Exploration via Frank-Wolfe Self-Play
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
Learning in Bayesian Stackelberg Games With Unknown Follower's Types
von: Bollini, Matteo, et al.
Veröffentlicht: (2026)
von: Bollini, Matteo, et al.
Veröffentlicht: (2026)
Conjectural Online Learning with First-order Beliefs in Asymmetric Information Stochastic Games
von: Li, Tao, et al.
Veröffentlicht: (2024)
von: Li, Tao, et al.
Veröffentlicht: (2024)
Aumann-SHAP: The Geometry of Counterfactual Interaction Explanations in Machine Learning
von: Belahcen, Adam, et al.
Veröffentlicht: (2026)
von: Belahcen, Adam, et al.
Veröffentlicht: (2026)
Incentivized Truthful Communication for Federated Bandits
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
On Randomized Algorithms in Online Strategic Classification
von: Hutton, Chase, et al.
Veröffentlicht: (2026)
von: Hutton, Chase, et al.
Veröffentlicht: (2026)
Learning to Play Multi-Follower Bayesian Stackelberg Games
von: Personnat, Gerson, et al.
Veröffentlicht: (2025)
von: Personnat, Gerson, et al.
Veröffentlicht: (2025)
State-Constrained Zero-Sum Differential Games with One-Sided Information
von: Ghimire, Mukesh, et al.
Veröffentlicht: (2024)
von: Ghimire, Mukesh, et al.
Veröffentlicht: (2024)
SpinGPT: A Large-Language-Model Approach to Playing Poker Correctly
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
von: Ren, Hang, et al.
Veröffentlicht: (2025)
von: Ren, Hang, et al.
Veröffentlicht: (2025)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
Beyond Nash Equilibrium: Achieving Bayesian Perfect Equilibrium with Belief Update Fictitious Play
von: Ju, Qi, et al.
Veröffentlicht: (2024)
von: Ju, Qi, et al.
Veröffentlicht: (2024)
Accelerating Nash Equilibrium Convergence in Monte Carlo Settings Through Counterfactual Value Based Fictitious Play
von: Qi, Ju, et al.
Veröffentlicht: (2023)
von: Qi, Ju, et al.
Veröffentlicht: (2023)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
Dynamic Reserve Price Design with Distributed Solving Algorithm
von: Li, Mang
Veröffentlicht: (2022)
von: Li, Mang
Veröffentlicht: (2022)
Data-Driven Revenue Management for Air Cargo
von: Eren, Ezgi, et al.
Veröffentlicht: (2024)
von: Eren, Ezgi, et al.
Veröffentlicht: (2024)
Improved Bandits in Many-to-one Matching Markets with Incentive Compatibility
von: Kong, Fang, et al.
Veröffentlicht: (2024)
von: Kong, Fang, et al.
Veröffentlicht: (2024)
Online Allocation with Replenishable Budgets: Worst Case and Beyond
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
ReLExS: Reinforcement Learning Explanations for Stackelberg No-Regret Learners
von: Huang, Xiangge, et al.
Veröffentlicht: (2024)
von: Huang, Xiangge, et al.
Veröffentlicht: (2024)
RL-CFR: Improving Action Abstraction for Imperfect Information Extensive-Form Games with Reinforcement Learning
von: Li, Boning, et al.
Veröffentlicht: (2024)
von: Li, Boning, et al.
Veröffentlicht: (2024)
Heterogeneous Data Game: Characterizing the Model Competition Across Multiple Data Sources
von: Xu, Renzhe, et al.
Veröffentlicht: (2025)
von: Xu, Renzhe, et al.
Veröffentlicht: (2025)
On the Decomposition of Differential Game
von: Zhou, Nanxiang, et al.
Veröffentlicht: (2024)
von: Zhou, Nanxiang, et al.
Veröffentlicht: (2024)
Last-iterate Convergence Separation between Extra-gradient and Optimism in Constrained Periodic Games
von: Feng, Yi, et al.
Veröffentlicht: (2024)
von: Feng, Yi, et al.
Veröffentlicht: (2024)
Intelligent Agents for Auction-based Federated Learning: A Survey
von: Tang, Xiaoli, et al.
Veröffentlicht: (2024)
von: Tang, Xiaoli, et al.
Veröffentlicht: (2024)
Attacking and Securing Community Detection: A Game-Theoretic Framework
von: Niu, Yifan, et al.
Veröffentlicht: (2025)
von: Niu, Yifan, et al.
Veröffentlicht: (2025)
Pontryagin Neural Operator for Solving Parametric General-Sum Differential Games
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
Value Approximation for Two-Player General-Sum Differential Games with State Constraints
von: Zhang, Lei, et al.
Veröffentlicht: (2023)
von: Zhang, Lei, et al.
Veröffentlicht: (2023)
Learning Nash Equilibrial Hamiltonian for Two-Player Collision-Avoiding Interactions
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Understanding Model Selection For Learning In Strategic Environments
von: Handina, Tinashe, et al.
Veröffentlicht: (2024) -
Online Budgeted Matching with General Bids
von: Yang, Jianyi, et al.
Veröffentlicht: (2024) -
Last-Iterate Convergence of Payoff-Based Independent Learning in Zero-Sum Stochastic Games
von: Chen, Zaiwei, et al.
Veröffentlicht: (2024) -
Breaking the Curse of Multiagency in Robust Multi-Agent Reinforcement Learning
von: Shi, Laixi, et al.
Veröffentlicht: (2024) -
Competitive Algorithms for Multi-Agent Ski-Rental Problems
von: Wang, Xuchuang, et al.
Veröffentlicht: (2025)