Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yingru, Liu, Liangqi, Pu, Wenqiang, Liang, Hao, Luo, Zhi-Quan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
von: Park, Chanwoo, et al.
Veröffentlicht: (2024)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
von: Erez, Liad, et al.
Veröffentlicht: (2022)
von: Erez, Liad, et al.
Veröffentlicht: (2022)
Tight Regret Upper and Lower Bounds for Optimistic Hedge in Two-Player Zero-Sum Games
von: Tsuchiya, Taira
Veröffentlicht: (2025)
von: Tsuchiya, Taira
Veröffentlicht: (2025)
Improving Sample Efficiency of Model-Free Algorithms for Zero-Sum Markov Games
von: Feng, Songtao, et al.
Veröffentlicht: (2023)
von: Feng, Songtao, et al.
Veröffentlicht: (2023)
Real-Time Parallel Counterfactual Regret Minimization
von: Li, Boning, et al.
Veröffentlicht: (2026)
von: Li, Boning, et al.
Veröffentlicht: (2026)
Deep (Predictive) Discounted Counterfactual Regret Minimization
von: Xu, Hang, et al.
Veröffentlicht: (2025)
von: Xu, Hang, et al.
Veröffentlicht: (2025)
Next-Token Prediction and Regret Minimization
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
Efficiently Training Neural Networks for Imperfect Information Games by Sampling Information Sets
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
Solving Pasur Using GPU-Accelerated Counterfactual Regret Minimization
von: Baghal, Sina
Veröffentlicht: (2025)
von: Baghal, Sina
Veröffentlicht: (2025)
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
von: Chen, Yang, et al.
Veröffentlicht: (2025)
von: Chen, Yang, et al.
Veröffentlicht: (2025)
From External to Swap Regret 2.0: An Efficient Reduction and Oblivious Adversary for Large Action Spaces
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
von: Dagan, Yuval, et al.
Veröffentlicht: (2023)
Gradient-based Learning in State-based Potential Games for Self-Learning Production Systems
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Online Learning for Uninformed Markov Games: Empirical Nash-Value Regret and Non-Stationarity Adaptation
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
von: Liu, Junyan, et al.
Veröffentlicht: (2026)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
Optimistic Online Learning in Symmetric Cone Games
von: Barakat, Anas, et al.
Veröffentlicht: (2025)
von: Barakat, Anas, et al.
Veröffentlicht: (2025)
Online Test Synthesis From Requirements: Enhancing Reinforcement Learning with Game Theory
von: Sankur, Ocan, et al.
Veröffentlicht: (2024)
von: Sankur, Ocan, et al.
Veröffentlicht: (2024)
Incentivizing Truthful Language Models via Peer Elicitation Games
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
Paths to Equilibrium in Games
von: Yongacoglu, Bora, et al.
Veröffentlicht: (2024)
von: Yongacoglu, Bora, et al.
Veröffentlicht: (2024)
The Hidden Game Problem
von: Buzaglo, Gon, et al.
Veröffentlicht: (2025)
von: Buzaglo, Gon, et al.
Veröffentlicht: (2025)
LiteEFG: An Efficient Python Library for Solving Extensive-form Games
von: Liu, Mingyang, et al.
Veröffentlicht: (2024)
von: Liu, Mingyang, et al.
Veröffentlicht: (2024)
The Intelligent Disobedience Game: Formulating Disobedience in Stackelberg Games and Markov Decision Processes
von: Hornig, Benedikt, et al.
Veröffentlicht: (2026)
von: Hornig, Benedikt, et al.
Veröffentlicht: (2026)
Meta-Computing Enhanced Federated Learning in IIoT: Satisfaction-Aware Incentive Scheme via DRL-Based Stackelberg Game
von: Li, Xiaohuan, et al.
Veröffentlicht: (2025)
von: Li, Xiaohuan, et al.
Veröffentlicht: (2025)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
von: Smid, Yanna Elizabeth, et al.
Veröffentlicht: (2025)
von: Smid, Yanna Elizabeth, et al.
Veröffentlicht: (2025)
Efficient Last-iterate Convergence Algorithms in Solving Games
von: Meng, Linjian, et al.
Veröffentlicht: (2023)
von: Meng, Linjian, et al.
Veröffentlicht: (2023)
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
von: Liu, Mingyang, et al.
Veröffentlicht: (2024)
von: Liu, Mingyang, et al.
Veröffentlicht: (2024)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
Last-Iterate Convergence of No-Regret Learning for Equilibria in Bargaining Games
von: Kamp, Serafina, et al.
Veröffentlicht: (2025)
von: Kamp, Serafina, et al.
Veröffentlicht: (2025)
Last-Iterate Convergence Properties of Regret-Matching Algorithms in Games
von: Cai, Yang, et al.
Veröffentlicht: (2023)
von: Cai, Yang, et al.
Veröffentlicht: (2023)
Proximal Regret and Proximal Correlated Equilibria: A New Tractable Solution Concept for Online Learning and Games
von: Cai, Yang, et al.
Veröffentlicht: (2025)
von: Cai, Yang, et al.
Veröffentlicht: (2025)
RuleSmith: Multi-Agent LLMs for Automated Game Balancing
von: Zeng, Ziyao, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyao, et al.
Veröffentlicht: (2026)
Learning in Bayesian Stackelberg Games With Unknown Follower's Types
von: Bollini, Matteo, et al.
Veröffentlicht: (2026)
von: Bollini, Matteo, et al.
Veröffentlicht: (2026)
Is Thompson Sampling Susceptible to Algorithmic Collusion?
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
The Optimal Sample Complexity of Linear Contracts
von: Høgsgaard, Mikael Møller
Veröffentlicht: (2026)
von: Høgsgaard, Mikael Møller
Veröffentlicht: (2026)
Monopoly Deal: A Benchmark Environment for Bounded One-Sided Response Games
von: Wolf, Will
Veröffentlicht: (2025)
von: Wolf, Will
Veröffentlicht: (2025)
ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
von: Lei, Shiqi, et al.
Veröffentlicht: (2024)
von: Lei, Shiqi, et al.
Veröffentlicht: (2024)
Antithetic Sampling for Top-k Shapley Identification
von: Kolpaczki, Patrick, et al.
Veröffentlicht: (2025)
von: Kolpaczki, Patrick, et al.
Veröffentlicht: (2025)
Self-optimization in distributed manufacturing systems using Modular State-based Stackelberg Games
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024) -
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
von: Park, Chanwoo, et al.
Veröffentlicht: (2024) -
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024) -
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
von: Erez, Liad, et al.
Veröffentlicht: (2022) -
Tight Regret Upper and Lower Bounds for Optimistic Hedge in Two-Player Zero-Sum Games
von: Tsuchiya, Taira
Veröffentlicht: (2025)