Pure Exploration via Frank-Wolfe Self-Play
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Xinyu, Qin, Chao, You, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bayesian Algorithms for Adversarial Online Learning: from Finite to Infinite Action Spaces
von: Terenin, Alexander, et al.
Veröffentlicht: (2025)
von: Terenin, Alexander, et al.
Veröffentlicht: (2025)
Optimal Scoring Rule Design under Partial Knowledge
von: Chen, Yiling, et al.
Veröffentlicht: (2021)
von: Chen, Yiling, et al.
Veröffentlicht: (2021)
Stochastic contextual bandits with graph feedback: from independence number to MAS number
von: Wen, Yuxiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuxiao, et al.
Veröffentlicht: (2024)
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
von: Yan, Yuling, et al.
Veröffentlicht: (2022)
von: Yan, Yuling, et al.
Veröffentlicht: (2022)
GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium
von: Liu, Kaizhao, et al.
Veröffentlicht: (2025)
von: Liu, Kaizhao, et al.
Veröffentlicht: (2025)
Instance-Adaptive Hypothesis Tests with Heterogeneous Agents
von: Shi, Flora C., et al.
Veröffentlicht: (2025)
von: Shi, Flora C., et al.
Veröffentlicht: (2025)
Sharp Results for Hypothesis Testing with Risk-Sensitive Agents
von: Shi, Flora C., et al.
Veröffentlicht: (2024)
von: Shi, Flora C., et al.
Veröffentlicht: (2024)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2026)
Learning to Play Against Unknown Opponents
von: Arunachaleswaran, Eshwar Ram, et al.
Veröffentlicht: (2024)
von: Arunachaleswaran, Eshwar Ram, et al.
Veröffentlicht: (2024)
Playing Markov Games Without Observing Payoffs
von: Ablin, Daniel, et al.
Veröffentlicht: (2025)
von: Ablin, Daniel, et al.
Veröffentlicht: (2025)
Safe Exploitative Play with Untrusted Type Beliefs
von: Li, Tongxin, et al.
Veröffentlicht: (2024)
von: Li, Tongxin, et al.
Veröffentlicht: (2024)
Grouped Satisficing Paths in Pure Strategy Games: a Topological Perspective
von: Fu, Yanqing, et al.
Veröffentlicht: (2025)
von: Fu, Yanqing, et al.
Veröffentlicht: (2025)
Isotonic Mechanism for Exponential Family Estimation in Machine Learning Peer Review
von: Yan, Yuling, et al.
Veröffentlicht: (2023)
von: Yan, Yuling, et al.
Veröffentlicht: (2023)
Computing Optimal Regularizers for Online Linear Optimization
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2024)
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2024)
Principal-Agent Hypothesis Testing
von: Bates, Stephen, et al.
Veröffentlicht: (2022)
von: Bates, Stephen, et al.
Veröffentlicht: (2022)
You Are the Best Reviewer of Your Own Papers: The Isotonic Mechanism
von: Su, Weijie
Veröffentlicht: (2022)
von: Su, Weijie
Veröffentlicht: (2022)
Computing Pure-Strategy Nash Equilibria in a Two-Party Policy Competition: Existence and Algorithmic Approaches
von: Lin, Chuang-Chieh, et al.
Veröffentlicht: (2025)
von: Lin, Chuang-Chieh, et al.
Veröffentlicht: (2025)
Prior-Agnostic Incentive-Compatible Exploration
von: Ramalingam, Ramya, et al.
Veröffentlicht: (2026)
von: Ramalingam, Ramya, et al.
Veröffentlicht: (2026)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
von: Lanier, JB, et al.
Veröffentlicht: (2026)
von: Lanier, JB, et al.
Veröffentlicht: (2026)
Incentivizing Exploration with Linear Contexts and Combinatorial Actions
von: Sellke, Mark
Veröffentlicht: (2023)
von: Sellke, Mark
Veröffentlicht: (2023)
Fast and Furious Symmetric Learning in Zero-Sum Games: Gradient Descent as Fictitious Play
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
von: Lazarsfeld, John, et al.
Veröffentlicht: (2025)
Geometry Meets Incentives: Sample-Efficient Incentivized Exploration with Linear Contexts
von: Schiffer, Benjamin, et al.
Veröffentlicht: (2025)
von: Schiffer, Benjamin, et al.
Veröffentlicht: (2025)
Offline Two-Player Zero-Sum Markov Games with KL Regularization
von: Chen, Claire, et al.
Veröffentlicht: (2026)
von: Chen, Claire, et al.
Veröffentlicht: (2026)
Fairness-aware Contextual Dynamic Pricing with Strategic Buyers
von: Liu, Pangpang, et al.
Veröffentlicht: (2025)
von: Liu, Pangpang, et al.
Veröffentlicht: (2025)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
Low-Rank Graphon Estimation: Theory and Applications to Graphon Games
von: Klopp, Olga, et al.
Veröffentlicht: (2025)
von: Klopp, Olga, et al.
Veröffentlicht: (2025)
Linear Functions to the Extended Reals
von: Waggoner, Bo
Veröffentlicht: (2021)
von: Waggoner, Bo
Veröffentlicht: (2021)
Peer Expectation in Robust Forecast Aggregation: The Possibility/Impossibility
von: Kong, Yuqing
Veröffentlicht: (2024)
von: Kong, Yuqing
Veröffentlicht: (2024)
Offline Fictitious Self-Play for Competitive Games
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
Exploration and Persuasion
von: Slivkins, Aleksandrs
Veröffentlicht: (2024)
von: Slivkins, Aleksandrs
Veröffentlicht: (2024)
Provable Policy Gradient Methods for Average-Reward Markov Potential Games
von: Cheng, Min, et al.
Veröffentlicht: (2024)
von: Cheng, Min, et al.
Veröffentlicht: (2024)
Incentive-Aware Synthetic Control: Accurate Counterfactual Estimation via Incentivized Exploration
von: Ngo, Daniel, et al.
Veröffentlicht: (2023)
von: Ngo, Daniel, et al.
Veröffentlicht: (2023)
Learning to Play Multi-Follower Bayesian Stackelberg Games
von: Personnat, Gerson, et al.
Veröffentlicht: (2025)
von: Personnat, Gerson, et al.
Veröffentlicht: (2025)
Strategic Federated Learning: Application to Smart Meter Data Clustering
von: Mohamad, Hassan, et al.
Veröffentlicht: (2024)
von: Mohamad, Hassan, et al.
Veröffentlicht: (2024)
Nash Equilibria via Stochastic Eigendecomposition
von: Gemp, Ian
Veröffentlicht: (2024)
von: Gemp, Ian
Veröffentlicht: (2024)
Dynamic Pricing and Learning with Long-term Reference Effects
von: Agrawal, Shipra, et al.
Veröffentlicht: (2024)
von: Agrawal, Shipra, et al.
Veröffentlicht: (2024)
Robust forecast aggregation via additional queries
von: Frongillo, Rafael, et al.
Veröffentlicht: (2025)
von: Frongillo, Rafael, et al.
Veröffentlicht: (2025)
Online Stackelberg Optimization via Nonlinear Control
von: Brown, William, et al.
Veröffentlicht: (2024)
von: Brown, William, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Bayesian Algorithms for Adversarial Online Learning: from Finite to Infinite Action Spaces
von: Terenin, Alexander, et al.
Veröffentlicht: (2025) -
Optimal Scoring Rule Design under Partial Knowledge
von: Chen, Yiling, et al.
Veröffentlicht: (2021) -
Stochastic contextual bandits with graph feedback: from independence number to MAS number
von: Wen, Yuxiao, et al.
Veröffentlicht: (2024) -
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025) -
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
von: Yan, Yuling, et al.
Veröffentlicht: (2022)