Learning Safely Without Knowing the World:COMPASS-Hedge
Fuente:
arXiv
Salvato in:
| Autori principali: | Hu, Ting, Cai, Luanda, Vlatakis-Gkaragkounis, Emmanouil-Vasileios |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Prudent-Banker: No Extra Fees for Baseline Safety in Adversarial Bandits With and Without Delays
di: Hu, Ting, et al.
Pubblicazione: (2026)
di: Hu, Ting, et al.
Pubblicazione: (2026)
No Coin Left Behind: Maximizing Strategic Surplus Against No-Regret Dynamics
di: Su, Yiheng, et al.
Pubblicazione: (2026)
di: Su, Yiheng, et al.
Pubblicazione: (2026)
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
di: Patel, Deep, et al.
Pubblicazione: (2025)
di: Patel, Deep, et al.
Pubblicazione: (2025)
Last-Iterate Convergence of Adaptive Riemannian Gradient Descent for Equilibrium Computation
di: Cai, Yang, et al.
Pubblicazione: (2023)
di: Cai, Yang, et al.
Pubblicazione: (2023)
Solving Zero-Sum Convex Markov Games
di: Kalogiannis, Fivos, et al.
Pubblicazione: (2025)
di: Kalogiannis, Fivos, et al.
Pubblicazione: (2025)
Contracting with a Learning Agent
di: Guruganesh, Guru, et al.
Pubblicazione: (2024)
di: Guruganesh, Guru, et al.
Pubblicazione: (2024)
Algorithms and Complexity for Computing Nash Equilibria in Adversarial Team Games
di: Anagnostides, Ioannis, et al.
Pubblicazione: (2023)
di: Anagnostides, Ioannis, et al.
Pubblicazione: (2023)
Breaking $1/ε$ Barrier in Quantum Zero-Sum Games: Generalizing Metric Subregularity for Spectraplexes
di: Su, Yiheng, et al.
Pubblicazione: (2025)
di: Su, Yiheng, et al.
Pubblicazione: (2025)
On the Universal Near Optimality of Hedge in Combinatorial Settings
di: Fan, Zhiyuan, et al.
Pubblicazione: (2025)
di: Fan, Zhiyuan, et al.
Pubblicazione: (2025)
Hedging and Approximate Truthfulness in Traditional Forecasting Competitions
di: Monroe, Mary, et al.
Pubblicazione: (2024)
di: Monroe, Mary, et al.
Pubblicazione: (2024)
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
di: Vasconcelos, Francisca, et al.
Pubblicazione: (2023)
di: Vasconcelos, Francisca, et al.
Pubblicazione: (2023)
Tight Regret Upper and Lower Bounds for Optimistic Hedge in Two-Player Zero-Sum Games
di: Tsuchiya, Taira
Pubblicazione: (2025)
di: Tsuchiya, Taira
Pubblicazione: (2025)
Leveraging Noisy Observations in Zero-Sum Games
di: Athanasakos, Emmanouil M, et al.
Pubblicazione: (2024)
di: Athanasakos, Emmanouil M, et al.
Pubblicazione: (2024)
Playing Markov Games Without Observing Payoffs
di: Ablin, Daniel, et al.
Pubblicazione: (2025)
di: Ablin, Daniel, et al.
Pubblicazione: (2025)
Safe Exploitative Play with Untrusted Type Beliefs
di: Li, Tongxin, et al.
Pubblicazione: (2024)
di: Li, Tongxin, et al.
Pubblicazione: (2024)
CHG Shapley: Efficient Data Valuation and Selection towards Trustworthy Machine Learning
di: Cai, Huaiguang
Pubblicazione: (2024)
di: Cai, Huaiguang
Pubblicazione: (2024)
Regularized Proportional Fairness Mechanism for Resource Allocation Without Money
di: Zeng, Sihan, et al.
Pubblicazione: (2025)
di: Zeng, Sihan, et al.
Pubblicazione: (2025)
Optimism Without Regularization: Constant Regret in Zero-Sum Games
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
Learning to Bid in FCR Markets: A Best-of-Both-Worlds Approach
di: Potfer, Marius, et al.
Pubblicazione: (2026)
di: Potfer, Marius, et al.
Pubblicazione: (2026)
MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems
di: Nöther, Jonathan, et al.
Pubblicazione: (2026)
di: Nöther, Jonathan, et al.
Pubblicazione: (2026)
Proximal Regret and Proximal Correlated Equilibria: A New Tractable Solution Concept for Online Learning and Games
di: Cai, Yang, et al.
Pubblicazione: (2025)
di: Cai, Yang, et al.
Pubblicazione: (2025)
Scale-Invariant Regret Matching and Online Learning with Optimal Convergence: Bridging Theory and Practice in Zero-Sum Games
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
Algorithmic Collusion Without Threats
di: Arunachaleswaran, Eshwar Ram, et al.
Pubblicazione: (2024)
di: Arunachaleswaran, Eshwar Ram, et al.
Pubblicazione: (2024)
Learning and Computation of $Φ$-Equilibria at the Frontier of Tractability
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
Calibration through the Lens of Indistinguishability
di: Gopalan, Parikshit, et al.
Pubblicazione: (2025)
di: Gopalan, Parikshit, et al.
Pubblicazione: (2025)
Truthful mechanisms for linear bandit games with private contexts
di: Hu, Yiting, et al.
Pubblicazione: (2025)
di: Hu, Yiting, et al.
Pubblicazione: (2025)
Is Online Linear Optimization Sufficient for Strategic Robustness?
di: Cai, Yang, et al.
Pubblicazione: (2026)
di: Cai, Yang, et al.
Pubblicazione: (2026)
Incentivizing High-Quality Human Annotations with Golden Questions
di: Liu, Shang, et al.
Pubblicazione: (2025)
di: Liu, Shang, et al.
Pubblicazione: (2025)
Nash Equilibrium Between Consumer Electronic Devices and DoS Attacker for Distributed IoT-enabled RSE Systems
di: Chen, Gengcan, et al.
Pubblicazione: (2025)
di: Chen, Gengcan, et al.
Pubblicazione: (2025)
User Response in Ad Auctions: An MDP Formulation of Long-Term Revenue Optimization
di: Cai, Yang, et al.
Pubblicazione: (2023)
di: Cai, Yang, et al.
Pubblicazione: (2023)
Learning not to Regret
di: Sychrovský, David, et al.
Pubblicazione: (2023)
di: Sychrovský, David, et al.
Pubblicazione: (2023)
On Tractable $Φ$-Equilibria in Non-Concave Games
di: Cai, Yang, et al.
Pubblicazione: (2024)
di: Cai, Yang, et al.
Pubblicazione: (2024)
Learning to Bid in Non-Stationary Repeated First-Price Auctions
di: Hu, Zihao, et al.
Pubblicazione: (2025)
di: Hu, Zihao, et al.
Pubblicazione: (2025)
Learning Safe Strategies for Value Maximizing Buyers in Uniform Price Auctions
di: Golrezaei, Negin, et al.
Pubblicazione: (2024)
di: Golrezaei, Negin, et al.
Pubblicazione: (2024)
PAC Learning with Improvements
di: Attias, Idan, et al.
Pubblicazione: (2025)
di: Attias, Idan, et al.
Pubblicazione: (2025)
Learning Local Stackelberg Equilibria from Repeated Interactions with a Learning Agent
di: Ananthakrishnan, Nivasini, et al.
Pubblicazione: (2025)
di: Ananthakrishnan, Nivasini, et al.
Pubblicazione: (2025)
Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models
di: Fang, Zeyu, et al.
Pubblicazione: (2024)
di: Fang, Zeyu, et al.
Pubblicazione: (2024)
Online Learning with Bounded Recall
di: Schneider, Jon, et al.
Pubblicazione: (2022)
di: Schneider, Jon, et al.
Pubblicazione: (2022)
Learning in Structured Stackelberg Games
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
Incentivized Collaboration in Active Learning
di: Cohen, Lee, et al.
Pubblicazione: (2023)
di: Cohen, Lee, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Prudent-Banker: No Extra Fees for Baseline Safety in Adversarial Bandits With and Without Delays
di: Hu, Ting, et al.
Pubblicazione: (2026) -
No Coin Left Behind: Maximizing Strategic Surplus Against No-Regret Dynamics
di: Su, Yiheng, et al.
Pubblicazione: (2026) -
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
di: Patel, Deep, et al.
Pubblicazione: (2025) -
Last-Iterate Convergence of Adaptive Riemannian Gradient Descent for Equilibrium Computation
di: Cai, Yang, et al.
Pubblicazione: (2023) -
Solving Zero-Sum Convex Markov Games
di: Kalogiannis, Fivos, et al.
Pubblicazione: (2025)