LOQA: Learning with Opponent Q-Learning Awareness
Fuente:
arXiv
Guardado en:
| Autores principales: | Aghajohari, Milad, Duque, Juan Agustin, Cooijmans, Tim, Courville, Aaron |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Best Response Shaping
por: Aghajohari, Milad, et al.
Publicado: (2024)
por: Aghajohari, Milad, et al.
Publicado: (2024)
Towards Sustainable Investment Policies Informed by Opponent Shaping
por: Duque, Juan Agustin, et al.
Publicado: (2026)
por: Duque, Juan Agustin, et al.
Publicado: (2026)
Analysing the Sample Complexity of Opponent Shaping
por: Fung, Kitty, et al.
Publicado: (2024)
por: Fung, Kitty, et al.
Publicado: (2024)
Learning to Play Against Unknown Opponents
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2024)
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2024)
Meta-Computing Enhanced Federated Learning in IIoT: Satisfaction-Aware Incentive Scheme via DRL-Based Stackelberg Game
por: Li, Xiaohuan, et al.
Publicado: (2025)
por: Li, Xiaohuan, et al.
Publicado: (2025)
A Bayesian Learning Algorithm for Unknown Zero-sum Stochastic Games with an Arbitrary Opponent
por: Jafarnia-Jahromi, Mehdi, et al.
Publicado: (2021)
por: Jafarnia-Jahromi, Mehdi, et al.
Publicado: (2021)
Decision Theory-Guided Deep Reinforcement Learning for Fast Learning
por: Wan, Zelin, et al.
Publicado: (2024)
por: Wan, Zelin, et al.
Publicado: (2024)
Gradient-based Learning in State-based Potential Games for Self-Learning Production Systems
por: Yuwono, Steve, et al.
Publicado: (2024)
por: Yuwono, Steve, et al.
Publicado: (2024)
Learning and Collusion in Multi-unit Auctions
por: Brânzei, Simina, et al.
Publicado: (2023)
por: Brânzei, Simina, et al.
Publicado: (2023)
Machine Learning-Powered Course Allocation
por: Soumalias, Ermis, et al.
Publicado: (2022)
por: Soumalias, Ermis, et al.
Publicado: (2022)
Learning Truthful Mechanisms without Discretization
por: Ma, Yunxuan, et al.
Publicado: (2025)
por: Ma, Yunxuan, et al.
Publicado: (2025)
Antithetic Sampling for Top-k Shapley Identification
por: Kolpaczki, Patrick, et al.
Publicado: (2025)
por: Kolpaczki, Patrick, et al.
Publicado: (2025)
Explore Reinforced: Equilibrium Approximation with Reinforcement Learning
por: Yu, Ryan, et al.
Publicado: (2024)
por: Yu, Ryan, et al.
Publicado: (2024)
Deep Reinforcement Learning for Sequential Combinatorial Auctions
por: Ravindranath, Sai Srivatsa, et al.
Publicado: (2024)
por: Ravindranath, Sai Srivatsa, et al.
Publicado: (2024)
An Auction-based Marketplace for Model Trading in Federated Learning
por: Cui, Yue, et al.
Publicado: (2024)
por: Cui, Yue, et al.
Publicado: (2024)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
por: Li, Yingru, et al.
Publicado: (2024)
por: Li, Yingru, et al.
Publicado: (2024)
Knowledge-Free Correlated Agreement for Incentivizing Federated Learning
por: Witt, Leon, et al.
Publicado: (2026)
por: Witt, Leon, et al.
Publicado: (2026)
Managing multiple agents by automatically adjusting incentives
por: Akatsuka, Shunichi, et al.
Publicado: (2024)
por: Akatsuka, Shunichi, et al.
Publicado: (2024)
Trustworthy Machine Learning under Social and Adversarial Data Sources
por: Shao, Han
Publicado: (2024)
por: Shao, Han
Publicado: (2024)
Rethinking Teacher-Student Curriculum Learning through the Cooperative Mechanics of Experience
por: Diaz, Manfred, et al.
Publicado: (2024)
por: Diaz, Manfred, et al.
Publicado: (2024)
Federated Learning for Data Market: Shapley-UCB for Seller Selection and Incentives
por: Chen, Kongyang, et al.
Publicado: (2024)
por: Chen, Kongyang, et al.
Publicado: (2024)
Online Test Synthesis From Requirements: Enhancing Reinforcement Learning with Game Theory
por: Sankur, Ocan, et al.
Publicado: (2024)
por: Sankur, Ocan, et al.
Publicado: (2024)
Optimizing Hard-to-Place Kidney Allocation: A Machine Learning Approach to Center Ranking
por: Berry, Sean, et al.
Publicado: (2024)
por: Berry, Sean, et al.
Publicado: (2024)
Agent-oriented Joint Decision Support for Data Owners in Auction-based Federated Learning
por: Tang, Xiaoli, et al.
Publicado: (2024)
por: Tang, Xiaoli, et al.
Publicado: (2024)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
por: Park, Chanwoo, et al.
Publicado: (2024)
por: Park, Chanwoo, et al.
Publicado: (2024)
Puzzle it Out: Local-to-Global World Model for Offline Multi-Agent Reinforcement Learning
por: Li, Sijia, et al.
Publicado: (2026)
por: Li, Sijia, et al.
Publicado: (2026)
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
por: Chen, Yang, et al.
Publicado: (2025)
por: Chen, Yang, et al.
Publicado: (2025)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
por: Smid, Yanna Elizabeth, et al.
Publicado: (2025)
por: Smid, Yanna Elizabeth, et al.
Publicado: (2025)
GemNet: Menu-Based, Strategy-Proof Multi-Bidder Auctions Through Deep Learning
por: Wang, Tonghan, et al.
Publicado: (2024)
por: Wang, Tonghan, et al.
Publicado: (2024)
How Can Incentives and Cut Layer Selection Influence Data Contribution in Split Federated Learning?
por: Lee, Joohyung, et al.
Publicado: (2024)
por: Lee, Joohyung, et al.
Publicado: (2024)
NePPO: Near-Potential Policy Optimization for General-Sum Multi-Agent Reinforcement Learning
por: Kalanther, Addison, et al.
Publicado: (2026)
por: Kalanther, Addison, et al.
Publicado: (2026)
Evaluating LLM Agent Collusion in Double Auctions
por: Agrawal, Kushal, et al.
Publicado: (2025)
por: Agrawal, Kushal, et al.
Publicado: (2025)
Learning Bilateral Team Formation in Cooperative Multi-Agent Reinforcement Learning
por: Moslemi, Koorosh, et al.
Publicado: (2025)
por: Moslemi, Koorosh, et al.
Publicado: (2025)
Learning Aggregation Rules in Participatory Budgeting: A Data-Driven Approach
por: Fairstein, Roy, et al.
Publicado: (2024)
por: Fairstein, Roy, et al.
Publicado: (2024)
Reinforcement Learning for Hanabi
por: Cohen, Nina, et al.
Publicado: (2025)
por: Cohen, Nina, et al.
Publicado: (2025)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
por: Zhang, Yuheng, et al.
Publicado: (2024)
por: Zhang, Yuheng, et al.
Publicado: (2024)
Contracting with a Learning Agent
por: Guruganesh, Guru, et al.
Publicado: (2024)
por: Guruganesh, Guru, et al.
Publicado: (2024)
Efficient Inverse Multiagent Learning
por: Goktas, Denizalp, et al.
Publicado: (2025)
por: Goktas, Denizalp, et al.
Publicado: (2025)
Nash Learning from Human Feedback
por: Munos, Rémi, et al.
Publicado: (2023)
por: Munos, Rémi, et al.
Publicado: (2023)
Ejemplares similares
-
Best Response Shaping
por: Aghajohari, Milad, et al.
Publicado: (2024) -
Towards Sustainable Investment Policies Informed by Opponent Shaping
por: Duque, Juan Agustin, et al.
Publicado: (2026) -
Analysing the Sample Complexity of Opponent Shaping
por: Fung, Kitty, et al.
Publicado: (2024) -
Learning to Play Against Unknown Opponents
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2024) -
Meta-Computing Enhanced Federated Learning in IIoT: Satisfaction-Aware Incentive Scheme via DRL-Based Stackelberg Game
por: Li, Xiaohuan, et al.
Publicado: (2025)