Self-Play Q-learners Can Provably Collude in the Iterated Prisoner's Dilemma
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bertrand, Quentin, Duque, Juan, Calvano, Emilio, Gidel, Gauthier |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Properties of Winning Iterated Prisoner's Dilemma Strategies
von: Glynatsi, Nikoleta E., et al.
Veröffentlicht: (2020)
von: Glynatsi, Nikoleta E., et al.
Veröffentlicht: (2020)
Forgiveness is an Adaptation in Iterated Prisoner's Dilemma with Memory
von: Turker, Meliksah, et al.
Veröffentlicht: (2021)
von: Turker, Meliksah, et al.
Veröffentlicht: (2021)
Inferring Strategies from Observations in Long Iterated Prisoner's Dilemma Experiments
von: Montero-Porras, Eladio, et al.
Veröffentlicht: (2022)
von: Montero-Porras, Eladio, et al.
Veröffentlicht: (2022)
Why Open Source? A Game-Theoretic Analysis of the AI Race
von: Mladenovic, Andjela, et al.
Veröffentlicht: (2026)
von: Mladenovic, Andjela, et al.
Veröffentlicht: (2026)
Emergence of Cooperation and Commitment in Optional Prisoner's Dilemma
von: Song, Zhao, et al.
Veröffentlicht: (2025)
von: Song, Zhao, et al.
Veröffentlicht: (2025)
A Persuasive Approach to Combating Misinformation
von: Hossain, Safwan, et al.
Veröffentlicht: (2023)
von: Hossain, Safwan, et al.
Veröffentlicht: (2023)
Performative Prediction with Neural Networks
von: Mofakhami, Mehrnaz, et al.
Veröffentlicht: (2023)
von: Mofakhami, Mehrnaz, et al.
Veröffentlicht: (2023)
Domination-Avoiding Learning Agents Cannot Collude
von: Nisan, Noam, et al.
Veröffentlicht: (2026)
von: Nisan, Noam, et al.
Veröffentlicht: (2026)
Move Over, Prisoner's Dilemma: Colonel Blotto has arrived
von: Paarporn, Keith, et al.
Veröffentlicht: (2026)
von: Paarporn, Keith, et al.
Veröffentlicht: (2026)
Nash equilibria in four-strategy quantum game extensions of the Prisoner's Dilemma
von: Frąckiewicz, Piotr, et al.
Veröffentlicht: (2024)
von: Frąckiewicz, Piotr, et al.
Veröffentlicht: (2024)
Formal Methods for An Iterated Volunteer's Dilemma
von: Dineen, Jacob, et al.
Veröffentlicht: (2020)
von: Dineen, Jacob, et al.
Veröffentlicht: (2020)
Preplay Losing Contracts: Inducing Strong Nash Equilibrium in the $n$-player Prisoner's Dilemma
von: Fligler, Ian
Veröffentlicht: (2026)
von: Fligler, Ian
Veröffentlicht: (2026)
Evolving Personalities in Chaos: An LLM-Augmented Framework for Character Discovery in the Iterated Prisoners Dilemma under Environmental Stress
von: Yildirim, Oguzhan
Veröffentlicht: (2026)
von: Yildirim, Oguzhan
Veröffentlicht: (2026)
Agent-based Modelling of Quantum Prisoner's Dilemma
von: Benjamin, Colin, et al.
Veröffentlicht: (2024)
von: Benjamin, Colin, et al.
Veröffentlicht: (2024)
Expected flow networks in stochastic environments and two-player zero-sum games
von: Jiralerspong, Marco, et al.
Veröffentlicht: (2023)
von: Jiralerspong, Marco, et al.
Veröffentlicht: (2023)
Words are not Wind -- How Public Joint Commitment and Reputation Solve the Prisoner's Dilemma
von: Krellner, Marcus, et al.
Veröffentlicht: (2023)
von: Krellner, Marcus, et al.
Veröffentlicht: (2023)
The Evolution of Lying in a Spatially-Explicit Prisoner's Dilemma Model
von: Hartvigsen, Gregg
Veröffentlicht: (2026)
von: Hartvigsen, Gregg
Veröffentlicht: (2026)
Performative Prediction on Games and Mechanism Design
von: Góis, António, et al.
Veröffentlicht: (2024)
von: Góis, António, et al.
Veröffentlicht: (2024)
Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?
von: Fontana, Nicoló, et al.
Veröffentlicht: (2024)
von: Fontana, Nicoló, et al.
Veröffentlicht: (2024)
Quantifying the Self-Interest Level of Markov Social Dilemmas
von: Willis, Richard, et al.
Veröffentlicht: (2025)
von: Willis, Richard, et al.
Veröffentlicht: (2025)
Can a Weaker Player Win? Adaptive Play in Repeated Games
von: ANSELMI, Jonatha, et al.
Veröffentlicht: (2026)
von: ANSELMI, Jonatha, et al.
Veröffentlicht: (2026)
Strategizing against Q-learners: A Control-theoretical Approach
von: Arslantas, Yuksel, et al.
Veröffentlicht: (2024)
von: Arslantas, Yuksel, et al.
Veröffentlicht: (2024)
Too Noisy to Collude? Algorithmic Collusion Under Laplacian Noise
von: Zhang, Niuniu
Veröffentlicht: (2025)
von: Zhang, Niuniu
Veröffentlicht: (2025)
A memory-based spatial evolutionary game with the dynamic interaction between learners and profiteers
von: Pi, Bin, et al.
Veröffentlicht: (2024)
von: Pi, Bin, et al.
Veröffentlicht: (2024)
Preference-Centric Route Recommendation: Equilibrium, Learning, and Provable Efficiency
von: Yang, Ya-Ting, et al.
Veröffentlicht: (2025)
von: Yang, Ya-Ting, et al.
Veröffentlicht: (2025)
Meta-Learning in Self-Play Regret Minimization
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
von: Sychrovský, David, et al.
Veröffentlicht: (2025)
Playing against a stationary opponent
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2025)
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2025)
Anytime Detection of Strategic Deviations in Multi-Agent Systems
von: Gauthier, Etienne, et al.
Veröffentlicht: (2026)
von: Gauthier, Etienne, et al.
Veröffentlicht: (2026)
LOQA: Learning with Opponent Q-Learning Awareness
von: Aghajohari, Milad, et al.
Veröffentlicht: (2024)
von: Aghajohari, Milad, et al.
Veröffentlicht: (2024)
Interactive and Iterative Peer Assessment
von: Dery, Lihi
Veröffentlicht: (2022)
von: Dery, Lihi
Veröffentlicht: (2022)
PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers
von: Li, Boning, et al.
Veröffentlicht: (2026)
von: Li, Boning, et al.
Veröffentlicht: (2026)
Tie-breaking Agnostic Lower Bound for Fictitious Play
von: Wang, Yuanhao
Veröffentlicht: (2025)
von: Wang, Yuanhao
Veröffentlicht: (2025)
Playing Stochastically in Weighted Timed Games to Emulate Memory
von: Monmege, Benjamin, et al.
Veröffentlicht: (2021)
von: Monmege, Benjamin, et al.
Veröffentlicht: (2021)
Pure Exploration via Frank-Wolfe Self-Play
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
Iterative Vickrey Auctions via Linear Programming
von: Lahaie, Sébastien, et al.
Veröffentlicht: (2025)
von: Lahaie, Sébastien, et al.
Veröffentlicht: (2025)
Contracting With a Reinforcement Learning Agent by Playing Trick or Treat
von: Bollini, Matteo, et al.
Veröffentlicht: (2024)
von: Bollini, Matteo, et al.
Veröffentlicht: (2024)
Zero-Sum Fictitious Play Cannot Converge to a Point
von: Moon, Jaehong
Veröffentlicht: (2026)
von: Moon, Jaehong
Veröffentlicht: (2026)
GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2026)
The Value Problem for Weighted Timed Games with Two Clocks is Undecidable
von: Guilmant, Quentin, et al.
Veröffentlicht: (2025)
von: Guilmant, Quentin, et al.
Veröffentlicht: (2025)
Language Self-Play For Data-Free Training
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2025)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Properties of Winning Iterated Prisoner's Dilemma Strategies
von: Glynatsi, Nikoleta E., et al.
Veröffentlicht: (2020) -
Forgiveness is an Adaptation in Iterated Prisoner's Dilemma with Memory
von: Turker, Meliksah, et al.
Veröffentlicht: (2021) -
Inferring Strategies from Observations in Long Iterated Prisoner's Dilemma Experiments
von: Montero-Porras, Eladio, et al.
Veröffentlicht: (2022) -
Why Open Source? A Game-Theoretic Analysis of the AI Race
von: Mladenovic, Andjela, et al.
Veröffentlicht: (2026) -
Emergence of Cooperation and Commitment in Optional Prisoner's Dilemma
von: Song, Zhao, et al.
Veröffentlicht: (2025)