Do you want to play a game? Learning to play Tic-Tac-Toe in Hypermedia Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Beaumont, Katharine, Collier, Rem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
astra-langchain4j: Experiences Combining LLMs and Agent Programming
von: Collier, Rem, et al.
Veröffentlicht: (2026)
von: Collier, Rem, et al.
Veröffentlicht: (2026)
Towards Hypermedia Environments for Adaptive Coordination in Industrial Automation
von: Ramanathan, Ganesh, et al.
Veröffentlicht: (2024)
von: Ramanathan, Ganesh, et al.
Veröffentlicht: (2024)
GenTac: Generative Modeling and Forecasting of Soccer Tactics
von: Rao, Jiayuan, et al.
Veröffentlicht: (2026)
von: Rao, Jiayuan, et al.
Veröffentlicht: (2026)
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders
von: Li, Cheng, et al.
Veröffentlicht: (2024)
von: Li, Cheng, et al.
Veröffentlicht: (2024)
Human-compatible driving partners through data-regularized self-play reinforcement learning
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2024)
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2024)
Tic-Tac-Toe on Designs
von: Danziger, Peter, et al.
Veröffentlicht: (2024)
von: Danziger, Peter, et al.
Veröffentlicht: (2024)
Tic‐Tac‐Toe on Designs
von: Peter Danziger, et al.
Veröffentlicht: (2024)
von: Peter Danziger, et al.
Veröffentlicht: (2024)
Spirals, Tic-Tac-Toe Partition, and Deep Diagonal Maps
von: Zou, Zhengyu
Veröffentlicht: (2024)
von: Zou, Zhengyu
Veröffentlicht: (2024)
Depending on yourself when you should: Mentoring LLM with RL agents to become the master in cybersecurity games
von: Yan, Yikuan, et al.
Veröffentlicht: (2024)
von: Yan, Yikuan, et al.
Veröffentlicht: (2024)
PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits
von: Bhuiya, Neeladri, et al.
Veröffentlicht: (2025)
von: Bhuiya, Neeladri, et al.
Veröffentlicht: (2025)
Finite-time convergence to an $ε$-efficient Nash equilibrium in potential games
von: Maddux, Anna, et al.
Veröffentlicht: (2024)
von: Maddux, Anna, et al.
Veröffentlicht: (2024)
Transformers as Game Players: Provable In-context Game-playing Capabilities of Pre-trained Models
von: Shi, Chengshuai, et al.
Veröffentlicht: (2024)
von: Shi, Chengshuai, et al.
Veröffentlicht: (2024)
Modeling human reputation-seeking behavior in a spatio-temporally complex public good provision game
von: Hughes, Edward, et al.
Veröffentlicht: (2025)
von: Hughes, Edward, et al.
Veröffentlicht: (2025)
Tic‐Tac‐Toe in Socioeducational Interviews in Open‐Environment Settings
von: Diana Galone Somer, et al.
Veröffentlicht: (2026)
von: Diana Galone Somer, et al.
Veröffentlicht: (2026)
Autonomous vehicles need social awareness to find optima in multi-agent reinforcement learning routing games
von: Psarou, Anastasia, et al.
Veröffentlicht: (2025)
von: Psarou, Anastasia, et al.
Veröffentlicht: (2025)
Adaptive Decision-Making for Autonomous Vehicles: A Learning-Enhanced Game-Theoretic Approach in Interactive Environments
von: Huang, Heye, et al.
Veröffentlicht: (2024)
von: Huang, Heye, et al.
Veröffentlicht: (2024)
MLC-Agent: Cognitive Model based on Memory-Learning Collaboration in LLM Empowered Agent Simulation Environment
von: Zhang, Ming, et al.
Veröffentlicht: (2025)
von: Zhang, Ming, et al.
Veröffentlicht: (2025)
Knowing What Not to Do: Leverage Language Model Insights for Action Space Pruning in Multi-agent Reinforcement Learning
von: Liu, Zhihao, et al.
Veröffentlicht: (2024)
von: Liu, Zhihao, et al.
Veröffentlicht: (2024)
Strategic Concealment of Environment Representations in Competitive Games
von: Guan, Yue, et al.
Veröffentlicht: (2025)
von: Guan, Yue, et al.
Veröffentlicht: (2025)
Zero-determinant strategies in repeated continuously-relaxed games
von: Ueda, Masahiko, et al.
Veröffentlicht: (2025)
von: Ueda, Masahiko, et al.
Veröffentlicht: (2025)
Collaborating in a competitive world: Heterogeneous Multi-Agent Decision Making in Symbiotic Supply Chain Environments
von: Wang, Wan, et al.
Veröffentlicht: (2025)
von: Wang, Wan, et al.
Veröffentlicht: (2025)
Adaptive bias for dissensus in nonlinear opinion dynamics with application to evolutionary division of labor games
von: Paine, Tyler M., et al.
Veröffentlicht: (2024)
von: Paine, Tyler M., et al.
Veröffentlicht: (2024)
What Do Agents Communicate? Characterizing Information Exchange in Multi-Agent Systems
von: Chun, Yong Jin, et al.
Veröffentlicht: (2026)
von: Chun, Yong Jin, et al.
Veröffentlicht: (2026)
Bearing-Distance Flocking with Zone-Based Interactions in Constrained Dynamic Environments
von: Jond, Hossein B.
Veröffentlicht: (2024)
von: Jond, Hossein B.
Veröffentlicht: (2024)
Policy Optimization in Multi-Agent Settings under Partially Observable Environments
von: Zhaikhan, Ainur, et al.
Veröffentlicht: (2025)
von: Zhaikhan, Ainur, et al.
Veröffentlicht: (2025)
Dynamic Strategy Adaptation in Multi-Agent Environments with Large Language Models
von: Mallampati, Shaurya, et al.
Veröffentlicht: (2025)
von: Mallampati, Shaurya, et al.
Veröffentlicht: (2025)
Quantized distributed Nash equilibrium seeking under DoS attacks
von: Feng, Shuai, et al.
Veröffentlicht: (2023)
von: Feng, Shuai, et al.
Veröffentlicht: (2023)
Nash equilibrium seeking for a class of quadratic-bilinear Wasserstein distributionally robust games
von: Pantazis, Georgios, et al.
Veröffentlicht: (2024)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2024)
Knowledge Graph-Based Multi-Agent Path Planning in Dynamic Environments using WAITR
von: Holmberg, Ted Edward, et al.
Veröffentlicht: (2024)
von: Holmberg, Ted Edward, et al.
Veröffentlicht: (2024)
Enhancing the Performance of Multi-Vehicle Navigation in Unstructured Environments using Hard Sample Mining
von: Ma, Yining, et al.
Veröffentlicht: (2024)
von: Ma, Yining, et al.
Veröffentlicht: (2024)
Congestion Mitigation Path Planning for Large-Scale Multi-Agent Navigation in Dense Environments
von: Kato, Takuro, et al.
Veröffentlicht: (2025)
von: Kato, Takuro, et al.
Veröffentlicht: (2025)
EconAI: Dynamic Persona Evolution and Memory-Aware Agents in Evolving Economic Environments
von: Liu, Annie, et al.
Veröffentlicht: (2026)
von: Liu, Annie, et al.
Veröffentlicht: (2026)
Decentralized Ergodic Coverage Control in Unknown Time-Varying Environments
von: Mendoza, Maria G., et al.
Veröffentlicht: (2026)
von: Mendoza, Maria G., et al.
Veröffentlicht: (2026)
What Contributes to Affective Polarization in Networked Online Environments? Evidence from an Agent-Based Model
von: Vedam, Narayani, et al.
Veröffentlicht: (2025)
von: Vedam, Narayani, et al.
Veröffentlicht: (2025)
On the existence of fair zero-determinant strategies in the periodic prisoner's dilemma game
von: Nakamura, Ken, et al.
Veröffentlicht: (2026)
von: Nakamura, Ken, et al.
Veröffentlicht: (2026)
DTPPO: Dual-Transformer Encoder-based Proximal Policy Optimization for Multi-UAV Navigation in Unseen Complex Environments
von: Wei, Anning, et al.
Veröffentlicht: (2024)
von: Wei, Anning, et al.
Veröffentlicht: (2024)
Tapas Are Free! Training-Free Adaptation of Programmatic Agents via LLM-Guided Program Synthesis in Dynamic Environments
von: Hu, Jinwei, et al.
Veröffentlicht: (2025)
von: Hu, Jinwei, et al.
Veröffentlicht: (2025)
OpenMines: A Light and Comprehensive Mining Simulation Environment for Truck Dispatching
von: Meng, Shi, et al.
Veröffentlicht: (2024)
von: Meng, Shi, et al.
Veröffentlicht: (2024)
Graph-Based Complexity Metrics for Multi-Agent Curriculum Learning: A Validated Approach to Task Ordering in Cooperative Coordination Environments
von: Ebadulla, Farhaan, et al.
Veröffentlicht: (2025)
von: Ebadulla, Farhaan, et al.
Veröffentlicht: (2025)
On the Benefits of Robot Platooning for Navigating Crowded Environments
von: Argote-Gerald, Jahir, et al.
Veröffentlicht: (2024)
von: Argote-Gerald, Jahir, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
astra-langchain4j: Experiences Combining LLMs and Agent Programming
von: Collier, Rem, et al.
Veröffentlicht: (2026) -
Towards Hypermedia Environments for Adaptive Coordination in Industrial Automation
von: Ramanathan, Ganesh, et al.
Veröffentlicht: (2024) -
GenTac: Generative Modeling and Forecasting of Soccer Tactics
von: Rao, Jiayuan, et al.
Veröffentlicht: (2026) -
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders
von: Li, Cheng, et al.
Veröffentlicht: (2024) -
Human-compatible driving partners through data-regularized self-play reinforcement learning
von: Cornelisse, Daphne, et al.
Veröffentlicht: (2024)