A Minimax Approach to Ad Hoc Teamwork
Fuente:
arXiv
Guardado en:
| Autores principales: | Villin, Victor, Buening, Thomas Kleine, Dimitrakakis, Christos |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Environment Design for Inverse Reinforcement Learning
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
Inference of Altruism and Intrinsic Rewards in Multi-Agent Systems
por: Villin, Victor, et al.
Publicado: (2025)
por: Villin, Victor, et al.
Publicado: (2025)
Symmetry-Breaking Augmentations for Ad Hoc Teamwork
por: Hammond, Ravi, et al.
Publicado: (2024)
por: Hammond, Ravi, et al.
Publicado: (2024)
Strategic Linear Contextual Bandits
por: Buening, Thomas Kleine, et al.
Publicado: (2024)
por: Buening, Thomas Kleine, et al.
Publicado: (2024)
Causal Imitation Learning under Expert-Observable and Expert-Unobservable Confounding
por: Shao, Daqian, et al.
Publicado: (2025)
por: Shao, Daqian, et al.
Publicado: (2025)
Minimum Coverage Sets for Training Robust Ad Hoc Teamwork Agents
por: Rahman, Arrasy, et al.
Publicado: (2023)
por: Rahman, Arrasy, et al.
Publicado: (2023)
PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc Teamwork
por: Chan, Hohei, et al.
Publicado: (2025)
por: Chan, Hohei, et al.
Publicado: (2025)
N-Agent Ad Hoc Teamwork
por: Wang, Caroline, et al.
Publicado: (2024)
por: Wang, Caroline, et al.
Publicado: (2024)
Fair Contracts in Principal-Agent Games with Heterogeneous Types
por: Tłuczek, Jakub, et al.
Publicado: (2025)
por: Tłuczek, Jakub, et al.
Publicado: (2025)
Generic-to-Specific Reasoning and Learning for Scalable Ad Hoc Teamwork
por: Dodampegama, Hasra, et al.
Publicado: (2025)
por: Dodampegama, Hasra, et al.
Publicado: (2025)
Multi-party Agent Relation Sampling for Multi-party Ad Hoc Teamwork
por: Zhang, Beiwen, et al.
Publicado: (2025)
por: Zhang, Beiwen, et al.
Publicado: (2025)
Stackelberg Learning from Human Feedback: Preference Optimization as a Sequential Game
por: Pásztor, Barna, et al.
Publicado: (2025)
por: Pásztor, Barna, et al.
Publicado: (2025)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
por: Ribeiro, João G., et al.
Publicado: (2025)
por: Ribeiro, João G., et al.
Publicado: (2025)
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
por: Jing, Yuheng, et al.
Publicado: (2026)
por: Jing, Yuheng, et al.
Publicado: (2026)
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
por: Wang, Caroline, et al.
Publicado: (2025)
por: Wang, Caroline, et al.
Publicado: (2025)
MAVRL: Learning Reward Functions from Multiple Feedback Types with Amortized Variational Inference
por: Baur, Raphaël, et al.
Publicado: (2026)
por: Baur, Raphaël, et al.
Publicado: (2026)
Aligning Language Models from User Interactions
por: Buening, Thomas Kleine, et al.
Publicado: (2026)
por: Buening, Thomas Kleine, et al.
Publicado: (2026)
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork
por: Ruhdorfer, Constantin, et al.
Publicado: (2025)
por: Ruhdorfer, Constantin, et al.
Publicado: (2025)
Proactive Memory for Ad-Hoc Recall over Streaming Dialogues
por: Wang, Bingbing, et al.
Publicado: (2026)
por: Wang, Bingbing, et al.
Publicado: (2026)
Ad-Hoc Human-AI Coordination Challenge
por: Dizdarević, Tin, et al.
Publicado: (2025)
por: Dizdarević, Tin, et al.
Publicado: (2025)
Seldonian Reinforcement Learning for Ad Hoc Teamwork
por: Zorzi, Edoardo, et al.
Publicado: (2025)
por: Zorzi, Edoardo, et al.
Publicado: (2025)
Reinforcement Learning via Self-Distillation
por: Hübotter, Jonas, et al.
Publicado: (2026)
por: Hübotter, Jonas, et al.
Publicado: (2026)
Traffic Simulation in Ad Hoc Network of Flying UAVs with Generative AI Adaptation
por: Grekhov, Andrii, et al.
Publicado: (2026)
por: Grekhov, Andrii, et al.
Publicado: (2026)
LLM-Ehnanced Holonic Architecture for Ad-Hoc Scalable SoS
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
Sequence Modeling for N-Agent Ad Hoc Teamwork
por: Wang, Caroline, et al.
Publicado: (2025)
por: Wang, Caroline, et al.
Publicado: (2025)
Open Ad Hoc Teamwork with Cooperative Game Theory
por: Wang, Jianhong, et al.
Publicado: (2024)
por: Wang, Jianhong, et al.
Publicado: (2024)
Towards Adaptive, Scalable, and Robust Coordination of LLM Agents: A Dynamic Ad-Hoc Networking Perspective
por: Li, Rui, et al.
Publicado: (2026)
por: Li, Rui, et al.
Publicado: (2026)
Zero-Shot Coordination in Ad Hoc Teams with Generalized Policy Improvement and Difference Rewards
por: Nigam, Rupal, et al.
Publicado: (2025)
por: Nigam, Rupal, et al.
Publicado: (2025)
Completeness of Unbounded Best-First Minimax and Descent Minimax
por: Cohen-Solal, Quentin
Publicado: (2026)
por: Cohen-Solal, Quentin
Publicado: (2026)
Current Practices for Building LLM-Powered Reasoning Tools Are Ad Hoc -- and We Can Do Better
por: Bembenek, Aaron
Publicado: (2025)
por: Bembenek, Aaron
Publicado: (2025)
Minimax Strikes Back
por: Cohen-Solal, Quentin, et al.
Publicado: (2020)
por: Cohen-Solal, Quentin, et al.
Publicado: (2020)
On some improvements to Unbounded Minimax
por: Cohen-Solal, Quentin, et al.
Publicado: (2025)
por: Cohen-Solal, Quentin, et al.
Publicado: (2025)
Collaborating with AI Agents: Field Experiments on Teamwork, Productivity, and Performance
por: Ju, Harang, et al.
Publicado: (2025)
por: Ju, Harang, et al.
Publicado: (2025)
Leveraging Large Language Model for Heterogeneous Ad Hoc Teamwork Collaboration
por: Liu, Xinzhu, et al.
Publicado: (2024)
por: Liu, Xinzhu, et al.
Publicado: (2024)
TeamMedAgents: Pareto-Efficient Multi-Agent Medical Reasoning Through Teamwork Theory
por: Mishra, Pranav Pushkar, et al.
Publicado: (2025)
por: Mishra, Pranav Pushkar, et al.
Publicado: (2025)
Measuring Successful Cooperation in Human-AI Teamwork: Development and Validation of the Perceived Cooperativity and Teaming Perception Scales
por: Attig, Christiane, et al.
Publicado: (2026)
por: Attig, Christiane, et al.
Publicado: (2026)
Cooperation on the Fly: Exploring Language Agents for Ad Hoc Teamwork in the Avalon Game
por: Shi, Zijing, et al.
Publicado: (2023)
por: Shi, Zijing, et al.
Publicado: (2023)
Strategyproof Reinforcement Learning from Human Feedback
por: Buening, Thomas Kleine, et al.
Publicado: (2025)
por: Buening, Thomas Kleine, et al.
Publicado: (2025)
Shapley Machine: A Game-Theoretic Framework for N-Agent Ad Hoc Teamwork
por: Wang, Jianhong, et al.
Publicado: (2025)
por: Wang, Jianhong, et al.
Publicado: (2025)
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
por: Fosong, Elliot, et al.
Publicado: (2023)
por: Fosong, Elliot, et al.
Publicado: (2023)
Ejemplares similares
-
Environment Design for Inverse Reinforcement Learning
por: Buening, Thomas Kleine, et al.
Publicado: (2022) -
Inference of Altruism and Intrinsic Rewards in Multi-Agent Systems
por: Villin, Victor, et al.
Publicado: (2025) -
Symmetry-Breaking Augmentations for Ad Hoc Teamwork
por: Hammond, Ravi, et al.
Publicado: (2024) -
Strategic Linear Contextual Bandits
por: Buening, Thomas Kleine, et al.
Publicado: (2024) -
Causal Imitation Learning under Expert-Observable and Expert-Unobservable Confounding
por: Shao, Daqian, et al.
Publicado: (2025)