Multi-agent cooperation through in-context co-player inference
Fuente:
arXiv
Guardado en:
| Autores principales: | Weis, Marissa A., Wołczyk, Maciej, Nasser, Rajai, Saurous, Rif A., Arcas, Blaise Agüera y, Sacramento, João, Meulemans, Alexander |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning
por: Meulemans, Alexander, et al.
Publicado: (2025)
por: Meulemans, Alexander, et al.
Publicado: (2025)
Multi-agent cooperation through learning-aware policy gradients
por: Meulemans, Alexander, et al.
Publicado: (2024)
por: Meulemans, Alexander, et al.
Publicado: (2024)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
por: Kobayashi, Seijin, et al.
Publicado: (2025)
por: Kobayashi, Seijin, et al.
Publicado: (2025)
The unreasonable effectiveness of pattern matching
por: Lupyan, Gary, et al.
Publicado: (2026)
por: Lupyan, Gary, et al.
Publicado: (2026)
MesaNet: Sequence Modeling by Locally Optimal Test-Time Training
por: von Oswald, Johannes, et al.
Publicado: (2025)
por: von Oswald, Johannes, et al.
Publicado: (2025)
Agentic AI and the next intelligence explosion
por: Evans, James, et al.
Publicado: (2026)
por: Evans, James, et al.
Publicado: (2026)
All Random Features Representations are Equivalent
por: Sernau, Luke, et al.
Publicado: (2024)
por: Sernau, Luke, et al.
Publicado: (2024)
What Lives? A meta-analysis of diverse opinions on the definition of life
por: Bender, Reed, et al.
Publicado: (2025)
por: Bender, Reed, et al.
Publicado: (2025)
State Soup: In-Context Skill Learning, Retrieval and Mixing
por: Pióro, Maciej, et al.
Publicado: (2024)
por: Pióro, Maciej, et al.
Publicado: (2024)
Uncovering mesa-optimization algorithms in Transformers
por: von Oswald, Johannes, et al.
Publicado: (2023)
por: von Oswald, Johannes, et al.
Publicado: (2023)
Reasoning Models Generate Societies of Thought
por: Kim, Junsol, et al.
Publicado: (2026)
por: Kim, Junsol, et al.
Publicado: (2026)
Can LLMs get help from other LLMs without revealing private information?
por: Hartmann, Florian, et al.
Publicado: (2024)
por: Hartmann, Florian, et al.
Publicado: (2024)
Social Learning: Towards Collaborative Learning with Large Language Models
por: Mohtashami, Amirkeivan, et al.
Publicado: (2023)
por: Mohtashami, Amirkeivan, et al.
Publicado: (2023)
Computational Life: How Well-formed, Self-replicating Programs Emerge from Simple Interaction
por: Arcas, Blaise Agüera y, et al.
Publicado: (2024)
por: Arcas, Blaise Agüera y, et al.
Publicado: (2024)
Scalable Spatiotemporal Prediction with Bayesian Neural Fields
por: Saad, Feras, et al.
Publicado: (2024)
por: Saad, Feras, et al.
Publicado: (2024)
Robust Inverse Graphics via Probabilistic Inference
por: Le, Tuan Anh, et al.
Publicado: (2024)
por: Le, Tuan Anh, et al.
Publicado: (2024)
Multi-View Stochastic Block Models
por: Cohen-Addad, Vincent, et al.
Publicado: (2024)
por: Cohen-Addad, Vincent, et al.
Publicado: (2024)
Discovering modular solutions that generalize compositionally
por: Schug, Simon, et al.
Publicado: (2023)
por: Schug, Simon, et al.
Publicado: (2023)
When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors
por: Emmons, Scott, et al.
Publicado: (2025)
por: Emmons, Scott, et al.
Publicado: (2025)
Towards a future space-based, highly scalable AI infrastructure system design
por: Arcas, Blaise Agüera y, et al.
Publicado: (2025)
por: Arcas, Blaise Agüera y, et al.
Publicado: (2025)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
por: Góral, Gracjan, et al.
Publicado: (2024)
por: Góral, Gracjan, et al.
Publicado: (2024)
Can LLMs make trade-offs involving stipulated pain and pleasure states?
por: Keeling, Geoff, et al.
Publicado: (2024)
por: Keeling, Geoff, et al.
Publicado: (2024)
The challenges of education in contexts of increasing migratory diversity: (mis)adjustments, adaptative practices and creativity in Portuguese schools
por: Octávio Sacramento
Publicado: (2023)
por: Octávio Sacramento
Publicado: (2023)
MNIST-Nd: a set of naturalistic datasets to benchmark clustering across dimensions
por: Turishcheva, Polina, et al.
Publicado: (2024)
por: Turishcheva, Polina, et al.
Publicado: (2024)
Interactive inference: a multi-agent model of cooperative joint actions
por: Maisto, Domenico, et al.
Publicado: (2022)
por: Maisto, Domenico, et al.
Publicado: (2022)
Formal Conjectures: An Open and Evolving Benchmark for Verified Discovery in Mathematics
por: Firsching, Moritz, et al.
Publicado: (2026)
por: Firsching, Moritz, et al.
Publicado: (2026)
LLMs achieve adult human performance on higher-order theory of mind tasks
por: Street, Winnie, et al.
Publicado: (2024)
por: Street, Winnie, et al.
Publicado: (2024)
Pushing Frontiers for Proteoglycans
por: Marissa L. Maciej‐Hulme
Publicado: (2026)
por: Marissa L. Maciej‐Hulme
Publicado: (2026)
Hierarchical clustering with maximum density paths and mixture models
por: Ritzert, Martin, et al.
Publicado: (2025)
por: Ritzert, Martin, et al.
Publicado: (2025)
FlySearch: Exploring how vision-language models explore
por: Pardyl, Adam, et al.
Publicado: (2025)
por: Pardyl, Adam, et al.
Publicado: (2025)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
por: Góral, Gracjan, et al.
Publicado: (2025)
por: Góral, Gracjan, et al.
Publicado: (2025)
AdaGlimpse: Active Visual Exploration with Arbitrary Glimpse Position and Scale
por: Pardyl, Adam, et al.
Publicado: (2024)
por: Pardyl, Adam, et al.
Publicado: (2024)
Sustainable cooperation on the hybrid pollution-control game with heterogeneous players
por: Wu, Yilun, et al.
Publicado: (2025)
por: Wu, Yilun, et al.
Publicado: (2025)
Charge‐Mediated Interactions Affect Enzymatic Reactions in Peptide Condensates
por: Rif Harris, et al.
Publicado: (2024)
por: Rif Harris, et al.
Publicado: (2024)
Can LLM-Augmented autonomous agents cooperate?, An evaluation of their cooperative capabilities through Melting Pot
por: Mosquera, Manuel, et al.
Publicado: (2024)
por: Mosquera, Manuel, et al.
Publicado: (2024)
Els Massaguer de Torroella: precursors de la fotografia a Girona
por: Dolors Agüera
Publicado: (2025)
por: Dolors Agüera
Publicado: (2025)
Comparación de rasgos de personalidad entre pacientes con trastorno de la conducta alimentaria y sus hermanas sanas
por: Zaida Agüera
Publicado: (2011)
por: Zaida Agüera
Publicado: (2011)
A logic of co-valuations
por: Malicki, Maciej
Publicado: (2025)
por: Malicki, Maciej
Publicado: (2025)
NMDA antagonists use in bipolar depression: A case report
por: Kirolos Ibrahim, et al.
Publicado: (2024)
por: Kirolos Ibrahim, et al.
Publicado: (2024)
Multi-agent Optimization of Non-cooperative Multimodal Mobility Systems
por: Rafi, Md Nafees Fuad, et al.
Publicado: (2026)
por: Rafi, Md Nafees Fuad, et al.
Publicado: (2026)
Ejemplares similares
-
Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning
por: Meulemans, Alexander, et al.
Publicado: (2025) -
Multi-agent cooperation through learning-aware policy gradients
por: Meulemans, Alexander, et al.
Publicado: (2024) -
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
por: Kobayashi, Seijin, et al.
Publicado: (2025) -
The unreasonable effectiveness of pattern matching
por: Lupyan, Gary, et al.
Publicado: (2026) -
MesaNet: Sequence Modeling by Locally Optimal Test-Time Training
por: von Oswald, Johannes, et al.
Publicado: (2025)