SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Shindo, Hikaru, Lin, Hanzhao, Helff, Lukas, Schramowski, Patrick, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ActivationReasoning: Logical Reasoning in Latent Activation Spaces
by: Helff, Lukas, et al.
Published: (2025)
by: Helff, Lukas, et al.
Published: (2025)
Agent-Oriented Planning in Multi-Agent Systems
by: Li, Ao, et al.
Published: (2024)
by: Li, Ao, et al.
Published: (2024)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
by: Helff, Lukas, et al.
Published: (2026)
by: Helff, Lukas, et al.
Published: (2026)
SLR: Automated Synthesis for Scalable Logical Reasoning
by: Helff, Lukas, et al.
Published: (2025)
by: Helff, Lukas, et al.
Published: (2025)
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Grid-Agent: An LLM-Powered Multi-Agent System for Power Grid Control
by: Zhang, Yan, et al.
Published: (2025)
by: Zhang, Yan, et al.
Published: (2025)
V-LoL: A Diagnostic Dataset for Visual Logical Learning
by: Helff, Lukas, et al.
Published: (2023)
by: Helff, Lukas, et al.
Published: (2023)
Discrete Diffusion for Complex and Congested Multi-Agent Path Finding with Sparse Social Attention
by: Wang, Yuanzhe, et al.
Published: (2026)
by: Wang, Yuanzhe, et al.
Published: (2026)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
by: Helff, Lukas, et al.
Published: (2024)
by: Helff, Lukas, et al.
Published: (2024)
PeerGuard: Defending Multi-Agent Systems Against Backdoor Attacks Through Mutual Reasoning
by: Fan, Falong, et al.
Published: (2025)
by: Fan, Falong, et al.
Published: (2025)
EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents
by: Juneja, Gurusha, et al.
Published: (2026)
by: Juneja, Gurusha, et al.
Published: (2026)
Verification-Aware Planning for Multi-Agent Systems
by: Xu, Tianyang, et al.
Published: (2025)
by: Xu, Tianyang, et al.
Published: (2025)
BenchMARL: Benchmarking Multi-Agent Reinforcement Learning
by: Bettini, Matteo, et al.
Published: (2023)
by: Bettini, Matteo, et al.
Published: (2023)
DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration
by: Nourzad, Narjes, et al.
Published: (2025)
by: Nourzad, Narjes, et al.
Published: (2025)
SocialGFs: Learning Social Gradient Fields for Multi-Agent Reinforcement Learning
by: Long, Qian, et al.
Published: (2024)
by: Long, Qian, et al.
Published: (2024)
POGEMA: A Benchmark Platform for Cooperative Multi-Agent Pathfinding
by: Skrynnik, Alexey, et al.
Published: (2024)
by: Skrynnik, Alexey, et al.
Published: (2024)
EvoMem: Improving Multi-Agent Planning with Dual-Evolving Memory
by: Fan, Wenzhe, et al.
Published: (2025)
by: Fan, Wenzhe, et al.
Published: (2025)
From Assumptions to Actions: Turning LLM Reasoning into Uncertainty-Aware Planning for Embodied Agents
by: Seo, SeungWon, et al.
Published: (2026)
by: Seo, SeungWon, et al.
Published: (2026)
FightLadder: A Benchmark for Competitive Multi-Agent Reinforcement Learning
by: Li, Wenzhe, et al.
Published: (2024)
by: Li, Wenzhe, et al.
Published: (2024)
How to Train a Leader: Hierarchical Reasoning in Multi-Agent LLMs
by: Estornell, Andrew, et al.
Published: (2025)
by: Estornell, Andrew, et al.
Published: (2025)
Agentic Neural Networks: Self-Evolving Multi-Agent Systems via Textual Backpropagation
by: Ma, Xiaowen, et al.
Published: (2025)
by: Ma, Xiaowen, et al.
Published: (2025)
The Role of Social Learning and Collective Norm Formation in Fostering Cooperation in LLM Multi-Agent Systems
by: Gupta, Prateek, et al.
Published: (2025)
by: Gupta, Prateek, et al.
Published: (2025)
Improvisational Games as a Benchmark for Social Intelligence of AI Agents: The Case of Connections
by: Parikh, Gaurav Rajesh, et al.
Published: (2026)
by: Parikh, Gaurav Rajesh, et al.
Published: (2026)
Understanding Individual Agent Importance in Multi-Agent System via Counterfactual Reasoning
by: Chen, Jianming, et al.
Published: (2024)
by: Chen, Jianming, et al.
Published: (2024)
Dynamic Speculative Agent Planning
by: Guan, Yilin, et al.
Published: (2025)
by: Guan, Yilin, et al.
Published: (2025)
TAMAS: Benchmarking Adversarial Risks in Multi-Agent LLM Systems
by: Kavathekar, Ishan, et al.
Published: (2025)
by: Kavathekar, Ishan, et al.
Published: (2025)
GoAgent: Group-of-Agents Communication Topology Generation for LLM-based Multi-Agent Systems
by: Chen, Hongjiang, et al.
Published: (2026)
by: Chen, Hongjiang, et al.
Published: (2026)
Integrating LLM in Agent-Based Social Simulation: Opportunities and Challenges
by: Taillandier, Patrick, et al.
Published: (2025)
by: Taillandier, Patrick, et al.
Published: (2025)
Challenges in Credit Assignment for Multi-Agent Reinforcement Learning in Open Agent Systems
by: Abadi, Alireza Saleh, et al.
Published: (2025)
by: Abadi, Alireza Saleh, et al.
Published: (2025)
Schema-Guided Scene-Graph Reasoning based on Multi-Agent Large Language Model System
by: Chen, Yiye, et al.
Published: (2025)
by: Chen, Yiye, et al.
Published: (2025)
NegotiationGym: Self-Optimizing Agents in a Multi-Agent Social Simulation Environment
by: Mangla, Shashank, et al.
Published: (2025)
by: Mangla, Shashank, et al.
Published: (2025)
Simultaneous Computation with Multiple Prioritizations in Multi-Agent Motion Planning
by: Scheffe, Patrick, et al.
Published: (2025)
by: Scheffe, Patrick, et al.
Published: (2025)
Multi-Agent Dynamic Relational Reasoning for Social Robot Navigation
by: Li, Jiachen, et al.
Published: (2024)
by: Li, Jiachen, et al.
Published: (2024)
HAMLET: A Hierarchical and Adaptive Multi-Agent Framework for Live Embodied Theatrics
by: Jiang, Shufan, et al.
Published: (2025)
by: Jiang, Shufan, et al.
Published: (2025)
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
by: Sarkar, Bidipta, et al.
Published: (2025)
by: Sarkar, Bidipta, et al.
Published: (2025)
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
by: Shlomov, Segev, et al.
Published: (2024)
by: Shlomov, Segev, et al.
Published: (2024)
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
by: Molinari, Gioele, et al.
Published: (2026)
by: Molinari, Gioele, et al.
Published: (2026)
Multi-Agent DRL for V2X Resource Allocation: Disentangling Challenges and Benchmarking Solutions
by: Wang, Siyuan, et al.
Published: (2026)
by: Wang, Siyuan, et al.
Published: (2026)
The Social Laboratory: A Psychometric Framework for Multi-Agent LLM Evaluation
by: Reza, Zarreen
Published: (2025)
by: Reza, Zarreen
Published: (2025)
Learn as Individuals, Evolve as a Team: Multi-agent LLMs Adaptation in Embodied Environments
by: Li, Xinran, et al.
Published: (2025)
by: Li, Xinran, et al.
Published: (2025)
Similar Items
-
ActivationReasoning: Logical Reasoning in Latent Activation Spaces
by: Helff, Lukas, et al.
Published: (2025) -
Agent-Oriented Planning in Multi-Agent Systems
by: Li, Ao, et al.
Published: (2024) -
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
by: Helff, Lukas, et al.
Published: (2026) -
SLR: Automated Synthesis for Scalable Logical Reasoning
by: Helff, Lukas, et al.
Published: (2025) -
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
by: Wu, Di, et al.
Published: (2025)