Improvisational Games as a Benchmark for Social Intelligence of AI Agents: The Case of Connections
Fuente:
arXiv
Guardado en:
| Autores principales: | Parikh, Gaurav Rajesh, Ghosal, Angikar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Agent Island: A Saturation- and Contamination-Resistant Benchmark from Multiagent Games
por: Murphy, Connacher
Publicado: (2026)
por: Murphy, Connacher
Publicado: (2026)
Multi-Actor Generative Artificial Intelligence as a Game Engine
por: Vezhnevets, Alexander Sasha, et al.
Publicado: (2025)
por: Vezhnevets, Alexander Sasha, et al.
Publicado: (2025)
The Orchestration of Multi-Agent Systems: Architectures, Protocols, and Enterprise Adoption
por: Adimulam, Apoorva, et al.
Publicado: (2026)
por: Adimulam, Apoorva, et al.
Publicado: (2026)
SALM: A Multi-Agent Framework for Language Model-Driven Social Network Simulation
por: Koley, Gaurav
Publicado: (2025)
por: Koley, Gaurav
Publicado: (2025)
Memory Intelligence Agent
por: Qiao, Jingyang, et al.
Publicado: (2026)
por: Qiao, Jingyang, et al.
Publicado: (2026)
Agentic Artificial Intelligence (AI): Architectures, Taxonomies, and Evaluation of Large Language Model Agents
por: V, Arunkumar, et al.
Publicado: (2026)
por: V, Arunkumar, et al.
Publicado: (2026)
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
por: Luo, Yinyi, et al.
Publicado: (2026)
por: Luo, Yinyi, et al.
Publicado: (2026)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
por: Hu, Shuyue, et al.
Publicado: (2025)
por: Hu, Shuyue, et al.
Publicado: (2025)
AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance
por: Patel, Dhaval, et al.
Publicado: (2025)
por: Patel, Dhaval, et al.
Publicado: (2025)
Coral Protocol: Open Infrastructure Connecting The Internet of Agents
por: Georgio, Roman J., et al.
Publicado: (2025)
por: Georgio, Roman J., et al.
Publicado: (2025)
LOKA Protocol: A Decentralized Framework for Trustworthy and Ethical AI Agent Ecosystems
por: Ranjan, Rajesh, et al.
Publicado: (2025)
por: Ranjan, Rajesh, et al.
Publicado: (2025)
Fairness in Agentic AI: A Unified Framework for Ethical and Equitable Multi-Agent System
por: Ranjan, Rajesh, et al.
Publicado: (2025)
por: Ranjan, Rajesh, et al.
Publicado: (2025)
Games Agents Play: Towards Transactional Analysis in LLM-based Multi-Agent Systems
por: Zamojska, Monika, et al.
Publicado: (2025)
por: Zamojska, Monika, et al.
Publicado: (2025)
Contextual Knowledge Sharing in Multi-Agent Reinforcement Learning with Decentralized Communication and Coordination
por: Du, Hung, et al.
Publicado: (2025)
por: Du, Hung, et al.
Publicado: (2025)
OpenGuanDan: A Large-Scale Imperfect Information Game Benchmark
por: Li, Chao, et al.
Publicado: (2026)
por: Li, Chao, et al.
Publicado: (2026)
Conceptual Logical Foundations of Artificial Social Intelligence
por: Werner, Eric
Publicado: (2025)
por: Werner, Eric
Publicado: (2025)
Foam-Agent: Towards Automated Intelligent CFD Workflows
por: Yue, Ling, et al.
Publicado: (2025)
por: Yue, Ling, et al.
Publicado: (2025)
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
por: Wu, Bin, et al.
Publicado: (2026)
por: Wu, Bin, et al.
Publicado: (2026)
Dynamic Intelligence Assessment: Benchmarking LLMs on the Road to AGI with a Focus on Model Confidence
por: Tihanyi, Norbert, et al.
Publicado: (2024)
por: Tihanyi, Norbert, et al.
Publicado: (2024)
LLM-enabled Social Agents
por: Gürcan, Önder, et al.
Publicado: (2026)
por: Gürcan, Önder, et al.
Publicado: (2026)
Agent Exchange: Shaping the Future of AI Agent Economics
por: Yang, Yingxuan, et al.
Publicado: (2025)
por: Yang, Yingxuan, et al.
Publicado: (2025)
NegotiationGym: Self-Optimizing Agents in a Multi-Agent Social Simulation Environment
por: Mangla, Shashank, et al.
Publicado: (2025)
por: Mangla, Shashank, et al.
Publicado: (2025)
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
por: Shlomov, Segev, et al.
Publicado: (2024)
por: Shlomov, Segev, et al.
Publicado: (2024)
Multi-Agent Intelligence for Multidisciplinary Decision-Making in Gastrointestinal Oncology
por: Zhang, Rongzhao, et al.
Publicado: (2025)
por: Zhang, Rongzhao, et al.
Publicado: (2025)
Domain-driven Metrics for Reinforcement Learning: A Case Study on Epidemic Control using Agent-based Simulation
por: Gaur, Rishabh, et al.
Publicado: (2025)
por: Gaur, Rishabh, et al.
Publicado: (2025)
Free Agent in Agent-Based Mixture-of-Experts Generative AI Framework
por: Liu, Jung-Hua
Publicado: (2025)
por: Liu, Jung-Hua
Publicado: (2025)
Enforcement Agents: Enhancing Accountability and Resilience in Multi-Agent AI Frameworks
por: Tamang, Sagar, et al.
Publicado: (2025)
por: Tamang, Sagar, et al.
Publicado: (2025)
Sentinel Agents for Secure and Trustworthy Agentic AI in Multi-Agent Systems
por: Gosmar, Diego, et al.
Publicado: (2025)
por: Gosmar, Diego, et al.
Publicado: (2025)
CREW-WILDFIRE: Benchmarking Agentic Multi-Agent Collaborations at Scale
por: Hyun, Jonathan, et al.
Publicado: (2025)
por: Hyun, Jonathan, et al.
Publicado: (2025)
TAMAS: Benchmarking Adversarial Risks in Multi-Agent LLM Systems
por: Kavathekar, Ishan, et al.
Publicado: (2025)
por: Kavathekar, Ishan, et al.
Publicado: (2025)
Rethinking Multi-Agent Intelligence Through the Lens of Small-World Networks
por: Wang, Boxuan, et al.
Publicado: (2025)
por: Wang, Boxuan, et al.
Publicado: (2025)
Towards Scientific Intelligence: A Survey of LLM-based Scientific Agents
por: Ren, Shuo, et al.
Publicado: (2025)
por: Ren, Shuo, et al.
Publicado: (2025)
If You Want Coherence, Orchestrate a Team of Rivals: Multi-Agent Models of Organizational Intelligence
por: Vijayaraghavan, Gopal, et al.
Publicado: (2026)
por: Vijayaraghavan, Gopal, et al.
Publicado: (2026)
SODE: Analyzing Social Dynamics in LLM Agents
por: Jung, Inseo, et al.
Publicado: (2026)
por: Jung, Inseo, et al.
Publicado: (2026)
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
por: Shindo, Hikaru, et al.
Publicado: (2026)
por: Shindo, Hikaru, et al.
Publicado: (2026)
CoMMa: Contribution-Aware Medical Multi-Agents From A Game-Theoretic Perspective
por: Wu, Yichen, et al.
Publicado: (2026)
por: Wu, Yichen, et al.
Publicado: (2026)
SocialGFs: Learning Social Gradient Fields for Multi-Agent Reinforcement Learning
por: Long, Qian, et al.
Publicado: (2024)
por: Long, Qian, et al.
Publicado: (2024)
Solving Context Window Overflow in AI Agents
por: Labate, Anton Bulle, et al.
Publicado: (2025)
por: Labate, Anton Bulle, et al.
Publicado: (2025)
EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents
por: Juneja, Gurusha, et al.
Publicado: (2026)
por: Juneja, Gurusha, et al.
Publicado: (2026)
Agentic SPARQL: Evaluating SPARQL-MCP-powered Intelligent Agents on the Federated KGQA Benchmark
por: Dobriy, Daniel, et al.
Publicado: (2026)
por: Dobriy, Daniel, et al.
Publicado: (2026)
Ejemplares similares
-
Agent Island: A Saturation- and Contamination-Resistant Benchmark from Multiagent Games
por: Murphy, Connacher
Publicado: (2026) -
Multi-Actor Generative Artificial Intelligence as a Game Engine
por: Vezhnevets, Alexander Sasha, et al.
Publicado: (2025) -
The Orchestration of Multi-Agent Systems: Architectures, Protocols, and Enterprise Adoption
por: Adimulam, Apoorva, et al.
Publicado: (2026) -
SALM: A Multi-Agent Framework for Language Model-Driven Social Network Simulation
por: Koley, Gaurav
Publicado: (2025) -
Memory Intelligence Agent
por: Qiao, Jingyang, et al.
Publicado: (2026)