Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents
Fuente:
arXiv
Guardado en:
| Autores principales: | Ramesh, Mahesh, Jayakumar, Kaousheik, Ramkumar, Aswinkumar, Thodima, Pavan, Rege, Aniket, Vlatakis-Gkaragkounis, Emmanouil-Vasileios |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
No Coin Left Behind: Maximizing Strategic Surplus Against No-Regret Dynamics
por: Su, Yiheng, et al.
Publicado: (2026)
por: Su, Yiheng, et al.
Publicado: (2026)
MABViT -- Modified Attention Block Enhances Vision Transformers
por: Ramesh, Mahesh, et al.
Publicado: (2023)
por: Ramesh, Mahesh, et al.
Publicado: (2023)
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
por: Patel, Deep, et al.
Publicado: (2025)
por: Patel, Deep, et al.
Publicado: (2025)
Prudent-Banker: No Extra Fees for Baseline Safety in Adversarial Bandits With and Without Delays
por: Hu, Ting, et al.
Publicado: (2026)
por: Hu, Ting, et al.
Publicado: (2026)
Learning Safely Without Knowing the World:COMPASS-Hedge
por: Hu, Ting, et al.
Publicado: (2026)
por: Hu, Ting, et al.
Publicado: (2026)
Breaking $1/ε$ Barrier in Quantum Zero-Sum Games: Generalizing Metric Subregularity for Spectraplexes
por: Su, Yiheng, et al.
Publicado: (2025)
por: Su, Yiheng, et al.
Publicado: (2025)
Building Robust and Scalable Multilingual ASR for Indian Languages
por: Gangwar, Arjun, et al.
Publicado: (2025)
por: Gangwar, Arjun, et al.
Publicado: (2025)
Solving Zero-Sum Convex Markov Games
por: Kalogiannis, Fivos, et al.
Publicado: (2025)
por: Kalogiannis, Fivos, et al.
Publicado: (2025)
Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD
por: Emmanouilidis, Konstantinos, et al.
Publicado: (2026)
por: Emmanouilidis, Konstantinos, et al.
Publicado: (2026)
Algorithms and Complexity for Computing Nash Equilibria in Adversarial Team Games
por: Anagnostides, Ioannis, et al.
Publicado: (2023)
por: Anagnostides, Ioannis, et al.
Publicado: (2023)
Last-Iterate Convergence of Adaptive Riemannian Gradient Descent for Equilibrium Computation
por: Cai, Yang, et al.
Publicado: (2023)
por: Cai, Yang, et al.
Publicado: (2023)
Armada: Memory-Efficient Distributed Training of Large-Scale Graph Neural Networks
por: Waleffe, Roger, et al.
Publicado: (2025)
por: Waleffe, Roger, et al.
Publicado: (2025)
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
por: Vasconcelos, Francisca, et al.
Publicado: (2023)
por: Vasconcelos, Francisca, et al.
Publicado: (2023)
Contracting with a Learning Agent
por: Guruganesh, Guru, et al.
Publicado: (2024)
por: Guruganesh, Guru, et al.
Publicado: (2024)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
por: Liang, Fangzhou, et al.
Publicado: (2025)
por: Liang, Fangzhou, et al.
Publicado: (2025)
RiddleBench: A New Generative Reasoning Benchmark for LLMs
por: Halder, Deepon, et al.
Publicado: (2025)
por: Halder, Deepon, et al.
Publicado: (2025)
Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest
por: O'Neill, Abigail, et al.
Publicado: (2026)
por: O'Neill, Abigail, et al.
Publicado: (2026)
Shakespearean Sparks: The Dance of Hallucination and Creativity in LLMs' Decoding Layers
por: He, Zicong, et al.
Publicado: (2025)
por: He, Zicong, et al.
Publicado: (2025)
Cooperative Strategic Planning Enhances Reasoning Capabilities in Large Language Models
por: Wang, Danqing, et al.
Publicado: (2024)
por: Wang, Danqing, et al.
Publicado: (2024)
Multi-Turn Puzzles: Evaluating Interactive Reasoning and Strategic Dialogue in LLMs
por: Badola, Kartikeya, et al.
Publicado: (2025)
por: Badola, Kartikeya, et al.
Publicado: (2025)
From Building Blocks to Planning: Multi-Step Spatial Reasoning in LLMs with Reinforcement Learning
por: Tahmasbi, Amir, et al.
Publicado: (2025)
por: Tahmasbi, Amir, et al.
Publicado: (2025)
Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning
por: Wu, Jinyang, et al.
Publicado: (2026)
por: Wu, Jinyang, et al.
Publicado: (2026)
EPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via Reinforcement Learning
por: Liu, Xiaoqian, et al.
Publicado: (2025)
por: Liu, Xiaoqian, et al.
Publicado: (2025)
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
por: Halder, Deepon, et al.
Publicado: (2025)
por: Halder, Deepon, et al.
Publicado: (2025)
CryptoLLM: Unleashing the Power of Prompted LLMs for SmartQnA and Classification of Crypto Posts
por: Deroy, Aniket, et al.
Publicado: (2024)
por: Deroy, Aniket, et al.
Publicado: (2024)
Improving Clinical Diagnosis with Counterfactual Multi-Agent Reasoning
por: You, Zhiwen, et al.
Publicado: (2026)
por: You, Zhiwen, et al.
Publicado: (2026)
Enhancing Language Agent Strategic Reasoning through Self-Play in Adversarial Games
por: Zhang, Yikai, et al.
Publicado: (2025)
por: Zhang, Yikai, et al.
Publicado: (2025)
Sparks of Tabular Reasoning via Text2SQL Reinforcement Learning
por: Stoisser, Josefa Lia, et al.
Publicado: (2025)
por: Stoisser, Josefa Lia, et al.
Publicado: (2025)
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents
por: Costarelli, Anthony, et al.
Publicado: (2024)
por: Costarelli, Anthony, et al.
Publicado: (2024)
EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution
por: He, Shiyu, et al.
Publicado: (2026)
por: He, Shiyu, et al.
Publicado: (2026)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
por: Tian, Yuan, et al.
Publicado: (2025)
por: Tian, Yuan, et al.
Publicado: (2025)
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification
por: Deroy, Aniket, et al.
Publicado: (2024)
por: Deroy, Aniket, et al.
Publicado: (2024)
Reinforcement Learning for Hanabi
por: Cohen, Nina, et al.
Publicado: (2025)
por: Cohen, Nina, et al.
Publicado: (2025)
SparkRA: A Retrieval-Augmented Knowledge Service System Based on Spark Large Language Model
por: Wu, Dayong, et al.
Publicado: (2024)
por: Wu, Dayong, et al.
Publicado: (2024)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
por: Liu, Hongtao, et al.
Publicado: (2025)
por: Liu, Hongtao, et al.
Publicado: (2025)
Strategic Chain-of-Thought: Guiding Accurate Reasoning in LLMs through Strategy Elicitation
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos
por: Goel, Arushi, et al.
Publicado: (2026)
por: Goel, Arushi, et al.
Publicado: (2026)
Losses that Cook: Topological Optimal Transport for Structured Recipe Generation
por: Ottoborgo, Mattia, et al.
Publicado: (2026)
por: Ottoborgo, Mattia, et al.
Publicado: (2026)
ADVOSYNTH: A Synthetic Multi-Advocate Dataset for Speaker Identification in Courtroom Scenarios
por: Deroy, Aniket
Publicado: (2026)
por: Deroy, Aniket
Publicado: (2026)
Synthesizing the Virtual Advocate: A Multi-Persona Speech Generation Framework for Diverse Linguistic Jurisdictions in Indic Languages
por: Deroy, Aniket
Publicado: (2026)
por: Deroy, Aniket
Publicado: (2026)
Ejemplares similares
-
No Coin Left Behind: Maximizing Strategic Surplus Against No-Regret Dynamics
por: Su, Yiheng, et al.
Publicado: (2026) -
MABViT -- Modified Attention Block Enhances Vision Transformers
por: Ramesh, Mahesh, et al.
Publicado: (2023) -
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
por: Patel, Deep, et al.
Publicado: (2025) -
Prudent-Banker: No Extra Fees for Baseline Safety in Adversarial Bandits With and Without Delays
por: Hu, Ting, et al.
Publicado: (2026) -
Learning Safely Without Knowing the World:COMPASS-Hedge
por: Hu, Ting, et al.
Publicado: (2026)