LLMs as Agentic Cooperative Players in Multiplayer UNO
Fuente:
arXiv
Guardado en:
| Autores principales: | Matinez, Yago Romano, Roberts, Jesse |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multiplayer Nash Preference Optimization
por: Wu, Fang, et al.
Publicado: (2025)
por: Wu, Fang, et al.
Publicado: (2025)
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
por: Yim, Yauwai, et al.
Publicado: (2024)
por: Yim, Yauwai, et al.
Publicado: (2024)
The Non-Determinism of Small LLMs: Evidence of Low Answer Consistency in Repetition Trials of Standard Multiple-Choice Benchmarks
por: Pinhanez, Claudio, et al.
Publicado: (2025)
por: Pinhanez, Claudio, et al.
Publicado: (2025)
Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models
por: Moore, Kyle, et al.
Publicado: (2025)
por: Moore, Kyle, et al.
Publicado: (2025)
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
por: Zhang, Guibin, et al.
Publicado: (2025)
por: Zhang, Guibin, et al.
Publicado: (2025)
Computer Environments Elicit General Agentic Intelligence in LLMs
por: Cheng, Daixuan, et al.
Publicado: (2026)
por: Cheng, Daixuan, et al.
Publicado: (2026)
Chain of Thought Still Thinks Fast: APriCoT Helps with Thinking Slow
por: Moore, Kyle, et al.
Publicado: (2024)
por: Moore, Kyle, et al.
Publicado: (2024)
Are LLMs complicated ethical dilemma analyzers?
por: Jiashen, et al.
Publicado: (2025)
por: Jiashen, et al.
Publicado: (2025)
GEM: A Gym for Agentic LLMs
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
The Base-Rate Effect on LLM Benchmark Performance: Disambiguating Test-Taking Strategies from Benchmark Performance
por: Moore, Kyle, et al.
Publicado: (2024)
por: Moore, Kyle, et al.
Publicado: (2024)
Large Language Model Recall Uncertainty is Modulated by the Fan Effect
por: Roberts, Jesse, et al.
Publicado: (2024)
por: Roberts, Jesse, et al.
Publicado: (2024)
Player-Driven Emergence in LLM-Driven Game Narrative
por: Peng, Xiangyu, et al.
Publicado: (2024)
por: Peng, Xiangyu, et al.
Publicado: (2024)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
por: Li, Yangning, et al.
Publicado: (2025)
por: Li, Yangning, et al.
Publicado: (2025)
TxGemma: Efficient and Agentic LLMs for Therapeutics
por: Wang, Eric, et al.
Publicado: (2025)
por: Wang, Eric, et al.
Publicado: (2025)
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
por: Cavalin, Paulo, et al.
Publicado: (2025)
por: Cavalin, Paulo, et al.
Publicado: (2025)
Collaborative Quest Completion with LLM-driven Non-Player Characters in Minecraft
por: Rao, Sudha, et al.
Publicado: (2024)
por: Rao, Sudha, et al.
Publicado: (2024)
Agentic Adversarial QA for Improving Domain-Specific LLMs
por: Grari, Vincent, et al.
Publicado: (2026)
por: Grari, Vincent, et al.
Publicado: (2026)
Mitigating Hallucination in Large Language Models (LLMs): An Application-Oriented Survey on RAG, Reasoning, and Agentic Systems
por: Li, Yihan, et al.
Publicado: (2025)
por: Li, Yihan, et al.
Publicado: (2025)
Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning
por: Fan, Wei, et al.
Publicado: (2026)
por: Fan, Wei, et al.
Publicado: (2026)
Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions
por: Zhao, Minda, et al.
Publicado: (2026)
por: Zhao, Minda, et al.
Publicado: (2026)
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset
por: Henkel, Owen, et al.
Publicado: (2023)
por: Henkel, Owen, et al.
Publicado: (2023)
Can "AI" Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs
por: Barone, Mariano, et al.
Publicado: (2026)
por: Barone, Mariano, et al.
Publicado: (2026)
PRISM: Agentic Retrieval with LLMs for Multi-Hop Question Answering
por: Nahid, Md Mahadi Hasan, et al.
Publicado: (2025)
por: Nahid, Md Mahadi Hasan, et al.
Publicado: (2025)
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education
por: Henkel, Owen, et al.
Publicado: (2024)
por: Henkel, Owen, et al.
Publicado: (2024)
Toward Optimal LLM Alignments Using Two-Player Games
por: Zheng, Rui, et al.
Publicado: (2024)
por: Zheng, Rui, et al.
Publicado: (2024)
Targeted Visualization of the Backbone of Encoder LLMs
por: Roberts, Isaac, et al.
Publicado: (2024)
por: Roberts, Isaac, et al.
Publicado: (2024)
Tool Preferences in Agentic LLMs are Unreliable
por: Faghih, Kazem, et al.
Publicado: (2025)
por: Faghih, Kazem, et al.
Publicado: (2025)
SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs
por: Wang, Sijia, et al.
Publicado: (2026)
por: Wang, Sijia, et al.
Publicado: (2026)
Agentic Confidence Calibration
por: Zhang, Jiaxin, et al.
Publicado: (2026)
por: Zhang, Jiaxin, et al.
Publicado: (2026)
Agentic Uncertainty Quantification
por: Zhang, Jiaxin, et al.
Publicado: (2026)
por: Zhang, Jiaxin, et al.
Publicado: (2026)
Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools
por: Wu, Junde, et al.
Publicado: (2025)
por: Wu, Junde, et al.
Publicado: (2025)
AgenticSum: An Agentic Inference-Time Framework for Faithful Clinical Text Summarization
por: Piya, Fahmida Liza, et al.
Publicado: (2026)
por: Piya, Fahmida Liza, et al.
Publicado: (2026)
AgenticMath: Enhancing LLM Reasoning via Agentic-based Math Data Generation
por: Liu, Xianyang, et al.
Publicado: (2025)
por: Liu, Xianyang, et al.
Publicado: (2025)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
por: Duan, Jinhao, et al.
Publicado: (2025)
por: Duan, Jinhao, et al.
Publicado: (2025)
Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework
por: Guan, Zihan, et al.
Publicado: (2026)
por: Guan, Zihan, et al.
Publicado: (2026)
Semantic Invariance in Agentic AI
por: de Zarzà, I., et al.
Publicado: (2026)
por: de Zarzà, I., et al.
Publicado: (2026)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
por: Dai, Hankun, et al.
Publicado: (2025)
por: Dai, Hankun, et al.
Publicado: (2025)
More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration
por: Yadav, Advait, et al.
Publicado: (2026)
por: Yadav, Advait, et al.
Publicado: (2026)
Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking
por: Xu, Qinwu, et al.
Publicado: (2026)
por: Xu, Qinwu, et al.
Publicado: (2026)
PokerGPT: An End-to-End Lightweight Solver for Multi-Player Texas Hold'em via Large Language Model
por: Huang, Chenghao, et al.
Publicado: (2024)
por: Huang, Chenghao, et al.
Publicado: (2024)
Ejemplares similares
-
Multiplayer Nash Preference Optimization
por: Wu, Fang, et al.
Publicado: (2025) -
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
por: Yim, Yauwai, et al.
Publicado: (2024) -
The Non-Determinism of Small LLMs: Evidence of Low Answer Consistency in Repetition Trials of Standard Multiple-Choice Benchmarks
por: Pinhanez, Claudio, et al.
Publicado: (2025) -
Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models
por: Moore, Kyle, et al.
Publicado: (2025) -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
por: Zhang, Guibin, et al.
Publicado: (2025)