Large Language Models as Pokémon Battle Agents: Strategic Play and Content Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Jain, Daksh, Jain, Aarya, Desai, Ashutosh, Verma, Avyakt, Bhanuka, Ishan, Narang, Pratik, Kumar, Dhruv |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Lights, Camera, Consistency: A Multistage Pipeline for Character-Stable AI Video Stories
por: Jain, Chayan, et al.
Publicado: (2025)
por: Jain, Chayan, et al.
Publicado: (2025)
PokeLLMon: A Human-Parity Agent for Pokemon Battles with Large Language Models
por: Hu, Sihao, et al.
Publicado: (2024)
por: Hu, Sihao, et al.
Publicado: (2024)
Beyond Sentiment: A Multi-Agent Pipeline for Actionable Business Advice from Reviews
por: Bhandari, Kartikey Singh, et al.
Publicado: (2026)
por: Bhandari, Kartikey Singh, et al.
Publicado: (2026)
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
por: Mishra, Shubham, et al.
Publicado: (2025)
por: Mishra, Shubham, et al.
Publicado: (2025)
A Hybrid Supervised-LLM Pipeline for Actionable Suggestion Mining in Unstructured Customer Reviews
por: Trivedi, Aakash, et al.
Publicado: (2026)
por: Trivedi, Aakash, et al.
Publicado: (2026)
Generation Z's Ability to Discriminate Between AI-generated and Human-Authored Text on Discord
por: Ramu, Dhruv, et al.
Publicado: (2023)
por: Ramu, Dhruv, et al.
Publicado: (2023)
SwissNYF: Tool Grounded LLM Agents for Black Box Setting
por: Kumar, Somnath Sendhil, et al.
Publicado: (2024)
por: Kumar, Somnath Sendhil, et al.
Publicado: (2024)
IndicMMLU-Pro: Benchmarking Indic Large Language Models on Multi-Task Language Understanding
por: KJ, Sankalp, et al.
Publicado: (2025)
por: KJ, Sankalp, et al.
Publicado: (2025)
Investigating Spatial Attention Bias in Vision-Language Models
por: Chaudhary, Aryan, et al.
Publicado: (2025)
por: Chaudhary, Aryan, et al.
Publicado: (2025)
Domain-Partitioned Hybrid RAG for Legal Reasoning: Toward Modular and Explainable Legal AI for India
por: Goel, Rakshita, et al.
Publicado: (2025)
por: Goel, Rakshita, et al.
Publicado: (2025)
LUDOBENCH: Evaluating LLM Behavioural Decision-Making Through Spot-Based Board Game Scenarios in Ludo
por: Jain, Ojas, et al.
Publicado: (2026)
por: Jain, Ojas, et al.
Publicado: (2026)
Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
por: Budagam, Devichand, et al.
Publicado: (2024)
por: Budagam, Devichand, et al.
Publicado: (2024)
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs
por: Kumar, Divyanshu, et al.
Publicado: (2024)
por: Kumar, Divyanshu, et al.
Publicado: (2024)
A Multi-Agent Pokemon Tournament for Evaluating Strategic Reasoning of Large Language Models
por: Yashwanth, Tadisetty Sai, et al.
Publicado: (2025)
por: Yashwanth, Tadisetty Sai, et al.
Publicado: (2025)
ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data
por: Shen, Junhong, et al.
Publicado: (2024)
por: Shen, Junhong, et al.
Publicado: (2024)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
por: Khoshnoodi, Mahsa, et al.
Publicado: (2024)
por: Khoshnoodi, Mahsa, et al.
Publicado: (2024)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
por: Sinha, Neelabh, et al.
Publicado: (2024)
por: Sinha, Neelabh, et al.
Publicado: (2024)
Hallucination is Inevitable: An Innate Limitation of Large Language Models
por: Xu, Ziwei, et al.
Publicado: (2024)
por: Xu, Ziwei, et al.
Publicado: (2024)
UNITYAI-GUARD: Pioneering Toxicity Detection Across Low-Resource Indian Languages
por: Beniwal, Himanshu, et al.
Publicado: (2025)
por: Beniwal, Himanshu, et al.
Publicado: (2025)
ALAS: Autonomous Learning Agent for Self-Updating Language Models
por: Atreja, Dhruv
Publicado: (2025)
por: Atreja, Dhruv
Publicado: (2025)
Strategic Prompting for Conversational Tasks: A Comparative Analysis of Large Language Models Across Diverse Conversational Tasks
por: Joshi, Ratnesh Kumar, et al.
Publicado: (2024)
por: Joshi, Ratnesh Kumar, et al.
Publicado: (2024)
Agent-Centric Projection of Prompting Techniques and Implications for Synthetic Training Data for Large Language Models
por: Dhamani, Dhruv, et al.
Publicado: (2025)
por: Dhamani, Dhruv, et al.
Publicado: (2025)
LLM-Gomoku: A Large Language Model-Based System for Strategic Gomoku with Self-Play and Reinforcement Learning
por: Wang, Hui
Publicado: (2025)
por: Wang, Hui
Publicado: (2025)
Plug-and-Play Policy Planner for Large Language Model Powered Dialogue Agents
por: Deng, Yang, et al.
Publicado: (2023)
por: Deng, Yang, et al.
Publicado: (2023)
VoiceAgentBench: Are Voice Assistants ready for agentic tasks?
por: Jain, Dhruv, et al.
Publicado: (2025)
por: Jain, Dhruv, et al.
Publicado: (2025)
Measuring Representation Robustness in Large Language Models for Geometry
por: Jawandhia, Vedant, et al.
Publicado: (2026)
por: Jawandhia, Vedant, et al.
Publicado: (2026)
SoK: Measuring What Matters for Closed-Loop Security Agents
por: Khurana, Mudita, et al.
Publicado: (2025)
por: Khurana, Mudita, et al.
Publicado: (2025)
Dissecting Persona-Driven Reasoning in Language Models via Activation Patching
por: Poonia, Ansh, et al.
Publicado: (2025)
por: Poonia, Ansh, et al.
Publicado: (2025)
Fourier Head: Helping Large Language Models Learn Complex Probability Distributions
por: Gillman, Nate, et al.
Publicado: (2024)
por: Gillman, Nate, et al.
Publicado: (2024)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
por: Singh, Smriti, et al.
Publicado: (2024)
por: Singh, Smriti, et al.
Publicado: (2024)
Agribot: agriculture-specific question answer system
por: Jain, Naman, et al.
Publicado: (2025)
por: Jain, Naman, et al.
Publicado: (2025)
Role-Playing Evaluation for Large Language Models
por: Boudouri, Yassine El, et al.
Publicado: (2025)
por: Boudouri, Yassine El, et al.
Publicado: (2025)
$\forall$uto$\exists$val: Autonomous Assessment of LLMs in Formal Synthesis and Interpretation Tasks
por: Karia, Rushang, et al.
Publicado: (2024)
por: Karia, Rushang, et al.
Publicado: (2024)
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
por: Sahoo, Pranab, et al.
Publicado: (2024)
por: Sahoo, Pranab, et al.
Publicado: (2024)
From Sycophancy to Sensemaking: Premise Governance for Human-AI Decision Making
por: Jain, Raunak
Publicado: (2026)
por: Jain, Raunak
Publicado: (2026)
$x$ Plays Pokemon, for Almost-Every $x$
por: Hedges, C. Evans
Publicado: (2025)
por: Hedges, C. Evans
Publicado: (2025)
PokéAI: A Goal-Generating, Battle-Optimizing Multi-agent System for Pokemon Red
por: Liu, Zihao, et al.
Publicado: (2025)
por: Liu, Zihao, et al.
Publicado: (2025)
Analyzing the Effectiveness of Large Language Models on Text-to-SQL Synthesis
por: Roberson, Richard, et al.
Publicado: (2024)
por: Roberson, Richard, et al.
Publicado: (2024)
CHATATC: Large Language Model-Driven Conversational Agents for Supporting Strategic Air Traffic Flow Management
por: Abdulhak, Sinan, et al.
Publicado: (2024)
por: Abdulhak, Sinan, et al.
Publicado: (2024)
Hypergame Rationalisability: Solving Agent Misalignment In Strategic Play
por: Trencsenyi, Vince
Publicado: (2025)
por: Trencsenyi, Vince
Publicado: (2025)
Ejemplares similares
-
Lights, Camera, Consistency: A Multistage Pipeline for Character-Stable AI Video Stories
por: Jain, Chayan, et al.
Publicado: (2025) -
PokeLLMon: A Human-Parity Agent for Pokemon Battles with Large Language Models
por: Hu, Sihao, et al.
Publicado: (2024) -
Beyond Sentiment: A Multi-Agent Pipeline for Actionable Business Advice from Reviews
por: Bhandari, Kartikey Singh, et al.
Publicado: (2026) -
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
por: Mishra, Shubham, et al.
Publicado: (2025) -
A Hybrid Supervised-LLM Pipeline for Actionable Suggestion Mining in Unstructured Customer Reviews
por: Trivedi, Aakash, et al.
Publicado: (2026)