Large Language Models as Agents in Two-Player Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yang, Sun, Peng, Li, Hang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
player2vec: A Language Modeling Approach to Understand Player Behavior in Games
von: Wang, Tianze, et al.
Veröffentlicht: (2024)
von: Wang, Tianze, et al.
Veröffentlicht: (2024)
DuoGuard: A Two-Player RL-Driven Framework for Multilingual LLM Guardrails
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
von: Zhou, Runlong, et al.
Veröffentlicht: (2024)
von: Zhou, Runlong, et al.
Veröffentlicht: (2024)
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts
von: Pang, Jing-Cheng, et al.
Veröffentlicht: (2024)
von: Pang, Jing-Cheng, et al.
Veröffentlicht: (2024)
Spatial-Temporal Large Language Model for Traffic Prediction
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
CoBa: Convergence Balancer for Multitask Finetuning of Large Language Models
von: Gong, Zi, et al.
Veröffentlicht: (2024)
von: Gong, Zi, et al.
Veröffentlicht: (2024)
Rethinking Machine Unlearning for Large Language Models
von: Liu, Sijia, et al.
Veröffentlicht: (2024)
von: Liu, Sijia, et al.
Veröffentlicht: (2024)
Robust and Scalable Model Editing for Large Language Models
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
Structured Agent Distillation for Large Language Model
von: Liu, Jun, et al.
Veröffentlicht: (2025)
von: Liu, Jun, et al.
Veröffentlicht: (2025)
Off-Policy Value-Based Reinforcement Learning for Large Language Models
von: Wang, Peng-Yuan, et al.
Veröffentlicht: (2026)
von: Wang, Peng-Yuan, et al.
Veröffentlicht: (2026)
CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning
von: Li, Ran, et al.
Veröffentlicht: (2026)
von: Li, Ran, et al.
Veröffentlicht: (2026)
Massive Activations in Large Language Models
von: Sun, Mingjie, et al.
Veröffentlicht: (2024)
von: Sun, Mingjie, et al.
Veröffentlicht: (2024)
Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial?
von: Li, Wenzhe, et al.
Veröffentlicht: (2025)
von: Li, Wenzhe, et al.
Veröffentlicht: (2025)
Game of LLMs: Discovering Structural Constructs in Activities using Large Language Models
von: Hiremath, Shruthi K., et al.
Veröffentlicht: (2024)
von: Hiremath, Shruthi K., et al.
Veröffentlicht: (2024)
Unveiling and Addressing Pseudo Forgetting in Large Language Models
von: Sun, Huashan, et al.
Veröffentlicht: (2024)
von: Sun, Huashan, et al.
Veröffentlicht: (2024)
Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
Optimizing Large Language Model Training Using FP4 Quantization
von: Wang, Ruizhe, et al.
Veröffentlicht: (2025)
von: Wang, Ruizhe, et al.
Veröffentlicht: (2025)
Discovering Decoupled Functional Modules in Large Language Models
von: Yu, Yanke, et al.
Veröffentlicht: (2026)
von: Yu, Yanke, et al.
Veröffentlicht: (2026)
Multi-Head Attention Is a Multi-Player Game
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2026)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2026)
On the Thinking-Language Modeling Gap in Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
ClinicalAgent: Clinical Trial Multi-Agent System with Large Language Model-based Reasoning
von: Yue, Ling, et al.
Veröffentlicht: (2024)
von: Yue, Ling, et al.
Veröffentlicht: (2024)
SparseEval: Efficient Evaluation of Large Language Models by Sparse Optimization
von: Zhang, Taolin, et al.
Veröffentlicht: (2026)
von: Zhang, Taolin, et al.
Veröffentlicht: (2026)
OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
von: Shao, Wenqi, et al.
Veröffentlicht: (2023)
von: Shao, Wenqi, et al.
Veröffentlicht: (2023)
CEB: Compositional Evaluation Benchmark for Fairness in Large Language Models
von: Wang, Song, et al.
Veröffentlicht: (2024)
von: Wang, Song, et al.
Veröffentlicht: (2024)
Large Language Models for Intent-Driven Session Recommendations
von: Sun, Zhu, et al.
Veröffentlicht: (2023)
von: Sun, Zhu, et al.
Veröffentlicht: (2023)
Bias in Large Language Models: Origin, Evaluation, and Mitigation
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
Controlling Large Language Model with Latent Actions
von: Jia, Chengxing, et al.
Veröffentlicht: (2025)
von: Jia, Chengxing, et al.
Veröffentlicht: (2025)
MUCAR: Benchmarking Multilingual Cross-Modal Ambiguity Resolution for Multimodal Large Language Models
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions
von: Tsai, Chen Feng, et al.
Veröffentlicht: (2023)
von: Tsai, Chen Feng, et al.
Veröffentlicht: (2023)
Concept Bottleneck Large Language Models
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
Min-K%++: Improved Baseline for Detecting Pre-Training Data from Large Language Models
von: Zhang, Jingyang, et al.
Veröffentlicht: (2024)
von: Zhang, Jingyang, et al.
Veröffentlicht: (2024)
KwaiAgents: Generalized Information-seeking Agent System with Large Language Models
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
Crafting Large Language Models for Enhanced Interpretability
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
How Attention Sinks Emerge in Large Language Models: An Interpretability Perspective
von: Peng, Runyu, et al.
Veröffentlicht: (2026)
von: Peng, Runyu, et al.
Veröffentlicht: (2026)
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
Prompting Large Language Models for Clinical Temporal Relation Extraction
von: He, Jianping, et al.
Veröffentlicht: (2024)
von: He, Jianping, et al.
Veröffentlicht: (2024)
Shuttle Between the Instructions and the Parameters of Large Language Models
von: Sun, Wangtao, et al.
Veröffentlicht: (2025)
von: Sun, Wangtao, et al.
Veröffentlicht: (2025)
Test-Time Training on Nearest Neighbors for Large Language Models
von: Hardt, Moritz, et al.
Veröffentlicht: (2023)
von: Hardt, Moritz, et al.
Veröffentlicht: (2023)
LCQ: Low-Rank Codebook based Quantization for Large Language Models
von: Cai, Wen-Pu, et al.
Veröffentlicht: (2024)
von: Cai, Wen-Pu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
player2vec: A Language Modeling Approach to Understand Player Behavior in Games
von: Wang, Tianze, et al.
Veröffentlicht: (2024) -
DuoGuard: A Two-Player RL-Driven Framework for Multilingual LLM Guardrails
von: Deng, Yihe, et al.
Veröffentlicht: (2025) -
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
von: Zhou, Runlong, et al.
Veröffentlicht: (2024) -
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts
von: Pang, Jing-Cheng, et al.
Veröffentlicht: (2024) -
Spatial-Temporal Large Language Model for Traffic Prediction
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)