Reasoning, Memorization, and Fine-Tuning Language Models for Non-Cooperative Games
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yunhao, Berthellemy, Leonard, Topcu, Ufuk |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human-Agent Cooperation in Games under Incomplete Information through Natural Language Communication
by: Chen, Shenghui, et al.
Published: (2024)
by: Chen, Shenghui, et al.
Published: (2024)
Fine-Tuning Language Models Using Formal Methods Feedback
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Memorization in Fine-Tuned Large Language Models
by: Savine, Danil
Published: (2025)
by: Savine, Danil
Published: (2025)
Joint Verification and Refinement of Language Models for Safety-Constrained Planning
by: Yang, Yunhao, et al.
Published: (2024)
by: Yang, Yunhao, et al.
Published: (2024)
IG-MCTS: Human-in-the-Loop Cooperative Navigation under Incomplete Information
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
by: Hou, Jie, et al.
Published: (2025)
by: Hou, Jie, et al.
Published: (2025)
Unintended Memorization of Sensitive Information in Fine-Tuned Language Models
by: Szep, Marton, et al.
Published: (2026)
by: Szep, Marton, et al.
Published: (2026)
Human-Agent Coordination in Games under Incomplete Information via Multi-Step Intent
by: Chen, Shenghui, et al.
Published: (2024)
by: Chen, Shenghui, et al.
Published: (2024)
Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
Assessing and Mitigating Data Memorization Risks in Fine-Tuned Large Language Models
by: Ramakrishnan, Badrinath, et al.
Published: (2025)
by: Ramakrishnan, Badrinath, et al.
Published: (2025)
Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game
by: Karabag, Mustafa O., et al.
Published: (2025)
by: Karabag, Mustafa O., et al.
Published: (2025)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
by: Bhatt, Neel P., et al.
Published: (2024)
by: Bhatt, Neel P., et al.
Published: (2024)
Foundation Models for Logistics: Toward Certifiable, Conversational Planning Interfaces
by: Yang, Yunhao, et al.
Published: (2025)
by: Yang, Yunhao, et al.
Published: (2025)
Zero-Shot Reinforcement Learning via Function Encoders
by: Ingebrand, Tyler, et al.
Published: (2024)
by: Ingebrand, Tyler, et al.
Published: (2024)
Large Language Models Reasoning Abilities Under Non-Ideal Conditions After RL-Fine-Tuning
by: Tian, Chang, et al.
Published: (2025)
by: Tian, Chang, et al.
Published: (2025)
Exploring Memorization in Fine-tuned Language Models
by: Zeng, Shenglai, et al.
Published: (2023)
by: Zeng, Shenglai, et al.
Published: (2023)
Neural Port-Hamiltonian Differential Algebraic Equations for Compositional Learning of Electrical Networks
by: Neary, Cyrus, et al.
Published: (2024)
by: Neary, Cyrus, et al.
Published: (2024)
RepV: Safety-Separable Latent Spaces for Scalable Neurosymbolic Plan Verification
by: Yang, Yunhao, et al.
Published: (2025)
by: Yang, Yunhao, et al.
Published: (2025)
VLN-Zero: Rapid Exploration and Cache-Enabled Neurosymbolic Vision-Language Planning for Zero-Shot Transfer in Robot Navigation
by: Bhatt, Neel P., et al.
Published: (2025)
by: Bhatt, Neel P., et al.
Published: (2025)
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
by: Bovy, Eline M., et al.
Published: (2025)
by: Bovy, Eline M., et al.
Published: (2025)
When Should a Leader Act Suboptimally? The Role of Inferability in Repeated Stackelberg Games
by: Karabag, Mustafa O., et al.
Published: (2023)
by: Karabag, Mustafa O., et al.
Published: (2023)
Learning to Coordinate without Communication under Incomplete Information
by: Chen, Shenghui, et al.
Published: (2024)
by: Chen, Shenghui, et al.
Published: (2024)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
by: Kwon, Minjae, et al.
Published: (2025)
by: Kwon, Minjae, et al.
Published: (2025)
Sequential Resource Trading Using Comparison-Based Gradient Estimation
by: Murthy, Surya, et al.
Published: (2024)
by: Murthy, Surya, et al.
Published: (2024)
Using Large Language Models to Automate and Expedite Reinforcement Learning with Reward Machine
by: Alsadat, Shayan Meshkat, et al.
Published: (2024)
by: Alsadat, Shayan Meshkat, et al.
Published: (2024)
Unveiling the Impact of Coding Data Instruction Fine-Tuning on Large Language Models Reasoning
by: Zhang, Xinlu, et al.
Published: (2024)
by: Zhang, Xinlu, et al.
Published: (2024)
Categorical semantics of compositional reinforcement learning
by: Bakirtzis, Georgios, et al.
Published: (2022)
by: Bakirtzis, Georgios, et al.
Published: (2022)
Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems
by: Sharifi, Iman, et al.
Published: (2026)
by: Sharifi, Iman, et al.
Published: (2026)
The Reasoning-Memorization Interplay in Language Models Is Mediated by a Single Direction
by: Hong, Yihuai, et al.
Published: (2025)
by: Hong, Yihuai, et al.
Published: (2025)
Why Do LLMs Struggle in Strategic Play? Broken Links Between Observations, Beliefs, and Actions
by: Sobotka, Jan, et al.
Published: (2026)
by: Sobotka, Jan, et al.
Published: (2026)
Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning of Vision Language Models
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
GRIP: In-Parameter Graph Reasoning through Fine-Tuning Large Language Models
by: Feng, Jiarui, et al.
Published: (2025)
by: Feng, Jiarui, et al.
Published: (2025)
Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing
by: Zongo, Alex, et al.
Published: (2026)
by: Zongo, Alex, et al.
Published: (2026)
Mitigating Memorization In Language Models
by: Sakarvadia, Mansi, et al.
Published: (2024)
by: Sakarvadia, Mansi, et al.
Published: (2024)
Online Foundation Model Selection in Robotics
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
On-Policy Supervised Fine-Tuning for Efficient Reasoning
by: Zhao, Anhao, et al.
Published: (2026)
by: Zhao, Anhao, et al.
Published: (2026)
Pramana: Fine-Tuning Large Language Models for Epistemic Reasoning through Navya-Nyaya
by: Sathish, Sharath
Published: (2026)
by: Sathish, Sharath
Published: (2026)
Think in Games: Learning to Reason in Games via Reinforcement Learning with Large Language Models
by: Liao, Yi, et al.
Published: (2025)
by: Liao, Yi, et al.
Published: (2025)
Enhance Reasoning for Large Language Models in the Game Werewolf
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
Similar Items
-
Human-Agent Cooperation in Games under Incomplete Information through Natural Language Communication
by: Chen, Shenghui, et al.
Published: (2024) -
Fine-Tuning Language Models Using Formal Methods Feedback
by: Yang, Yunhao, et al.
Published: (2023) -
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023) -
Memorization in Fine-Tuned Large Language Models
by: Savine, Danil
Published: (2025) -
Joint Verification and Refinement of Language Models for Safety-Constrained Planning
by: Yang, Yunhao, et al.
Published: (2024)