Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization
Fuente:
arXiv
Salvato in:
| Autori principali: | Liao, Junyi, Zhu, Zihan, Fang, Ethan, Yang, Zhuoran, Tarokh, Vahid |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models
di: Tan, Chee Wei, et al.
Pubblicazione: (2026)
di: Tan, Chee Wei, et al.
Pubblicazione: (2026)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
di: Zeytuncu, Yunus E.
Pubblicazione: (2026)
di: Zeytuncu, Yunus E.
Pubblicazione: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
di: Hong, Yoosung
Pubblicazione: (2026)
di: Hong, Yoosung
Pubblicazione: (2026)
SUN Team's Contribution to ABAW 2024 Competition: Audio-visual Valence-Arousal Estimation and Expression Recognition
di: Dresvyanskiy, Denis, et al.
Pubblicazione: (2024)
di: Dresvyanskiy, Denis, et al.
Pubblicazione: (2024)
Shrinkage Initialization for Smooth Learning of Neural Networks
di: Cheng, Miao, et al.
Pubblicazione: (2025)
di: Cheng, Miao, et al.
Pubblicazione: (2025)
Amortized Molecular Optimization via Group Relative Policy Optimization
di: Javaid, Muhammad bin, et al.
Pubblicazione: (2026)
di: Javaid, Muhammad bin, et al.
Pubblicazione: (2026)
Kolmogorov Arnold Networks and Multi-Layer Perceptrons: A Paradigm Shift in Neural Modelling
di: Gaonkar, Aradhya, et al.
Pubblicazione: (2026)
di: Gaonkar, Aradhya, et al.
Pubblicazione: (2026)
Distributional Reinforcement Learning for Condition-Based Maintenance of Multi-Pump Equipment
di: Yasuno, Takato
Pubblicazione: (2026)
di: Yasuno, Takato
Pubblicazione: (2026)
SQARL: A Size-Agnostic Reinforcement Learning approach for Circuit Allocation in Distributed Quantum Architectures
di: Carballo, Víctor, et al.
Pubblicazione: (2026)
di: Carballo, Víctor, et al.
Pubblicazione: (2026)
Learning Controllable and Diverse Player Behaviors in Multi-Agent Environments
di: Cilan, Atahan, et al.
Pubblicazione: (2025)
di: Cilan, Atahan, et al.
Pubblicazione: (2025)
Multi-Agent Pathfinding with Non-Unit Integer Edge Costs via Enhanced Conflict-Based Search and Graph Discretization
di: Fan, Hongkai, et al.
Pubblicazione: (2026)
di: Fan, Hongkai, et al.
Pubblicazione: (2026)
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
di: Silue, Bram, et al.
Pubblicazione: (2025)
di: Silue, Bram, et al.
Pubblicazione: (2025)
SRLAgent: Enhancing Self-Regulated Learning Skills through Gamification and LLM Assistance
di: Ge, Wentao, et al.
Pubblicazione: (2025)
di: Ge, Wentao, et al.
Pubblicazione: (2025)
Imitation learning for sim-to-real transfer of robotic cutting policies based on residual Gaussian process disturbance force model
di: Hathaway, Jamie, et al.
Pubblicazione: (2023)
di: Hathaway, Jamie, et al.
Pubblicazione: (2023)
End-to-end example-based sim-to-real RL policy transfer based on neural stylisation with application to robotic cutting
di: Hathaway, Jamie, et al.
Pubblicazione: (2026)
di: Hathaway, Jamie, et al.
Pubblicazione: (2026)
Augmenting deep neural networks with symbolic knowledge: Towards trustworthy and interpretable AI for education
di: Hooshyar, Danial, et al.
Pubblicazione: (2023)
di: Hooshyar, Danial, et al.
Pubblicazione: (2023)
Perfecting Aircraft Maneuvers with Reinforcement Learning
di: Cilan, Atahan, et al.
Pubblicazione: (2026)
di: Cilan, Atahan, et al.
Pubblicazione: (2026)
TelePlanNet: An AI-Driven Framework for Efficient Telecom Network Planning
di: Deng, Zongyuan, et al.
Pubblicazione: (2025)
di: Deng, Zongyuan, et al.
Pubblicazione: (2025)
BandiK: Efficient Multi-Task Decomposition Using a Multi-Bandit Framework
di: Millinghoffer, András, et al.
Pubblicazione: (2025)
di: Millinghoffer, András, et al.
Pubblicazione: (2025)
Semi-overlapping Multi-bandit Best Arm Identification for Sequential Support Network Learning
di: Antos, András, et al.
Pubblicazione: (2025)
di: Antos, András, et al.
Pubblicazione: (2025)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
A survey of air combat behavior modeling using machine learning
di: Gorton, Patrick Ribu, et al.
Pubblicazione: (2024)
di: Gorton, Patrick Ribu, et al.
Pubblicazione: (2024)
Model Fusion via Retrofitting
di: Luenam, Phoomraphee, et al.
Pubblicazione: (2025)
di: Luenam, Phoomraphee, et al.
Pubblicazione: (2025)
End-to-end deep learning-based framework for path planning and collision checking: bin picking application
di: Tamizi, Mehran Ghafarian, et al.
Pubblicazione: (2023)
di: Tamizi, Mehran Ghafarian, et al.
Pubblicazione: (2023)
Application of Sensitivity Analysis Methods for Studying Neural Network Models
di: Miao, Jiaxuan, et al.
Pubblicazione: (2025)
di: Miao, Jiaxuan, et al.
Pubblicazione: (2025)
N-Agent Ad Hoc Teamwork
di: Wang, Caroline, et al.
Pubblicazione: (2024)
di: Wang, Caroline, et al.
Pubblicazione: (2024)
Pseudoconvex Problems in Operational Decision Systems: Algorithms for Joint Learning and Optimization
di: Li, Zijun, et al.
Pubblicazione: (2026)
di: Li, Zijun, et al.
Pubblicazione: (2026)
AI Agents for the Dhumbal Card Game: A Comparative Study
di: Malla, Sahaj Raj
Pubblicazione: (2025)
di: Malla, Sahaj Raj
Pubblicazione: (2025)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
di: Sauter, Andreas, et al.
Pubblicazione: (2026)
di: Sauter, Andreas, et al.
Pubblicazione: (2026)
Leveraging Large Language Models to Extract and Translate Medical Information in Doctors' Notes for Health Records and Diagnostic Billing Codes
di: Hartnett, Peter, et al.
Pubblicazione: (2026)
di: Hartnett, Peter, et al.
Pubblicazione: (2026)
The Arrival of AGI? When Expert Personas Exceed Expert Benchmarks
di: Mullens, Drake, et al.
Pubblicazione: (2026)
di: Mullens, Drake, et al.
Pubblicazione: (2026)
Compositional Concept-Based Neuron-Level Interpretability for Deep Reinforcement Learning
di: Jiang, Zeyu, et al.
Pubblicazione: (2025)
di: Jiang, Zeyu, et al.
Pubblicazione: (2025)
Hierarchical Pooling and Explainability in Graph Neural Networks for Tumor and Tissue-of-Origin Classification Using RNA-seq Data
di: Fontanari, Thomas Vaitses, et al.
Pubblicazione: (2026)
di: Fontanari, Thomas Vaitses, et al.
Pubblicazione: (2026)
Unsupervised Discovery of Clinical Disease Signatures Using Probabilistic Independence
di: Lasko, Thomas A., et al.
Pubblicazione: (2024)
di: Lasko, Thomas A., et al.
Pubblicazione: (2024)
Atom dimension adaptation for infinite set dictionary learning
di: Băltoiu, Andra, et al.
Pubblicazione: (2024)
di: Băltoiu, Andra, et al.
Pubblicazione: (2024)
Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting
di: Lettner, Jan Niklas, et al.
Pubblicazione: (2026)
di: Lettner, Jan Niklas, et al.
Pubblicazione: (2026)
Fixing the Double Penalty in Data-Driven Weather Forecasting Through a Modified Spherical Harmonic Loss Function
di: Subich, Christopher, et al.
Pubblicazione: (2025)
di: Subich, Christopher, et al.
Pubblicazione: (2025)
Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning
di: Kim, Kyuhee, et al.
Pubblicazione: (2026)
di: Kim, Kyuhee, et al.
Pubblicazione: (2026)
When Reasoning Fails: Evaluating 'Thinking' LLMs for Stock Prediction
di: Sodha, Rakeshkumar H
Pubblicazione: (2025)
di: Sodha, Rakeshkumar H
Pubblicazione: (2025)
On the Compatibility of Generative AI and Generative Linguistics
di: Portelance, Eva, et al.
Pubblicazione: (2024)
di: Portelance, Eva, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models
di: Tan, Chee Wei, et al.
Pubblicazione: (2026) -
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
di: Zeytuncu, Yunus E.
Pubblicazione: (2026) -
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
di: Hong, Yoosung
Pubblicazione: (2026) -
SUN Team's Contribution to ABAW 2024 Competition: Audio-visual Valence-Arousal Estimation and Expression Recognition
di: Dresvyanskiy, Denis, et al.
Pubblicazione: (2024) -
Shrinkage Initialization for Smooth Learning of Neural Networks
di: Cheng, Miao, et al.
Pubblicazione: (2025)