For How Long Should We Be Punching? Learning Action Duration in Fighting Games
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Hoang Hai, Driessens, Kurt, Soemers, Dennis J. N. J. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Research Agenda for Usability and Generalisation in Reinforcement Learning
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
Exploring RL-based LLM Training for Formal Language Tasks with Programmed Rewards
by: Padula, Alexander G., et al.
Published: (2024)
by: Padula, Alexander G., et al.
Published: (2024)
Anytime Sequential Halving in Monte-Carlo Tree Search
by: Sagers, Dominic, et al.
Published: (2024)
by: Sagers, Dominic, et al.
Published: (2024)
Best Agent Identification for General Game Playing
by: Stephenson, Matthew, et al.
Published: (2025)
by: Stephenson, Matthew, et al.
Published: (2025)
How Should We Represent History in Interpretable Models of Clinical Policies?
by: Matsson, Anton, et al.
Published: (2024)
by: Matsson, Anton, et al.
Published: (2024)
Towards a Characterisation of Monte-Carlo Tree Search Performance in Different Games
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
by: Soemers, Dennis J. N. J., et al.
Published: (2024)
Should We Ever Prefer Decision Transformer for Offline Reinforcement Learning?
by: Omori, Yumi, et al.
Published: (2025)
by: Omori, Yumi, et al.
Published: (2025)
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
Games of Knightian Uncertainty as AGI testbeds
by: Samothrakis, Spyridon, et al.
Published: (2024)
by: Samothrakis, Spyridon, et al.
Published: (2024)
Unsupervisedly Learned Representations: Should the Quest be Over?
by: Nissani, Daniel N.
Published: (2020)
by: Nissani, Daniel N.
Published: (2020)
ECG-RAMBA: Zero-Shot ECG Generalization by Morphology-Rhythm Disentanglement and Long-Range Modeling
by: Nguyen, Hai Duong, et al.
Published: (2025)
by: Nguyen, Hai Duong, et al.
Published: (2025)
Enhancing Player Enjoyment with a Two-Tier DRL and LLM-Based Agent System for Fighting Games
by: Wang, Shouren, et al.
Published: (2025)
by: Wang, Shouren, et al.
Published: (2025)
We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems
by: Fang, Junfeng, et al.
Published: (2025)
by: Fang, Junfeng, et al.
Published: (2025)
Not Every AI Problem is a Data Problem: We Should Be Intentional About Data Scaling
by: Rodchenko, Tanya, et al.
Published: (2025)
by: Rodchenko, Tanya, et al.
Published: (2025)
Using Synthetic Data to estimate the True Error is theoretically and practically doable
by: Thanh, Hai Hoang, et al.
Published: (2025)
by: Thanh, Hai Hoang, et al.
Published: (2025)
An Advantage-based Optimization Method for Reinforcement Learning in Large Action Space
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Toward Cost-efficient Adaptive Clinical Trials in Knee Osteoarthritis with Reinforcement Learning
by: Nguyen, Khanh, et al.
Published: (2024)
by: Nguyen, Khanh, et al.
Published: (2024)
Variable-Agnostic Causal Exploration for Reinforcement Learning
by: Nguyen, Minh Hoang, et al.
Published: (2024)
by: Nguyen, Minh Hoang, et al.
Published: (2024)
StratFormer: Adaptive Opponent Modeling and Exploitation in Imperfect-Information Games
by: Caen, Andy, et al.
Published: (2026)
by: Caen, Andy, et al.
Published: (2026)
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
How Far Are We from True Unlearnability?
by: Ye, Kai, et al.
Published: (2025)
by: Ye, Kai, et al.
Published: (2025)
Learning from Teaching Regularization: Generalizable Correlations Should be Easy to Imitate
by: Jin, Can, et al.
Published: (2024)
by: Jin, Can, et al.
Published: (2024)
How Transformers Learn In-Context Recall Tasks? Optimality, Training Dynamics and Generalization
by: Nguyen, Quan, et al.
Published: (2025)
by: Nguyen, Quan, et al.
Published: (2025)
Learning Reconfigurable Representations for Multimodal Federated Learning with Missing Data
by: Nguyen, Duong M., et al.
Published: (2025)
by: Nguyen, Duong M., et al.
Published: (2025)
Causal Graph Learning via Distributional Invariance of Cause-Effect Relationship
by: Nguyen, Nang Hung, et al.
Published: (2026)
by: Nguyen, Nang Hung, et al.
Published: (2026)
Ludax: A GPU-Accelerated Domain Specific Language for Board Games
by: Todd, Graham, et al.
Published: (2025)
by: Todd, Graham, et al.
Published: (2025)
Should We Attend More or Less? Modulating Attention for Fairness
by: Zayed, Abdelrahman, et al.
Published: (2023)
by: Zayed, Abdelrahman, et al.
Published: (2023)
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
by: Zhang, Chen, et al.
Published: (2024)
by: Zhang, Chen, et al.
Published: (2024)
Spectral Flattening Is All Muon Needs: How Orthogonalization Controls Learning Rate and Convergence
by: Nguyen, Tien-Phat, et al.
Published: (2026)
by: Nguyen, Tien-Phat, et al.
Published: (2026)
Causal-Aware Generative Adversarial Networks with Reinforcement Learning
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
by: Nguyen, Tu Anh Hoang, et al.
Published: (2025)
Spectral Text Fusion: A Frequency-Aware Approach to Multimodal Time-Series Forecasting
by: Nguyen, Huu Hiep, et al.
Published: (2026)
by: Nguyen, Huu Hiep, et al.
Published: (2026)
Scrutinize What We Ignore: Reining In Task Representation Shift Of Context-Based Offline Meta Reinforcement Learning
by: Zhang, Hai, et al.
Published: (2024)
by: Zhang, Hai, et al.
Published: (2024)
Transformers Learn Robust In-Context Regression under Distributional Uncertainty
by: Cao, Hoang T. H., et al.
Published: (2026)
by: Cao, Hoang T. H., et al.
Published: (2026)
The Ludii Game Description Language is Universal
by: Soemers, Dennis J. N. J., et al.
Published: (2022)
by: Soemers, Dennis J. N. J., et al.
Published: (2022)
How Fast Should a Model Commit to Supervision? Training Reasoning Models on the Tsallis Loss Continuum
by: Lin, Chu-Cheng, et al.
Published: (2026)
by: Lin, Chu-Cheng, et al.
Published: (2026)
When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?
by: Mu, Tongzhou, et al.
Published: (2024)
by: Mu, Tongzhou, et al.
Published: (2024)
Multiple-Input Variational Auto-Encoder for Anomaly Detection in Heterogeneous Data
by: Dinh, Phai Vu, et al.
Published: (2025)
by: Dinh, Phai Vu, et al.
Published: (2025)
A Compact LSTM-SVM Fusion Model for Long-Duration Cardiovascular Diseases Detection
by: Wu, Siyang
Published: (2023)
by: Wu, Siyang
Published: (2023)
When, How Long and How Much? Interpretable Neural Networks for Time Series Regression by Learning to Mask and Aggregate
by: Forest, Florent, et al.
Published: (2025)
by: Forest, Florent, et al.
Published: (2025)
Similar Items
-
A Research Agenda for Usability and Generalisation in Reinforcement Learning
by: Soemers, Dennis J. N. J., et al.
Published: (2024) -
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025) -
Exploring RL-based LLM Training for Formal Language Tasks with Programmed Rewards
by: Padula, Alexander G., et al.
Published: (2024) -
Anytime Sequential Halving in Monte-Carlo Tree Search
by: Sagers, Dominic, et al.
Published: (2024) -
Best Agent Identification for General Game Playing
by: Stephenson, Matthew, et al.
Published: (2025)