SMAC-Hard: Enabling Mixed Opponent Strategy Script and Self-play on SMAC
Fuente:
arXiv
Saved in:
| Main Authors: | Deng, Yue, Yu, Yan, Ma, Weiyu, Wang, Zirui, Zhu, Wenhui, Zhao, Jian, Zhang, Yin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SMAC-R1: The Emergence of Intelligence in Decision-Making Tasks
by: Deng, Yue, et al.
Published: (2024)
by: Deng, Yue, et al.
Published: (2024)
SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer
by: de Lara, Nathan Samuel, et al.
Published: (2026)
by: de Lara, Nathan Samuel, et al.
Published: (2026)
EvoCurr: Self-evolving Curriculum with Behavior Code Generation for Complex Decision-making
by: Cheng, Yang, et al.
Published: (2025)
by: Cheng, Yang, et al.
Published: (2025)
Strategy-Augmented Planning for Large Language Models via Opponent Exploitation
by: Xu, Shuai, et al.
Published: (2025)
by: Xu, Shuai, et al.
Published: (2025)
Efficient Adaptation in Mixed-Motive Environments via Hierarchical Opponent Modeling and Planning
by: Huang, Yizhe, et al.
Published: (2024)
by: Huang, Yizhe, et al.
Published: (2024)
Decision-making with Speculative Opponent Models
by: Sun, Jing, et al.
Published: (2022)
by: Sun, Jing, et al.
Published: (2022)
A Survey on Self-play Methods in Reinforcement Learning
by: Zhang, Ruize, et al.
Published: (2024)
by: Zhang, Ruize, et al.
Published: (2024)
Scaling Opponent Shaping to High Dimensional Games
by: Khan, Akbir, et al.
Published: (2023)
by: Khan, Akbir, et al.
Published: (2023)
Differentiable Belief-based Opponent Shaping
by: Sane, Aarav G, et al.
Published: (2026)
by: Sane, Aarav G, et al.
Published: (2026)
Script Sensitivity: Benchmarking Language Models on Unicode, Romanized and Mixed-Script Sinhala
by: Rajapakse, Minuri, et al.
Published: (2026)
by: Rajapakse, Minuri, et al.
Published: (2026)
The wedge effect in the beam column joint of the Multipurpose Construction System for Cuba (SMAC system)
by: Leonardo Ruiz Alejo
Published: (2015)
by: Leonardo Ruiz Alejo
Published: (2015)
Minibal: Balanced Game-Playing Without Opponent Modeling
by: Cohen-Solal, Quentin, et al.
Published: (2026)
by: Cohen-Solal, Quentin, et al.
Published: (2026)
Absolute Zero: Reinforced Self-play Reasoning with Zero Data
by: Zhao, Andrew, et al.
Published: (2025)
by: Zhao, Andrew, et al.
Published: (2025)
DecisionHoldem: Safe Depth-Limited Solving With Diverse Opponents for Imperfect-Information Games
by: Zhou, Qibin, et al.
Published: (2022)
by: Zhou, Qibin, et al.
Published: (2022)
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
Script-Strategy Aligned Generation: Aligning LLMs with Expert-Crafted Dialogue Scripts and Therapeutic Strategies for Psychotherapy
by: Sun, Xin, et al.
Published: (2024)
by: Sun, Xin, et al.
Published: (2024)
StratFormer: Adaptive Opponent Modeling and Exploitation in Imperfect-Information Games
by: Caen, Andy, et al.
Published: (2026)
by: Caen, Andy, et al.
Published: (2026)
Mastering the Game of Go with Self-play Experience Replay
by: Liu, Jingbin, et al.
Published: (2026)
by: Liu, Jingbin, et al.
Published: (2026)
Analysing the Sample Complexity of Opponent Shaping
by: Fung, Kitty, et al.
Published: (2024)
by: Fung, Kitty, et al.
Published: (2024)
SOM: Structured Opponent Modeling for LLM-based Agents via Structural Causal Model
by: Cao, Shiyue, et al.
Published: (2026)
by: Cao, Shiyue, et al.
Published: (2026)
Opponent Modeling in Multiplayer Imperfect-Information Games
by: Ganzfried, Sam, et al.
Published: (2022)
by: Ganzfried, Sam, et al.
Published: (2022)
Adaptive Opponent Policy Detection in Multi-Agent MDPs: Real-Time Strategy Switch Identification Using Running Error Estimation
by: Mridul, Mohidul Haque, et al.
Published: (2024)
by: Mridul, Mohidul Haque, et al.
Published: (2024)
Improving Rationality in the Reasoning Process of Language Models through Self-playing Game
by: Wang, Pinzheng, et al.
Published: (2025)
by: Wang, Pinzheng, et al.
Published: (2025)
TacticCraft: Natural Language-Driven Tactical Adaptation for StarCraft II
by: Ma, Weiyu, et al.
Published: (2025)
by: Ma, Weiyu, et al.
Published: (2025)
TacEleven: generative tactic discovery for football open play
by: Zhao, Siyao, et al.
Published: (2025)
by: Zhao, Siyao, et al.
Published: (2025)
LOQA: Learning with Opponent Q-Learning Awareness
by: Aghajohari, Milad, et al.
Published: (2024)
by: Aghajohari, Milad, et al.
Published: (2024)
STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving
by: Dong, Kefan, et al.
Published: (2025)
by: Dong, Kefan, et al.
Published: (2025)
Anticipating Oblivious Opponents in Stochastic Games
by: Kalat, Shadi Tasdighi, et al.
Published: (2024)
by: Kalat, Shadi Tasdighi, et al.
Published: (2024)
S^2tory: Story Spine Distillation for Movie Script Summarization
by: Lu, Mingzhe, et al.
Published: (2026)
by: Lu, Mingzhe, et al.
Published: (2026)
2K-Characters-10K-Stories: A Quality-Gated Stylized Narrative Dataset with Disentangled Control and Sequence Consistency
by: Yin, Xingxi, et al.
Published: (2025)
by: Yin, Xingxi, et al.
Published: (2025)
Improving LLM-based Recommendation with Self-Hard Negatives from Intermediate Layers
by: Li, Bingqian, et al.
Published: (2026)
by: Li, Bingqian, et al.
Published: (2026)
IRIS: Interpolative Rényi Iterative Self-play for Large Language Model Fine-Tuning
by: Liao, Wenjie, et al.
Published: (2026)
by: Liao, Wenjie, et al.
Published: (2026)
Pruning Large Language Models with Semi-Structural Adaptive Sparse Training
by: Huang, Weiyu, et al.
Published: (2024)
by: Huang, Weiyu, et al.
Published: (2024)
Narrative-Driven Travel Planning: Geoculturally-Grounded Script Generation with Evolutionary Itinerary Optimization
by: Zhang, Ziyu, et al.
Published: (2025)
by: Zhang, Ziyu, et al.
Published: (2025)
Consistent Opponent Modeling in Imperfect-Information Games
by: Ganzfried, Sam
Published: (2025)
by: Ganzfried, Sam
Published: (2025)
Smart Help: Strategic Opponent Modeling for Proactive and Adaptive Robot Assistance in Households
by: Cao, Zhihao, et al.
Published: (2024)
by: Cao, Zhihao, et al.
Published: (2024)
Plan before Solving: Problem-Aware Strategy Routing for Mathematical Reasoning with LLMs
by: Qi, Shihao, et al.
Published: (2025)
by: Qi, Shihao, et al.
Published: (2025)
Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization Approach
by: Ma, Weiyu, et al.
Published: (2023)
by: Ma, Weiyu, et al.
Published: (2023)
Task-oriented Prompt Enhancement via Script Generation
by: Wang, Chung-Yu, et al.
Published: (2024)
by: Wang, Chung-Yu, et al.
Published: (2024)
AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment
by: Wei, Zhenlin, et al.
Published: (2026)
by: Wei, Zhenlin, et al.
Published: (2026)
Similar Items
-
SMAC-R1: The Emergence of Intelligence in Decision-Making Tasks
by: Deng, Yue, et al.
Published: (2024) -
SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer
by: de Lara, Nathan Samuel, et al.
Published: (2026) -
EvoCurr: Self-evolving Curriculum with Behavior Code Generation for Complex Decision-making
by: Cheng, Yang, et al.
Published: (2025) -
Strategy-Augmented Planning for Large Language Models via Opponent Exploitation
by: Xu, Shuai, et al.
Published: (2025) -
Efficient Adaptation in Mixed-Motive Environments via Hierarchical Opponent Modeling and Planning
by: Huang, Yizhe, et al.
Published: (2024)