The State-Action-Reward-State-Action Algorithm in Spatial Prisoner's Dilemma Game
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Lanyu, Jiang, Dongchun, Guo, Fuqiang, Fu, Mingjian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Perspective Attention Mechanism for Bias-Aware Sequential Recommendation
por: Fu, Mingjian, et al.
Publicado: (2025)
por: Fu, Mingjian, et al.
Publicado: (2025)
Dilution, Diffusion and Symbiosis in Spatial Prisoner's Dilemma with Reinforcement Learning
por: Mangold, Gustavo C., et al.
Publicado: (2025)
por: Mangold, Gustavo C., et al.
Publicado: (2025)
Learning Lifted Action Models From Traces of Incomplete Actions and States
por: Jansen, Niklas, et al.
Publicado: (2025)
por: Jansen, Niklas, et al.
Publicado: (2025)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
por: Jiang, Guanxuan, et al.
Publicado: (2025)
por: Jiang, Guanxuan, et al.
Publicado: (2025)
The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
por: Cuadron, Alejandro, et al.
Publicado: (2025)
por: Cuadron, Alejandro, et al.
Publicado: (2025)
Learning Lifted Action Models from Traces with Minimal Information About Actions and States
por: Gösgens, Jonas, et al.
Publicado: (2026)
por: Gösgens, Jonas, et al.
Publicado: (2026)
Steganography in Game Actions
por: Chang, Ching-Chun, et al.
Publicado: (2024)
por: Chang, Ching-Chun, et al.
Publicado: (2024)
LLMs for High-Frequency Decision-Making: Normalized Action Reward-Guided Consistency Policy Optimization
por: Zhao, Yang, et al.
Publicado: (2026)
por: Zhao, Yang, et al.
Publicado: (2026)
Action-to-Action Flow Matching
por: Jia, Jindou, et al.
Publicado: (2026)
por: Jia, Jindou, et al.
Publicado: (2026)
SAJA: A State-Action Joint Attack Framework on Multi-Agent Deep Reinforcement Learning
por: Guo, Weiqi, et al.
Publicado: (2025)
por: Guo, Weiqi, et al.
Publicado: (2025)
Discovering State Equivalences in UCT Search Trees By Action Pruning
por: Schmöcker, Robin, et al.
Publicado: (2025)
por: Schmöcker, Robin, et al.
Publicado: (2025)
Algorithmic Collective Action with Multiple Collectives
por: Battiloro, Claudio, et al.
Publicado: (2025)
por: Battiloro, Claudio, et al.
Publicado: (2025)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
por: Tiwari, Saket, et al.
Publicado: (2022)
por: Tiwari, Saket, et al.
Publicado: (2022)
State-Action Inpainting Diffuser for Continuous Control with Delay
por: Han, Dongqi, et al.
Publicado: (2026)
por: Han, Dongqi, et al.
Publicado: (2026)
Actor-Critic for Continuous Action Chunks: A Reinforcement Learning Framework for Long-Horizon Robotic Manipulation with Sparse Reward
por: Yang, Jiarui, et al.
Publicado: (2025)
por: Yang, Jiarui, et al.
Publicado: (2025)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
por: Nie, Buqing, et al.
Publicado: (2025)
por: Nie, Buqing, et al.
Publicado: (2025)
Evaluate-as-Action: Self-Evaluated Process Rewards for Retrieval-Augmented Agents
por: Shu, Jiangming, et al.
Publicado: (2026)
por: Shu, Jiangming, et al.
Publicado: (2026)
Android Coach: Improve Online Agentic Training Efficiency with Single State Multiple Actions
por: Gan, Guo, et al.
Publicado: (2026)
por: Gan, Guo, et al.
Publicado: (2026)
State-Novelty Guided Action Persistence in Deep Reinforcement Learning
por: Hu, Jianshu, et al.
Publicado: (2024)
por: Hu, Jianshu, et al.
Publicado: (2024)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
por: Tiwari, Saket, et al.
Publicado: (2025)
por: Tiwari, Saket, et al.
Publicado: (2025)
Planning Domain Model Acquisition from State Traces without Action Parameters
por: Balyo, Tomáš, et al.
Publicado: (2024)
por: Balyo, Tomáš, et al.
Publicado: (2024)
GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents
por: Wu, Xiongbin, et al.
Publicado: (2026)
por: Wu, Xiongbin, et al.
Publicado: (2026)
CombatVLA: An Efficient Vision-Language-Action Model for Combat Tasks in 3D Action Role-Playing Games
por: Chen, Peng, et al.
Publicado: (2025)
por: Chen, Peng, et al.
Publicado: (2025)
NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards
por: Hung, Chia-Yu, et al.
Publicado: (2025)
por: Hung, Chia-Yu, et al.
Publicado: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
por: Qu, Delin, et al.
Publicado: (2025)
por: Qu, Delin, et al.
Publicado: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
por: Zhang, Yizhe, et al.
Publicado: (2025)
por: Zhang, Yizhe, et al.
Publicado: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
GroundAct: Can LLM Agents Ground Actions in Environmental States?
por: Wang, Zixuan, et al.
Publicado: (2025)
por: Wang, Zixuan, et al.
Publicado: (2025)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
por: Daley, Brett, et al.
Publicado: (2025)
por: Daley, Brett, et al.
Publicado: (2025)
Low Dimensional State Representation Learning with Robotics Priors in Continuous Action Spaces
por: Botteghi, Nicolò, et al.
Publicado: (2021)
por: Botteghi, Nicolò, et al.
Publicado: (2021)
ActionParty: Multi-Subject Action Binding in Generative Video Games
por: Pondaven, Alexander, et al.
Publicado: (2026)
por: Pondaven, Alexander, et al.
Publicado: (2026)
From Forecasting to Planning: Policy World Model for Collaborative State-Action Prediction
por: Zhao, Zhida, et al.
Publicado: (2025)
por: Zhao, Zhida, et al.
Publicado: (2025)
Evolution of Rewards for Food and Motor Action by Simulating Birth and Death
por: Kanagawa, Yuji, et al.
Publicado: (2024)
por: Kanagawa, Yuji, et al.
Publicado: (2024)
Beyond World-Frame Action Heads: Motion-Centric Action Frames for Vision-Language-Action Models
por: Yang, Huoren, et al.
Publicado: (2026)
por: Yang, Huoren, et al.
Publicado: (2026)
Reinforcement Learning for AMR Charging Decisions: The Impact of Reward and Action Space Design
por: Bischoff, Janik, et al.
Publicado: (2025)
por: Bischoff, Janik, et al.
Publicado: (2025)
CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations
por: Dang, Xuzhe, et al.
Publicado: (2023)
por: Dang, Xuzhe, et al.
Publicado: (2023)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
por: Kim, Jeonghye, et al.
Publicado: (2025)
por: Kim, Jeonghye, et al.
Publicado: (2025)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
por: Chen, Xiaoyu, et al.
Publicado: (2025)
por: Chen, Xiaoyu, et al.
Publicado: (2025)
Ejemplares similares
-
Multi-Perspective Attention Mechanism for Bias-Aware Sequential Recommendation
por: Fu, Mingjian, et al.
Publicado: (2025) -
Dilution, Diffusion and Symbiosis in Spatial Prisoner's Dilemma with Reinforcement Learning
por: Mangold, Gustavo C., et al.
Publicado: (2025) -
Learning Lifted Action Models From Traces of Incomplete Actions and States
por: Jansen, Niklas, et al.
Publicado: (2025) -
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
por: Jiang, Guanxuan, et al.
Publicado: (2025) -
The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
por: Cuadron, Alejandro, et al.
Publicado: (2025)