Multi-State TD Target for Model-Free Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Wuhao, Chen, Zhiyong, Zhang, Lepeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
by: Klačan, Ján, et al.
Published: (2026)
by: Klačan, Ján, et al.
Published: (2026)
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025)
by: Varricchione, Giovanni, et al.
Published: (2025)
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability
by: Wang, Wuhao, et al.
Published: (2025)
by: Wang, Wuhao, et al.
Published: (2025)
Symbolic Branch Networks: Tree-Inherited Neural Models for Interpretable Multiclass Classification
by: Rodríguez-Salas, Dalia
Published: (2025)
by: Rodríguez-Salas, Dalia
Published: (2025)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025)
by: Yilmaz, Berk, et al.
Published: (2025)
A Framework for Scalable Heterogeneous Multi-Agent Adversarial Reinforcement Learning in IsaacLab
by: Peterson, Isaac, et al.
Published: (2025)
by: Peterson, Isaac, et al.
Published: (2025)
Clinical Data Goes MEDS? Let's OWL make sense of it
by: Marfoglia, Alberto, et al.
Published: (2026)
by: Marfoglia, Alberto, et al.
Published: (2026)
Unsupervised Ensemble Learning Through Deep Energy-based Models
by: Maymon, Ariel, et al.
Published: (2026)
by: Maymon, Ariel, et al.
Published: (2026)
Topological Foundations of Reinforcement Learning
by: Kadurha, David Krame
Published: (2024)
by: Kadurha, David Krame
Published: (2024)
Assessing Long-Term Electricity Market Design for Ambitious Decarbonization Targets using Multi-Agent Reinforcement Learning
by: Gonzalez-Ruiz, Javier, et al.
Published: (2025)
by: Gonzalez-Ruiz, Javier, et al.
Published: (2025)
BatteryML:An Open-source platform for Machine Learning on Battery Degradation
by: Zhang, Han, et al.
Published: (2023)
by: Zhang, Han, et al.
Published: (2023)
Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation
by: Osman, Asim, et al.
Published: (2026)
by: Osman, Asim, et al.
Published: (2026)
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
by: Gao, Weiguo, et al.
Published: (2025)
by: Gao, Weiguo, et al.
Published: (2025)
Grouped Sequential Optimization Strategy -- the Application of Hyperparameter Importance Assessment in Deep Learning
by: Wang, Ruinan, et al.
Published: (2025)
by: Wang, Ruinan, et al.
Published: (2025)
Large Language Model Meets Graph Neural Network in Knowledge Distillation
by: Hu, Shengxiang, et al.
Published: (2024)
by: Hu, Shengxiang, et al.
Published: (2024)
Reinforcement Learning with Action-Triggered Observations
by: Ryabchenko, Alexander, et al.
Published: (2025)
by: Ryabchenko, Alexander, et al.
Published: (2025)
Comparing Deep Reinforcement Learning Algorithms in Two-Echelon Supply Chains
by: Stranieri, Francesco, et al.
Published: (2022)
by: Stranieri, Francesco, et al.
Published: (2022)
FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
Improving Credit Card Fraud Detection with an Optimized Explainable Boosting Machine
by: Fazel, Reza E., et al.
Published: (2026)
by: Fazel, Reza E., et al.
Published: (2026)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
by: Park, Junseok, et al.
Published: (2025)
by: Park, Junseok, et al.
Published: (2025)
Graph Neural Networks for Brain Graph Learning: A Survey
by: Luo, Xuexiong, et al.
Published: (2024)
by: Luo, Xuexiong, et al.
Published: (2024)
Reciprocal Learning
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
Technical Report for the Forgotten-by-Design Project: Targeted Obfuscation for Machine Learning
by: Brännvall, Rickard, et al.
Published: (2025)
by: Brännvall, Rickard, et al.
Published: (2025)
Generation of Geodesics with Actor-Critic Reinforcement Learning to Predict Midpoints
by: Kasaura, Kazumi
Published: (2024)
by: Kasaura, Kazumi
Published: (2024)
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024)
by: Tyukin, Ivan Y., et al.
Published: (2024)
When Do Credal Sets Stabilize? Fixed-Point Theorems for Credal Set Updates
by: Caprio, Michele, et al.
Published: (2025)
by: Caprio, Michele, et al.
Published: (2025)
Computational Hardness of Reinforcement Learning with Partial $q^π$-Realizability
by: Karimi, Shayan, et al.
Published: (2025)
by: Karimi, Shayan, et al.
Published: (2025)
Credal Bayesian Deep Learning
by: Caprio, Michele, et al.
Published: (2023)
by: Caprio, Michele, et al.
Published: (2023)
Maximally Permissive Reward Machines
by: Varricchione, Giovanni, et al.
Published: (2024)
by: Varricchione, Giovanni, et al.
Published: (2024)
DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
by: Perkins, Daniel, et al.
Published: (2025)
by: Perkins, Daniel, et al.
Published: (2025)
Subset Selection for Fine-Tuning: A Utility-Diversity Balanced Approach for Mathematical Domain Adaptation
by: Kotecha, Madhav, et al.
Published: (2025)
by: Kotecha, Madhav, et al.
Published: (2025)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
CircuitBuilder: From Polynomials to Circuits via Reinforcement Learning
by: Zhang, Weikun K., et al.
Published: (2026)
by: Zhang, Weikun K., et al.
Published: (2026)
Time Series Analysis by State Space Learning
by: Ramos, André, et al.
Published: (2024)
by: Ramos, André, et al.
Published: (2024)
Learning Decentralized Swarms Using Rotation Equivariant Graph Neural Networks
by: Transue, Taos, et al.
Published: (2025)
by: Transue, Taos, et al.
Published: (2025)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
by: Fan, Yingda, et al.
Published: (2025)
by: Fan, Yingda, et al.
Published: (2025)
FedUNet: A Lightweight Additive U-Net Module for Federated Learning with Heterogeneous Models
by: Seo, Beomseok, et al.
Published: (2025)
by: Seo, Beomseok, et al.
Published: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
by: Wang, Youkang, et al.
Published: (2025)
by: Wang, Youkang, et al.
Published: (2025)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
by: Alege, Aliyu Agboola
Published: (2026)
by: Alege, Aliyu Agboola
Published: (2026)
MACS: Multi-Agent Reinforcement Learning for Optimization of Crystal Structures
by: Zamaraeva, Elena, et al.
Published: (2025)
by: Zamaraeva, Elena, et al.
Published: (2025)
Similar Items
-
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
by: Klačan, Ján, et al.
Published: (2026) -
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025) -
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability
by: Wang, Wuhao, et al.
Published: (2025) -
Symbolic Branch Networks: Tree-Inherited Neural Models for Interpretable Multiclass Classification
by: Rodríguez-Salas, Dalia
Published: (2025) -
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025)