DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
Fuente:
arXiv
Saved in:
| Main Authors: | Perkins, Daniel, Escobar, Oscar J., Green, Luke |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024)
by: Tyukin, Ivan Y., et al.
Published: (2024)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
by: Wang, Youkang, et al.
Published: (2025)
by: Wang, Youkang, et al.
Published: (2025)
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025)
by: Yilmaz, Berk, et al.
Published: (2025)
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
by: Gao, Weiguo, et al.
Published: (2025)
by: Gao, Weiguo, et al.
Published: (2025)
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025)
by: Varricchione, Giovanni, et al.
Published: (2025)
Subset Selection for Fine-Tuning: A Utility-Diversity Balanced Approach for Mathematical Domain Adaptation
by: Kotecha, Madhav, et al.
Published: (2025)
by: Kotecha, Madhav, et al.
Published: (2025)
Unsupervised Ensemble Learning Through Deep Energy-based Models
by: Maymon, Ariel, et al.
Published: (2026)
by: Maymon, Ariel, et al.
Published: (2026)
Maximally Permissive Reward Machines
by: Varricchione, Giovanni, et al.
Published: (2024)
by: Varricchione, Giovanni, et al.
Published: (2024)
BatteryML:An Open-source platform for Machine Learning on Battery Degradation
by: Zhang, Han, et al.
Published: (2023)
by: Zhang, Han, et al.
Published: (2023)
ASNN: Learning to Suggest Neural Architectures from Performance Distributions
by: Hong, Jinwook
Published: (2025)
by: Hong, Jinwook
Published: (2025)
Grouped Sequential Optimization Strategy -- the Application of Hyperparameter Importance Assessment in Deep Learning
by: Wang, Ruinan, et al.
Published: (2025)
by: Wang, Ruinan, et al.
Published: (2025)
Reciprocal Learning
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
Large Language Model Meets Graph Neural Network in Knowledge Distillation
by: Hu, Shengxiang, et al.
Published: (2024)
by: Hu, Shengxiang, et al.
Published: (2024)
Multi-State TD Target for Model-Free Reinforcement Learning
by: Wang, Wuhao, et al.
Published: (2024)
by: Wang, Wuhao, et al.
Published: (2024)
A Neural Affinity Framework for Abstract Reasoning: Diagnosing the Compositional Gap in Transformer Architectures via Procedural Task Taxonomy
by: Ingram, Miguel, et al.
Published: (2025)
by: Ingram, Miguel, et al.
Published: (2025)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
by: Hanika, Tom, et al.
Published: (2024)
by: Hanika, Tom, et al.
Published: (2024)
EcoTransformer: Attention without Multiplication
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
Noradrenergic-inspired gain modulation attenuates the stability gap in joint training
by: Rodriguez-Garcia, Alejandro, et al.
Published: (2025)
by: Rodriguez-Garcia, Alejandro, et al.
Published: (2025)
Topological Foundations of Reinforcement Learning
by: Kadurha, David Krame
Published: (2024)
by: Kadurha, David Krame
Published: (2024)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
by: Park, Junseok, et al.
Published: (2025)
by: Park, Junseok, et al.
Published: (2025)
Investigating the Interplay of Prioritized Replay and Generalization
by: Panahi, Parham Mohammad, et al.
Published: (2024)
by: Panahi, Parham Mohammad, et al.
Published: (2024)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
Censored Sampling for Topology Design: Guiding Diffusion with Human Preferences
by: Kim, Euihyun, et al.
Published: (2025)
by: Kim, Euihyun, et al.
Published: (2025)
Insights into Schizophrenia: Leveraging Machine Learning for Early Identification via EEG, ERP, and Demographic Attributes
by: Alkhalifa, Sara
Published: (2025)
by: Alkhalifa, Sara
Published: (2025)
Neurosymbolic Association Rule Mining from Tabular Data
by: Karabulut, Erkan, et al.
Published: (2025)
by: Karabulut, Erkan, et al.
Published: (2025)
Car Sensors Health Monitoring by Verification Based on Autoencoder and Random Forest Regression
by: Torkhesari, Sahar, et al.
Published: (2025)
by: Torkhesari, Sahar, et al.
Published: (2025)
FedSWA: Improving Generalization in Federated Learning with Highly Heterogeneous Data via Momentum-Based Stochastic Controlled Weight Averaging
by: junkang, Liu, et al.
Published: (2025)
by: junkang, Liu, et al.
Published: (2025)
PAC-MCTS: Bias-Aware Pruning for Robust LLM-Guided Search and Planning
by: Qian, Tianhao
Published: (2026)
by: Qian, Tianhao
Published: (2026)
A social path to human-like artificial intelligence
by: Duéñez-Guzmán, Edgar A., et al.
Published: (2024)
by: Duéñez-Guzmán, Edgar A., et al.
Published: (2024)
Memory-efficient Continual Learning with Prototypical Exemplar Condensation
by: Nguyen, Minh-Duong, et al.
Published: (2026)
by: Nguyen, Minh-Duong, et al.
Published: (2026)
Tracking Changing Probabilities via Dynamic Learners
by: Madani, Omid
Published: (2024)
by: Madani, Omid
Published: (2024)
Self-Directed Task Identification
by: Gould, Timothy, et al.
Published: (2026)
by: Gould, Timothy, et al.
Published: (2026)
QGraphLIME - Explaining Quantum Graph Neural Networks
by: Jena, Haribandhu, et al.
Published: (2025)
by: Jena, Haribandhu, et al.
Published: (2025)
Graph Transformers: A Survey
by: Shehzad, Ahsan, et al.
Published: (2024)
by: Shehzad, Ahsan, et al.
Published: (2024)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
by: Fan, Yingda, et al.
Published: (2025)
by: Fan, Yingda, et al.
Published: (2025)
Enhancing LLM Agents for Code Generation with Possibility and Pass-rate Prioritized Experience Replay
by: Chen, Yuyang, et al.
Published: (2024)
by: Chen, Yuyang, et al.
Published: (2024)
GoldenStart: Q-Guided Priors and Entropy Control for Distilling Flow Policies
by: Zhang, He, et al.
Published: (2026)
by: Zhang, He, et al.
Published: (2026)
Why Does Differential Privacy with Large Epsilon Defend Against Practical Membership Inference Attacks?
by: Lowy, Andrew, et al.
Published: (2024)
by: Lowy, Andrew, et al.
Published: (2024)
Order-Robust Class Incremental Learning: Graph-Driven Dynamic Similarity Grouping
by: Lai, Guannan, et al.
Published: (2025)
by: Lai, Guannan, et al.
Published: (2025)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
by: Liu, Jinyi, et al.
Published: (2023)
by: Liu, Jinyi, et al.
Published: (2023)
Similar Items
-
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024) -
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
by: Wang, Youkang, et al.
Published: (2025) -
Deep Reinforcement Learning Xiangqi Player with Monte Carlo Tree Search
by: Yilmaz, Berk, et al.
Published: (2025) -
Convergence Dynamics and Stabilization Strategies of Co-Evolving Generative Models
by: Gao, Weiguo, et al.
Published: (2025) -
Pushdown Reward Machines for Reinforcement Learning
by: Varricchione, Giovanni, et al.
Published: (2025)