Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Turcato, Niccolò, Sinigaglia, Alberto, Libera, Alberto Dalla, Carli, Ruggero, Susto, Gian Antonio |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Edge Delayed Deep Deterministic Policy Gradient: efficient continuous control for edge scenarios
por: Sinigaglia, Alberto, et al.
Publicado: (2024)
por: Sinigaglia, Alberto, et al.
Publicado: (2024)
AI Olympics challenge with Evolutionary Soft Actor Critic
por: Calì, Marco, et al.
Publicado: (2024)
por: Calì, Marco, et al.
Publicado: (2024)
Learning global control of underactuated systems with Model-Based Reinforcement Learning
por: Turcato, Niccolò, et al.
Publicado: (2025)
por: Turcato, Niccolò, et al.
Publicado: (2025)
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models
por: Turcato, Niccolò, et al.
Publicado: (2025)
por: Turcato, Niccolò, et al.
Publicado: (2025)
Finetuning Deep Reinforcement Learning Policies with Evolutionary Strategies for Control of Underactuated Robots
por: Calì, Marco, et al.
Publicado: (2025)
por: Calì, Marco, et al.
Publicado: (2025)
A Black-Box Physics-Informed Estimator based on Gaussian Process Regression for Robot Inverse Dynamics Identification
por: Giacomuzzos, Giulio, et al.
Publicado: (2023)
por: Giacomuzzos, Giulio, et al.
Publicado: (2023)
Learning control of underactuated double pendulum with Model-Based Reinforcement Learning
por: Turcato, Niccolò, et al.
Publicado: (2024)
por: Turcato, Niccolò, et al.
Publicado: (2024)
Advancing Constrained Monotonic Neural Networks: Achieving Universal Approximation Beyond Bounded Activations
por: Sartor, Davide, et al.
Publicado: (2025)
por: Sartor, Davide, et al.
Publicado: (2025)
Simple and Effective Specialized Representations for Fair Classifiers
por: Sinigaglia, Alberto, et al.
Publicado: (2025)
por: Sinigaglia, Alberto, et al.
Publicado: (2025)
Data efficient Robotic Object Throwing with Model-Based Reinforcement Learning
por: Turcato, Niccolò, et al.
Publicado: (2025)
por: Turcato, Niccolò, et al.
Publicado: (2025)
Accelerating Model-Based Reinforcement Learning using Non-Linear Trajectory Optimization
por: Calì, Marco, et al.
Publicado: (2025)
por: Calì, Marco, et al.
Publicado: (2025)
Towards Batch-to-Streaming Deep Reinforcement Learning for Continuous Control
por: De Monte, Riccardo, et al.
Publicado: (2026)
por: De Monte, Riccardo, et al.
Publicado: (2026)
Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability
por: Kiram, Firas Mohamed Elamine, et al.
Publicado: (2026)
por: Kiram, Firas Mohamed Elamine, et al.
Publicado: (2026)
Reinforcement Learning for Durable Algorithmic Recourse
por: Ceccon, Marina, et al.
Publicado: (2025)
por: Ceccon, Marina, et al.
Publicado: (2025)
Fully Dynamic Rebalancing in Dockless Bike-Sharing Systems via Deep Reinforcement Learning
por: Scarpel, Edoardo, et al.
Publicado: (2026)
por: Scarpel, Edoardo, et al.
Publicado: (2026)
A Quantile Regression Approach for Remaining Useful Life Estimation with State Space Models
por: Frizzo, Davide, et al.
Publicado: (2025)
por: Frizzo, Davide, et al.
Publicado: (2025)
Multi-layer Abstraction for Nested Generation of Options (MANGO) in Hierarchical Reinforcement Learning
por: Arcudi, Alessio, et al.
Publicado: (2025)
por: Arcudi, Alessio, et al.
Publicado: (2025)
A Distributed Approach to Autonomous Intersection Management via Multi-Agent Reinforcement Learning
por: Cederle, Matteo, et al.
Publicado: (2024)
por: Cederle, Matteo, et al.
Publicado: (2024)
Weight Clipping for Deep Continual and Reinforcement Learning
por: Elsayed, Mohamed, et al.
Publicado: (2024)
por: Elsayed, Mohamed, et al.
Publicado: (2024)
Adaptive Robust Controller for handling Unknown Uncertainty of Robotic Manipulators
por: Abdelwahab, Mohamed, et al.
Publicado: (2024)
por: Abdelwahab, Mohamed, et al.
Publicado: (2024)
Double Successive Over-Relaxation Q-Learning with an Extension to Deep Reinforcement Learning
por: R, Shreyas S
Publicado: (2024)
por: R, Shreyas S
Publicado: (2024)
Towards Explainable Anomaly Detection in Shared Mobility Systems
por: Isgandarov, Elnur, et al.
Publicado: (2025)
por: Isgandarov, Elnur, et al.
Publicado: (2025)
Continual Learning for Behavior-based Driver Identification
por: Fanan, Mattia, et al.
Publicado: (2024)
por: Fanan, Mattia, et al.
Publicado: (2024)
ProDER: A Continual Learning Approach for Fault Prediction in Evolving Smart Grids
por: Efatinasab, Emad, et al.
Publicado: (2025)
por: Efatinasab, Emad, et al.
Publicado: (2025)
Towards Model-Free Learning in Dynamic Population Games: An Application to Karma Economies
por: Cederle, Matteo, et al.
Publicado: (2026)
por: Cederle, Matteo, et al.
Publicado: (2026)
Towards Adapting Reinforcement Learning Agents to New Tasks: Insights from Q-Values
por: Ramaswamy, Ashwin, et al.
Publicado: (2024)
por: Ramaswamy, Ashwin, et al.
Publicado: (2024)
Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning
por: Cederle, Matteo, et al.
Publicado: (2026)
por: Cederle, Matteo, et al.
Publicado: (2026)
Bayesian Deep Learning for Remaining Useful Life Estimation via Stein Variational Gradient Descent
por: Della Libera, Luca, et al.
Publicado: (2024)
por: Della Libera, Luca, et al.
Publicado: (2024)
Towards Transparent and Efficient Anomaly Detection in Industrial Processes through ExIFFI
por: Frizzo, Davide, et al.
Publicado: (2024)
por: Frizzo, Davide, et al.
Publicado: (2024)
A Robust Controller based on Gaussian Processes for Robotic Manipulators with Unknown Uncertainty
por: Giacomuzzo, Giulio, et al.
Publicado: (2025)
por: Giacomuzzo, Giulio, et al.
Publicado: (2025)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
por: Xu, Qiushui, et al.
Publicado: (2025)
por: Xu, Qiushui, et al.
Publicado: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
por: Liu, Wenhui, et al.
Publicado: (2025)
por: Liu, Wenhui, et al.
Publicado: (2025)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
por: Lyu, Jiafei, et al.
Publicado: (2022)
por: Lyu, Jiafei, et al.
Publicado: (2022)
Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning
por: Castagna, Alberto
Publicado: (2025)
por: Castagna, Alberto
Publicado: (2025)
SPQR: Controlling Q-ensemble Independence with Spiked Random Model for Reinforcement Learning
por: Lee, Dohyeok, et al.
Publicado: (2024)
por: Lee, Dohyeok, et al.
Publicado: (2024)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
por: Zhou, Zehao
Publicado: (2024)
por: Zhou, Zehao
Publicado: (2024)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
por: Mao, Yixiu, et al.
Publicado: (2025)
por: Mao, Yixiu, et al.
Publicado: (2025)
Meta Reinforcement Learning with Finite Training Tasks -- a Density Estimation Approach
por: Rimon, Zohar, et al.
Publicado: (2022)
por: Rimon, Zohar, et al.
Publicado: (2022)
On The Presence of Double-Descent in Deep Reinforcement Learning
por: Veselý, Viktor, et al.
Publicado: (2025)
por: Veselý, Viktor, et al.
Publicado: (2025)
BandPO: Bridging Trust Regions and Ratio Clipping via Probability-Aware Bounds for LLM Reinforcement Learning
por: Li, Yuan, et al.
Publicado: (2026)
por: Li, Yuan, et al.
Publicado: (2026)
Ejemplares similares
-
Edge Delayed Deep Deterministic Policy Gradient: efficient continuous control for edge scenarios
por: Sinigaglia, Alberto, et al.
Publicado: (2024) -
AI Olympics challenge with Evolutionary Soft Actor Critic
por: Calì, Marco, et al.
Publicado: (2024) -
Learning global control of underactuated systems with Model-Based Reinforcement Learning
por: Turcato, Niccolò, et al.
Publicado: (2025) -
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models
por: Turcato, Niccolò, et al.
Publicado: (2025) -
Finetuning Deep Reinforcement Learning Policies with Evolutionary Strategies for Control of Underactuated Robots
por: Calì, Marco, et al.
Publicado: (2025)