Enhanced-FQL($λ$), an Efficient and Interpretable RL with novel Fuzzy Eligibility Traces and Segmented Experience Replay
Fuente:
arXiv
Salvato in:
| Autori principali: | Jalaeian-Farimani, Mohsen, Xiong, Xiong, Bascetta, Luca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks
di: Xu, Zirui, et al.
Pubblicazione: (2024)
di: Xu, Zirui, et al.
Pubblicazione: (2024)
Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
di: Amirabadi, Roya Khalili, et al.
Pubblicazione: (2026)
di: Amirabadi, Roya Khalili, et al.
Pubblicazione: (2026)
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
di: Studt, Max, et al.
Pubblicazione: (2025)
di: Studt, Max, et al.
Pubblicazione: (2025)
Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU
di: Fang, Jeffrey, et al.
Pubblicazione: (2026)
di: Fang, Jeffrey, et al.
Pubblicazione: (2026)
Aero-Promptness: Drag-Aware Aerodynamic Manipulability for Propeller-driven Vehicles
di: Franchi, Antonio
Pubblicazione: (2026)
di: Franchi, Antonio
Pubblicazione: (2026)
Who Plays First? Optimizing the Order of Play in Stackelberg Games with Many Robots
di: Hu, Haimin, et al.
Pubblicazione: (2024)
di: Hu, Haimin, et al.
Pubblicazione: (2024)
A Tricycle Model to Accurately Control an Autonomous Racecar with Locked Differential
di: Raji, Ayoub, et al.
Pubblicazione: (2023)
di: Raji, Ayoub, et al.
Pubblicazione: (2023)
Blending Data-Driven Priors in Dynamic Games
di: Lidard, Justin, et al.
Pubblicazione: (2024)
di: Lidard, Justin, et al.
Pubblicazione: (2024)
A novel agent with formal goal-reaching guarantees: an experimental study with a mobile robot
di: Yaremenko, Grigory, et al.
Pubblicazione: (2024)
di: Yaremenko, Grigory, et al.
Pubblicazione: (2024)
Model Predictive Control of Hybrid Dynamical Systems
di: Sanfelice, Ricardo G., et al.
Pubblicazione: (2026)
di: Sanfelice, Ricardo G., et al.
Pubblicazione: (2026)
Design and Realization of a Benchmarking Testbed for Evaluating Autonomous Platooning Algorithms
di: Shaham, Michael, et al.
Pubblicazione: (2024)
di: Shaham, Michael, et al.
Pubblicazione: (2024)
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making
di: Xu, Zirui, et al.
Pubblicazione: (2022)
di: Xu, Zirui, et al.
Pubblicazione: (2022)
Performance-Aware Self-Configurable Multi-Agent Networks: A Distributed Submodular Approach for Simultaneous Coordination and Network Design
di: Xu, Zirui, et al.
Pubblicazione: (2024)
di: Xu, Zirui, et al.
Pubblicazione: (2024)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
di: Fabbro, Nicolò Dal, et al.
Pubblicazione: (2024)
di: Fabbro, Nicolò Dal, et al.
Pubblicazione: (2024)
Scalable Data-Driven Reachability Analysis and Control via Koopman Operators with Conformal Coverage Guarantees
di: Nath, Devesh, et al.
Pubblicazione: (2026)
di: Nath, Devesh, et al.
Pubblicazione: (2026)
ViSTR-GP: Online Cyberattack Detection via Vision-to-State Tensor Regression and Gaussian Processes in Automated Robotic Operations
di: Aftabi, Navid, et al.
Pubblicazione: (2025)
di: Aftabi, Navid, et al.
Pubblicazione: (2025)
Conformal Prediction in The Loop: A Feedback-Based Uncertainty Model for Trajectory Optimization
di: Wang, Han, et al.
Pubblicazione: (2025)
di: Wang, Han, et al.
Pubblicazione: (2025)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
di: Shen, Keyi, et al.
Pubblicazione: (2026)
di: Shen, Keyi, et al.
Pubblicazione: (2026)
Neural-Rendezvous: Provably Robust Guidance and Control to Encounter Interstellar Objects
di: Tsukamoto, Hiroyasu, et al.
Pubblicazione: (2022)
di: Tsukamoto, Hiroyasu, et al.
Pubblicazione: (2022)
Lyapunov-stable Neural Control for State and Output Feedback: A Novel Formulation
di: Yang, Lujie, et al.
Pubblicazione: (2024)
di: Yang, Lujie, et al.
Pubblicazione: (2024)
Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization
di: Murillo-Gonzalez, Alejandro, et al.
Pubblicazione: (2026)
di: Murillo-Gonzalez, Alejandro, et al.
Pubblicazione: (2026)
Active Constraint Learning in High Dimensions from Demonstrations
di: Qiu, Zheng, et al.
Pubblicazione: (2025)
di: Qiu, Zheng, et al.
Pubblicazione: (2025)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
di: Malaniya, Georgiy, et al.
Pubblicazione: (2025)
di: Malaniya, Georgiy, et al.
Pubblicazione: (2025)
A universal policy wrapper with guarantees
di: Bolychev, Anton, et al.
Pubblicazione: (2025)
di: Bolychev, Anton, et al.
Pubblicazione: (2025)
Aerial Inspection Behaviors via RL-based Quadrotor Control for Under-canopy Forest Environments
di: Suarez, Fausto Mauricio Lagos, et al.
Pubblicazione: (2026)
di: Suarez, Fausto Mauricio Lagos, et al.
Pubblicazione: (2026)
Cost-Effective Robotic Handwriting System with AI Integration
di: Huang, Tianyi, et al.
Pubblicazione: (2025)
di: Huang, Tianyi, et al.
Pubblicazione: (2025)
Safe Control and Learning Using Generalized Action Governor
di: Fang, Peiyuan, et al.
Pubblicazione: (2022)
di: Fang, Peiyuan, et al.
Pubblicazione: (2022)
Interpreting and Improving Optimal Control Problems with Directional Corrections
di: Barron, Trevor, et al.
Pubblicazione: (2025)
di: Barron, Trevor, et al.
Pubblicazione: (2025)
Stability-Preserving Online Adaptation of Neural Closed-loop Maps
di: Saccani, Danilo, et al.
Pubblicazione: (2026)
di: Saccani, Danilo, et al.
Pubblicazione: (2026)
Eigendecomposition Parameterization of Penalty Matrices for Enhanced Control Design: Aerospace Applications
di: Nurre, Nicholas P., et al.
Pubblicazione: (2025)
di: Nurre, Nicholas P., et al.
Pubblicazione: (2025)
Enhanced Trust Region Sequential Convex Optimization for Multi-Drone Thermal Screening Trajectory Planning in Urban Environments
di: Chen, Kaiyuan, et al.
Pubblicazione: (2025)
di: Chen, Kaiyuan, et al.
Pubblicazione: (2025)
Maximal Controlled Invariant-MPC: Enhancing Feasibility and Reducing Conservatism through Terminal CBF Constraint in Safety-Critical Control
di: Dokania, Tanmay, et al.
Pubblicazione: (2026)
di: Dokania, Tanmay, et al.
Pubblicazione: (2026)
Human-in-the-Loop AI for HVAC Management Enhancing Comfort and Energy Efficiency
di: Liang, Xinyu, et al.
Pubblicazione: (2025)
di: Liang, Xinyu, et al.
Pubblicazione: (2025)
Hereditary Geometric Meta-RL: Nonlocal Generalization via Task Symmetries
di: Nitschke, Paul, et al.
Pubblicazione: (2026)
di: Nitschke, Paul, et al.
Pubblicazione: (2026)
Deep Reinforcement Learning for Multi-Objective Optimization: Enhancing Wind Turbine Energy Generation while Mitigating Noise Emissions
di: de Frutos, Martín, et al.
Pubblicazione: (2024)
di: de Frutos, Martín, et al.
Pubblicazione: (2024)
Efficient Optimal Path Planning in Dynamic Environments Using Koopman MPC
di: Abtahi, Mohammad, et al.
Pubblicazione: (2025)
di: Abtahi, Mohammad, et al.
Pubblicazione: (2025)
A Genetic Fuzzy-Enabled Framework on Robotic Manipulation for In-Space Servicing
di: Steffen, Nathan, et al.
Pubblicazione: (2025)
di: Steffen, Nathan, et al.
Pubblicazione: (2025)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
di: Faradonbeh, Mohamad Kazem Shirani, et al.
Pubblicazione: (2022)
di: Faradonbeh, Mohamad Kazem Shirani, et al.
Pubblicazione: (2022)
Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
di: Zhang, Songyuan, et al.
Pubblicazione: (2025)
di: Zhang, Songyuan, et al.
Pubblicazione: (2025)
A Nonasymptotic Theory of Gain-Dependent Error Dynamics in Behavior Cloning
di: Seo, Junghoon
Pubblicazione: (2026)
di: Seo, Junghoon
Pubblicazione: (2026)
Documenti analoghi
-
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks
di: Xu, Zirui, et al.
Pubblicazione: (2024) -
Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
di: Amirabadi, Roya Khalili, et al.
Pubblicazione: (2026) -
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
di: Studt, Max, et al.
Pubblicazione: (2025) -
Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU
di: Fang, Jeffrey, et al.
Pubblicazione: (2026) -
Aero-Promptness: Drag-Aware Aerodynamic Manipulability for Propeller-driven Vehicles
di: Franchi, Antonio
Pubblicazione: (2026)