Enhanced-FQL($λ$), an Efficient and Interpretable RL with novel Fuzzy Eligibility Traces and Segmented Experience Replay
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jalaeian-Farimani, Mohsen, Xiong, Xiong, Bascetta, Luca |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
von: Amirabadi, Roya Khalili, et al.
Veröffentlicht: (2026)
von: Amirabadi, Roya Khalili, et al.
Veröffentlicht: (2026)
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
von: Studt, Max, et al.
Veröffentlicht: (2025)
von: Studt, Max, et al.
Veröffentlicht: (2025)
Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU
von: Fang, Jeffrey, et al.
Veröffentlicht: (2026)
von: Fang, Jeffrey, et al.
Veröffentlicht: (2026)
Aero-Promptness: Drag-Aware Aerodynamic Manipulability for Propeller-driven Vehicles
von: Franchi, Antonio
Veröffentlicht: (2026)
von: Franchi, Antonio
Veröffentlicht: (2026)
Who Plays First? Optimizing the Order of Play in Stackelberg Games with Many Robots
von: Hu, Haimin, et al.
Veröffentlicht: (2024)
von: Hu, Haimin, et al.
Veröffentlicht: (2024)
A Tricycle Model to Accurately Control an Autonomous Racecar with Locked Differential
von: Raji, Ayoub, et al.
Veröffentlicht: (2023)
von: Raji, Ayoub, et al.
Veröffentlicht: (2023)
Blending Data-Driven Priors in Dynamic Games
von: Lidard, Justin, et al.
Veröffentlicht: (2024)
von: Lidard, Justin, et al.
Veröffentlicht: (2024)
A novel agent with formal goal-reaching guarantees: an experimental study with a mobile robot
von: Yaremenko, Grigory, et al.
Veröffentlicht: (2024)
von: Yaremenko, Grigory, et al.
Veröffentlicht: (2024)
Model Predictive Control of Hybrid Dynamical Systems
von: Sanfelice, Ricardo G., et al.
Veröffentlicht: (2026)
von: Sanfelice, Ricardo G., et al.
Veröffentlicht: (2026)
Design and Realization of a Benchmarking Testbed for Evaluating Autonomous Platooning Algorithms
von: Shaham, Michael, et al.
Veröffentlicht: (2024)
von: Shaham, Michael, et al.
Veröffentlicht: (2024)
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making
von: Xu, Zirui, et al.
Veröffentlicht: (2022)
von: Xu, Zirui, et al.
Veröffentlicht: (2022)
Performance-Aware Self-Configurable Multi-Agent Networks: A Distributed Submodular Approach for Simultaneous Coordination and Network Design
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
von: Xu, Zirui, et al.
Veröffentlicht: (2024)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
von: Fabbro, Nicolò Dal, et al.
Veröffentlicht: (2024)
von: Fabbro, Nicolò Dal, et al.
Veröffentlicht: (2024)
Scalable Data-Driven Reachability Analysis and Control via Koopman Operators with Conformal Coverage Guarantees
von: Nath, Devesh, et al.
Veröffentlicht: (2026)
von: Nath, Devesh, et al.
Veröffentlicht: (2026)
ViSTR-GP: Online Cyberattack Detection via Vision-to-State Tensor Regression and Gaussian Processes in Automated Robotic Operations
von: Aftabi, Navid, et al.
Veröffentlicht: (2025)
von: Aftabi, Navid, et al.
Veröffentlicht: (2025)
Conformal Prediction in The Loop: A Feedback-Based Uncertainty Model for Trajectory Optimization
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
von: Shen, Keyi, et al.
Veröffentlicht: (2026)
von: Shen, Keyi, et al.
Veröffentlicht: (2026)
Neural-Rendezvous: Provably Robust Guidance and Control to Encounter Interstellar Objects
von: Tsukamoto, Hiroyasu, et al.
Veröffentlicht: (2022)
von: Tsukamoto, Hiroyasu, et al.
Veröffentlicht: (2022)
Lyapunov-stable Neural Control for State and Output Feedback: A Novel Formulation
von: Yang, Lujie, et al.
Veröffentlicht: (2024)
von: Yang, Lujie, et al.
Veröffentlicht: (2024)
Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization
von: Murillo-Gonzalez, Alejandro, et al.
Veröffentlicht: (2026)
von: Murillo-Gonzalez, Alejandro, et al.
Veröffentlicht: (2026)
Active Constraint Learning in High Dimensions from Demonstrations
von: Qiu, Zheng, et al.
Veröffentlicht: (2025)
von: Qiu, Zheng, et al.
Veröffentlicht: (2025)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
von: Malaniya, Georgiy, et al.
Veröffentlicht: (2025)
von: Malaniya, Georgiy, et al.
Veröffentlicht: (2025)
A universal policy wrapper with guarantees
von: Bolychev, Anton, et al.
Veröffentlicht: (2025)
von: Bolychev, Anton, et al.
Veröffentlicht: (2025)
Aerial Inspection Behaviors via RL-based Quadrotor Control for Under-canopy Forest Environments
von: Suarez, Fausto Mauricio Lagos, et al.
Veröffentlicht: (2026)
von: Suarez, Fausto Mauricio Lagos, et al.
Veröffentlicht: (2026)
Cost-Effective Robotic Handwriting System with AI Integration
von: Huang, Tianyi, et al.
Veröffentlicht: (2025)
von: Huang, Tianyi, et al.
Veröffentlicht: (2025)
Safe Control and Learning Using Generalized Action Governor
von: Fang, Peiyuan, et al.
Veröffentlicht: (2022)
von: Fang, Peiyuan, et al.
Veröffentlicht: (2022)
Interpreting and Improving Optimal Control Problems with Directional Corrections
von: Barron, Trevor, et al.
Veröffentlicht: (2025)
von: Barron, Trevor, et al.
Veröffentlicht: (2025)
Stability-Preserving Online Adaptation of Neural Closed-loop Maps
von: Saccani, Danilo, et al.
Veröffentlicht: (2026)
von: Saccani, Danilo, et al.
Veröffentlicht: (2026)
Eigendecomposition Parameterization of Penalty Matrices for Enhanced Control Design: Aerospace Applications
von: Nurre, Nicholas P., et al.
Veröffentlicht: (2025)
von: Nurre, Nicholas P., et al.
Veröffentlicht: (2025)
Enhanced Trust Region Sequential Convex Optimization for Multi-Drone Thermal Screening Trajectory Planning in Urban Environments
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
Maximal Controlled Invariant-MPC: Enhancing Feasibility and Reducing Conservatism through Terminal CBF Constraint in Safety-Critical Control
von: Dokania, Tanmay, et al.
Veröffentlicht: (2026)
von: Dokania, Tanmay, et al.
Veröffentlicht: (2026)
Human-in-the-Loop AI for HVAC Management Enhancing Comfort and Energy Efficiency
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
Hereditary Geometric Meta-RL: Nonlocal Generalization via Task Symmetries
von: Nitschke, Paul, et al.
Veröffentlicht: (2026)
von: Nitschke, Paul, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning for Multi-Objective Optimization: Enhancing Wind Turbine Energy Generation while Mitigating Noise Emissions
von: de Frutos, Martín, et al.
Veröffentlicht: (2024)
von: de Frutos, Martín, et al.
Veröffentlicht: (2024)
Efficient Optimal Path Planning in Dynamic Environments Using Koopman MPC
von: Abtahi, Mohammad, et al.
Veröffentlicht: (2025)
von: Abtahi, Mohammad, et al.
Veröffentlicht: (2025)
A Genetic Fuzzy-Enabled Framework on Robotic Manipulation for In-Space Servicing
von: Steffen, Nathan, et al.
Veröffentlicht: (2025)
von: Steffen, Nathan, et al.
Veröffentlicht: (2025)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
Solving Multi-Agent Safe Optimal Control with Distributed Epigraph Form MARL
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
A Nonasymptotic Theory of Gain-Dependent Error Dynamics in Behavior Cloning
von: Seo, Junghoon
Veröffentlicht: (2026)
von: Seo, Junghoon
Veröffentlicht: (2026)
Ähnliche Einträge
-
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks
von: Xu, Zirui, et al.
Veröffentlicht: (2024) -
Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
von: Amirabadi, Roya Khalili, et al.
Veröffentlicht: (2026) -
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
von: Studt, Max, et al.
Veröffentlicht: (2025) -
Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU
von: Fang, Jeffrey, et al.
Veröffentlicht: (2026) -
Aero-Promptness: Drag-Aware Aerodynamic Manipulability for Propeller-driven Vehicles
von: Franchi, Antonio
Veröffentlicht: (2026)