ISAACS: Iterative Soft Adversarial Actor-Critic for Safety
Fuente:
arXiv
Guardado en:
| Autores principales: | Hsu, Kai-Chieh, Nguyen, Duy Phuong, Fisac, Jaime Fernández |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MAGICS: Adversarial RL with Minimax Actors Guided by Implicit Critic Stackelberg for Convergent Neural Synthesis of Robot Safety
por: Wang, Justin, et al.
Publicado: (2024)
por: Wang, Justin, et al.
Publicado: (2024)
Provably Optimal Reinforcement Learning under Safety Filtering
por: Oh, Donggeon David, et al.
Publicado: (2025)
por: Oh, Donggeon David, et al.
Publicado: (2025)
Gameplay Filters: Robust Zero-Shot Safety through Adversarial Imagination
por: Nguyen, Duy P., et al.
Publicado: (2024)
por: Nguyen, Duy P., et al.
Publicado: (2024)
Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning
por: Oh, Donggeon David, et al.
Publicado: (2026)
por: Oh, Donggeon David, et al.
Publicado: (2026)
Actor-Critic Physics-informed Neural Lyapunov Control
por: Wang, Jiarui, et al.
Publicado: (2024)
por: Wang, Jiarui, et al.
Publicado: (2024)
Fast, Smooth, and Safe: Implicit Control Barrier Functions through Reach-Avoid Differential Dynamic Programming
por: Kumar, Athindran Ramesh, et al.
Publicado: (2023)
por: Kumar, Athindran Ramesh, et al.
Publicado: (2023)
Learning for Layered Safety-Critical Control with Predictive Control Barrier Functions
por: Compton, William D., et al.
Publicado: (2024)
por: Compton, William D., et al.
Publicado: (2024)
Wasserstein Barycenter Soft Actor-Critic
por: Shahrooei, Zahra, et al.
Publicado: (2025)
por: Shahrooei, Zahra, et al.
Publicado: (2025)
Unifying Entropy Regularization in Optimal Control: From and Back to Classical Objectives via Iterated Soft Policies and Path Integral Solutions
por: Bhole, Ajinkya, et al.
Publicado: (2025)
por: Bhole, Ajinkya, et al.
Publicado: (2025)
MSACL: Multi-Step Actor-Critic Learning with Lyapunov Certificates for Exponentially Stabilizing Control
por: Zhang, Yongwei, et al.
Publicado: (2025)
por: Zhang, Yongwei, et al.
Publicado: (2025)
Distributional Soft Actor-Critic with Three Refinements
por: Duan, Jingliang, et al.
Publicado: (2023)
por: Duan, Jingliang, et al.
Publicado: (2023)
Dynamic High-Order Control Barrier Functions with Diffuser for Safety-Critical Trajectory Planning at Signal-Free Intersections
por: Chen, Di, et al.
Publicado: (2024)
por: Chen, Di, et al.
Publicado: (2024)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
por: Singh, Nikhil Kumar, et al.
Publicado: (2024)
por: Singh, Nikhil Kumar, et al.
Publicado: (2024)
Value Iteration for Learning Concurrently Executable Robotic Control Tasks
por: Tahmid, Sheikh A., et al.
Publicado: (2025)
por: Tahmid, Sheikh A., et al.
Publicado: (2025)
Quasi-Periodic Gaussian Process Predictive Iterative Learning Control
por: Nigam, Unnati, et al.
Publicado: (2026)
por: Nigam, Unnati, et al.
Publicado: (2026)
Iterative Tuning of Nonlinear Model Predictive Control for Robotic Manufacturing Tasks
por: Ingole, Deepak, et al.
Publicado: (2025)
por: Ingole, Deepak, et al.
Publicado: (2025)
Lyapunov Constrained Soft Actor-Critic (LC-SAC) using Koopman Operator Theory for Quadrotor Trajectory Tracking
por: Kushwaha, Dhruv S., et al.
Publicado: (2026)
por: Kushwaha, Dhruv S., et al.
Publicado: (2026)
Data-driven Kinematic Modeling in Soft Robots: System Identification and Uncertainty Quantification
por: Jiang, Zhanhong, et al.
Publicado: (2025)
por: Jiang, Zhanhong, et al.
Publicado: (2025)
DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems
por: Yoo, Se-Wook, et al.
Publicado: (2025)
por: Yoo, Se-Wook, et al.
Publicado: (2025)
Two-Stage Learning of Stabilizing Neural Controllers via Zubov Sampling and Iterative Domain Expansion
por: Li, Haoyu, et al.
Publicado: (2025)
por: Li, Haoyu, et al.
Publicado: (2025)
Control Pneumatic Soft Bending Actuator with Online Learning Pneumatic Physical Reservoir Computing
por: Shen, Junyi, et al.
Publicado: (2025)
por: Shen, Junyi, et al.
Publicado: (2025)
An Iterative LQR Controller for Off-Road and On-Road Vehicles using a Neural Network Dynamics Model
por: Nagariya, Akhil, et al.
Publicado: (2020)
por: Nagariya, Akhil, et al.
Publicado: (2020)
Domain Adaptive Safety Filters via Deep Operator Learning
por: Manda, Lakshmideepakreddy, et al.
Publicado: (2024)
por: Manda, Lakshmideepakreddy, et al.
Publicado: (2024)
Constraint-Aware Refinement for Safety Verification of Neural Feedback Loops
por: Rober, Nicholas, et al.
Publicado: (2024)
por: Rober, Nicholas, et al.
Publicado: (2024)
Hovering Flight of Soft-Actuated Insect-Scale Micro Aerial Vehicles using Deep Reinforcement Learning
por: Hsiao, Yi-Hsuan, et al.
Publicado: (2025)
por: Hsiao, Yi-Hsuan, et al.
Publicado: (2025)
Safe Control using Learned Safety Filters and Adaptive Conformal Inference
por: Huriot, Sacha, et al.
Publicado: (2026)
por: Huriot, Sacha, et al.
Publicado: (2026)
Uncertainty-aware Latent Safety Filters for Avoiding Out-of-Distribution Failures
por: Seo, Junwon, et al.
Publicado: (2025)
por: Seo, Junwon, et al.
Publicado: (2025)
Deep Reinforcement Learning for Time-Critical Wilderness Search And Rescue Using Drones
por: Ewers, Jan-Hendrik, et al.
Publicado: (2024)
por: Ewers, Jan-Hendrik, et al.
Publicado: (2024)
Discretionary Lane-Change Decision and Control via Parameterized Soft Actor-Critic for Hybrid Action Space
por: Lin, Yuan, et al.
Publicado: (2024)
por: Lin, Yuan, et al.
Publicado: (2024)
Safety Filtering While Training: Improving the Performance and Sample Efficiency of Reinforcement Learning Agents
por: Bejarano, Federico Pizarro, et al.
Publicado: (2024)
por: Bejarano, Federico Pizarro, et al.
Publicado: (2024)
Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm
por: Qiao, Ting, et al.
Publicado: (2024)
por: Qiao, Ting, et al.
Publicado: (2024)
Dynamic Tube MPC: Learning Tube Dynamics with Massively Parallel Simulation for Robust Safety in Practice
por: Compton, William D., et al.
Publicado: (2024)
por: Compton, William D., et al.
Publicado: (2024)
Adaptive Model-Predictive Control of a Soft Continuum Robot Using a Physics-Informed Neural Network Based on Cosserat Rod Theory
por: Licher, Johann, et al.
Publicado: (2025)
por: Licher, Johann, et al.
Publicado: (2025)
Enhancing Safety in Mixed Traffic: Learning-Based Modeling and Efficient Control of Autonomous and Human-Driven Vehicles
por: Wang, Jie, et al.
Publicado: (2024)
por: Wang, Jie, et al.
Publicado: (2024)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
por: Omidi, Saber, et al.
Publicado: (2025)
por: Omidi, Saber, et al.
Publicado: (2025)
Formal Safety Verification and Refinement for Generative Motion Planners via Certified Local Stabilization
por: Nath, Devesh, et al.
Publicado: (2025)
por: Nath, Devesh, et al.
Publicado: (2025)
Risk-Sensitive Soft Actor-Critic for Robust Deep Reinforcement Learning under Distribution Shifts
por: Enders, Tobias, et al.
Publicado: (2024)
por: Enders, Tobias, et al.
Publicado: (2024)
Safety with Agency: Human-Centered Safety Filter with Application to AI-Assisted Motorsports
por: Oh, Donggeon David, et al.
Publicado: (2025)
por: Oh, Donggeon David, et al.
Publicado: (2025)
SAFE--MA--RRT: Multi-Agent Motion Planning with Data-Driven Safety Certificates
por: Esmaeili, Babak, et al.
Publicado: (2025)
por: Esmaeili, Babak, et al.
Publicado: (2025)
Safety Beyond the Training Data: Robust Out-of-Distribution MPC via Conformalized System Level Synthesis
por: Srinivasan, Anutam, et al.
Publicado: (2026)
por: Srinivasan, Anutam, et al.
Publicado: (2026)
Ejemplares similares
-
MAGICS: Adversarial RL with Minimax Actors Guided by Implicit Critic Stackelberg for Convergent Neural Synthesis of Robot Safety
por: Wang, Justin, et al.
Publicado: (2024) -
Provably Optimal Reinforcement Learning under Safety Filtering
por: Oh, Donggeon David, et al.
Publicado: (2025) -
Gameplay Filters: Robust Zero-Shot Safety through Adversarial Imagination
por: Nguyen, Duy P., et al.
Publicado: (2024) -
Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning
por: Oh, Donggeon David, et al.
Publicado: (2026) -
Actor-Critic Physics-informed Neural Lyapunov Control
por: Wang, Jiarui, et al.
Publicado: (2024)