Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Bossens, David M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mirror Descent Policy Optimisation for Robust Constrained Markov Decision Processes
von: Bossens, David M., et al.
Veröffentlicht: (2025)
von: Bossens, David M., et al.
Veröffentlicht: (2025)
Adversarially Robust Spiking Neural Networks Through Conversion
von: Özdenizci, Ozan, et al.
Veröffentlicht: (2023)
von: Özdenizci, Ozan, et al.
Veröffentlicht: (2023)
Exploring Layerwise Adversarial Robustness Through the Lens of t-SNE
von: Valentim, Inês, et al.
Veröffentlicht: (2024)
von: Valentim, Inês, et al.
Veröffentlicht: (2024)
NERO-Net: A Neuroevolutionary Approach for the Design of Adversarially Robust CNNs
von: Valentim, Inês, et al.
Veröffentlicht: (2026)
von: Valentim, Inês, et al.
Veröffentlicht: (2026)
Bridging Models to Defend: A Population-Based Strategy for Robust Adversarial Defense
von: Wang, Ren, et al.
Veröffentlicht: (2023)
von: Wang, Ren, et al.
Veröffentlicht: (2023)
Evaluating the Robustness of Deep-Learning Algorithm-Selection Models by Evolving Adversarial Instances
von: Hart, Emma, et al.
Veröffentlicht: (2024)
von: Hart, Emma, et al.
Veröffentlicht: (2024)
Enabling Robust In-Context Memory and Rapid Task Adaptation in Transformers with Hebbian and Gradient-Based Plasticity
von: Chaudhary, Siddharth
Veröffentlicht: (2025)
von: Chaudhary, Siddharth
Veröffentlicht: (2025)
Quality-Diversity Meta-Evolution: customising behaviour spaces to a meta-objective
von: Bossens, David M., et al.
Veröffentlicht: (2021)
von: Bossens, David M., et al.
Veröffentlicht: (2021)
On the use of feature-maps and parameter control for improved quality-diversity meta-evolution
von: Bossens, David M., et al.
Veröffentlicht: (2021)
von: Bossens, David M., et al.
Veröffentlicht: (2021)
Unveiling the Decision-Making Process in Reinforcement Learning with Genetic Programming
von: Eberhardinger, Manuel, et al.
Veröffentlicht: (2024)
von: Eberhardinger, Manuel, et al.
Veröffentlicht: (2024)
Seemingly Redundant Modules Enhance Robust Odor Learning in Fruit Flies
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
Multi-Timescale Conductance Spiking Networks: A Sparse, Gradient-Trainable Framework with Rich Firing Dynamics for Enhanced Temporal Processing
von: Fulleda-Garcia, Alex, et al.
Veröffentlicht: (2026)
von: Fulleda-Garcia, Alex, et al.
Veröffentlicht: (2026)
Decision SpikeFormer: Spike-Driven Transformer for Decision Making
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
On the Markov Property of Neural Algorithmic Reasoning: Analyses and Methods
von: Bohde, Montgomery, et al.
Veröffentlicht: (2024)
von: Bohde, Montgomery, et al.
Veröffentlicht: (2024)
The Robustness of Spiking Neural Networks in Communication and its Application towards Network Efficiency in Federated Learning
von: Nguyen, Manh V., et al.
Veröffentlicht: (2024)
von: Nguyen, Manh V., et al.
Veröffentlicht: (2024)
Scaling Policy Gradient Quality-Diversity with Massive Parallelization via Behavioral Variations
von: Mitsides, Konstantinos, et al.
Veröffentlicht: (2025)
von: Mitsides, Konstantinos, et al.
Veröffentlicht: (2025)
The Alpha-Alternator: Dynamic Adaptation To Varying Noise Levels In Sequences Using The Vendi Score For Improved Robustness and Performance
von: Rezaei, Mohammad Reza, et al.
Veröffentlicht: (2025)
von: Rezaei, Mohammad Reza, et al.
Veröffentlicht: (2025)
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
von: Sander, Jacob, et al.
Veröffentlicht: (2025)
von: Sander, Jacob, et al.
Veröffentlicht: (2025)
Understanding Transformer Optimization via Gradient Heterogeneity
von: Tomihari, Akiyoshi, et al.
Veröffentlicht: (2025)
von: Tomihari, Akiyoshi, et al.
Veröffentlicht: (2025)
AlphaGrad: Non-Linear Gradient Normalization Optimizer
von: Sane, Soham
Veröffentlicht: (2025)
von: Sane, Soham
Veröffentlicht: (2025)
An Inverse Modeling Constrained Multi-Objective Evolutionary Algorithm Based on Decomposition
von: Farias, Lucas R. C., et al.
Veröffentlicht: (2024)
von: Farias, Lucas R. C., et al.
Veröffentlicht: (2024)
Rapidly adapting robot swarms with Swarm Map-based Bayesian Optimisation
von: Bossens, David M., et al.
Veröffentlicht: (2020)
von: Bossens, David M., et al.
Veröffentlicht: (2020)
A Truly Sparse and General Implementation of Gradient-Based Synaptic Plasticity
von: Lohoff, Jamie, et al.
Veröffentlicht: (2025)
von: Lohoff, Jamie, et al.
Veröffentlicht: (2025)
Perceptual Motor Learning with Active Inference Framework for Robust Lateral Control
von: Delavari, Elahe, et al.
Veröffentlicht: (2025)
von: Delavari, Elahe, et al.
Veröffentlicht: (2025)
Scalable Event-by-event Processing of Neuromorphic Sensory Signals With Deep State-Space Models
von: Schöne, Mark, et al.
Veröffentlicht: (2024)
von: Schöne, Mark, et al.
Veröffentlicht: (2024)
Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies
von: Patankar, Dhruv, et al.
Veröffentlicht: (2026)
von: Patankar, Dhruv, et al.
Veröffentlicht: (2026)
Advancing Direct Training for Spiking Neural Networks with Circulate-Firing Neurons and Learnable Gradients
von: Zhou, Feifan, et al.
Veröffentlicht: (2026)
von: Zhou, Feifan, et al.
Veröffentlicht: (2026)
Discovering Effective Policies for Land-Use Planning with Neuroevolution
von: Young, Daniel, et al.
Veröffentlicht: (2023)
von: Young, Daniel, et al.
Veröffentlicht: (2023)
On-line Policy Improvement using Monte-Carlo Search
von: Tesauro, Gerald, et al.
Veröffentlicht: (2025)
von: Tesauro, Gerald, et al.
Veröffentlicht: (2025)
Gradient-Free Continual Learning in Spiking Neural Networks via Inter-Spike Interval Regularization
von: Roy, Samrendra, et al.
Veröffentlicht: (2026)
von: Roy, Samrendra, et al.
Veröffentlicht: (2026)
t-DGR: A Trajectory-Based Deep Generative Replay Method for Continual Learning in Decision Making
von: Yue, William, et al.
Veröffentlicht: (2024)
von: Yue, William, et al.
Veröffentlicht: (2024)
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
von: Sane, Soham
Veröffentlicht: (2025)
von: Sane, Soham
Veröffentlicht: (2025)
Neural Policy Style Transfer
von: Fernandez-Fernandez, Raul, et al.
Veröffentlicht: (2024)
von: Fernandez-Fernandez, Raul, et al.
Veröffentlicht: (2024)
Can We Optimize Deep RL Policy Weights as Trajectory Modeling?
von: Tang, Hongyao
Veröffentlicht: (2025)
von: Tang, Hongyao
Veröffentlicht: (2025)
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling
von: Musto, Henry, et al.
Veröffentlicht: (2024)
von: Musto, Henry, et al.
Veröffentlicht: (2024)
An Efficient Reconstructed Differential Evolution Variant by Some of the Current State-of-the-art Strategies for Solving Single Objective Bound Constrained Problems
von: Tao, Sichen, et al.
Veröffentlicht: (2024)
von: Tao, Sichen, et al.
Veröffentlicht: (2024)
Combining Large Language Models and Gradient-Free Optimization for Automatic Control Policy Synthesis
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
Bridge the Inference Gaps of Neural Processes via Expectation Maximization
von: Wang, Qi, et al.
Veröffentlicht: (2025)
von: Wang, Qi, et al.
Veröffentlicht: (2025)
The Digital Ecosystem of Beliefs: does evolution favour AI over humans?
von: Bossens, David M., et al.
Veröffentlicht: (2024)
von: Bossens, David M., et al.
Veröffentlicht: (2024)
Robust Dynamic Material Handling via Adaptive Constrained Evolutionary Reinforcement Learning
von: Hu, Chengpeng, et al.
Veröffentlicht: (2025)
von: Hu, Chengpeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mirror Descent Policy Optimisation for Robust Constrained Markov Decision Processes
von: Bossens, David M., et al.
Veröffentlicht: (2025) -
Adversarially Robust Spiking Neural Networks Through Conversion
von: Özdenizci, Ozan, et al.
Veröffentlicht: (2023) -
Exploring Layerwise Adversarial Robustness Through the Lens of t-SNE
von: Valentim, Inês, et al.
Veröffentlicht: (2024) -
NERO-Net: A Neuroevolutionary Approach for the Design of Adversarially Robust CNNs
von: Valentim, Inês, et al.
Veröffentlicht: (2026) -
Bridging Models to Defend: A Population-Based Strategy for Robust Adversarial Defense
von: Wang, Ren, et al.
Veröffentlicht: (2023)