Criticality and Safety Margins for Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Grushin, Alexander, Woods, Walt, Velasquez, Alvaro, Khan, Simon |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safety Margins for Reinforcement Learning
by: Grushin, Alexander, et al.
Published: (2023)
by: Grushin, Alexander, et al.
Published: (2023)
Combining AI Control Systems and Human Decision Support via Robustness and Criticality
by: Woods, Walt, et al.
Published: (2024)
by: Woods, Walt, et al.
Published: (2024)
Reliable Grid Forecasting: State Space Models for Safety-Critical Energy Systems
by: Hong, Sunki, et al.
Published: (2026)
by: Hong, Sunki, et al.
Published: (2026)
Upside Down Reinforcement Learning with Policy Generators
by: Di Ventura, Jacopo, et al.
Published: (2025)
by: Di Ventura, Jacopo, et al.
Published: (2025)
Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
by: Ricciardi, Antonio Pio, et al.
Published: (2025)
by: Ricciardi, Antonio Pio, et al.
Published: (2025)
On the Convergence and Stability of Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning, and Online Decision Transformers
by: Štrupl, Miroslav, et al.
Published: (2025)
by: Štrupl, Miroslav, et al.
Published: (2025)
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
by: Yamchote, Phaphontee, et al.
Published: (2025)
by: Yamchote, Phaphontee, et al.
Published: (2025)
Simple Network Graph Comparative Learning
by: Yu, Qiang, et al.
Published: (2026)
by: Yu, Qiang, et al.
Published: (2026)
HyperMask: Adaptive Hypernetwork-based Masks for Continual Learning
by: Książek, Kamil, et al.
Published: (2023)
by: Książek, Kamil, et al.
Published: (2023)
CGLearn: Consistent Gradient-Based Learning for Out-of-Distribution Generalization
by: Chowdhury, Jawad, et al.
Published: (2024)
by: Chowdhury, Jawad, et al.
Published: (2024)
A Machine Learning Framework for Turbofan Health Estimation via Inverse Problem Formulation
by: Leyli-Abadi, Milad, et al.
Published: (2026)
by: Leyli-Abadi, Milad, et al.
Published: (2026)
Lost or Hidden? A Concept-Level Forgetting in Supervised Continual Learning
by: Filus, Katarzyna, et al.
Published: (2026)
by: Filus, Katarzyna, et al.
Published: (2026)
A Practical Guide to Streaming Continual Learning
by: Cossu, Andrea, et al.
Published: (2026)
by: Cossu, Andrea, et al.
Published: (2026)
Reliable Unlearning Harmful Information in LLMs with Metamorphosis Representation Projection
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
HGTUL: A Hypergraph-based Model For Trajectory User Linking
by: Chang, Fengjie, et al.
Published: (2025)
by: Chang, Fengjie, et al.
Published: (2025)
Is ReLU Adversarially Robust?
by: Sooksatra, Korn, et al.
Published: (2024)
by: Sooksatra, Korn, et al.
Published: (2024)
Multi-Level Fusion Graph Neural Network for Molecule Property Prediction
by: Liu, XiaYu, et al.
Published: (2025)
by: Liu, XiaYu, et al.
Published: (2025)
Concept Prerequisite Relation Prediction by Using Permutation-Equivariant Directed Graph Neural Networks
by: Qu, Xiran, et al.
Published: (2023)
by: Qu, Xiran, et al.
Published: (2023)
Analyzing Closed-loop Training Techniques for Realistic Traffic Agent Models in Autonomous Highway Driving Simulations
by: Bitzer, Matthias, et al.
Published: (2024)
by: Bitzer, Matthias, et al.
Published: (2024)
Don't Look Back in Anger: MAGIC Net for Streaming Continual Learning with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
by: Turan, Berkant, et al.
Published: (2025)
by: Turan, Berkant, et al.
Published: (2025)
ProactBench: Beyond What The User Asked For
by: Harfi, Sepehr, et al.
Published: (2026)
by: Harfi, Sepehr, et al.
Published: (2026)
Decentralized Time Series Classification with ROCKET Features
by: Casella, Bruno, et al.
Published: (2025)
by: Casella, Bruno, et al.
Published: (2025)
Matryoshka Policy Gradient for Entropy-Regularized RL: Convergence and Global Optimality
by: Ged, François, et al.
Published: (2023)
by: Ged, François, et al.
Published: (2023)
Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams
by: Henry, James
Published: (2026)
by: Henry, James
Published: (2026)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
by: Wesselink, Wieger, et al.
Published: (2025)
by: Wesselink, Wieger, et al.
Published: (2025)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
by: Henry, James
Published: (2026)
by: Henry, James
Published: (2026)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
by: Shafieinejad, Masoumeh, et al.
Published: (2026)
cPNN: Continuous Progressive Neural Networks for Evolving Streaming Time Series
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
by: Klačan, Ján, et al.
Published: (2026)
by: Klačan, Ján, et al.
Published: (2026)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Deep Learning-Based Forecasting of Boarding Patient Counts to Address ED Overcrowding
by: Vural, Orhun, et al.
Published: (2025)
by: Vural, Orhun, et al.
Published: (2025)
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
by: Brännvall, Rickard, et al.
Published: (2023)
by: Brännvall, Rickard, et al.
Published: (2023)
On the Origin of Algorithmic Progress in AI
by: Gundlach, Hans, et al.
Published: (2025)
by: Gundlach, Hans, et al.
Published: (2025)
Safe Continual Reinforcement Learning Methods for Nonstationary Environments. Towards a Survey of the State of the Art
by: Tomashevskiy, Timofey
Published: (2026)
by: Tomashevskiy, Timofey
Published: (2026)
multivariateGPT: a decoder-only transformer for multivariate categorical and numeric data
by: Loza, Andrew J., et al.
Published: (2025)
by: Loza, Andrew J., et al.
Published: (2025)
DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting
by: Mysore, Naveen
Published: (2026)
by: Mysore, Naveen
Published: (2026)
Scaling Laws in the Tiny Regime: How Small Models Change Their Mistakes
by: Alnemari, Mohammed, et al.
Published: (2026)
by: Alnemari, Mohammed, et al.
Published: (2026)
R3L: Relative Representations for Reinforcement Learning
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
Similar Items
-
Safety Margins for Reinforcement Learning
by: Grushin, Alexander, et al.
Published: (2023) -
Combining AI Control Systems and Human Decision Support via Robustness and Criticality
by: Woods, Walt, et al.
Published: (2024) -
Reliable Grid Forecasting: State Space Models for Safety-Critical Energy Systems
by: Hong, Sunki, et al.
Published: (2026) -
Upside Down Reinforcement Learning with Policy Generators
by: Di Ventura, Jacopo, et al.
Published: (2025) -
Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
by: Ricciardi, Antonio Pio, et al.
Published: (2025)