Gespeichert in:
| Hauptverfasser: | Wu, Haochen, Sharma, Shubham, Patra, Sunandita, Gopalakrishnan, Sriram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2308.12367 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
QBD-RankedDataGen: Generating Custom Ranked Datasets for Improving Query-By-Document Search Using LLM-Reranking with Reduced Human Effort
von: Gopalakrishnan, Sriram, et al.
Veröffentlicht: (2025)
von: Gopalakrishnan, Sriram, et al.
Veröffentlicht: (2025)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
The Importance of Time in Causal Algorithmic Recourse
von: Beretta, Isacco, et al.
Veröffentlicht: (2023)
von: Beretta, Isacco, et al.
Veröffentlicht: (2023)
Reinforcement Learning for Durable Algorithmic Recourse
von: Ceccon, Marina, et al.
Veröffentlicht: (2025)
von: Ceccon, Marina, et al.
Veröffentlicht: (2025)
Personalized Algorithmic Recourse with Preference Elicitation
von: De Toni, Giovanni, et al.
Veröffentlicht: (2022)
von: De Toni, Giovanni, et al.
Veröffentlicht: (2022)
Causal Algorithmic Recourse: Foundations and Methods
von: Plecko, Drago, et al.
Veröffentlicht: (2026)
von: Plecko, Drago, et al.
Veröffentlicht: (2026)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
Safe Deep Policy Adaptation
von: Xiao, Wenli, et al.
Veröffentlicht: (2023)
von: Xiao, Wenli, et al.
Veröffentlicht: (2023)
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Verification-Guided Falsification for Safe RL via Explainable Abstraction and Risk-Aware Exploration
von: Le, Tuan, et al.
Veröffentlicht: (2025)
von: Le, Tuan, et al.
Veröffentlicht: (2025)
Rating Multi-Modal Time-Series Forecasting Models (MM-TSFM) for Robustness Through a Causal Lens
von: Lakkaraju, Kausik, et al.
Veröffentlicht: (2024)
von: Lakkaraju, Kausik, et al.
Veröffentlicht: (2024)
From Universal to Individualized Actionability: Revisiting Personalization in Algorithmic Recourse
von: Budde, Lena Marie, et al.
Veröffentlicht: (2026)
von: Budde, Lena Marie, et al.
Veröffentlicht: (2026)
Skill-based Safe Reinforcement Learning with Risk Planning
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
von: Kim, Dohyeong, et al.
Veröffentlicht: (2024)
von: Kim, Dohyeong, et al.
Veröffentlicht: (2024)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
von: Yao, Yihang, et al.
Veröffentlicht: (2023)
von: Yao, Yihang, et al.
Veröffentlicht: (2023)
Deep SPI: Safe Policy Improvement via World Models
von: Delgrange, Florent, et al.
Veröffentlicht: (2025)
von: Delgrange, Florent, et al.
Veröffentlicht: (2025)
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
von: Chemingui, Yassine, et al.
Veröffentlicht: (2024)
von: Chemingui, Yassine, et al.
Veröffentlicht: (2024)
Creating a Causally Grounded Rating Method for Assessing the Robustness of AI Models for Time-Series Forecasting
von: Lakkaraju, Kausik, et al.
Veröffentlicht: (2025)
von: Lakkaraju, Kausik, et al.
Veröffentlicht: (2025)
Ethics-Aware Safe Reinforcement Learning for Rare-Event Risk Control in Interactive Urban Driving
von: Li, Dianzhao, et al.
Veröffentlicht: (2025)
von: Li, Dianzhao, et al.
Veröffentlicht: (2025)
Revisiting Safe Exploration in Safe Reinforcement learning
von: Eckel, David, et al.
Veröffentlicht: (2024)
von: Eckel, David, et al.
Veröffentlicht: (2024)
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
von: Chen, Keru, et al.
Veröffentlicht: (2024)
von: Chen, Keru, et al.
Veröffentlicht: (2024)
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
von: Wei, Zeming, et al.
Veröffentlicht: (2026)
von: Wei, Zeming, et al.
Veröffentlicht: (2026)
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
Pareto Optimal Algorithmic Recourse in Multi-cost Function
von: Chen, Wen-Ling, et al.
Veröffentlicht: (2025)
von: Chen, Wen-Ling, et al.
Veröffentlicht: (2025)
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks
von: Geng, Xue, et al.
Veröffentlicht: (2024)
von: Geng, Xue, et al.
Veröffentlicht: (2024)
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2026)
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2026)
Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
von: Ji, Jiaming, et al.
Veröffentlicht: (2025)
von: Ji, Jiaming, et al.
Veröffentlicht: (2025)
CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning
von: Narava, Rahul, et al.
Veröffentlicht: (2026)
von: Narava, Rahul, et al.
Veröffentlicht: (2026)
Iterative Batch Reinforcement Learning via Safe Diversified Model-based Policy Search
von: Najib, Amna, et al.
Veröffentlicht: (2024)
von: Najib, Amna, et al.
Veröffentlicht: (2024)
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2025)
Verified Safe Reinforcement Learning for Neural Network Dynamic Models
von: Wu, Junlin, et al.
Veröffentlicht: (2024)
von: Wu, Junlin, et al.
Veröffentlicht: (2024)
Personalized Path Recourse for Reinforcement Learning Agents
von: Hong, Dat, et al.
Veröffentlicht: (2023)
von: Hong, Dat, et al.
Veröffentlicht: (2023)
Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data
von: Xue, Ruiqi, et al.
Veröffentlicht: (2026)
von: Xue, Ruiqi, et al.
Veröffentlicht: (2026)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
LIBRA: Language Model Informed Bandit Recourse Algorithm for Personalized Treatment Planning
von: Cao, Junyu, et al.
Veröffentlicht: (2026)
von: Cao, Junyu, et al.
Veröffentlicht: (2026)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Information-Theoretic Safe Bayesian Optimization
von: Bottero, Alessandro G., et al.
Veröffentlicht: (2024)
von: Bottero, Alessandro G., et al.
Veröffentlicht: (2024)
On the Mathematical Impossibility of Safe Universal Approximators
von: Yao, Jasper
Veröffentlicht: (2025)
von: Yao, Jasper
Veröffentlicht: (2025)
Ähnliche Einträge
-
QBD-RankedDataGen: Generating Custom Ranked Datasets for Improving Query-By-Document Search Using LLM-Reranking with Reduced Human Effort
von: Gopalakrishnan, Sriram, et al.
Veröffentlicht: (2025) -
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026) -
The Importance of Time in Causal Algorithmic Recourse
von: Beretta, Isacco, et al.
Veröffentlicht: (2023) -
Reinforcement Learning for Durable Algorithmic Recourse
von: Ceccon, Marina, et al.
Veröffentlicht: (2025) -
Personalized Algorithmic Recourse with Preference Elicitation
von: De Toni, Giovanni, et al.
Veröffentlicht: (2022)