On Discovering Algorithms for Adversarial Imitation Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Chirra, Shashank Reddy, Teoh, Jayden, Paruchuri, Praveen, Varakantham, Pradeep |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Preserving the Privacy of Reward Functions in MDPs through Deception
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024)
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024)
Safety through feedback in Constrained RL
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024)
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
di: Hoang, Huy, et al.
Pubblicazione: (2024)
di: Hoang, Huy, et al.
Pubblicazione: (2024)
Imitating Cost-Constrained Behaviors in Reinforcement Learning
di: Shao, Qian, et al.
Pubblicazione: (2024)
di: Shao, Qian, et al.
Pubblicazione: (2024)
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
di: Hoang, Huy, et al.
Pubblicazione: (2023)
di: Hoang, Huy, et al.
Pubblicazione: (2023)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
di: Teoh, Jayden, et al.
Pubblicazione: (2025)
di: Teoh, Jayden, et al.
Pubblicazione: (2025)
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression
di: Ge, Zichang, et al.
Pubblicazione: (2025)
di: Ge, Zichang, et al.
Pubblicazione: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
di: Belaire, Roman, et al.
Pubblicazione: (2024)
di: Belaire, Roman, et al.
Pubblicazione: (2024)
Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
di: Teoh, Jayden, et al.
Pubblicazione: (2025)
di: Teoh, Jayden, et al.
Pubblicazione: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
di: Hoang, Huy, et al.
Pubblicazione: (2025)
di: Hoang, Huy, et al.
Pubblicazione: (2025)
Regret-Based Defense in Adversarial Reinforcement Learning
di: Belaire, Roman, et al.
Pubblicazione: (2023)
di: Belaire, Roman, et al.
Pubblicazione: (2023)
PyVRP$^+$: LLM-Driven Metacognitive Heuristic Evolution for Hybrid Genetic Search in Vehicle Routing Problems
di: Malik, Manuj, et al.
Pubblicazione: (2026)
di: Malik, Manuj, et al.
Pubblicazione: (2026)
Efficient Unsupervised Environment Design through Hierarchical Policy Representation Learning
di: Li, Dexun, et al.
Pubblicazione: (2026)
di: Li, Dexun, et al.
Pubblicazione: (2026)
Enhancing the Hierarchical Environment Design via Generative Trajectory Modeling
di: Li, Dexun, et al.
Pubblicazione: (2023)
di: Li, Dexun, et al.
Pubblicazione: (2023)
Offline Safe Reinforcement Learning Using Trajectory Classification
di: Gong, Ze, et al.
Pubblicazione: (2024)
di: Gong, Ze, et al.
Pubblicazione: (2024)
Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning
di: Lu, Yuxiao, et al.
Pubblicazione: (2023)
di: Lu, Yuxiao, et al.
Pubblicazione: (2023)
Offline Safe Policy Optimization From Heterogeneous Feedback
di: Gong, Ze, et al.
Pubblicazione: (2025)
di: Gong, Ze, et al.
Pubblicazione: (2025)
Unlocking Large Language Model's Planning Capabilities with Maximum Diversity Fine-tuning
di: Li, Wenjun, et al.
Pubblicazione: (2024)
di: Li, Wenjun, et al.
Pubblicazione: (2024)
Optimizing Ride-Pooling Operations with Extended Pickup and Drop-Off Flexibility
di: Jiang, Hao, et al.
Pubblicazione: (2025)
di: Jiang, Hao, et al.
Pubblicazione: (2025)
Automatic LLM Red Teaming
di: Belaire, Roman, et al.
Pubblicazione: (2025)
di: Belaire, Roman, et al.
Pubblicazione: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
di: Hoang, Huy, et al.
Pubblicazione: (2024)
di: Hoang, Huy, et al.
Pubblicazione: (2024)
Semantic Loss Guided Data Efficient Supervised Fine Tuning for Safe Responses in LLMs
di: Lu, Yuxiao, et al.
Pubblicazione: (2024)
di: Lu, Yuxiao, et al.
Pubblicazione: (2024)
A Factored MDP Approach To Moving Target Defense With Dynamic Threat Modeling and Cost Efficiency
di: Bose, Megha, et al.
Pubblicazione: (2024)
di: Bose, Megha, et al.
Pubblicazione: (2024)
EduQate: Generating Adaptive Curricula through RMABs in Education Settings
di: Tio, Sidney, et al.
Pubblicazione: (2024)
di: Tio, Sidney, et al.
Pubblicazione: (2024)
Towards Neural Network based Cognitive Models of Dynamic Decision-Making by Humans
di: Chen, Changyu, et al.
Pubblicazione: (2024)
di: Chen, Changyu, et al.
Pubblicazione: (2024)
Adversarial Imitation Learning via Boosting
di: Chang, Jonathan D., et al.
Pubblicazione: (2024)
di: Chang, Jonathan D., et al.
Pubblicazione: (2024)
Sample-efficient Adversarial Imitation Learning
di: Jung, Dahuin, et al.
Pubblicazione: (2023)
di: Jung, Dahuin, et al.
Pubblicazione: (2023)
Neuro-Symbolic Imitation Learning: Discovering Symbolic Abstractions for Skill Learning
di: Keller, Leon, et al.
Pubblicazione: (2025)
di: Keller, Leon, et al.
Pubblicazione: (2025)
Diffusion-Reward Adversarial Imitation Learning
di: Lai, Chun-Mao, et al.
Pubblicazione: (2024)
di: Lai, Chun-Mao, et al.
Pubblicazione: (2024)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
di: Jain, Gauri, et al.
Pubblicazione: (2024)
di: Jain, Gauri, et al.
Pubblicazione: (2024)
Model Predictive Adversarial Imitation Learning for Planning from Observation
di: Han, Tyler, et al.
Pubblicazione: (2025)
di: Han, Tyler, et al.
Pubblicazione: (2025)
The Imitation Game for Educational AI
di: Sonkar, Shashank, et al.
Pubblicazione: (2025)
di: Sonkar, Shashank, et al.
Pubblicazione: (2025)
Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs
di: Panfilov, Alexander, et al.
Pubblicazione: (2026)
di: Panfilov, Alexander, et al.
Pubblicazione: (2026)
Discovering Temporally-Aware Reinforcement Learning Algorithms
di: Jackson, Matthew Thomas, et al.
Pubblicazione: (2024)
di: Jackson, Matthew Thomas, et al.
Pubblicazione: (2024)
The Elicitation Game: Evaluating Capability Elicitation Techniques
di: Hofstätter, Felix, et al.
Pubblicazione: (2025)
di: Hofstätter, Felix, et al.
Pubblicazione: (2025)
Learning to Discover Iterative Spectral Algorithms
di: Liu, Zihang, et al.
Pubblicazione: (2026)
di: Liu, Zihang, et al.
Pubblicazione: (2026)
PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization
di: Ye, Yuyang, et al.
Pubblicazione: (2024)
di: Ye, Yuyang, et al.
Pubblicazione: (2024)
A Survey of Imitation Learning: Algorithms, Recent Developments, and Challenges
di: Zare, Maryam, et al.
Pubblicazione: (2023)
di: Zare, Maryam, et al.
Pubblicazione: (2023)
Imitation Game: A Model-based and Imitation Learning Deep Reinforcement Learning Hybrid
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
PCIA: A Path Construction Imitation Algorithm for Global Optimization
di: Rezaei, Mohammad-Javad, et al.
Pubblicazione: (2025)
di: Rezaei, Mohammad-Javad, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Preserving the Privacy of Reward Functions in MDPs through Deception
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024) -
Safety through feedback in Constrained RL
di: Chirra, Shashank Reddy, et al.
Pubblicazione: (2024) -
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
di: Hoang, Huy, et al.
Pubblicazione: (2024) -
Imitating Cost-Constrained Behaviors in Reinforcement Learning
di: Shao, Qian, et al.
Pubblicazione: (2024) -
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
di: Hoang, Huy, et al.
Pubblicazione: (2023)