MILE: Model-based Intervention Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Korkmaz, Yigit, Bıyık, Erdem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
von: Hwang, Minjune, et al.
Veröffentlicht: (2026)
von: Hwang, Minjune, et al.
Veröffentlicht: (2026)
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
von: Korkmaz, Yigit, et al.
Veröffentlicht: (2025)
von: Korkmaz, Yigit, et al.
Veröffentlicht: (2025)
When a Robot is More Capable than a Human: Learning from Constrained Demonstrators
von: Li, Xinhu, et al.
Veröffentlicht: (2025)
von: Li, Xinhu, et al.
Veröffentlicht: (2025)
Batch Active Learning of Reward Functions from Human Preferences
von: Bıyık, Erdem, et al.
Veröffentlicht: (2024)
von: Bıyık, Erdem, et al.
Veröffentlicht: (2024)
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
A Generalized Acquisition Function for Preference-based Reward Learning
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
von: Ellis, Evan, et al.
Veröffentlicht: (2024)
CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
IMPACT: Intelligent Motion Planning with Acceptable Contact Trajectories via Vision-Language Models
von: Ling, Yiyang, et al.
Veröffentlicht: (2025)
von: Ling, Yiyang, et al.
Veröffentlicht: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
Multi-Agent Inverse Q-Learning from Demonstrations
von: Haynam, Nathaniel, et al.
Veröffentlicht: (2025)
von: Haynam, Nathaniel, et al.
Veröffentlicht: (2025)
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
von: Zhang, Jesse, et al.
Veröffentlicht: (2024)
von: Zhang, Jesse, et al.
Veröffentlicht: (2024)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
von: Jain, Ayush, et al.
Veröffentlicht: (2024)
von: Jain, Ayush, et al.
Veröffentlicht: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
Developing Driving Strategies Efficiently: A Skill-Based Hierarchical Reinforcement Learning Approach
von: Gurses, Yigit, et al.
Veröffentlicht: (2023)
von: Gurses, Yigit, et al.
Veröffentlicht: (2023)
ViSaRL: Visual Reinforcement Learning Guided by Human Saliency
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
von: Liang, Anthony, et al.
Veröffentlicht: (2024)
Object and Relation Centric Representations for Push Effect Prediction
von: Tekden, Ahmet E., et al.
Veröffentlicht: (2021)
von: Tekden, Ahmet E., et al.
Veröffentlicht: (2021)
Predictive Preference Learning from Human Interventions
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
Data-Efficient Learning from Human Interventions for Mobile Robots
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
Robot-Gated Interactive Imitation Learning with Adaptive Intervention Mechanism
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
Distilling On-device Language Models for Robot Planning with Minimal Human Intervention
von: Ravichandran, Zachary, et al.
Veröffentlicht: (2025)
von: Ravichandran, Zachary, et al.
Veröffentlicht: (2025)
Neuro-Symbolic Skill Discovery for Conditional Multi-Level Planning
von: Aktas, Hakan, et al.
Veröffentlicht: (2024)
von: Aktas, Hakan, et al.
Veröffentlicht: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
von: Cao, Chenyang, et al.
Veröffentlicht: (2025)
Value Explicit Pretraining for Learning Transferable Representations
von: Lekkala, Kiran, et al.
Veröffentlicht: (2023)
von: Lekkala, Kiran, et al.
Veröffentlicht: (2023)
Innate-Values-driven Reinforcement Learning based Cognitive Modeling
von: Yang, Qin
Veröffentlicht: (2024)
von: Yang, Qin
Veröffentlicht: (2024)
Interactive Double Deep Q-network: Integrating Human Interventions and Evaluative Predictions in Reinforcement Learning of Autonomous Driving
von: Sygkounas, Alkis, et al.
Veröffentlicht: (2025)
von: Sygkounas, Alkis, et al.
Veröffentlicht: (2025)
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning
von: Lee, Dongsu, et al.
Veröffentlicht: (2025)
von: Lee, Dongsu, et al.
Veröffentlicht: (2025)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
von: Lee, Vint, et al.
Veröffentlicht: (2023)
von: Lee, Vint, et al.
Veröffentlicht: (2023)
MOTO: Offline Pre-training to Online Fine-tuning for Model-based Robot Learning
von: Rafailov, Rafael, et al.
Veröffentlicht: (2024)
von: Rafailov, Rafael, et al.
Veröffentlicht: (2024)
WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
Rating-based Reinforcement Learning
von: White, Devin, et al.
Veröffentlicht: (2023)
von: White, Devin, et al.
Veröffentlicht: (2023)
Coprocessor Actor Critic: A Model-Based Reinforcement Learning Approach For Adaptive Brain Stimulation
von: Pan, Michelle, et al.
Veröffentlicht: (2024)
von: Pan, Michelle, et al.
Veröffentlicht: (2024)
CAnDOIT: Causal Discovery with Observational and Interventional Data from Time-Series
von: Castri, Luca, et al.
Veröffentlicht: (2024)
von: Castri, Luca, et al.
Veröffentlicht: (2024)
HeteroMILE: a Multi-Level Graph Representation Learning Framework for Heterogeneous Graphs
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
Learning-based Autonomous Oversteer Control and Collision Avoidance
von: Lee, Seokjun, et al.
Veröffentlicht: (2025)
von: Lee, Seokjun, et al.
Veröffentlicht: (2025)
Research on Autonomous Robots Navigation based on Reinforcement Learning
von: Wang, Zixiang, et al.
Veröffentlicht: (2024)
von: Wang, Zixiang, et al.
Veröffentlicht: (2024)
RAILGUN: A Unified Convolutional Policy for Multi-Agent Path Finding Across Different Environments and Tasks
von: Tang, Yimin, et al.
Veröffentlicht: (2025)
von: Tang, Yimin, et al.
Veröffentlicht: (2025)
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
Ark: An Open-source Python-based Framework for Robot Learning
von: Dierking, Magnus, et al.
Veröffentlicht: (2025)
von: Dierking, Magnus, et al.
Veröffentlicht: (2025)
Flow-based Domain Randomization for Learning and Sequencing Robotic Skills
von: Curtis, Aidan, et al.
Veröffentlicht: (2025)
von: Curtis, Aidan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
von: Hwang, Minjune, et al.
Veröffentlicht: (2026) -
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
von: Korkmaz, Yigit, et al.
Veröffentlicht: (2025) -
When a Robot is More Capable than a Human: Learning from Constrained Demonstrators
von: Li, Xinhu, et al.
Veröffentlicht: (2025) -
Batch Active Learning of Reward Functions from Human Preferences
von: Bıyık, Erdem, et al.
Veröffentlicht: (2024) -
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)