Efficient Imitation Without Demonstrations via Value-Penalized Auxiliary Control from Examples
Fuente:
arXiv
Guardado en:
| Autores principales: | Ablett, Trevor, Chan, Bryan, Wang, Jayce Haoran, Kelly, Jonathan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multimodal and Force-Matched Imitation Learning with a See-Through Visuotactile Sensor
por: Ablett, Trevor, et al.
Publicado: (2023)
por: Ablett, Trevor, et al.
Publicado: (2023)
Working Backwards: Learning to Place by Picking
por: Limoyo, Oliver, et al.
Publicado: (2023)
por: Limoyo, Oliver, et al.
Publicado: (2023)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
por: Chan, Bryan, et al.
Publicado: (2024)
por: Chan, Bryan, et al.
Publicado: (2024)
Sample-Efficient Expert Query Control in Active Imitation Learning via Conformal Prediction
por: Firouzkouhi, Arad, et al.
Publicado: (2025)
por: Firouzkouhi, Arad, et al.
Publicado: (2025)
Demonstration-Free Robotic Control via LLM Agents
por: Tsui, Brian Y., et al.
Publicado: (2026)
por: Tsui, Brian Y., et al.
Publicado: (2026)
Tube-NeRF: Efficient Imitation Learning of Visuomotor Policies from MPC using Tube-Guided Data Augmentation and NeRFs
por: Tagliabue, Andrea, et al.
Publicado: (2023)
por: Tagliabue, Andrea, et al.
Publicado: (2023)
TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
por: Zhuang, Zifeng, et al.
Publicado: (2025)
por: Zhuang, Zifeng, et al.
Publicado: (2025)
An Integrated Imitation and Reinforcement Learning Methodology for Robust Agile Aircraft Control with Limited Pilot Demonstration Data
por: Sever, Gulay Goktas, et al.
Publicado: (2023)
por: Sever, Gulay Goktas, et al.
Publicado: (2023)
SoftMimic: Learning Compliant Whole-body Control from Examples
por: Margolis, Gabriel B., et al.
Publicado: (2025)
por: Margolis, Gabriel B., et al.
Publicado: (2025)
Reinforcement Learning via Auxiliary Task Distillation
por: Harish, Abhinav Narayan, et al.
Publicado: (2024)
por: Harish, Abhinav Narayan, et al.
Publicado: (2024)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
por: Xu, Chen, et al.
Publicado: (2025)
por: Xu, Chen, et al.
Publicado: (2025)
Generalization Capability for Imitation Learning
por: Wang, Yixiao
Publicado: (2025)
por: Wang, Yixiao
Publicado: (2025)
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
por: Cho, Seongwoong, et al.
Publicado: (2024)
por: Cho, Seongwoong, et al.
Publicado: (2024)
Cross-Domain Imitation Learning via Optimal Transport
por: Fickinger, Arnaud, et al.
Publicado: (2021)
por: Fickinger, Arnaud, et al.
Publicado: (2021)
Reverse Forward Curriculum Learning for Extreme Sample and Demonstration Efficiency in Reinforcement Learning
por: Tao, Stone, et al.
Publicado: (2024)
por: Tao, Stone, et al.
Publicado: (2024)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
por: Guo, Yihong, et al.
Publicado: (2024)
por: Guo, Yihong, et al.
Publicado: (2024)
PRIME: Scaffolding Manipulation Tasks with Behavior Primitives for Data-Efficient Imitation Learning
por: Gao, Tian, et al.
Publicado: (2024)
por: Gao, Tian, et al.
Publicado: (2024)
Diffusion-Reward Adversarial Imitation Learning
por: Lai, Chun-Mao, et al.
Publicado: (2024)
por: Lai, Chun-Mao, et al.
Publicado: (2024)
Error-Feedback Model for Output Correction in Bilateral Control-Based Imitation Learning
por: Sato, Hiroshi, et al.
Publicado: (2024)
por: Sato, Hiroshi, et al.
Publicado: (2024)
SENSOR: Imitate Third-Person Expert's Behaviors via Active Sensoring
por: Huang, Kaichen, et al.
Publicado: (2024)
por: Huang, Kaichen, et al.
Publicado: (2024)
RILe: Reinforced Imitation Learning
por: Albaba, Mert, et al.
Publicado: (2024)
por: Albaba, Mert, et al.
Publicado: (2024)
Imitation Learning from Observation through Optimal Transport
por: Chang, Wei-Di, et al.
Publicado: (2023)
por: Chang, Wei-Di, et al.
Publicado: (2023)
Imitation Learning from Observation with Automatic Discount Scheduling
por: Liu, Yuyang, et al.
Publicado: (2023)
por: Liu, Yuyang, et al.
Publicado: (2023)
Learning Constraint Network from Demonstrations via Positive-Unlabeled Learning with Memory Replay
por: Peng, Baiyu, et al.
Publicado: (2024)
por: Peng, Baiyu, et al.
Publicado: (2024)
Learning Parameterized Skills from Demonstrations
por: Gupta, Vedant, et al.
Publicado: (2025)
por: Gupta, Vedant, et al.
Publicado: (2025)
Exploiting Contextual Structure to Generate Useful Auxiliary Tasks
por: Quartey, Benedict, et al.
Publicado: (2023)
por: Quartey, Benedict, et al.
Publicado: (2023)
DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation Learning
por: Wan, Weikang, et al.
Publicado: (2024)
por: Wan, Weikang, et al.
Publicado: (2024)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
RIZE: Adaptive Regularization for Imitation Learning
por: Karimi, Adib, et al.
Publicado: (2025)
por: Karimi, Adib, et al.
Publicado: (2025)
Learning to Drive by Imitating Surrounding Vehicles
por: Sonmez, Yasin, et al.
Publicado: (2025)
por: Sonmez, Yasin, et al.
Publicado: (2025)
Learning Novel Skills from Language-Generated Demonstrations
por: Jin, Ao-Qun, et al.
Publicado: (2024)
por: Jin, Ao-Qun, et al.
Publicado: (2024)
LodeStar: Long-horizon Dexterity via Synthetic Data Augmentation from Human Demonstrations
por: Wan, Weikang, et al.
Publicado: (2025)
por: Wan, Weikang, et al.
Publicado: (2025)
Generalized Animal Imitator: Agile Locomotion with Versatile Motion Prior
por: Yang, Ruihan, et al.
Publicado: (2023)
por: Yang, Ruihan, et al.
Publicado: (2023)
Efficient and Generalized end-to-end Autonomous Driving System with Latent Deep Reinforcement Learning and Demonstrations
por: Tang, Zuojin, et al.
Publicado: (2024)
por: Tang, Zuojin, et al.
Publicado: (2024)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
por: Yamada, Jun, et al.
Publicado: (2025)
por: Yamada, Jun, et al.
Publicado: (2025)
CAML: Collaborative Auxiliary Modality Learning for Multi-Agent Systems
por: Liu, Rui, et al.
Publicado: (2025)
por: Liu, Rui, et al.
Publicado: (2025)
Offline Diversity Maximization Under Imitation Constraints
por: Vlastelica, Marin, et al.
Publicado: (2023)
por: Vlastelica, Marin, et al.
Publicado: (2023)
Memory-Consistent Neural Networks for Imitation Learning
por: Sridhar, Kaustubh, et al.
Publicado: (2023)
por: Sridhar, Kaustubh, et al.
Publicado: (2023)
One-Shot Imitation under Mismatched Execution
por: Kedia, Kushal, et al.
Publicado: (2024)
por: Kedia, Kushal, et al.
Publicado: (2024)
Unsupervised Motion Retargeting for Human-Robot Imitation
por: Annabi, Louis, et al.
Publicado: (2024)
por: Annabi, Louis, et al.
Publicado: (2024)
Ejemplares similares
-
Multimodal and Force-Matched Imitation Learning with a See-Through Visuotactile Sensor
por: Ablett, Trevor, et al.
Publicado: (2023) -
Working Backwards: Learning to Place by Picking
por: Limoyo, Oliver, et al.
Publicado: (2023) -
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
por: Chan, Bryan, et al.
Publicado: (2024) -
Sample-Efficient Expert Query Control in Active Imitation Learning via Conformal Prediction
por: Firouzkouhi, Arad, et al.
Publicado: (2025) -
Demonstration-Free Robotic Control via LLM Agents
por: Tsui, Brian Y., et al.
Publicado: (2026)