Learning from Observation: A Survey of Recent Advances
Fuente:
arXiv
Saved in:
| Main Authors: | Burnwal, Returaj, Mehta, Hriday, Bhatt, Nirav Pravinbhai, Ravindran, Balaraman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2025)
by: Burnwal, Returaj, et al.
Published: (2025)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2026)
by: Burnwal, Returaj, et al.
Published: (2026)
A Survey of Imitation Learning: Algorithms, Recent Developments, and Challenges
by: Zare, Maryam, et al.
Published: (2023)
by: Zare, Maryam, et al.
Published: (2023)
Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics
by: Acharjee, Jashaswimalya, et al.
Published: (2026)
by: Acharjee, Jashaswimalya, et al.
Published: (2026)
Generalized Adaptive Transfer Network: Enhancing Transfer Learning in Reinforcement Learning Across Domains
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
Imitation Learning from Observation through Optimal Transport
by: Chang, Wei-Di, et al.
Published: (2023)
by: Chang, Wei-Di, et al.
Published: (2023)
Imitation Learning from Observation with Automatic Discount Scheduling
by: Liu, Yuyang, et al.
Published: (2023)
by: Liu, Yuyang, et al.
Published: (2023)
CuRLA: Curriculum Learning Based Deep Reinforcement Learning for Autonomous Driving
by: Uppuluri, Bhargava, et al.
Published: (2025)
by: Uppuluri, Bhargava, et al.
Published: (2025)
Functional Groups are All you Need for Chemically Interpretable Molecular Property Prediction
by: Balaji, Roshan, et al.
Published: (2025)
by: Balaji, Roshan, et al.
Published: (2025)
SWAN: Sparse Winnowed Attention for Reduced Inference Memory via Decompression-Free KV-Cache Compression
by: S, Santhosh G, et al.
Published: (2025)
by: S, Santhosh G, et al.
Published: (2025)
AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs
by: S, Santhosh G, et al.
Published: (2025)
by: S, Santhosh G, et al.
Published: (2025)
PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment
by: Verma, Richa, et al.
Published: (2026)
by: Verma, Richa, et al.
Published: (2026)
MuJoCo MPC for Humanoid Control: Evaluation on HumanoidBench
by: Meser, Moritz, et al.
Published: (2024)
by: Meser, Moritz, et al.
Published: (2024)
Discrete Variational Autoencoding via Policy Search
by: Drolet, Michael, et al.
Published: (2025)
by: Drolet, Michael, et al.
Published: (2025)
CAnDOIT: Causal Discovery with Observational and Interventional Data from Time-Series
by: Castri, Luca, et al.
Published: (2024)
by: Castri, Luca, et al.
Published: (2024)
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability
by: Wang, Kevin, et al.
Published: (2024)
by: Wang, Kevin, et al.
Published: (2024)
Neural-Network-Driven Reward Prediction as a Heuristic: Advancing Q-Learning for Mobile Robot Path Planning
by: Ji, Yiming, et al.
Published: (2024)
by: Ji, Yiming, et al.
Published: (2024)
A Comprehensive Survey of Cross-Domain Policy Transfer for Embodied Agents
by: Niu, Haoyi, et al.
Published: (2024)
by: Niu, Haoyi, et al.
Published: (2024)
A Survey on Approximate Edge AI for Energy Efficient Autonomous Driving Services
by: Katare, Dewant, et al.
Published: (2023)
by: Katare, Dewant, et al.
Published: (2023)
Learning Control Barrier Functions and their application in Reinforcement Learning: A Survey
by: Guerrier, Maeva, et al.
Published: (2024)
by: Guerrier, Maeva, et al.
Published: (2024)
World Models for Autonomous Driving: An Initial Survey
by: Guan, Yanchen, et al.
Published: (2024)
by: Guan, Yanchen, et al.
Published: (2024)
An energy-efficient learning solution for the Agile Earth Observation Satellite Scheduling Problem
by: Mercado-Martínez, Antonio M., et al.
Published: (2025)
by: Mercado-Martínez, Antonio M., et al.
Published: (2025)
A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics
by: Kojima, Takeshi, et al.
Published: (2025)
by: Kojima, Takeshi, et al.
Published: (2025)
Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI
by: Shu, Bo, et al.
Published: (2024)
by: Shu, Bo, et al.
Published: (2024)
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
by: Zhang, Yunuo, et al.
Published: (2025)
by: Zhang, Yunuo, et al.
Published: (2025)
ExBody2: Advanced Expressive Humanoid Whole-Body Control
by: Ji, Mazeyu, et al.
Published: (2024)
by: Ji, Mazeyu, et al.
Published: (2024)
Recent Advances of Multimodal Continual Learning: A Comprehensive Survey
by: Yu, Dianzhi, et al.
Published: (2024)
by: Yu, Dianzhi, et al.
Published: (2024)
Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
by: Nakanishi, Kosuke, et al.
Published: (2025)
by: Nakanishi, Kosuke, et al.
Published: (2025)
A Survey of Machine Learning Techniques for Improving Global Navigation Satellite Systems
by: Mohanty, Adyasha, et al.
Published: (2024)
by: Mohanty, Adyasha, et al.
Published: (2024)
Inertial Navigation Meets Deep Learning: A Survey of Current Trends and Future Directions
by: Cohen, Nadav, et al.
Published: (2023)
by: Cohen, Nadav, et al.
Published: (2023)
Learning Parameterized Skills from Demonstrations
by: Gupta, Vedant, et al.
Published: (2025)
by: Gupta, Vedant, et al.
Published: (2025)
Lessons from Learning to Spin "Pens"
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Learning Constraint Network from Demonstrations via Positive-Unlabeled Learning with Memory Replay
by: Peng, Baiyu, et al.
Published: (2024)
by: Peng, Baiyu, et al.
Published: (2024)
Predictive Preference Learning from Human Interventions
by: Cai, Haoyuan, et al.
Published: (2025)
by: Cai, Haoyuan, et al.
Published: (2025)
Learning Adaptive Dexterous Grasping from Single Demonstrations
by: Shi, Liangzhi, et al.
Published: (2025)
by: Shi, Liangzhi, et al.
Published: (2025)
Learning Novel Skills from Language-Generated Demonstrations
by: Jin, Ao-Qun, et al.
Published: (2024)
by: Jin, Ao-Qun, et al.
Published: (2024)
Adaptive Querying for Reward Learning from Human Feedback
by: Anand, Yashwanthi, et al.
Published: (2024)
by: Anand, Yashwanthi, et al.
Published: (2024)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
Similar Items
-
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2025) -
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2026) -
A Survey of Imitation Learning: Algorithms, Recent Developments, and Challenges
by: Zare, Maryam, et al.
Published: (2023) -
Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics
by: Acharjee, Jashaswimalya, et al.
Published: (2026) -
Generalized Adaptive Transfer Network: Enhancing Transfer Learning in Reinforcement Learning Across Domains
by: Verma, Abhishek, et al.
Published: (2025)