A Smooth Sea Never Made a Skilled SAILOR: Robust Imitation via Learning to Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jain, Arnav Kumar, Mohta, Vibhakar, Kim, Subin, Bhardwaj, Atiksh, Ren, Juntao, Feng, Yunhai, Choudhury, Sanjiban, Swamy, Gokul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Imitation under Misspecification
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
Hybrid Inverse Reinforcement Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning without Reinforcement Learning
von: Swamy, Gokul, et al.
Veröffentlicht: (2023)
von: Swamy, Gokul, et al.
Veröffentlicht: (2023)
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2025)
von: Ren, Juntao, et al.
Veröffentlicht: (2025)
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
von: Swamy, Gokul, et al.
Veröffentlicht: (2025)
von: Swamy, Gokul, et al.
Veröffentlicht: (2025)
The Virtues of Pessimism in Inverse Reinforcement Learning
von: Wu, David, et al.
Veröffentlicht: (2024)
von: Wu, David, et al.
Veröffentlicht: (2024)
Process Reward Models for LLM Agents: Practical Framework and Directions
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
von: Choudhury, Sanjiban
Veröffentlicht: (2025)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
Distilling Realizable Students from Unrealizable Teachers
von: Kim, Yujin, et al.
Veröffentlicht: (2025)
von: Kim, Yujin, et al.
Veröffentlicht: (2025)
Multi-Turn Code Generation Through Single-Step Rewards
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2025)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2025)
Imitation Learning from a Single Temporally Misaligned Video
von: Huey, William, et al.
Veröffentlicht: (2025)
von: Huey, William, et al.
Veröffentlicht: (2025)
One-Shot Imitation under Mismatched Execution
von: Kedia, Kushal, et al.
Veröffentlicht: (2024)
von: Kedia, Kushal, et al.
Veröffentlicht: (2024)
Multi-Agent Imitation Learning: Value is Easy, Regret is Hard
von: Tang, Jingwu, et al.
Veröffentlicht: (2024)
von: Tang, Jingwu, et al.
Veröffentlicht: (2024)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
von: Choudhury, Sanjiban, et al.
Veröffentlicht: (2024)
von: Choudhury, Sanjiban, et al.
Veröffentlicht: (2024)
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
von: Wu, David, et al.
Veröffentlicht: (2024)
von: Wu, David, et al.
Veröffentlicht: (2024)
Aligning LLMs with Domain Invariant Reward Models
von: Wu, David, et al.
Veröffentlicht: (2025)
von: Wu, David, et al.
Veröffentlicht: (2025)
Imitation Learning via Focused Satisficing
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
EvIL: Evolution Strategies for Generalisable Imitation Learning
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
von: Sapora, Silvia, et al.
Veröffentlicht: (2024)
SAILOR: Perceptual Anchoring For Robotic Cognitive Architectures
von: González-Santamarta, Miguel Á., et al.
Veröffentlicht: (2023)
von: González-Santamarta, Miguel Á., et al.
Veröffentlicht: (2023)
In Search of the Never-Never
Veröffentlicht: (2019)
Veröffentlicht: (2019)
Neural Network-Driven Resume Skill Inflation Detection Using NLP and Source Code Repositories
von: Krishna Swamy, Tejesh Kumar, et al.
Veröffentlicht: (2026)
von: Krishna Swamy, Tejesh Kumar, et al.
Veröffentlicht: (2026)
MOSAIC: Modular Foundation Models for Assistive and Interactive Cooking
von: Wang, Huaxiaoyue, et al.
Veröffentlicht: (2024)
von: Wang, Huaxiaoyue, et al.
Veröffentlicht: (2024)
SAILOR: A Scalable and Energy-Efficient Ultra-Lightweight RISC-V for IoT Security
von: Ewert, Christian, et al.
Veröffentlicht: (2026)
von: Ewert, Christian, et al.
Veröffentlicht: (2026)
Query-Efficient Planning with Language Models
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2024)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2024)
Never-ending Search for Innovation
von: Benkert, Jean-Michel, et al.
Veröffentlicht: (2024)
von: Benkert, Jean-Michel, et al.
Veröffentlicht: (2024)
Never‐Ending Search for Innovation
von: Jean‐Michel Benkert, et al.
Veröffentlicht: (2025)
von: Jean‐Michel Benkert, et al.
Veröffentlicht: (2025)
Hybrid Grey Wolf Optimization and Cuckoo Search Algorithm for Extracting Unknown Parameters of PEM Fuel Cell
von: Banaja Mohanty, et al.
Veröffentlicht: (2026)
von: Banaja Mohanty, et al.
Veröffentlicht: (2026)
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Eight-shot measurement of spatially non-stationary complex coherence function
von: Mohta, Pranay, et al.
Veröffentlicht: (2024)
von: Mohta, Pranay, et al.
Veröffentlicht: (2024)
DEF-YOLO: Leveraging YOLO for Concealed Weapon Detection in Thermal Imagin
von: Bhardwaj, Divya, et al.
Veröffentlicht: (2025)
von: Bhardwaj, Divya, et al.
Veröffentlicht: (2025)
Hybrid BeeHive Algorithm: Proposed Ensemble Model for Multiclass Multi‐Label Ophthalmological Eye Diseases Prediction
von: Akanksha Bali, et al.
Veröffentlicht: (2026)
von: Akanksha Bali, et al.
Veröffentlicht: (2026)
Catalytic Hydrogenation of CO2 by Direct Air Capture to Valuable C1 Products Using Homogenous Catalysts
von: Ritu Bhardwaj, et al.
Veröffentlicht: (2025)
von: Ritu Bhardwaj, et al.
Veröffentlicht: (2025)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
Correction to “Never‐Ending Search for Innovation”
Veröffentlicht: (2025)
Veröffentlicht: (2025)
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
von: Sutawika, Lintang, et al.
Veröffentlicht: (2026)
Accelerated Smoothing: A Scalable Approach to Randomized Smoothing
von: Bhardwaj, Devansh, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Devansh, et al.
Veröffentlicht: (2024)
Learning Prehensile Dexterity by Imitating and Emulating State-only Observations
von: Han, Yunhai, et al.
Veröffentlicht: (2024)
von: Han, Yunhai, et al.
Veröffentlicht: (2024)
Noise-Resilient Imaging through Coherence Filtering
von: Mohta, Pranay, et al.
Veröffentlicht: (2026)
von: Mohta, Pranay, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Efficient Imitation under Misspecification
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025) -
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
von: Kedia, Kushal, et al.
Veröffentlicht: (2023) -
Hybrid Inverse Reinforcement Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2024) -
Inverse Reinforcement Learning without Reinforcement Learning
von: Swamy, Gokul, et al.
Veröffentlicht: (2023) -
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2025)