Beyond Static Datasets: Robust Offline Policy Optimization via Vetted Synthetic Transitions
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Agand, Pedram, Chen, Mo |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TransforMerger: Transformer-based Voice-Gesture Fusion for Robust Human-Robot Communication
par: Vanc, Petr, et autres
Publié: (2025)
par: Vanc, Petr, et autres
Publié: (2025)
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
par: Farhadi, MH, et autres
Publié: (2026)
par: Farhadi, MH, et autres
Publié: (2026)
Sensory-Motor Control with Large Language Models via Iterative Policy Refinement
par: Carvalho, Jônata Tyska, et autres
Publié: (2025)
par: Carvalho, Jônata Tyska, et autres
Publié: (2025)
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations
par: Chen, Jingxiao, et autres
Publié: (2025)
par: Chen, Jingxiao, et autres
Publié: (2025)
Inference-Time Policy Steering through Human Interactions
par: Wang, Yanwei, et autres
Publié: (2024)
par: Wang, Yanwei, et autres
Publié: (2024)
Proximal State Nudging: Reducing Skill Atrophy from AI Assistance
par: Srivastava, Megha, et autres
Publié: (2026)
par: Srivastava, Megha, et autres
Publié: (2026)
Chat Failures and Troubles: Reasons and Solutions
par: Helal, Manal, et autres
Publié: (2023)
par: Helal, Manal, et autres
Publié: (2023)
Kinematically Constrained Human-like Bimanual Robot-to-Human Handovers
par: Göksu, Yasemin, et autres
Publié: (2024)
par: Göksu, Yasemin, et autres
Publié: (2024)
THÖR-MAGNI Act: Actions for Human Motion Modeling in Robot-Shared Industrial Spaces
par: de Almeida, Tiago Rodrigues, et autres
Publié: (2024)
par: de Almeida, Tiago Rodrigues, et autres
Publié: (2024)
MoVEInt: Mixture of Variational Experts for Learning Human-Robot Interactions from Demonstrations
par: Prasad, Vignesh, et autres
Publié: (2024)
par: Prasad, Vignesh, et autres
Publié: (2024)
NeRF-enabled Analysis-Through-Synthesis for ISAR Imaging of Small Everyday Objects with Sparse and Noisy UWB Radar Data
par: Oshim, Md Farhan Tasnim, et autres
Publié: (2024)
par: Oshim, Md Farhan Tasnim, et autres
Publié: (2024)
The Role of Embodiment in Intuitive Whole-Body Teleoperation for Mobile Manipulation
par: Moyen, Sophia Bianchi, et autres
Publié: (2025)
par: Moyen, Sophia Bianchi, et autres
Publié: (2025)
Development and Evaluation of a Learning-based Model for Real-time Haptic Texture Rendering
par: Heravi, Negin, et autres
Publié: (2022)
par: Heravi, Negin, et autres
Publié: (2022)
Learning Multimodal Latent Dynamics for Human-Robot Interaction
par: Prasad, Vignesh, et autres
Publié: (2023)
par: Prasad, Vignesh, et autres
Publié: (2023)
Siamese Network with Dual Attention for EEG-Driven Social Learning: Bridging the Human-Robot Gap in Long-Tail Autonomous Driving
par: Zhou, Xiaoshan, et autres
Publié: (2025)
par: Zhou, Xiaoshan, et autres
Publié: (2025)
Learning Multimodal Confidence for Intention Recognition in Human-Robot Interaction
par: Zhao, Xiyuan, et autres
Publié: (2024)
par: Zhao, Xiyuan, et autres
Publié: (2024)
Human Comfortability Index Estimation in Industrial Human-Robot Collaboration Task
par: Savur, Celal, et autres
Publié: (2023)
par: Savur, Celal, et autres
Publié: (2023)
CaptAinGlove: Capacitive and Inertial Fusion-Based Glove for Real-Time on Edge Hand Gesture Recognition for Drone Control
par: Bello, Hymalai, et autres
Publié: (2023)
par: Bello, Hymalai, et autres
Publié: (2023)
Towards Privacy-Aware and Personalised Assistive Robots: A User-Centred Approach
par: Casado, Fernando E.
Publié: (2024)
par: Casado, Fernando E.
Publié: (2024)
FRAC-Q-Learning: A Reinforcement Learning with Boredom Avoidance Processes for Social Robots
par: Onishi, Akinari
Publié: (2023)
par: Onishi, Akinari
Publié: (2023)
Classifying Subjective Time Perception in a Multi-robot Control Scenario Using Eye-tracking Information
par: Aust, Till, et autres
Publié: (2025)
par: Aust, Till, et autres
Publié: (2025)
Improving Low-Cost Teleoperation: Augmenting GELLO with Force
par: Sujit, Shivakanth, et autres
Publié: (2025)
par: Sujit, Shivakanth, et autres
Publié: (2025)
Online Learning of Human Constraints from Feedback in Shared Autonomy
par: Zhu, Shibei, et autres
Publié: (2024)
par: Zhu, Shibei, et autres
Publié: (2024)
Automating RT Planning at Scale: High Quality Data For AI Training
par: Gao, Riqiang, et autres
Publié: (2025)
par: Gao, Riqiang, et autres
Publié: (2025)
Using Causal Trees to Estimate Personalized Task Difficulty in Post-Stroke Individuals
par: Dennler, Nathaniel, et autres
Publié: (2024)
par: Dennler, Nathaniel, et autres
Publié: (2024)
Continuous ErrP detections during multimodal human-robot interaction
par: Kim, Su Kyoung, et autres
Publié: (2022)
par: Kim, Su Kyoung, et autres
Publié: (2022)
Collecting Human Motion Data in Large and Occlusion-Prone Environments using Ultra-Wideband Localization
par: Kaden, Janik, et autres
Publié: (2025)
par: Kaden, Janik, et autres
Publié: (2025)
Extended Reality for Enhanced Human-Robot Collaboration: a Human-in-the-Loop Approach
par: Karpichev, Yehor, et autres
Publié: (2024)
par: Karpichev, Yehor, et autres
Publié: (2024)
RAMPA: Robotic Augmented Reality for Machine Programming by DemonstrAtion
par: Dogangun, Fatih, et autres
Publié: (2024)
par: Dogangun, Fatih, et autres
Publié: (2024)
A Study on Domain Generalization for Failure Detection through Human Reactions in HRI
par: Parreira, Maria Teresa, et autres
Publié: (2024)
par: Parreira, Maria Teresa, et autres
Publié: (2024)
Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
par: Cheng, Xuxin, et autres
Publié: (2024)
par: Cheng, Xuxin, et autres
Publié: (2024)
HRI-SA: A Multimodal Dataset for Online Assessment of Human Situational Awareness during Remote Human-Robot Teaming
par: Senaratne, Hashini, et autres
Publié: (2026)
par: Senaratne, Hashini, et autres
Publié: (2026)
Offline Risk-sensitive RL with Partial Observability to Enhance Performance in Human-Robot Teaming
par: Angelotti, Giorgio, et autres
Publié: (2024)
par: Angelotti, Giorgio, et autres
Publié: (2024)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
par: Dennler, Nathaniel, et autres
Publié: (2024)
par: Dennler, Nathaniel, et autres
Publié: (2024)
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
par: Feith, Nikolaus, et autres
Publié: (2024)
par: Feith, Nikolaus, et autres
Publié: (2024)
ForceGrip: Reference-Free Curriculum Learning for Realistic Grip Force Control in VR Hand Manipulation
par: Han, DongHeun, et autres
Publié: (2025)
par: Han, DongHeun, et autres
Publié: (2025)
Towards Probabilistic Inference of Human Motor Intentions by Assistive Mobile Robots Controlled via a Brain-Computer Interface
par: Zhou, Xiaoshan, et autres
Publié: (2025)
par: Zhou, Xiaoshan, et autres
Publié: (2025)
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
par: Huang, Zilin, et autres
Publié: (2024)
par: Huang, Zilin, et autres
Publié: (2024)
Enabling Multi-Robot Collaboration from Single-Human Guidance
par: Ji, Zhengran, et autres
Publié: (2024)
par: Ji, Zhengran, et autres
Publié: (2024)
Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
par: Sasabuchi, Kazuhiro, et autres
Publié: (2025)
par: Sasabuchi, Kazuhiro, et autres
Publié: (2025)
Documents similaires
-
TransforMerger: Transformer-based Voice-Gesture Fusion for Robust Human-Robot Communication
par: Vanc, Petr, et autres
Publié: (2025) -
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
par: Farhadi, MH, et autres
Publié: (2026) -
Sensory-Motor Control with Large Language Models via Iterative Policy Refinement
par: Carvalho, Jônata Tyska, et autres
Publié: (2025) -
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations
par: Chen, Jingxiao, et autres
Publié: (2025) -
Inference-Time Policy Steering through Human Interactions
par: Wang, Yanwei, et autres
Publié: (2024)