Beyond Static Datasets: Robust Offline Policy Optimization via Vetted Synthetic Transitions
Fuente:
arXiv
Guardado en:
| Autores principales: | Agand, Pedram, Chen, Mo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TransforMerger: Transformer-based Voice-Gesture Fusion for Robust Human-Robot Communication
por: Vanc, Petr, et al.
Publicado: (2025)
por: Vanc, Petr, et al.
Publicado: (2025)
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
por: Farhadi, MH, et al.
Publicado: (2026)
por: Farhadi, MH, et al.
Publicado: (2026)
Sensory-Motor Control with Large Language Models via Iterative Policy Refinement
por: Carvalho, Jônata Tyska, et al.
Publicado: (2025)
por: Carvalho, Jônata Tyska, et al.
Publicado: (2025)
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations
por: Chen, Jingxiao, et al.
Publicado: (2025)
por: Chen, Jingxiao, et al.
Publicado: (2025)
Inference-Time Policy Steering through Human Interactions
por: Wang, Yanwei, et al.
Publicado: (2024)
por: Wang, Yanwei, et al.
Publicado: (2024)
Proximal State Nudging: Reducing Skill Atrophy from AI Assistance
por: Srivastava, Megha, et al.
Publicado: (2026)
por: Srivastava, Megha, et al.
Publicado: (2026)
Chat Failures and Troubles: Reasons and Solutions
por: Helal, Manal, et al.
Publicado: (2023)
por: Helal, Manal, et al.
Publicado: (2023)
Kinematically Constrained Human-like Bimanual Robot-to-Human Handovers
por: Göksu, Yasemin, et al.
Publicado: (2024)
por: Göksu, Yasemin, et al.
Publicado: (2024)
THÖR-MAGNI Act: Actions for Human Motion Modeling in Robot-Shared Industrial Spaces
por: de Almeida, Tiago Rodrigues, et al.
Publicado: (2024)
por: de Almeida, Tiago Rodrigues, et al.
Publicado: (2024)
MoVEInt: Mixture of Variational Experts for Learning Human-Robot Interactions from Demonstrations
por: Prasad, Vignesh, et al.
Publicado: (2024)
por: Prasad, Vignesh, et al.
Publicado: (2024)
NeRF-enabled Analysis-Through-Synthesis for ISAR Imaging of Small Everyday Objects with Sparse and Noisy UWB Radar Data
por: Oshim, Md Farhan Tasnim, et al.
Publicado: (2024)
por: Oshim, Md Farhan Tasnim, et al.
Publicado: (2024)
The Role of Embodiment in Intuitive Whole-Body Teleoperation for Mobile Manipulation
por: Moyen, Sophia Bianchi, et al.
Publicado: (2025)
por: Moyen, Sophia Bianchi, et al.
Publicado: (2025)
Development and Evaluation of a Learning-based Model for Real-time Haptic Texture Rendering
por: Heravi, Negin, et al.
Publicado: (2022)
por: Heravi, Negin, et al.
Publicado: (2022)
Learning Multimodal Latent Dynamics for Human-Robot Interaction
por: Prasad, Vignesh, et al.
Publicado: (2023)
por: Prasad, Vignesh, et al.
Publicado: (2023)
Siamese Network with Dual Attention for EEG-Driven Social Learning: Bridging the Human-Robot Gap in Long-Tail Autonomous Driving
por: Zhou, Xiaoshan, et al.
Publicado: (2025)
por: Zhou, Xiaoshan, et al.
Publicado: (2025)
Learning Multimodal Confidence for Intention Recognition in Human-Robot Interaction
por: Zhao, Xiyuan, et al.
Publicado: (2024)
por: Zhao, Xiyuan, et al.
Publicado: (2024)
Human Comfortability Index Estimation in Industrial Human-Robot Collaboration Task
por: Savur, Celal, et al.
Publicado: (2023)
por: Savur, Celal, et al.
Publicado: (2023)
CaptAinGlove: Capacitive and Inertial Fusion-Based Glove for Real-Time on Edge Hand Gesture Recognition for Drone Control
por: Bello, Hymalai, et al.
Publicado: (2023)
por: Bello, Hymalai, et al.
Publicado: (2023)
Towards Privacy-Aware and Personalised Assistive Robots: A User-Centred Approach
por: Casado, Fernando E.
Publicado: (2024)
por: Casado, Fernando E.
Publicado: (2024)
FRAC-Q-Learning: A Reinforcement Learning with Boredom Avoidance Processes for Social Robots
por: Onishi, Akinari
Publicado: (2023)
por: Onishi, Akinari
Publicado: (2023)
Classifying Subjective Time Perception in a Multi-robot Control Scenario Using Eye-tracking Information
por: Aust, Till, et al.
Publicado: (2025)
por: Aust, Till, et al.
Publicado: (2025)
Improving Low-Cost Teleoperation: Augmenting GELLO with Force
por: Sujit, Shivakanth, et al.
Publicado: (2025)
por: Sujit, Shivakanth, et al.
Publicado: (2025)
Online Learning of Human Constraints from Feedback in Shared Autonomy
por: Zhu, Shibei, et al.
Publicado: (2024)
por: Zhu, Shibei, et al.
Publicado: (2024)
Automating RT Planning at Scale: High Quality Data For AI Training
por: Gao, Riqiang, et al.
Publicado: (2025)
por: Gao, Riqiang, et al.
Publicado: (2025)
Using Causal Trees to Estimate Personalized Task Difficulty in Post-Stroke Individuals
por: Dennler, Nathaniel, et al.
Publicado: (2024)
por: Dennler, Nathaniel, et al.
Publicado: (2024)
Continuous ErrP detections during multimodal human-robot interaction
por: Kim, Su Kyoung, et al.
Publicado: (2022)
por: Kim, Su Kyoung, et al.
Publicado: (2022)
Collecting Human Motion Data in Large and Occlusion-Prone Environments using Ultra-Wideband Localization
por: Kaden, Janik, et al.
Publicado: (2025)
por: Kaden, Janik, et al.
Publicado: (2025)
Extended Reality for Enhanced Human-Robot Collaboration: a Human-in-the-Loop Approach
por: Karpichev, Yehor, et al.
Publicado: (2024)
por: Karpichev, Yehor, et al.
Publicado: (2024)
RAMPA: Robotic Augmented Reality for Machine Programming by DemonstrAtion
por: Dogangun, Fatih, et al.
Publicado: (2024)
por: Dogangun, Fatih, et al.
Publicado: (2024)
A Study on Domain Generalization for Failure Detection through Human Reactions in HRI
por: Parreira, Maria Teresa, et al.
Publicado: (2024)
por: Parreira, Maria Teresa, et al.
Publicado: (2024)
Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
por: Cheng, Xuxin, et al.
Publicado: (2024)
por: Cheng, Xuxin, et al.
Publicado: (2024)
HRI-SA: A Multimodal Dataset for Online Assessment of Human Situational Awareness during Remote Human-Robot Teaming
por: Senaratne, Hashini, et al.
Publicado: (2026)
por: Senaratne, Hashini, et al.
Publicado: (2026)
Offline Risk-sensitive RL with Partial Observability to Enhance Performance in Human-Robot Teaming
por: Angelotti, Giorgio, et al.
Publicado: (2024)
por: Angelotti, Giorgio, et al.
Publicado: (2024)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
por: Dennler, Nathaniel, et al.
Publicado: (2024)
por: Dennler, Nathaniel, et al.
Publicado: (2024)
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
por: Feith, Nikolaus, et al.
Publicado: (2024)
por: Feith, Nikolaus, et al.
Publicado: (2024)
ForceGrip: Reference-Free Curriculum Learning for Realistic Grip Force Control in VR Hand Manipulation
por: Han, DongHeun, et al.
Publicado: (2025)
por: Han, DongHeun, et al.
Publicado: (2025)
Towards Probabilistic Inference of Human Motor Intentions by Assistive Mobile Robots Controlled via a Brain-Computer Interface
por: Zhou, Xiaoshan, et al.
Publicado: (2025)
por: Zhou, Xiaoshan, et al.
Publicado: (2025)
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
por: Huang, Zilin, et al.
Publicado: (2024)
por: Huang, Zilin, et al.
Publicado: (2024)
Enabling Multi-Robot Collaboration from Single-Human Guidance
por: Ji, Zhengran, et al.
Publicado: (2024)
por: Ji, Zhengran, et al.
Publicado: (2024)
Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
por: Sasabuchi, Kazuhiro, et al.
Publicado: (2025)
por: Sasabuchi, Kazuhiro, et al.
Publicado: (2025)
Ejemplares similares
-
TransforMerger: Transformer-based Voice-Gesture Fusion for Robust Human-Robot Communication
por: Vanc, Petr, et al.
Publicado: (2025) -
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
por: Farhadi, MH, et al.
Publicado: (2026) -
Sensory-Motor Control with Large Language Models via Iterative Policy Refinement
por: Carvalho, Jônata Tyska, et al.
Publicado: (2025) -
RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations
por: Chen, Jingxiao, et al.
Publicado: (2025) -
Inference-Time Policy Steering through Human Interactions
por: Wang, Yanwei, et al.
Publicado: (2024)