RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms
Fuente:
arXiv
Saved in:
| Main Authors: | Asri, Zakariae El, Laiche, Ibrahim, Rambour, Clément, Sigaud, Olivier, Thome, Nicolas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
by: Asri, Zakariae El, et al.
Published: (2024)
by: Asri, Zakariae El, et al.
Published: (2024)
Energy Correction Model in the Feature Space for Out-of-Distribution Detection
by: Lafon, Marc, et al.
Published: (2024)
by: Lafon, Marc, et al.
Published: (2024)
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
by: Aissi, Mohamed Salim, et al.
Published: (2024)
by: Aissi, Mohamed Salim, et al.
Published: (2024)
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
by: Castanet, Nicolas, et al.
Published: (2025)
by: Castanet, Nicolas, et al.
Published: (2025)
Learning Physically Consistent Lagrangian Control Models Without Acceleration Measurements
by: Laiche, Ibrahim, et al.
Published: (2025)
by: Laiche, Ibrahim, et al.
Published: (2025)
VIPER: Visual Perception and Explainable Reasoning for Sequential Decision-Making
by: Aissi, Mohamed Salim, et al.
Published: (2025)
by: Aissi, Mohamed Salim, et al.
Published: (2025)
Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning
by: Carta, Thomas, et al.
Published: (2023)
by: Carta, Thomas, et al.
Published: (2023)
A tale of two goals: leveraging sequentiality in multi-goal scenarios
by: Serris, Olivier, et al.
Published: (2025)
by: Serris, Olivier, et al.
Published: (2025)
HERAKLES: Hierarchical Skill Compilation for Open-ended LLM Agents
by: Carta, Thomas, et al.
Published: (2025)
by: Carta, Thomas, et al.
Published: (2025)
GalLoP: Learning Global and Local Prompts for Vision-Language Models
by: Lafon, Marc, et al.
Published: (2024)
by: Lafon, Marc, et al.
Published: (2024)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
by: Colas, Cédric, et al.
Published: (2020)
by: Colas, Cédric, et al.
Published: (2020)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024)
by: Gaven, Loris, et al.
Published: (2024)
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
by: Couairon, Paul, et al.
Published: (2023)
by: Couairon, Paul, et al.
Published: (2023)
Counterfactual Inference under Thompson Sampling
by: Jeunen, Olivier
Published: (2025)
by: Jeunen, Olivier
Published: (2025)
Directly Forecasting Belief for Reinforcement Learning with Delays
by: Wu, Qingyuan, et al.
Published: (2025)
by: Wu, Qingyuan, et al.
Published: (2025)
Temporal receptive field in dynamic graph learning: A comprehensive analysis
by: Karmim, Yannis, et al.
Published: (2024)
by: Karmim, Yannis, et al.
Published: (2024)
Dealing with Synthetic Data Contamination in Online Continual Learning
by: Wang, Maorong, et al.
Published: (2024)
by: Wang, Maorong, et al.
Published: (2024)
A Human-Centered Privacy Approach (HCP) to AI
by: Sun, Luyi, et al.
Published: (2026)
by: Sun, Luyi, et al.
Published: (2026)
DAFTED: Decoupled Asymmetric Fusion of Tabular and Echocardiographic Data for Cardiac Hypertension Diagnosis
by: Stym-Popper, Jérémie, et al.
Published: (2025)
by: Stym-Popper, Jérémie, et al.
Published: (2025)
SemEnrich: Self-Supervised Semantic Enrichment of Radiology Reports for Vision-Language Learning
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
by: Gulluk, Halil Ibrahim, et al.
Published: (2026)
Temperature-Aware Recurrent Neural Operator for Temperature-Dependent Anisotropic Plasticity in HCP Materials
by: Hollenweger, Yannick, et al.
Published: (2025)
by: Hollenweger, Yannick, et al.
Published: (2025)
Assessing Per-Sample Membership Inference Vulnerability without Retraining
by: Dorseuil, Valentin, et al.
Published: (2026)
by: Dorseuil, Valentin, et al.
Published: (2026)
Revisiting the Learning Objectives of Vision-Language Reward Models
by: Roy, Simon, et al.
Published: (2025)
by: Roy, Simon, et al.
Published: (2025)
Deal: Distributed End-to-End GNN Inference for All Nodes
by: Chen, Shiyang, et al.
Published: (2025)
by: Chen, Shiyang, et al.
Published: (2025)
Hardware and Software Platform Inference
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
The Sampling Complexity of Condorcet Winner Identification in Dueling Bandits
by: Saad, El Mehdi, et al.
Published: (2026)
by: Saad, El Mehdi, et al.
Published: (2026)
HCP-DCNet: A Hierarchical Causal Primitive Dynamic Composition Network for Self-Improving Causal Understanding
by: Lei, Ming, et al.
Published: (2026)
by: Lei, Ming, et al.
Published: (2026)
Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning
by: Bourdrez, Constant, et al.
Published: (2026)
by: Bourdrez, Constant, et al.
Published: (2026)
CLIPTTA: Robust Contrastive Vision-Language Test-Time Adaptation
by: Lafon, Marc, et al.
Published: (2025)
by: Lafon, Marc, et al.
Published: (2025)
UP-dROM : Uncertainty-Aware and Parametrised dynamic Reduced-Order Model, application to unsteady flows
by: Zighed, Ismaël, et al.
Published: (2025)
by: Zighed, Ismaël, et al.
Published: (2025)
Improving Reinforcement Learning Sample-Efficiency using Local Approximation
by: Prashant, Mohit, et al.
Published: (2025)
by: Prashant, Mohit, et al.
Published: (2025)
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
by: Gu, Shangding, et al.
Published: (2024)
by: Gu, Shangding, et al.
Published: (2024)
Supra-Laplacian Encoding for Transformer on Dynamic Graphs
by: Karmim, Yannis, et al.
Published: (2024)
by: Karmim, Yannis, et al.
Published: (2024)
Permutation-based Inference for Variational Learning of Directed Acyclic Graphs
by: Bonilla, Edwin V., et al.
Published: (2024)
by: Bonilla, Edwin V., et al.
Published: (2024)
Sample-Efficient Reinforcement Learning from Human Feedback via Information-Directed Sampling
by: Qi, Han, et al.
Published: (2025)
by: Qi, Han, et al.
Published: (2025)
On the Efficiency of ERM in Feature Learning
by: Hanchi, Ayoub El, et al.
Published: (2024)
by: Hanchi, Ayoub El, et al.
Published: (2024)
On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning
by: Morin, Sacha, et al.
Published: (2026)
by: Morin, Sacha, et al.
Published: (2026)
Hyperparameter Optimization Can Even be Harmful in Off-Policy Learning and How to Deal with It
by: Saito, Yuta, et al.
Published: (2024)
by: Saito, Yuta, et al.
Published: (2024)
RLingua: Improving Reinforcement Learning Sample Efficiency in Robotic Manipulations With Large Language Models
by: Chen, Liangliang, et al.
Published: (2024)
by: Chen, Liangliang, et al.
Published: (2024)
Learning Parametric Distributions from Samples and Preferences
by: Jourdan, Marc, et al.
Published: (2025)
by: Jourdan, Marc, et al.
Published: (2025)
Similar Items
-
Physics-Informed Model and Hybrid Planning for Efficient Dyna-Style Reinforcement Learning
by: Asri, Zakariae El, et al.
Published: (2024) -
Energy Correction Model in the Feature Space for Out-of-Distribution Detection
by: Lafon, Marc, et al.
Published: (2024) -
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
by: Aissi, Mohamed Salim, et al.
Published: (2024) -
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
by: Castanet, Nicolas, et al.
Published: (2025) -
Learning Physically Consistent Lagrangian Control Models Without Acceleration Measurements
by: Laiche, Ibrahim, et al.
Published: (2025)