R$\times$R: Rapid eXploration for Reinforcement Learning via Sampling-based Reset Distributions and Imitation Pre-training
Fuente:
arXiv
Salvato in:
| Autori principali: | Khandate, Gagan, Saidi, Tristan L., Shang, Siqi, Chang, Eric T., Liu, Yang, Dennis, Seth, Adams, Johnson, Ciocarlie, Matei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Human-level Dexterity via Robot Learning
di: Khandate, Gagan
Pubblicazione: (2025)
di: Khandate, Gagan
Pubblicazione: (2025)
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
di: Zhu, Zhiyu, et al.
Pubblicazione: (2024)
di: Zhu, Zhiyu, et al.
Pubblicazione: (2024)
DynamiX: Dynamic Resource eXploration for Personalized Ad-Recommendations
di: Roychowdhury, Sohini, et al.
Pubblicazione: (2025)
di: Roychowdhury, Sohini, et al.
Pubblicazione: (2025)
Curation protocol of Phobos sample returned by Martian Moons eXploration
di: Ryota Fukai, et al.
Pubblicazione: (2024)
di: Ryota Fukai, et al.
Pubblicazione: (2024)
NMIXX: Domain-Adapted Neural Embeddings for Cross-Lingual eXploration of Finance
di: Lee, Hanwool, et al.
Pubblicazione: (2025)
di: Lee, Hanwool, et al.
Pubblicazione: (2025)
Train Robots in a JIF: Joint Inverse and Forward Dynamics with Human and Robot Demonstrations
di: Khandate, Gagan, et al.
Pubblicazione: (2025)
di: Khandate, Gagan, et al.
Pubblicazione: (2025)
aims-PAX: Parallel Active eXploration for the automated construction of Machine Learning Force Fields
di: Henkes, Tobias, et al.
Pubblicazione: (2025)
di: Henkes, Tobias, et al.
Pubblicazione: (2025)
Diurnal Variations of Lee Wave Clouds on Mars using Emirates eXploration Imager (EXI)
di: Alshamsi, Mariam R., et al.
Pubblicazione: (2025)
di: Alshamsi, Mariam R., et al.
Pubblicazione: (2025)
Uncertainty Comes for Free: Human-in-the-Loop Policies with Diffusion Models
di: He, Zhanpeng, et al.
Pubblicazione: (2025)
di: He, Zhanpeng, et al.
Pubblicazione: (2025)
Task-Based Design and Policy Co-Optimization for Tendon-driven Underactuated Kinematic Chains
di: Islam, Sharfin, et al.
Pubblicazione: (2024)
di: Islam, Sharfin, et al.
Pubblicazione: (2024)
Leveraging Vision-Language Pre-training for Human Activity Recognition in Still Images
di: Mahanta, Cristina, et al.
Pubblicazione: (2025)
di: Mahanta, Cristina, et al.
Pubblicazione: (2025)
Single-Reset Divide & Conquer Imitation Learning
di: Chenu, Alexandre, et al.
Pubblicazione: (2024)
di: Chenu, Alexandre, et al.
Pubblicazione: (2024)
An Investigation of Multi-feature Extraction and Super-resolution with Fast Microphone Arrays
di: Chang, Eric T., et al.
Pubblicazione: (2023)
di: Chang, Eric T., et al.
Pubblicazione: (2023)
FRaN-X: FRaming and Narratives-eXplorer
di: Muratov, Artur, et al.
Pubblicazione: (2025)
di: Muratov, Artur, et al.
Pubblicazione: (2025)
Compact LED-Based Displacement Sensing for Robot Fingers
di: El-Azizi, Amr, et al.
Pubblicazione: (2024)
di: El-Azizi, Amr, et al.
Pubblicazione: (2024)
LEDA: Latent Semantic Distribution Alignment for Multi-domain Graph Pre-training
di: Shan, Lianze, et al.
Pubblicazione: (2026)
di: Shan, Lianze, et al.
Pubblicazione: (2026)
VibeCheck: Using Active Acoustic Tactile Sensing for Contact-Rich Manipulation
di: Zhang, Kaidi, et al.
Pubblicazione: (2025)
di: Zhang, Kaidi, et al.
Pubblicazione: (2025)
Understanding Quantization of Optimizer States in LLM Pre-training: Dynamics of State Staleness and Effectiveness of State Resets
di: Topollai, Kristi, et al.
Pubblicazione: (2026)
di: Topollai, Kristi, et al.
Pubblicazione: (2026)
Structure-Based Drug Design via 3D Molecular Generative Pre-training and Sampling
di: Yang, Yuwei, et al.
Pubblicazione: (2024)
di: Yang, Yuwei, et al.
Pubblicazione: (2024)
TrajGPT-R: Generating Urban Mobility Trajectory with Reinforcement Learning-Enhanced Generative Pre-trained Transformer
di: Wang, Jiawei, et al.
Pubblicazione: (2026)
di: Wang, Jiawei, et al.
Pubblicazione: (2026)
The Power of Resets in Online Reinforcement Learning
di: Mhammedi, Zakaria, et al.
Pubblicazione: (2024)
di: Mhammedi, Zakaria, et al.
Pubblicazione: (2024)
Uncertainty-Aware Deployment of Pre-trained Language-Conditioned Imitation Learning Policies
di: Wu, Bo, et al.
Pubblicazione: (2024)
di: Wu, Bo, et al.
Pubblicazione: (2024)
NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models
di: Albaba, Mert, et al.
Pubblicazione: (2025)
di: Albaba, Mert, et al.
Pubblicazione: (2025)
Education and training of accountants in Sub-Saharan anglophone Africa / Sonia R. Johnson
di: Johnson, Sonia R
Pubblicazione: (1996)
di: Johnson, Sonia R
Pubblicazione: (1996)
Spectral Duality and Reset-Neutral Distributions in Random Walks with Multi-Site Geometric Resetting
di: Coso, Juan Antonio Vega
Pubblicazione: (2026)
di: Coso, Juan Antonio Vega
Pubblicazione: (2026)
ReactEMG Stroke: Healthy-to-Stroke Few-shot Adaptation for sEMG-Based Intent Detection
di: Wang, Runsheng, et al.
Pubblicazione: (2026)
di: Wang, Runsheng, et al.
Pubblicazione: (2026)
Sample-efficient LLM Optimization with Reset Replay
di: Liu, Zichuan, et al.
Pubblicazione: (2025)
di: Liu, Zichuan, et al.
Pubblicazione: (2025)
Generalized Rapid Action Value Estimation in Memory-Constrained Environments
di: Rautureau, Aloïs, et al.
Pubblicazione: (2026)
di: Rautureau, Aloïs, et al.
Pubblicazione: (2026)
Reset-free Reinforcement Learning with World Models
di: Yang, Zhao, et al.
Pubblicazione: (2024)
di: Yang, Zhao, et al.
Pubblicazione: (2024)
What a Big R Reset of the Public Service of Canada Needs to Do (and Not to Do)
di: Toby Fyfe
Pubblicazione: (2025)
di: Toby Fyfe
Pubblicazione: (2025)
Generalizable Imitation Learning Through Pre-Trained Representations
di: Chang, Wei-Di, et al.
Pubblicazione: (2023)
di: Chang, Wei-Di, et al.
Pubblicazione: (2023)
Imitation Game: A Model-based and Imitation Learning Deep Reinforcement Learning Hybrid
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
Grasp Force Assistance via Throttle-based Wrist Angle Control on a Robotic Hand Orthosis for C6-C7 Spinal Cord Injury
di: Palacios, Joaquin, et al.
Pubblicazione: (2024)
di: Palacios, Joaquin, et al.
Pubblicazione: (2024)
Tactile-based Object Retrieval From Granular Media
di: Xu, Jingxi, et al.
Pubblicazione: (2024)
di: Xu, Jingxi, et al.
Pubblicazione: (2024)
Reinforcement Learning for Constraint Satisfaction Game Agents
di: Praveen, Amritesh, et al.
Pubblicazione: (2026)
di: Praveen, Amritesh, et al.
Pubblicazione: (2026)
DeepRV: Accelerating Spatiotemporal Inference with Pre-trained Neural Priors
di: Navott, Jhonathan, et al.
Pubblicazione: (2025)
di: Navott, Jhonathan, et al.
Pubblicazione: (2025)
SpikeATac: A Multimodal Tactile Finger with Taxelized Dynamic Sensing for Dexterous Manipulation
di: Chang, Eric T., et al.
Pubblicazione: (2025)
di: Chang, Eric T., et al.
Pubblicazione: (2025)
ChatEMG: Synthetic Data Generation to Control a Robotic Hand Orthosis for Stroke
di: Xu, Jingxi, et al.
Pubblicazione: (2024)
di: Xu, Jingxi, et al.
Pubblicazione: (2024)
Stochastic Resetting Accelerates Policy Convergence in Reinforcement Learning
di: Zhou, Jello, et al.
Pubblicazione: (2026)
di: Zhou, Jello, et al.
Pubblicazione: (2026)
Recovering Manifold Structure Using Ollivier-Ricci Curvature
di: Saidi, Tristan Luca, et al.
Pubblicazione: (2024)
di: Saidi, Tristan Luca, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Human-level Dexterity via Robot Learning
di: Khandate, Gagan
Pubblicazione: (2025) -
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
di: Zhu, Zhiyu, et al.
Pubblicazione: (2024) -
DynamiX: Dynamic Resource eXploration for Personalized Ad-Recommendations
di: Roychowdhury, Sohini, et al.
Pubblicazione: (2025) -
Curation protocol of Phobos sample returned by Martian Moons eXploration
di: Ryota Fukai, et al.
Pubblicazione: (2024) -
NMIXX: Domain-Adapted Neural Embeddings for Cross-Lingual eXploration of Finance
di: Lee, Hanwool, et al.
Pubblicazione: (2025)