SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jesse, Pertsch, Karl, Zhang, Jiahui, Lim, Joseph J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
von: Zhang, Jesse, et al.
Veröffentlicht: (2024)
von: Zhang, Jesse, et al.
Veröffentlicht: (2024)
Affordance-Guided Reinforcement Learning via Visual Prompting
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
QMP: Q-switch Mixture of Policies for Multi-Task Behavior Sharing
von: Zhang, Grace, et al.
Veröffentlicht: (2023)
von: Zhang, Grace, et al.
Veröffentlicht: (2023)
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
von: Polubarov, Andrei, et al.
Veröffentlicht: (2026)
von: Polubarov, Andrei, et al.
Veröffentlicht: (2026)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
von: Höpner, Niklas, et al.
Veröffentlicht: (2025)
von: Höpner, Niklas, et al.
Veröffentlicht: (2025)
Information-Theoretic Policy Pre-Training with Empowerment
von: Schneider, Moritz, et al.
Veröffentlicht: (2025)
von: Schneider, Moritz, et al.
Veröffentlicht: (2025)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
von: Jain, Ayush, et al.
Veröffentlicht: (2024)
von: Jain, Ayush, et al.
Veröffentlicht: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation
von: Ren, Hanxiang, et al.
Veröffentlicht: (2026)
von: Ren, Hanxiang, et al.
Veröffentlicht: (2026)
Robust Finetuning of Vision-Language-Action Robot Policies via Parameter Merging
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
Unsupervised Learning of Efficient Exploration: Pre-training Adaptive Policies via Self-Imposed Goals
von: Pappalardo, Octavio
Veröffentlicht: (2026)
von: Pappalardo, Octavio
Veröffentlicht: (2026)
Training Agents Inside of Scalable World Models
von: Hafner, Danijar, et al.
Veröffentlicht: (2025)
von: Hafner, Danijar, et al.
Veröffentlicht: (2025)
Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2025)
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2025)
Don't Start from Scratch: Behavioral Refinement via Interpolant-based Policy Diffusion
von: Chen, Kaiqi, et al.
Veröffentlicht: (2024)
von: Chen, Kaiqi, et al.
Veröffentlicht: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2025)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
von: Zhang, Chen Bo Calvin, et al.
Veröffentlicht: (2024)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2026)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2026)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
von: Xiao, Wei, et al.
Veröffentlicht: (2025)
von: Xiao, Wei, et al.
Veröffentlicht: (2025)
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
von: Anwar, Abrar, et al.
Veröffentlicht: (2025)
von: Anwar, Abrar, et al.
Veröffentlicht: (2025)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
von: Liang, Anthony, et al.
Veröffentlicht: (2026)
D-SPEAR: Dual-Stream Prioritized Experience Adaptive Replay for Stable Reinforcement Learning in Robotic Manipulation
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
When Engineering Outruns Intelligence: Rethinking Instruction-Guided Navigation
von: Aghaei, Matin, et al.
Veröffentlicht: (2025)
von: Aghaei, Matin, et al.
Veröffentlicht: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training
von: Dave, Vedant, et al.
Veröffentlicht: (2024)
von: Dave, Vedant, et al.
Veröffentlicht: (2024)
Imagination Policy: Using Generative Point Cloud Models for Learning Manipulation Policies
von: Huang, Haojie, et al.
Veröffentlicht: (2024)
von: Huang, Haojie, et al.
Veröffentlicht: (2024)
Text2Motion: From Natural Language Instructions to Feasible Plans
von: Lin, Kevin, et al.
Veröffentlicht: (2023)
von: Lin, Kevin, et al.
Veröffentlicht: (2023)
A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies
von: Lei, Yu, et al.
Veröffentlicht: (2026)
von: Lei, Yu, et al.
Veröffentlicht: (2026)
Relabeling Minimal Training Subset to Flip a Prediction
von: Yang, Jinghan, et al.
Veröffentlicht: (2023)
von: Yang, Jinghan, et al.
Veröffentlicht: (2023)
Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model
von: Yuan, Xiu, et al.
Veröffentlicht: (2024)
von: Yuan, Xiu, et al.
Veröffentlicht: (2024)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Discrete Variational Autoencoding via Policy Search
von: Drolet, Michael, et al.
Veröffentlicht: (2025)
von: Drolet, Michael, et al.
Veröffentlicht: (2025)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
von: Ma, Yunchang, et al.
Veröffentlicht: (2025)
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2025)
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2025)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
von: Koo, Juil, et al.
Veröffentlicht: (2026)
von: Koo, Juil, et al.
Veröffentlicht: (2026)
GROOT-2: Weakly Supervised Multi-Modal Instruction Following Agents
von: Cai, Shaofei, et al.
Veröffentlicht: (2024)
von: Cai, Shaofei, et al.
Veröffentlicht: (2024)
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
von: Zhang, Yixue, et al.
Veröffentlicht: (2026)
IMLE Policy: Fast and Sample Efficient Visuomotor Policy Learning via Implicit Maximum Likelihood Estimation
von: Rana, Krishan, et al.
Veröffentlicht: (2025)
von: Rana, Krishan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
von: Zhang, Jesse, et al.
Veröffentlicht: (2024) -
Affordance-Guided Reinforcement Learning via Visual Prompting
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024) -
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025) -
QMP: Q-switch Mixture of Policies for Multi-Task Behavior Sharing
von: Zhang, Grace, et al.
Veröffentlicht: (2023) -
Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner
von: Polubarov, Andrei, et al.
Veröffentlicht: (2026)