Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
Fuente:
arXiv
Guardado en:
| Autores principales: | Ye, Weirui, Liu, Fangchen, Ding, Zheng, Gao, Yang, Rybkin, Oleh, Abbeel, Pieter |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
por: Liu, Fangchen, et al.
Publicado: (2024)
por: Liu, Fangchen, et al.
Publicado: (2024)
Scaling Tasks, Not Samples: Mastering Humanoid Control through Multi-Task Model-Based Reinforcement Learning
por: Liu, Shaohuai, et al.
Publicado: (2026)
por: Liu, Shaohuai, et al.
Publicado: (2026)
Body Transformer: Leveraging Robot Embodiment for Policy Learning
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Latent Diffusion Planning for Imitation Learning
por: Xie, Amber, et al.
Publicado: (2025)
por: Xie, Amber, et al.
Publicado: (2025)
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
por: Ye, Weirui, et al.
Publicado: (2023)
por: Ye, Weirui, et al.
Publicado: (2023)
HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
Privileged Sensing Scaffolds Reinforcement Learning
por: Hu, Edward S., et al.
Publicado: (2024)
por: Hu, Edward S., et al.
Publicado: (2024)
EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data
por: Wang, Shengjie, et al.
Publicado: (2024)
por: Wang, Shengjie, et al.
Publicado: (2024)
Object-centric 3D Motion Field for Robot Learning from Human Videos
por: Yin, Zhao-Heng, et al.
Publicado: (2025)
por: Yin, Zhao-Heng, et al.
Publicado: (2025)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
por: Seo, Younggyo, et al.
Publicado: (2024)
por: Seo, Younggyo, et al.
Publicado: (2024)
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
por: Jones, Joshua, et al.
Publicado: (2025)
por: Jones, Joshua, et al.
Publicado: (2025)
Offline Imitation Learning Through Graph Search and Retrieval
por: Yin, Zhao-Heng, et al.
Publicado: (2024)
por: Yin, Zhao-Heng, et al.
Publicado: (2024)
ViTaMIn: Learning Contact-Rich Tasks Through Robot-Free Visuo-Tactile Manipulation Interface
por: Liu, Fangchen, et al.
Publicado: (2025)
por: Liu, Fangchen, et al.
Publicado: (2025)
Solving New Tasks by Adapting Internet Video Knowledge
por: Luo, Calvin, et al.
Publicado: (2025)
por: Luo, Calvin, et al.
Publicado: (2025)
Interactive Task Planning with Language Models
por: Li, Boyi, et al.
Publicado: (2023)
por: Li, Boyi, et al.
Publicado: (2023)
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
por: Shentu, Yide, et al.
Publicado: (2024)
por: Shentu, Yide, et al.
Publicado: (2024)
Hand-Object Interaction Pretraining from Videos
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
por: Singh, Himanshu Gaurav, et al.
Publicado: (2024)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
por: Lee, Vint, et al.
Publicado: (2023)
por: Lee, Vint, et al.
Publicado: (2023)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
por: Sukhija, Bhavya, et al.
Publicado: (2024)
por: Sukhija, Bhavya, et al.
Publicado: (2024)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
por: Shen, Yichao, et al.
Publicado: (2025)
por: Shen, Yichao, et al.
Publicado: (2025)
FP3: A 3D Foundation Policy for Robotic Manipulation
por: Yang, Rujia, et al.
Publicado: (2025)
por: Yang, Rujia, et al.
Publicado: (2025)
Flow Policy Gradients for Robot Control
por: Yi, Brent, et al.
Publicado: (2026)
por: Yi, Brent, et al.
Publicado: (2026)
Learning the Generalizable Manipulation Skills on Soft-body Tasks via Guided Self-attention Behavior Cloning Policy
por: Li, Xuetao, et al.
Publicado: (2024)
por: Li, Xuetao, et al.
Publicado: (2024)
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
por: Kim, Moo Jin, et al.
Publicado: (2026)
por: Kim, Moo Jin, et al.
Publicado: (2026)
Real-World Reinforcement Learning of Active Perception Behaviors
por: Hu, Edward S., et al.
Publicado: (2025)
por: Hu, Edward S., et al.
Publicado: (2025)
EgoZero: Robot Learning from Smart Glasses
por: Liu, Vincent, et al.
Publicado: (2025)
por: Liu, Vincent, et al.
Publicado: (2025)
ArticuBot: Learning Universal Articulated Object Manipulation Policy via Large Scale Simulation
por: Wang, Yufei, et al.
Publicado: (2025)
por: Wang, Yufei, et al.
Publicado: (2025)
RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
por: Yang, Xuning, et al.
Publicado: (2026)
por: Yang, Xuning, et al.
Publicado: (2026)
ManiCM: Real-time 3D Diffusion Policy via Consistency Model for Robotic Manipulation
por: Lu, Guanxing, et al.
Publicado: (2024)
por: Lu, Guanxing, et al.
Publicado: (2024)
SDP: Spiking Diffusion Policy for Robotic Manipulation with Learnable Channel-Wise Membrane Thresholds
por: Hou, Zhixing, et al.
Publicado: (2024)
por: Hou, Zhixing, et al.
Publicado: (2024)
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
por: Yang, Lujie, et al.
Publicado: (2025)
por: Yang, Lujie, et al.
Publicado: (2025)
Feel the Force: Contact-Driven Learning from Humans
por: Adeniji, Ademi, et al.
Publicado: (2025)
por: Adeniji, Ademi, et al.
Publicado: (2025)
Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models
por: Soleymanzadeh, Davood, et al.
Publicado: (2026)
por: Soleymanzadeh, Davood, et al.
Publicado: (2026)
UrbanVerse: Scaling Urban Simulation by Watching City-Tour Videos
por: Liu, Mingxuan, et al.
Publicado: (2025)
por: Liu, Mingxuan, et al.
Publicado: (2025)
Language-Model-Assisted Bi-Level Programming for Reward Learning from Internet Videos
por: Mahesheka, Harsh, et al.
Publicado: (2024)
por: Mahesheka, Harsh, et al.
Publicado: (2024)
EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow
por: Chen, Yixiang, et al.
Publicado: (2025)
por: Chen, Yixiang, et al.
Publicado: (2025)
PRIME: Scaffolding Manipulation Tasks with Behavior Primitives for Data-Efficient Imitation Learning
por: Gao, Tian, et al.
Publicado: (2024)
por: Gao, Tian, et al.
Publicado: (2024)
Learning Manipulation Skills through Robot Chain-of-Thought with Sparse Failure Guidance
por: Zhang, Kaifeng, et al.
Publicado: (2024)
por: Zhang, Kaifeng, et al.
Publicado: (2024)
Ejemplares similares
-
MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
por: Liu, Fangchen, et al.
Publicado: (2024) -
Scaling Tasks, Not Samples: Mastering Humanoid Control through Multi-Task Model-Based Reinforcement Learning
por: Liu, Shaohuai, et al.
Publicado: (2026) -
Body Transformer: Leveraging Robot Embodiment for Policy Learning
por: Sferrazza, Carmelo, et al.
Publicado: (2024) -
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023) -
Latent Diffusion Planning for Imitation Learning
por: Xie, Amber, et al.
Publicado: (2025)