Humanoid Locomotion as Next Token Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Radosavovic, Ilija, Zhang, Bike, Shi, Baifeng, Rajasegaran, Jathushan, Kamat, Sarthak, Darrell, Trevor, Sreenath, Koushil, Malik, Jitendra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Humanoid Locomotion over Challenging Terrain
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
An Empirical Study of Autoregressive Pre-training from Videos
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
Tracking by Predicting 3-D Gaussians Over Time
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025)
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025)
Poly-Autoregressive Prediction for Modeling Interactions
von: Thakkar, Neerja, et al.
Veröffentlicht: (2025)
von: Thakkar, Neerja, et al.
Veröffentlicht: (2025)
EgoPet: Egomotion and Interaction Data from an Animal's Perspective
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
Scaling Properties of Diffusion Models for Perceptual Tasks
von: Ravishankar, Rahul, et al.
Veröffentlicht: (2024)
von: Ravishankar, Rahul, et al.
Veröffentlicht: (2024)
Whole-Body Conditioned Egocentric Video Prediction
von: Bai, Yutong, et al.
Veröffentlicht: (2025)
von: Bai, Yutong, et al.
Veröffentlicht: (2025)
Visual Imitation Enables Contextual Humanoid Control
von: Allshire, Arthur, et al.
Veröffentlicht: (2025)
von: Allshire, Arthur, et al.
Veröffentlicht: (2025)
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
von: Niu, Dantong, et al.
Veröffentlicht: (2024)
von: Niu, Dantong, et al.
Veröffentlicht: (2024)
Learning to Grasp Anything by Playing with Random Toys
von: Niu, Dantong, et al.
Veröffentlicht: (2025)
von: Niu, Dantong, et al.
Veröffentlicht: (2025)
Gaussian Masked Autoencoders
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids
von: Lin, Toru, et al.
Veröffentlicht: (2025)
von: Lin, Toru, et al.
Veröffentlicht: (2025)
Synergy and Synchrony in Couple Dances
von: Maluleke, Vongani, et al.
Veröffentlicht: (2024)
von: Maluleke, Vongani, et al.
Veröffentlicht: (2024)
From Generated Human Videos to Physically Plausible Robot Trajectories
von: Ni, James, et al.
Veröffentlicht: (2025)
von: Ni, James, et al.
Veröffentlicht: (2025)
Synthesizing Moving People with 3D Control
von: Li, Boyi, et al.
Veröffentlicht: (2024)
von: Li, Boyi, et al.
Veröffentlicht: (2024)
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
von: Hu, Yucheng, et al.
Veröffentlicht: (2024)
von: Hu, Yucheng, et al.
Veröffentlicht: (2024)
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
Coordinated Humanoid Manipulation with Choice Policies
von: Qi, Haozhi, et al.
Veröffentlicht: (2025)
von: Qi, Haozhi, et al.
Veröffentlicht: (2025)
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
von: Wang, Renhao, et al.
Veröffentlicht: (2025)
von: Wang, Renhao, et al.
Veröffentlicht: (2025)
RoboMirror: Understand Before You Imitate for Video to Humanoid Locomotion
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
Filling Missing Values Matters for Range Image-Based Point Cloud Segmentation
von: Chen, Bike, et al.
Veröffentlicht: (2024)
von: Chen, Bike, et al.
Veröffentlicht: (2024)
FewShotNeRF: Meta-Learning-based Novel View Synthesis for Rapid Scene-Specific Adaptation
von: Sivakumar, Piraveen, et al.
Veröffentlicht: (2024)
von: Sivakumar, Piraveen, et al.
Veröffentlicht: (2024)
Prompt a Robot to Walk with Large Language Models
von: Wang, Yen-Jen, et al.
Veröffentlicht: (2023)
von: Wang, Yen-Jen, et al.
Veröffentlicht: (2023)
From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
Berkeley Humanoid: A Research Platform for Learning-based Control
von: Liao, Qiayuan, et al.
Veröffentlicht: (2024)
von: Liao, Qiayuan, et al.
Veröffentlicht: (2024)
xT: Nested Tokenization for Larger Context in Large Images
von: Gupta, Ritwik, et al.
Veröffentlicht: (2024)
von: Gupta, Ritwik, et al.
Veröffentlicht: (2024)
SPARK: Skeleton-Parameter Aligned Retargeting on Humanoid Robots with Kinodynamic Trajectory Optimization
von: Wang, Hanwen, et al.
Veröffentlicht: (2026)
von: Wang, Hanwen, et al.
Veröffentlicht: (2026)
Continuous Locomotive Crowd Behavior Generation
von: Bae, Inhwan, et al.
Veröffentlicht: (2025)
von: Bae, Inhwan, et al.
Veröffentlicht: (2025)
TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
von: Ze, Yanjie, et al.
Veröffentlicht: (2025)
von: Ze, Yanjie, et al.
Veröffentlicht: (2025)
When Do We Not Need Larger Vision Models?
von: Shi, Baifeng, et al.
Veröffentlicht: (2024)
von: Shi, Baifeng, et al.
Veröffentlicht: (2024)
Navigation World Models
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
Twisting Lids Off with Two Hands
von: Lin, Toru, et al.
Veröffentlicht: (2024)
von: Lin, Toru, et al.
Veröffentlicht: (2024)
Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
von: Chen, Boyuan, et al.
Veröffentlicht: (2024)
von: Chen, Boyuan, et al.
Veröffentlicht: (2024)
Learning from Massive Human Videos for Universal Humanoid Pose Control
von: Mao, Jiageng, et al.
Veröffentlicht: (2024)
von: Mao, Jiageng, et al.
Veröffentlicht: (2024)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
von: Tian, Ran, et al.
Veröffentlicht: (2024)
von: Tian, Ran, et al.
Veröffentlicht: (2024)
MomaGraph: State-Aware Unified Scene Graphs with Vision-Language Model for Embodied Task Planning
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
Hierarchical World Models as Visual Whole-Body Humanoid Controllers
von: Hansen, Nicklas, et al.
Veröffentlicht: (2024)
von: Hansen, Nicklas, et al.
Veröffentlicht: (2024)
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience
von: Chane-Sane, Elliot, et al.
Veröffentlicht: (2024)
von: Chane-Sane, Elliot, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Humanoid Locomotion over Challenging Terrain
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024) -
An Empirical Study of Autoregressive Pre-training from Videos
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025) -
Tracking by Predicting 3-D Gaussians Over Time
von: Baranwal, Tanish, et al.
Veröffentlicht: (2025) -
Poly-Autoregressive Prediction for Modeling Interactions
von: Thakkar, Neerja, et al.
Veröffentlicht: (2025) -
EgoPet: Egomotion and Interaction Data from an Animal's Perspective
von: Bar, Amir, et al.
Veröffentlicht: (2024)