From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Jusheng, Tang, Jinzhou, Liu, Sidi, Li, Mingyan, Zhang, Sheng, Wang, Jian, Wang, Keze |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance
por: Li, Zhe, et al.
Publicado: (2025)
por: Li, Zhe, et al.
Publicado: (2025)
STORM: Search-Guided Generative World Models for Robotic Manipulation
por: Lin, Wenjun, et al.
Publicado: (2025)
por: Lin, Wenjun, et al.
Publicado: (2025)
LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map
por: Tang, Jinzhou, et al.
Publicado: (2026)
por: Tang, Jinzhou, et al.
Publicado: (2026)
Iterative Closed-Loop Motion Synthesis for Scaling the Capabilities of Humanoid Control
por: Xu, Weisheng, et al.
Publicado: (2026)
por: Xu, Weisheng, et al.
Publicado: (2026)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
por: Jiang, Nan, et al.
Publicado: (2025)
por: Jiang, Nan, et al.
Publicado: (2025)
E0: Enhancing Generalization and Fine-Grained Control in VLA Models via Tweedie Discrete Diffusion
por: Zhan, Zhihao, et al.
Publicado: (2025)
por: Zhan, Zhihao, et al.
Publicado: (2025)
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration
por: Ding, Pengxiang, et al.
Publicado: (2025)
por: Ding, Pengxiang, et al.
Publicado: (2025)
Hierarchical World Models as Visual Whole-Body Humanoid Controllers
por: Hansen, Nicklas, et al.
Publicado: (2024)
por: Hansen, Nicklas, et al.
Publicado: (2024)
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
por: Tang, Jinzhou, et al.
Publicado: (2025)
por: Tang, Jinzhou, et al.
Publicado: (2025)
Universal Humanoid Motion Representations for Physics-Based Control
por: Luo, Zhengyi, et al.
Publicado: (2023)
por: Luo, Zhengyi, et al.
Publicado: (2023)
Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots
por: Cui, Wei, et al.
Publicado: (2025)
por: Cui, Wei, et al.
Publicado: (2025)
PhyGile: Physics-Prefix Guided Motion Generation for Agile General Humanoid Motion Tracking
por: Bao, Jiacheng, et al.
Publicado: (2026)
por: Bao, Jiacheng, et al.
Publicado: (2026)
Visual Imitation Enables Contextual Humanoid Control
por: Allshire, Arthur, et al.
Publicado: (2025)
por: Allshire, Arthur, et al.
Publicado: (2025)
MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control
por: Li, Bin, et al.
Publicado: (2026)
por: Li, Bin, et al.
Publicado: (2026)
3D-Agent:Tri-Modal Multi-Agent Collaboration for Scalable 3D Object Annotation
por: Zhang, Jusheng, et al.
Publicado: (2026)
por: Zhang, Jusheng, et al.
Publicado: (2026)
Top-Down Semantic Refinement for Image Captioning
por: Zhang, Jusheng, et al.
Publicado: (2025)
por: Zhang, Jusheng, et al.
Publicado: (2025)
Before the Body Moves: Learning Anticipatory Joint Intent for Language-Conditioned Humanoid Control
por: Jia, Haozhe, et al.
Publicado: (2026)
por: Jia, Haozhe, et al.
Publicado: (2026)
ResAgent: Entropy-based Prior Point Discovery and Visual Reasoning for Referring Expression Segmentation
por: Wang, Yihao, et al.
Publicado: (2026)
por: Wang, Yihao, et al.
Publicado: (2026)
3DAlign-DAER: Dynamic Attention Policy and Efficient Retrieval Strategy for Fine-grained 3D-Text Alignment at Scale
por: Fan, Yijia, et al.
Publicado: (2025)
por: Fan, Yijia, et al.
Publicado: (2025)
FlashVLM: Text-Guided Visual Token Selection for Large Multimodal Models
por: Cai, Kaitong, et al.
Publicado: (2025)
por: Cai, Kaitong, et al.
Publicado: (2025)
ULTRA: Unified Multimodal Control for Autonomous Humanoid Whole-Body Loco-Manipulation
por: He, Xialin, et al.
Publicado: (2026)
por: He, Xialin, et al.
Publicado: (2026)
VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
por: Yin, Shaofeng, et al.
Publicado: (2025)
por: Yin, Shaofeng, et al.
Publicado: (2025)
TAG: Target-Agnostic Guidance for Stable Object-Centric Inference in Vision-Language-Action Models
por: Zhou, Jiaying, et al.
Publicado: (2026)
por: Zhou, Jiaying, et al.
Publicado: (2026)
Solving Motion Planning Tasks with a Scalable Generative Model
por: Hu, Yihan, et al.
Publicado: (2024)
por: Hu, Yihan, et al.
Publicado: (2024)
SMPLOlympics: Sports Environments for Physically Simulated Humanoids
por: Luo, Zhengyi, et al.
Publicado: (2024)
por: Luo, Zhengyi, et al.
Publicado: (2024)
FRoM-W1: Towards General Humanoid Whole-Body Control with Language Instructions
por: Li, Peng, et al.
Publicado: (2026)
por: Li, Peng, et al.
Publicado: (2026)
Decoupling Ego-Motion from Target Dynamics via Dual-Interval Motion Cues for UAV Detection
por: Wang, Liuyang, et al.
Publicado: (2026)
por: Wang, Liuyang, et al.
Publicado: (2026)
Visually-grounded Humanoid Agents
por: Ye, Hang, et al.
Publicado: (2026)
por: Ye, Hang, et al.
Publicado: (2026)
RoboMirror: Understand Before You Imitate for Video to Humanoid Locomotion
por: Li, Zhe, et al.
Publicado: (2025)
por: Li, Zhe, et al.
Publicado: (2025)
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
por: Dong, Runpei, et al.
Publicado: (2026)
por: Dong, Runpei, et al.
Publicado: (2026)
Kimodo: Scaling Controllable Human Motion Generation
por: Rempe, Davis, et al.
Publicado: (2026)
por: Rempe, Davis, et al.
Publicado: (2026)
Mimicking-Bench: A Benchmark for Generalizable Humanoid-Scene Interaction Learning via Human Mimicking
por: Liu, Yun, et al.
Publicado: (2024)
por: Liu, Yun, et al.
Publicado: (2024)
Robot Interaction Behavior Generation based on Social Motion Forecasting for Human-Robot Interaction
por: Mascaro, Esteve Valls, et al.
Publicado: (2024)
por: Mascaro, Esteve Valls, et al.
Publicado: (2024)
Pixel Motion Diffusion is What We Need for Robot Control
por: Nguyen, E-Ro, et al.
Publicado: (2025)
por: Nguyen, E-Ro, et al.
Publicado: (2025)
Humanoid Policy ~ Human Policy
por: Qiu, Ri-Zhao, et al.
Publicado: (2025)
por: Qiu, Ri-Zhao, et al.
Publicado: (2025)
Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
por: Chen, Jiahe, et al.
Publicado: (2026)
por: Chen, Jiahe, et al.
Publicado: (2026)
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects
por: Xue, Jialong, et al.
Publicado: (2025)
por: Xue, Jialong, et al.
Publicado: (2025)
Post-Training and Test-Time Scaling of Generative Agent Behavior Models for Interactive Autonomous Driving
por: Seong, Hyunki, et al.
Publicado: (2025)
por: Seong, Hyunki, et al.
Publicado: (2025)
TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning
por: Liu, Jiacheng, et al.
Publicado: (2025)
por: Liu, Jiacheng, et al.
Publicado: (2025)
Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs
por: Catalano, Iacopo, et al.
Publicado: (2026)
por: Catalano, Iacopo, et al.
Publicado: (2026)
Ejemplares similares
-
From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance
por: Li, Zhe, et al.
Publicado: (2025) -
STORM: Search-Guided Generative World Models for Robotic Manipulation
por: Lin, Wenjun, et al.
Publicado: (2025) -
LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map
por: Tang, Jinzhou, et al.
Publicado: (2026) -
Iterative Closed-Loop Motion Synthesis for Scaling the Capabilities of Humanoid Control
por: Xu, Weisheng, et al.
Publicado: (2026) -
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
por: Jiang, Nan, et al.
Publicado: (2025)