HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Xinyu, Zhang, Yizheng, Li, Yong-Lu, Han, Lei, Lu, Cewu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
di: Zhang, Yizheng, et al.
Pubblicazione: (2025)
di: Zhang, Yizheng, et al.
Pubblicazione: (2025)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
di: Su, Yue, et al.
Pubblicazione: (2024)
di: Su, Yue, et al.
Pubblicazione: (2024)
Towards Human-Like Manipulation through RL-Augmented Teleoperation and Mixture-of-Dexterous-Experts VLA
di: Tang, Tutian, et al.
Pubblicazione: (2026)
di: Tang, Tutian, et al.
Pubblicazione: (2026)
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration
di: Ding, Pengxiang, et al.
Pubblicazione: (2025)
di: Ding, Pengxiang, et al.
Pubblicazione: (2025)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
di: Xu, Yifu, et al.
Pubblicazione: (2026)
di: Xu, Yifu, et al.
Pubblicazione: (2026)
UniPlan: Vision-Language Task Planning for Mobile Manipulation with Unified PDDL Formulation
di: Ye, Haoming, et al.
Pubblicazione: (2026)
di: Ye, Haoming, et al.
Pubblicazione: (2026)
Physically Ground Commonsense Knowledge for Articulated Object Manipulation with Analytic Concepts
di: Wei, Jiude, et al.
Pubblicazione: (2025)
di: Wei, Jiude, et al.
Pubblicazione: (2025)
UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling
di: Chen, Boyu, et al.
Pubblicazione: (2026)
di: Chen, Boyu, et al.
Pubblicazione: (2026)
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
di: Li, Boyu, et al.
Pubblicazione: (2026)
di: Li, Boyu, et al.
Pubblicazione: (2026)
NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models
di: Zhu, Ziyue, et al.
Pubblicazione: (2026)
di: Zhu, Ziyue, et al.
Pubblicazione: (2026)
SmoothVLA: Aligning Vision-Language-Action Models with Physical Constraints via Intrinsic Smoothness Optimization
di: Li, Jiashun, et al.
Pubblicazione: (2026)
di: Li, Jiashun, et al.
Pubblicazione: (2026)
ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration
di: Zou, Yanwen, et al.
Pubblicazione: (2026)
di: Zou, Yanwen, et al.
Pubblicazione: (2026)
RedVLA: Physical Red Teaming for Vision-Language-Action Models
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
di: Jin, Yang, et al.
Pubblicazione: (2024)
di: Jin, Yang, et al.
Pubblicazione: (2024)
Digital Gene: Learning about the Physical World through Analytic Concepts
di: Sun, Jianhua, et al.
Pubblicazione: (2025)
di: Sun, Jianhua, et al.
Pubblicazione: (2025)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
di: Sun, Nan, et al.
Pubblicazione: (2025)
di: Sun, Nan, et al.
Pubblicazione: (2025)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
di: Peng, Xiongfeng, et al.
Pubblicazione: (2026)
di: Peng, Xiongfeng, et al.
Pubblicazione: (2026)
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
di: Li, Puhao, et al.
Pubblicazione: (2025)
di: Li, Puhao, et al.
Pubblicazione: (2025)
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning
di: Zhang, Borong, et al.
Pubblicazione: (2025)
di: Zhang, Borong, et al.
Pubblicazione: (2025)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
di: Guo, Xinyu, et al.
Pubblicazione: (2026)
di: Guo, Xinyu, et al.
Pubblicazione: (2026)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
di: Zhang, Kaidi, et al.
Pubblicazione: (2026)
di: Zhang, Kaidi, et al.
Pubblicazione: (2026)
Unified Noise Steering for Efficient Human-Guided VLA Adaptation
di: Lu, Junjie, et al.
Pubblicazione: (2026)
di: Lu, Junjie, et al.
Pubblicazione: (2026)
Empathetic Motion Generation for Humanoid Educational Robots via Reasoning-Guided Vision--Language--Motion Diffusion Architecture
di: Sun, Fuze, et al.
Pubblicazione: (2026)
di: Sun, Fuze, et al.
Pubblicazione: (2026)
L1 Sample Flow for Efficient Visuomotor Learning
di: Song, Weixi, et al.
Pubblicazione: (2025)
di: Song, Weixi, et al.
Pubblicazione: (2025)
Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos
di: Feng, Yicheng, et al.
Pubblicazione: (2025)
di: Feng, Yicheng, et al.
Pubblicazione: (2025)
Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks
di: Liu, Zhihong, et al.
Pubblicazione: (2026)
di: Liu, Zhihong, et al.
Pubblicazione: (2026)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
di: Wang, Zixuan, et al.
Pubblicazione: (2026)
di: Wang, Zixuan, et al.
Pubblicazione: (2026)
GazeVLA: Learning Human Intention for Robotic Manipulation
di: Li, Chengyang, et al.
Pubblicazione: (2026)
di: Li, Chengyang, et al.
Pubblicazione: (2026)
FutureVLA: Joint Visuomotor Prediction for Vision-Language-Action Model
di: Xu, Xiaoxu, et al.
Pubblicazione: (2026)
di: Xu, Xiaoxu, et al.
Pubblicazione: (2026)
DexTOG: Learning Task-Oriented Dexterous Grasp with Language
di: Zhang, Jieyi, et al.
Pubblicazione: (2025)
di: Zhang, Jieyi, et al.
Pubblicazione: (2025)
NanoVLA: Routing Decoupled Vision-Language Understanding for Nano-sized Generalist Robotic Policies
di: Chen, Jiahong, et al.
Pubblicazione: (2025)
di: Chen, Jiahong, et al.
Pubblicazione: (2025)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
di: Hu, Yucheng, et al.
Pubblicazione: (2026)
di: Hu, Yucheng, et al.
Pubblicazione: (2026)
FSGlove: An Inertial-Based Hand Tracking System with Shape-Aware Calibration
di: Li, Yutong, et al.
Pubblicazione: (2025)
di: Li, Yutong, et al.
Pubblicazione: (2025)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
di: Yin, Zhenhan, et al.
Pubblicazione: (2025)
di: Yin, Zhenhan, et al.
Pubblicazione: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
di: Zeng, Qixin, et al.
Pubblicazione: (2026)
di: Zeng, Qixin, et al.
Pubblicazione: (2026)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
di: Deng, Shengliang, et al.
Pubblicazione: (2025)
di: Deng, Shengliang, et al.
Pubblicazione: (2025)
Grasp, See, and Place: Efficient Unknown Object Rearrangement with Policy Structure Prior
di: Xu, Kechun, et al.
Pubblicazione: (2024)
di: Xu, Kechun, et al.
Pubblicazione: (2024)
Discovering Conceptual Knowledge with Analytic Ontology Templates for Articulated Objects
di: Sun, Jianhua, et al.
Pubblicazione: (2024)
di: Sun, Jianhua, et al.
Pubblicazione: (2024)
Towards Effective Utilization of Mixed-Quality Demonstrations in Robotic Manipulation via Segment-Level Selection and Optimization
di: Chen, Jingjing, et al.
Pubblicazione: (2024)
di: Chen, Jingjing, et al.
Pubblicazione: (2024)
Joint-Aligned Latent Action: Towards Scalable VLA Pretraining in the Wild
di: Luo, Hao, et al.
Pubblicazione: (2026)
di: Luo, Hao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
di: Zhang, Yizheng, et al.
Pubblicazione: (2025) -
Motion Before Action: Diffusing Object Motion as Manipulation Condition
di: Su, Yue, et al.
Pubblicazione: (2024) -
Towards Human-Like Manipulation through RL-Augmented Teleoperation and Mixture-of-Dexterous-Experts VLA
di: Tang, Tutian, et al.
Pubblicazione: (2026) -
Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration
di: Ding, Pengxiang, et al.
Pubblicazione: (2025) -
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
di: Xu, Yifu, et al.
Pubblicazione: (2026)