Saved in:
| Main Authors: | Wang, Hanqing, Chen, Jiahe, Huang, Wensi, Ben, Qingwei, Wang, Tai, Mi, Boyu, Huang, Tao, Zhao, Siheng, Chen, Yilun, Yang, Sizhe, Cao, Peizhou, Yu, Wenye, Ye, Zichao, Li, Jialun, Long, Junfeng, Wang, Zirui, Wang, Huiling, Zhao, Ying, Tu, Zhongying, Qiao, Yu, Lin, Dahua, Pang, Jiangmiao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.10943 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning H-Infinity Locomotion Control
by: Long, Junfeng, et al.
Published: (2024)
by: Long, Junfeng, et al.
Published: (2024)
Language-to-Space Programming for Training-Free 3D Visual Grounding
by: Mi, Boyu, et al.
Published: (2025)
by: Mi, Boyu, et al.
Published: (2025)
Novel Demonstration Generation with Gaussian Splatting Enables Robust One-Shot Manipulation
by: Yang, Sizhe, et al.
Published: (2025)
by: Yang, Sizhe, et al.
Published: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
by: Xu, Runsen, et al.
Published: (2024)
by: Xu, Runsen, et al.
Published: (2024)
VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs
by: Huang, Wensi, et al.
Published: (2025)
by: Huang, Wensi, et al.
Published: (2025)
Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM
by: Huang, Haifeng, et al.
Published: (2026)
by: Huang, Haifeng, et al.
Published: (2026)
PointLLM: Empowering Large Language Models to Understand Point Clouds
by: Xu, Runsen, et al.
Published: (2023)
by: Xu, Runsen, et al.
Published: (2023)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
VB-Com: Learning Vision-Blind Composite Humanoid Locomotion Against Deficient Perception
by: Ren, Junli, et al.
Published: (2025)
by: Ren, Junli, et al.
Published: (2025)
OVExp: Open Vocabulary Exploration for Object-Oriented Navigation
by: Wei, Meng, et al.
Published: (2024)
by: Wei, Meng, et al.
Published: (2024)
BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds
by: Wang, Huayi, et al.
Published: (2025)
by: Wang, Huayi, et al.
Published: (2025)
Rethinking the Embodied Gap in Vision-and-Language Navigation: A Holistic Study of Physical and Visual Disparities
by: Wang, Liuyi, et al.
Published: (2025)
by: Wang, Liuyi, et al.
Published: (2025)
MGF: Mixed Gaussian Flow for Diverse Trajectory Prediction
by: Chen, Jiahe, et al.
Published: (2024)
by: Chen, Jiahe, et al.
Published: (2024)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
by: Huang, Linghan, et al.
Published: (2025)
by: Huang, Linghan, et al.
Published: (2025)
Grounded 3D-LLM with Referent Tokens
by: Chen, Yilun, et al.
Published: (2024)
by: Chen, Yilun, et al.
Published: (2024)
Learning Humanoid Standing-up Control across Diverse Postures
by: Huang, Tao, et al.
Published: (2025)
by: Huang, Tao, et al.
Published: (2025)
Feel Robot Feels: Tactile Feedback Array Glove for Dexterous Manipulation
by: Jia, Feiyu, et al.
Published: (2026)
by: Jia, Feiyu, et al.
Published: (2026)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
by: Yang, Shuai, et al.
Published: (2025)
by: Yang, Shuai, et al.
Published: (2025)
GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation
by: Gao, Ning, et al.
Published: (2025)
by: Gao, Ning, et al.
Published: (2025)
Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
by: Tian, Yang, et al.
Published: (2024)
by: Tian, Yang, et al.
Published: (2024)
PhysHSI: Towards a Real-World Generalizable and Natural Humanoid-Scene Interaction System
by: Wang, Huayi, et al.
Published: (2025)
by: Wang, Huayi, et al.
Published: (2025)
Learning Humanoid Locomotion with Perceptive Internal Model
by: Long, Junfeng, et al.
Published: (2024)
by: Long, Junfeng, et al.
Published: (2024)
Cloud-OpsBench: A Reproducible Benchmark for Agentic Root Cause Analysis in Cloud Systems
by: Wang, Yilun, et al.
Published: (2026)
by: Wang, Yilun, et al.
Published: (2026)
On the Challenges of Fuzzing Techniques via Large Language Models
by: Huang, Linghan, et al.
Published: (2024)
by: Huang, Linghan, et al.
Published: (2024)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
Hybrid Internal Model: Learning Agile Legged Locomotion with Simulated Robot Response
by: Long, Junfeng, et al.
Published: (2023)
by: Long, Junfeng, et al.
Published: (2023)
NavDP: Learning Sim-to-Real Navigation Diffusion Policy with Privileged Information Guidance
by: Cai, Wenzhe, et al.
Published: (2025)
by: Cai, Wenzhe, et al.
Published: (2025)
HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit
by: Ben, Qingwei, et al.
Published: (2025)
by: Ben, Qingwei, et al.
Published: (2025)
Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers
by: Huang, Haifeng, et al.
Published: (2023)
by: Huang, Haifeng, et al.
Published: (2023)
Gallant: Voxel Grid-based Humanoid Locomotion and Local-navigation across 3D Constrained Terrains
by: Ben, Qingwei, et al.
Published: (2025)
by: Ben, Qingwei, et al.
Published: (2025)
Towards Adaptable Humanoid Control via Adaptive Motion Tracking
by: Huang, Tao, et al.
Published: (2025)
by: Huang, Tao, et al.
Published: (2025)
Humanoid Goalkeeper: Learning from Position Conditioned Task-Motion Constraints
by: Ren, Junli, et al.
Published: (2025)
by: Ren, Junli, et al.
Published: (2025)
What Makes CLIP More Robust to Long-Tailed Pre-Training Data? A Controlled Study for Transferable Insights
by: Wen, Xin, et al.
Published: (2024)
by: Wen, Xin, et al.
Published: (2024)
A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning
by: Wen, Xin, et al.
Published: (2025)
by: Wen, Xin, et al.
Published: (2025)
MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations
by: Lyu, Ruiyuan, et al.
Published: (2024)
by: Lyu, Ruiyuan, et al.
Published: (2024)
InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts
by: Zhong, Weipeng, et al.
Published: (2025)
by: Zhong, Weipeng, et al.
Published: (2025)
Robo3R: Enhancing Robotic Manipulation with Accurate Feed-Forward 3D Reconstruction
by: Yang, Sizhe, et al.
Published: (2026)
by: Yang, Sizhe, et al.
Published: (2026)
UltraDexGrasp: Learning Universal Dexterous Grasping for Bimanual Robots with Synthetic Data
by: Yang, Sizhe, et al.
Published: (2026)
by: Yang, Sizhe, et al.
Published: (2026)
StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling
by: Wei, Meng, et al.
Published: (2025)
by: Wei, Meng, et al.
Published: (2025)
DreamArt: Generating Interactable Articulated Objects from a Single Image
by: Lu, Ruijie, et al.
Published: (2025)
by: Lu, Ruijie, et al.
Published: (2025)
Similar Items
-
Learning H-Infinity Locomotion Control
by: Long, Junfeng, et al.
Published: (2024) -
Language-to-Space Programming for Training-Free 3D Visual Grounding
by: Mi, Boyu, et al.
Published: (2025) -
Novel Demonstration Generation with Gaussian Splatting Enables Robust One-Shot Manipulation
by: Yang, Sizhe, et al.
Published: (2025) -
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
by: Xu, Runsen, et al.
Published: (2024) -
VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs
by: Huang, Wensi, et al.
Published: (2025)