World In Your Hands: A Large-Scale and Open-Source Ecosystem for Learning Human-Centric Manipulation in the Wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Yupeng, Peng, Jichao, Li, Weize, Zheng, Yuhang, Li, Xiang, Jin, Yujie, Wei, Julong, Zhang, Guanhua, Zheng, Ruiling, Cao, Ming, Gu, Songen, Zou, Zhenhong, Li, Kaige, Wu, Ke, Yang, Mingmin, Liu, Jiahao, Li, Pengfei, Si, Hengjie, Zhu, Feiyu, Fu, Wang, Wang, Likun, Yao, Ruiwen, Zhao, Jieru, Chen, Yilun, Ding, Wenchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis
von: Gu, Songen, et al.
Veröffentlicht: (2026)
von: Gu, Songen, et al.
Veröffentlicht: (2026)
UniArt: Unified 3D Representation for Generating 3D Articulated Objects with Open-Set Articulation
von: Jin, Bu, et al.
Veröffentlicht: (2025)
von: Jin, Bu, et al.
Veröffentlicht: (2025)
OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation
von: Zheng, Yuhang, et al.
Veröffentlicht: (2026)
von: Zheng, Yuhang, et al.
Veröffentlicht: (2026)
Learning High-Frequency Continuous Action Chunks in Latent Space
von: Wang, Kunyun, et al.
Veröffentlicht: (2026)
von: Wang, Kunyun, et al.
Veröffentlicht: (2026)
PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance
von: Zheng, Yupeng, et al.
Veröffentlicht: (2026)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2026)
Semi-Supervised Vision-Centric 3D Occupancy World Model for Autonomous Driving
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Enhancing Indoor Occupancy Prediction via Sparse Query-Based Multi-Level Consistent Knowledge Distillation
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Rhythm: Learning Interactive Whole-Body Control for Dual Humanoids
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
Spectral extrema of graphs of given even size forbidding H(4,3)
von: Zheng, Ruiling, et al.
Veröffentlicht: (2025)
von: Zheng, Ruiling, et al.
Veröffentlicht: (2025)
TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes
von: Jin, Bu, et al.
Veröffentlicht: (2024)
von: Jin, Bu, et al.
Veröffentlicht: (2024)
The Conjugacy Relation on One-sided Subshifts is Non-treeable
von: Li, Ruiwen
Veröffentlicht: (2026)
von: Li, Ruiwen
Veröffentlicht: (2026)
OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving
von: Wei, Julong, et al.
Veröffentlicht: (2024)
von: Wei, Julong, et al.
Veröffentlicht: (2024)
Acoustic Volume Rendering for Neural Impulse Response Fields
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
Data Scaling Laws for Imitation Learning-Based End-to-End Autonomous Driving
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
MonoOcc: Digging into Monocular Semantic Occupancy Prediction
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
GaussianGrasper: 3D Language Gaussian Splatting for Open-vocabulary Robotic Grasping
von: Zheng, Yuhang, et al.
Veröffentlicht: (2024)
von: Zheng, Yuhang, et al.
Veröffentlicht: (2024)
Self‐propulsion of a droplet induced by combined diffusiophoresis and Marangoni effects
von: Yuhang Wang, et al.
Veröffentlicht: (2024)
von: Yuhang Wang, et al.
Veröffentlicht: (2024)
Mimir: Hierarchical Goal-Driven Diffusion with Uncertainty Propagation for End-to-End Autonomous Driving
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
OccTENS: 3D Occupancy World Model via Temporal Next-Scale Prediction
von: Jin, Bu, et al.
Veröffentlicht: (2025)
von: Jin, Bu, et al.
Veröffentlicht: (2025)
Additive‐Regulated Interface Chemistry Enables Depolarization for Ultra‐High Capacity LiCoO 2
von: Guorui Zheng, et al.
Veröffentlicht: (2025)
von: Guorui Zheng, et al.
Veröffentlicht: (2025)
AI Code in the Wild: Measuring Security Risks and Ecosystem Shifts of AI-Generated Code in Modern Software
von: Wang, Bin, et al.
Veröffentlicht: (2025)
von: Wang, Bin, et al.
Veröffentlicht: (2025)
Decentralized Intelligence in GameFi: Embodied AI Agents and the Convergence of DeFi and Virtual Ecosystems
von: Jia, Fernando, et al.
Veröffentlicht: (2024)
von: Jia, Fernando, et al.
Veröffentlicht: (2024)
On Explicit Tuning Laws of Active Disturbance Rejection Control for Nonlinear Uncertain Systems Under Quantized Sampled‐Data Measurements
von: Feiyu Xiang, et al.
Veröffentlicht: (2025)
von: Feiyu Xiang, et al.
Veröffentlicht: (2025)
HGS-Mapping: Online Dense Mapping Using Hybrid Gaussian Representation in Urban Scenes
von: Wu, Ke, et al.
Veröffentlicht: (2024)
von: Wu, Ke, et al.
Veröffentlicht: (2024)
O2V-Mapping: Online Open-Vocabulary Mapping with Neural Implicit Representation
von: Tie, Muer, et al.
Veröffentlicht: (2024)
von: Tie, Muer, et al.
Veröffentlicht: (2024)
knot-experiment
von: Zheng, Yujie
Veröffentlicht: (2025)
von: Zheng, Yujie
Veröffentlicht: (2025)
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
Controllability Test for Nonlinear Datatic Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
On the Equilibrium between Feasible Zone and Uncertain Model in Safe Exploration
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
On the Stability of Datatic Control Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
Integrated Production and Transportation Problem With Order Waiting
von: Yuejuan Zhu, et al.
Veröffentlicht: (2025)
von: Yuejuan Zhu, et al.
Veröffentlicht: (2025)
MemRL: Self-Evolving Agents via Runtime Reinforcement Learning on Episodic Memory
von: Zhang, Shengtao, et al.
Veröffentlicht: (2026)
von: Zhang, Shengtao, et al.
Veröffentlicht: (2026)
Isomorphism of pointed minimal systems is not classifiable by countable structures
von: Li, Ruiwen, et al.
Veröffentlicht: (2024)
von: Li, Ruiwen, et al.
Veröffentlicht: (2024)
Subshifts of finite symbolic rank
von: Gao, Su, et al.
Veröffentlicht: (2023)
von: Gao, Su, et al.
Veröffentlicht: (2023)
Performance analysis of satellite-terrestrial integrated radio access networks based on stochastic geometry
von: Sun, Yaohua, et al.
Veröffentlicht: (2024)
von: Sun, Yaohua, et al.
Veröffentlicht: (2024)
A container theorem for general digraphs with forbidden subdigraphs
von: Liang, Meili, et al.
Veröffentlicht: (2026)
von: Liang, Meili, et al.
Veröffentlicht: (2026)
Constructing BERT Models: How Team Dynamics and Focus Shape AI Model Impact
von: Cao, Likun, et al.
Veröffentlicht: (2026)
von: Cao, Likun, et al.
Veröffentlicht: (2026)
LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2025)
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2025)
Extreme Thermal Insulation and Tradeoff of Thermal Transport Mechanisms between Graphene and WS2 Monolayers
von: Ruiling Zhang, et al.
Veröffentlicht: (2024)
von: Ruiling Zhang, et al.
Veröffentlicht: (2024)
Next-Scale Autoregressive Models for Text-to-Motion Generation
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis
von: Gu, Songen, et al.
Veröffentlicht: (2026) -
UniArt: Unified 3D Representation for Generating 3D Articulated Objects with Open-Set Articulation
von: Jin, Bu, et al.
Veröffentlicht: (2025) -
OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation
von: Zheng, Yuhang, et al.
Veröffentlicht: (2026) -
Learning High-Frequency Continuous Action Chunks in Latent Space
von: Wang, Kunyun, et al.
Veröffentlicht: (2026) -
PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance
von: Zheng, Yupeng, et al.
Veröffentlicht: (2026)