Action-Geometry Prediction with 3D Geometric Prior for Bimanual Manipulation
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Chongyang, Li, Haipeng, Cheng, Shen, Hu, Jingyu, Fan, Haoqiang, Feng, Ziliang, Liu, Shuaicheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HeRO: Hierarchical 3D Semantic Representation for Pose-aware Object Manipulation
por: Xu, Chongyang, et al.
Publicado: (2026)
por: Xu, Chongyang, et al.
Publicado: (2026)
Ada3Drift: Adaptive Training-Time Drifting for One-Step 3D Visuomotor Robotic Manipulation
por: Xu, Chongyang, et al.
Publicado: (2026)
por: Xu, Chongyang, et al.
Publicado: (2026)
You Only Look Around: Learning Illumination Invariant Feature for Low-light Object Detection
por: Hong, Mingbo, et al.
Publicado: (2024)
por: Hong, Mingbo, et al.
Publicado: (2024)
Hierarchical Neural Semantic Representation for 3D Semantic Correspondence
por: Du, Keyu, et al.
Publicado: (2025)
por: Du, Keyu, et al.
Publicado: (2025)
LaS-Comp: Zero-shot 3D Completion with Latent-Spatial Consistency
por: Yan, Weilong, et al.
Publicado: (2026)
por: Yan, Weilong, et al.
Publicado: (2026)
Neural Spectral Decomposition for Dataset Distillation
por: Yang, Shaolei, et al.
Publicado: (2024)
por: Yang, Shaolei, et al.
Publicado: (2024)
Coding-Prior Guided Diffusion Network for Video Deblurring
por: Liu, Yike, et al.
Publicado: (2025)
por: Liu, Yike, et al.
Publicado: (2025)
PEGAsus: 3D Personalization of Geometry and Appearance
por: Hu, Jingyu, et al.
Publicado: (2026)
por: Hu, Jingyu, et al.
Publicado: (2026)
Blind-Spot Guided Diffusion for Self-supervised Real-World Denoising
por: Cheng, Shen, et al.
Publicado: (2025)
por: Cheng, Shen, et al.
Publicado: (2025)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
por: Xu, Hao, et al.
Publicado: (2024)
por: Xu, Hao, et al.
Publicado: (2024)
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
por: Tan, Hengkai, et al.
Publicado: (2025)
por: Tan, Hengkai, et al.
Publicado: (2025)
HybridReg: Robust 3D Point Cloud Registration with Hybrid Motions
por: Du, Keyu, et al.
Publicado: (2025)
por: Du, Keyu, et al.
Publicado: (2025)
PEAfowl: Perception-Enhanced Multi-View Vision-Language-Action for Bimanual Manipulation
por: Fan, Qingyu, et al.
Publicado: (2026)
por: Fan, Qingyu, et al.
Publicado: (2026)
CodingHomo: Bootstrapping Deep Homography With Video Coding
por: Liu, Yike, et al.
Publicado: (2025)
por: Liu, Yike, et al.
Publicado: (2025)
Adapting Human Mesh Recovery with Vision-Language Feedback
por: Xu, Chongyang, et al.
Publicado: (2025)
por: Xu, Chongyang, et al.
Publicado: (2025)
Estimating 2D Camera Motion with Hybrid Motion Basis
por: Li, Haipeng, et al.
Publicado: (2025)
por: Li, Haipeng, et al.
Publicado: (2025)
Large Pre-Trained Models for Bimanual Manipulation in 3D
por: Yurchyk, Hanna, et al.
Publicado: (2025)
por: Yurchyk, Hanna, et al.
Publicado: (2025)
GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction
por: Lin, Weiquan, et al.
Publicado: (2026)
por: Lin, Weiquan, et al.
Publicado: (2026)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
por: Shi, Hao, et al.
Publicado: (2025)
por: Shi, Hao, et al.
Publicado: (2025)
Diff-Shadow: Global-guided Diffusion Model for Shadow Removal
por: Luo, Jinting, et al.
Publicado: (2024)
por: Luo, Jinting, et al.
Publicado: (2024)
VTAO-BiManip: Masked Visual-Tactile-Action Pre-training with Object Understanding for Bimanual Dexterous Manipulation
por: Sun, Zhengnan, et al.
Publicado: (2025)
por: Sun, Zhengnan, et al.
Publicado: (2025)
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
por: Chen, Cheng, et al.
Publicado: (2024)
por: Chen, Cheng, et al.
Publicado: (2024)
GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
por: Fu, Xiao, et al.
Publicado: (2024)
por: Fu, Xiao, et al.
Publicado: (2024)
StableMotion: Repurposing Diffusion-Based Image Priors for Motion Estimation
por: Wang, Ziyi, et al.
Publicado: (2025)
por: Wang, Ziyi, et al.
Publicado: (2025)
AnchorSplat: Feed-Forward 3D Gaussian Splatting with 3D Geometric Priors
por: Zhang, Xiaoxue, et al.
Publicado: (2026)
por: Zhang, Xiaoxue, et al.
Publicado: (2026)
GeoTeacher: Geometry-Guided Semi-Supervised 3D Object Detection
por: Li, Jingyu, et al.
Publicado: (2025)
por: Li, Jingyu, et al.
Publicado: (2025)
GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
por: Qian, Jingjing, et al.
Publicado: (2025)
por: Qian, Jingjing, et al.
Publicado: (2025)
OAKINK2: A Dataset of Bimanual Hands-Object Manipulation in Complex Task Completion
por: Zhan, Xinyu, et al.
Publicado: (2024)
por: Zhan, Xinyu, et al.
Publicado: (2024)
Unleashing Semantic and Geometric Priors for 3D Scene Completion
por: Chen, Shiyuan, et al.
Publicado: (2025)
por: Chen, Shiyuan, et al.
Publicado: (2025)
InterACT: Inter-dependency Aware Action Chunking with Hierarchical Attention Transformers for Bimanual Manipulation
por: Lee, Andrew, et al.
Publicado: (2024)
por: Lee, Andrew, et al.
Publicado: (2024)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
por: Yang, Feng, et al.
Publicado: (2025)
por: Yang, Feng, et al.
Publicado: (2025)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
por: Wang, Kuanning, et al.
Publicado: (2026)
por: Wang, Kuanning, et al.
Publicado: (2026)
Let Geometry GUIDE: Layer-wise Unrolling of Geometric Priors in Multimodal LLMs
por: Wang, Chongyu, et al.
Publicado: (2026)
por: Wang, Chongyu, et al.
Publicado: (2026)
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
por: Bi, Hongzhe, et al.
Publicado: (2025)
por: Bi, Hongzhe, et al.
Publicado: (2025)
Leveraging Geometric Priors for Unaligned Scene Change Detection
por: Liu, Ziling, et al.
Publicado: (2025)
por: Liu, Ziling, et al.
Publicado: (2025)
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
por: Dai, Yixiang, et al.
Publicado: (2025)
por: Dai, Yixiang, et al.
Publicado: (2025)
Leveraging 3D Geometric Priors in 2D Rotation Symmetry Detection
por: Seo, Ahyun, et al.
Publicado: (2025)
por: Seo, Ahyun, et al.
Publicado: (2025)
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
por: Li, Kailin, et al.
Publicado: (2025)
por: Li, Kailin, et al.
Publicado: (2025)
Geometric Context Transformer for Streaming 3D Reconstruction
por: Chen, Lin-Zhuo, et al.
Publicado: (2026)
por: Chen, Lin-Zhuo, et al.
Publicado: (2026)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
por: Liu, Xin, et al.
Publicado: (2024)
por: Liu, Xin, et al.
Publicado: (2024)
Ejemplares similares
-
HeRO: Hierarchical 3D Semantic Representation for Pose-aware Object Manipulation
por: Xu, Chongyang, et al.
Publicado: (2026) -
Ada3Drift: Adaptive Training-Time Drifting for One-Step 3D Visuomotor Robotic Manipulation
por: Xu, Chongyang, et al.
Publicado: (2026) -
You Only Look Around: Learning Illumination Invariant Feature for Low-light Object Detection
por: Hong, Mingbo, et al.
Publicado: (2024) -
Hierarchical Neural Semantic Representation for 3D Semantic Correspondence
por: Du, Keyu, et al.
Publicado: (2025) -
LaS-Comp: Zero-shot 3D Completion with Latent-Spatial Consistency
por: Yan, Weilong, et al.
Publicado: (2026)