Coarse-to-Fine 3D Keyframe Transporter
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Xupeng, Klee, David, Wang, Dian, Hu, Boce, Huang, Haojie, Tangri, Arsh, Walters, Robin, Platt, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3D Equivariant Visuomotor Policy Learning via Spherical Projection
by: Hu, Boce, et al.
Published: (2025)
by: Hu, Boce, et al.
Published: (2025)
MATCH POLICY: A Simple Pipeline from Point Cloud Registration to Manipulation Policies
by: Huang, Haojie, et al.
Published: (2024)
by: Huang, Haojie, et al.
Published: (2024)
Push-Grasp Policy Learning Using Equivariant Models and Grasp Score Optimization
by: Hu, Boce, et al.
Published: (2025)
by: Hu, Boce, et al.
Published: (2025)
Fourier Transporter: Bi-Equivariant Robotic Manipulation in 3D
by: Huang, Haojie, et al.
Published: (2024)
by: Huang, Haojie, et al.
Published: (2024)
OrbitGrasp: $SE(3)$-Equivariant Grasp Learning
by: Hu, Boce, et al.
Published: (2024)
by: Hu, Boce, et al.
Published: (2024)
Clebsch-Gordan Transformer: Fast and Global Equivariant Attention
by: Howell, Owen Lewis, et al.
Published: (2025)
by: Howell, Owen Lewis, et al.
Published: (2025)
Equivariant Offline Reinforcement Learning
by: Tangri, Arsh, et al.
Published: (2024)
by: Tangri, Arsh, et al.
Published: (2024)
Equivariant Reinforcement Learning under Partial Observability
by: Nguyen, Hai, et al.
Published: (2024)
by: Nguyen, Hai, et al.
Published: (2024)
Equivariant Goal Conditioned Contrastive Reinforcement Learning
by: Tangri, Arsh, et al.
Published: (2025)
by: Tangri, Arsh, et al.
Published: (2025)
A Practical Guide for Incorporating Symmetry in Diffusion Policy
by: Wang, Dian, et al.
Published: (2025)
by: Wang, Dian, et al.
Published: (2025)
A Coarse-to-Fine Approach to Multi-Modality 3D Occupancy Grounding
by: Shi, Zhan, et al.
Published: (2025)
by: Shi, Zhan, et al.
Published: (2025)
KeySG: Hierarchical Keyframe-Based 3D Scene Graphs
by: Werby, Abdelrhman, et al.
Published: (2025)
by: Werby, Abdelrhman, et al.
Published: (2025)
Adaptive Keyframe Selection for Scalable 3D Scene Reconstruction in Dynamic Environments
by: Jha, Raman, et al.
Published: (2025)
by: Jha, Raman, et al.
Published: (2025)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
by: Jiang, Zebin, et al.
Published: (2025)
by: Jiang, Zebin, et al.
Published: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
by: He, Ziheng, et al.
Published: (2026)
by: He, Ziheng, et al.
Published: (2026)
Keyframe-Based Feed-Forward Visual Odometry
by: Dai, Weichen, et al.
Published: (2026)
by: Dai, Weichen, et al.
Published: (2026)
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
by: Gong, Zhefei, et al.
Published: (2024)
by: Gong, Zhefei, et al.
Published: (2024)
Generalizable Hierarchical Skill Learning via Object-Centric Representation
by: Zhao, Haibo, et al.
Published: (2025)
by: Zhao, Haibo, et al.
Published: (2025)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
Compact Keyframe-Optimized Multi-Agent Gaussian Splatting SLAM
by: Li, Monica M. Q., et al.
Published: (2026)
by: Li, Monica M. Q., et al.
Published: (2026)
Keyframe-based Dense Mapping with the Graph of View-Dependent Local Maps
by: Zielinski, Krzysztof, et al.
Published: (2026)
by: Zielinski, Krzysztof, et al.
Published: (2026)
TripleMixer: A 3D Point Cloud Denoising Model for Adverse Weather
by: Zhao, Xiongwei, et al.
Published: (2024)
by: Zhao, Xiongwei, et al.
Published: (2024)
A Minimal Subset Approach for Informed Keyframe Sampling in Large-Scale SLAM
by: Stathoulopoulos, Nikolaos, et al.
Published: (2025)
by: Stathoulopoulos, Nikolaos, et al.
Published: (2025)
EquAct: An SE(3)-Equivariant Multi-Task Transformer for Open-Loop Robotic Manipulation
by: Zhu, Xupeng, et al.
Published: (2025)
by: Zhu, Xupeng, et al.
Published: (2025)
Dissecting Embodied Abilities in Multimodal Language Models through Skill-level Evaluation and Diagnosis
by: Qi, Yu, et al.
Published: (2025)
by: Qi, Yu, et al.
Published: (2025)
A Coarse-to-Fine Place Recognition Approach using Attention-guided Descriptors and Overlap Estimation
by: Fu, Chencan, et al.
Published: (2023)
by: Fu, Chencan, et al.
Published: (2023)
FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation
by: Shao, Dian, et al.
Published: (2026)
by: Shao, Dian, et al.
Published: (2026)
Why Sample Space Matters: Keyframe Sampling Optimization for LiDAR-based Place Recognition
by: Stathoulopoulos, Nikolaos, et al.
Published: (2024)
by: Stathoulopoulos, Nikolaos, et al.
Published: (2024)
Solving Short-Term Relocalization Problems In Monocular Keyframe Visual SLAM Using Spatial And Semantic Data
by: Kamal, Azmyin Md., et al.
Published: (2024)
by: Kamal, Azmyin Md., et al.
Published: (2024)
LiDAR-VGGT: Cross-Modal Coarse-to-Fine Fusion for Globally Consistent and Metric-Scale Dense Mapping
by: Wang, Lijie, et al.
Published: (2025)
by: Wang, Lijie, et al.
Published: (2025)
WAM-Flow: Parallel Coarse-to-Fine Motion Planning via Discrete Flow Matching for Autonomous Driving
by: Xu, Yifang, et al.
Published: (2025)
by: Xu, Yifang, et al.
Published: (2025)
CoFi: Coarse-to-Fine ICP for LiDAR Localization in an Efficient Long-lasting Point Cloud Map
by: Lyu, Yecheng, et al.
Published: (2021)
by: Lyu, Yecheng, et al.
Published: (2021)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
CoFiI2P: Coarse-to-Fine Correspondences for Image-to-Point Cloud Registration
by: Kang, Shuhao, et al.
Published: (2023)
by: Kang, Shuhao, et al.
Published: (2023)
Is Your LiDAR Placement Optimized for 3D Scene Understanding?
by: Li, Ye, et al.
Published: (2024)
by: Li, Ye, et al.
Published: (2024)
LoopSplat: Loop Closure by Registering 3D Gaussian Splats
by: Zhu, Liyuan, et al.
Published: (2024)
by: Zhu, Liyuan, et al.
Published: (2024)
SGFormer: Satellite-Ground Fusion for 3D Semantic Scene Completion
by: Guo, Xiyue, et al.
Published: (2025)
by: Guo, Xiyue, et al.
Published: (2025)
REACT3D: Recovering Articulations for Interactive Physical 3D Scenes
by: Huang, Zhao, et al.
Published: (2025)
by: Huang, Zhao, et al.
Published: (2025)
3D and 4D World Modeling: A Survey
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
More than Segmentation: Benchmarking SAM 3 for Segmentation, 3D Perception, and Reconstruction in Robotic Surgery
by: Dong, Wenzhen, et al.
Published: (2025)
by: Dong, Wenzhen, et al.
Published: (2025)
Similar Items
-
3D Equivariant Visuomotor Policy Learning via Spherical Projection
by: Hu, Boce, et al.
Published: (2025) -
MATCH POLICY: A Simple Pipeline from Point Cloud Registration to Manipulation Policies
by: Huang, Haojie, et al.
Published: (2024) -
Push-Grasp Policy Learning Using Equivariant Models and Grasp Score Optimization
by: Hu, Boce, et al.
Published: (2025) -
Fourier Transporter: Bi-Equivariant Robotic Manipulation in 3D
by: Huang, Haojie, et al.
Published: (2024) -
OrbitGrasp: $SE(3)$-Equivariant Grasp Learning
by: Hu, Boce, et al.
Published: (2024)