D$^3$Fields: Dynamic 3D Descriptor Fields for Zero-Shot Generalizable Rearrangement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yixuan, Zhang, Mingtong, Li, Zhuoran, Kelestemur, Tarik, Driggs-Campbell, Katherine, Wu, Jiajun, Fei-Fei, Li, Li, Yunzhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
CuriousBot: Interactive Mobile Exploration via Actionable 3D Relational Object Graph
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
Dynamic 3D Gaussian Tracking for Graph-Based Neural Dynamics Modeling
von: Zhang, Mingtong, et al.
Veröffentlicht: (2024)
von: Zhang, Mingtong, et al.
Veröffentlicht: (2024)
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
Dream2Real: Zero-Shot 3D Object Rearrangement with Vision-Language Models
von: Kapelyukh, Ivan, et al.
Veröffentlicht: (2023)
von: Kapelyukh, Ivan, et al.
Veröffentlicht: (2023)
Zero-Shot 3D Visual Grounding from Vision-Language Models
von: Li, Rong, et al.
Veröffentlicht: (2025)
von: Li, Rong, et al.
Veröffentlicht: (2025)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
von: Li, Rong, et al.
Veröffentlicht: (2024)
von: Li, Rong, et al.
Veröffentlicht: (2024)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Particle-Grid Neural Dynamics for Learning Deformable Object Models from RGB-D Videos
von: Zhang, Kaifeng, et al.
Veröffentlicht: (2025)
von: Zhang, Kaifeng, et al.
Veröffentlicht: (2025)
VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation
von: Chen, Hanzhi, et al.
Veröffentlicht: (2025)
von: Chen, Hanzhi, et al.
Veröffentlicht: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
von: Huang, Wenlong, et al.
Veröffentlicht: (2024)
von: Huang, Wenlong, et al.
Veröffentlicht: (2024)
LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
von: Ze, Yanjie, et al.
Veröffentlicht: (2024)
Learning Part-Aware Dense 3D Feature Field for Generalizable Articulated Object Manipulation
von: Chen, Yue, et al.
Veröffentlicht: (2026)
von: Chen, Yue, et al.
Veröffentlicht: (2026)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
von: Chai, Ying, et al.
Veröffentlicht: (2025)
von: Chai, Ying, et al.
Veröffentlicht: (2025)
PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement
von: Jin, Shutong, et al.
Veröffentlicht: (2024)
von: Jin, Shutong, et al.
Veröffentlicht: (2024)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
von: Tian, Tongxuan, et al.
Veröffentlicht: (2025)
von: Tian, Tongxuan, et al.
Veröffentlicht: (2025)
MSGNav: Unleashing the Power of Multi-modal 3D Scene Graph for Zero-Shot Embodied Navigation
von: Huang, Xun, et al.
Veröffentlicht: (2025)
von: Huang, Xun, et al.
Veröffentlicht: (2025)
Neural Attention Field: Emerging Point Relevance in 3D Scenes for One-Shot Dexterous Grasping
von: Wang, Qianxu, et al.
Veröffentlicht: (2024)
von: Wang, Qianxu, et al.
Veröffentlicht: (2024)
Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
von: Ma, Ji, et al.
Veröffentlicht: (2024)
von: Ma, Ji, et al.
Veröffentlicht: (2024)
DegustaBot: Zero-Shot Visual Preference Estimation for Personalized Multi-Object Rearrangement
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
Zero-Shot UAV Navigation in Forests via Relightable 3D Gaussian Splatting
von: Lv, Zinan, et al.
Veröffentlicht: (2026)
von: Lv, Zinan, et al.
Veröffentlicht: (2026)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
CG-SLAM: Efficient Dense RGB-D SLAM in a Consistent Uncertainty-aware 3D Gaussian Field
von: Hu, Jiarui, et al.
Veröffentlicht: (2024)
von: Hu, Jiarui, et al.
Veröffentlicht: (2024)
AdaptiGraph: Material-Adaptive Graph-Based Neural Dynamics for Robotic Manipulation
von: Zhang, Kaifeng, et al.
Veröffentlicht: (2024)
von: Zhang, Kaifeng, et al.
Veröffentlicht: (2024)
An Expert Ensemble for Detecting Anomalous Scenes, Interactions, and Behaviors in Autonomous Driving
von: Ji, Tianchen, et al.
Veröffentlicht: (2025)
von: Ji, Tianchen, et al.
Veröffentlicht: (2025)
Anyview: Generalizable Indoor 3D Object Detection with Variable Frames
von: Wu, Zhenyu, et al.
Veröffentlicht: (2023)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2023)
FUSELOC: Fusing Global and Local Descriptors to Disambiguate 2D-3D Matching in Visual Localization
von: Nguyen, Son Tung, et al.
Veröffentlicht: (2024)
von: Nguyen, Son Tung, et al.
Veröffentlicht: (2024)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
Learning Generalizable 3D Manipulation With 10 Demonstrations
von: Ren, Yu, et al.
Veröffentlicht: (2024)
von: Ren, Yu, et al.
Veröffentlicht: (2024)
KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation
von: Liu, Zixian, et al.
Veröffentlicht: (2025)
von: Liu, Zixian, et al.
Veröffentlicht: (2025)
MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
von: Zhao, Wang, et al.
Veröffentlicht: (2024)
von: Zhao, Wang, et al.
Veröffentlicht: (2024)
FetchBot: Learning Generalizable Object Fetching in Cluttered Scenes via Zero-Shot Sim2Real
von: Liu, Weiheng, et al.
Veröffentlicht: (2025)
von: Liu, Weiheng, et al.
Veröffentlicht: (2025)
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
Click to Grasp: Zero-Shot Precise Manipulation via Visual Diffusion Descriptors
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2024)
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2024)
GNFactor: Multi-Task Real Robot Learning with Generalizable Neural Feature Fields
von: Ze, Yanjie, et al.
Veröffentlicht: (2023)
von: Ze, Yanjie, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
von: Wang, Yixuan, et al.
Veröffentlicht: (2024) -
CuriousBot: Interactive Mobile Exploration via Actionable 3D Relational Object Graph
von: Wang, Yixuan, et al.
Veröffentlicht: (2025) -
Dynamic 3D Gaussian Tracking for Graph-Based Neural Dynamics Modeling
von: Zhang, Mingtong, et al.
Veröffentlicht: (2024) -
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
von: Kamboj, Abhi, et al.
Veröffentlicht: (2024) -
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)