DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations
Fuente:
arXiv
Saved in:
| Main Authors: | Son, Dongwon, Son, Sanghyeon, Kim, Jaehyung, Kim, Beomjoon |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An intuitive multi-frequency feature representation for SO(3)-equivariant networks
by: Son, Dongwon, et al.
Published: (2024)
by: Son, Dongwon, et al.
Published: (2024)
NeuralSVCD for Efficient Swept Volume Collision Detection
by: Son, Dongwon, et al.
Published: (2025)
by: Son, Dongwon, et al.
Published: (2025)
FUSELOC: Fusing Global and Local Descriptors to Disambiguate 2D-3D Matching in Visual Localization
by: Nguyen, Son Tung, et al.
Published: (2024)
by: Nguyen, Son Tung, et al.
Published: (2024)
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Multi-step manipulation task and motion planning guided by video demonstration
by: Zorina, Kateryna, et al.
Published: (2025)
by: Zorina, Kateryna, et al.
Published: (2025)
Material-informed Gaussian Splatting for 3D World Reconstruction in a Digital Twin
by: Huynh, Andy, et al.
Published: (2025)
by: Huynh, Andy, et al.
Published: (2025)
Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
by: Kim, Heecheol, et al.
Published: (2022)
by: Kim, Heecheol, et al.
Published: (2022)
A low-cost and lightweight 6 DoF bimanual arm for dynamic and contact-rich manipulation
by: Kim, Jaehyung, et al.
Published: (2025)
by: Kim, Jaehyung, et al.
Published: (2025)
SegVec3D: A Method for Vector Embedding of 3D Objects Oriented Towards Robot manipulation
by: Kang, Zhihan, et al.
Published: (2025)
by: Kang, Zhihan, et al.
Published: (2025)
MessyKitchens: Contact-rich object-level 3D scene reconstruction
by: Ansari, Junaid Ahmed, et al.
Published: (2026)
by: Ansari, Junaid Ahmed, et al.
Published: (2026)
Functionality understanding and segmentation in 3D scenes
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
RoEL: Robust Event-based 3D Line Reconstruction
by: Bae, Gwangtak, et al.
Published: (2026)
by: Bae, Gwangtak, et al.
Published: (2026)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
by: Kim, Dongwon, et al.
Published: (2026)
by: Kim, Dongwon, et al.
Published: (2026)
LIVE-GS: Online LiDAR-Inertial-Visual State Estimation and Globally Consistent Mapping with 3D Gaussian Splatting
by: Park, Jaeseok, et al.
Published: (2025)
by: Park, Jaeseok, et al.
Published: (2025)
TranSplat: Surface Embedding-guided 3D Gaussian Splatting for Transparent Object Manipulation
by: Kim, Jeongyun, et al.
Published: (2025)
by: Kim, Jeongyun, et al.
Published: (2025)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
Quantifying Context Bias in Domain Adaptation for Object Detection
by: Son, Hojun, et al.
Published: (2024)
by: Son, Hojun, et al.
Published: (2024)
Clutt3R-Seg: Sparse-view 3D Instance Segmentation for Language-grounded Grasping in Cluttered Scenes
by: Noh, Jeongho, et al.
Published: (2026)
by: Noh, Jeongho, et al.
Published: (2026)
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025)
by: Kim, Young Hun, et al.
Published: (2025)
TRAN-D: 2D Gaussian Splatting-based Sparse-view Transparent Object Depth Reconstruction via Physics Simulation for Scene Update
by: Kim, Jeongyun, et al.
Published: (2025)
by: Kim, Jeongyun, et al.
Published: (2025)
RE-TRIP : Reflectivity Instance Augmented Triangle Descriptor for 3D Place Recognition
by: Park, Yechan, et al.
Published: (2025)
by: Park, Yechan, et al.
Published: (2025)
Context-Nav: Context-Driven Exploration and Viewpoint-Aware 3D Spatial Reasoning for Instance Navigation
by: Jang, Won Shik, et al.
Published: (2026)
by: Jang, Won Shik, et al.
Published: (2026)
Implicit 3D scene reconstruction using deep learning towards efficient collision understanding in autonomous driving
by: Ramanayake, Akarshani, et al.
Published: (2025)
by: Ramanayake, Akarshani, et al.
Published: (2025)
VPOcc: Exploiting Vanishing Point for 3D Semantic Occupancy Prediction
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
WHU-PCPR: A cross-platform heterogeneous point cloud dataset for place recognition in complex urban scenes
by: Zou, Xianghong, et al.
Published: (2026)
by: Zou, Xianghong, et al.
Published: (2026)
MrGS: Multi-modal Radiance Fields with 3D Gaussian Splatting for RGB-Thermal Novel View Synthesis
by: Kweon, Minseong, et al.
Published: (2025)
by: Kweon, Minseong, et al.
Published: (2025)
Point2Pose: Occlusion-Recovering 6D Pose Tracking and 3D Reconstruction for Multiple Unknown Objects Via 2D Point Trackers
by: Lin, Tzu-Yuan, et al.
Published: (2026)
by: Lin, Tzu-Yuan, et al.
Published: (2026)
3D Reconstruction-Based Seed Counting of Sorghum Panicles for Agricultural Inspection
by: Freeman, Harry, et al.
Published: (2022)
by: Freeman, Harry, et al.
Published: (2022)
Reliable-loc: Robust sequential LiDAR global localization in large-scale street scenes based on verifiable cues
by: Zou, Xianghong, et al.
Published: (2024)
by: Zou, Xianghong, et al.
Published: (2024)
LEXI-SG: Monocular 3D Scene Graph Mapping with Room-Guided Feed-Forward Reconstruction
by: Kassab, Christina, et al.
Published: (2026)
by: Kassab, Christina, et al.
Published: (2026)
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
by: Zhao, Zifan, et al.
Published: (2025)
by: Zhao, Zifan, et al.
Published: (2025)
Vision-based robot manipulation of transparent liquid containers in a laboratory setting
by: Schober, Daniel, et al.
Published: (2024)
by: Schober, Daniel, et al.
Published: (2024)
DynaWeightPnP: Toward global real-time 3D-2D solver in PnP without correspondences
by: Song, Jingwei, et al.
Published: (2024)
by: Song, Jingwei, et al.
Published: (2024)
Beyond the Patch: Exploring Vulnerabilities of Visuomotor Policies via Viewpoint-Consistent 3D Adversarial Object
by: Lee, Chanmi, et al.
Published: (2026)
by: Lee, Chanmi, et al.
Published: (2026)
MixSup: Mixed-grained Supervision for Label-efficient LiDAR-based 3D Object Detection
by: Yang, Yuxue, et al.
Published: (2024)
by: Yang, Yuxue, et al.
Published: (2024)
Real-time 3D semantic occupancy prediction for autonomous vehicles using memory-efficient sparse convolution
by: Sze, Samuel, et al.
Published: (2024)
by: Sze, Samuel, et al.
Published: (2024)
Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference
by: Kang, Beomseok, et al.
Published: (2026)
by: Kang, Beomseok, et al.
Published: (2026)
VIGS SLAM: IMU-based Large-Scale 3D Gaussian Splatting SLAM
by: Pak, Gyuhyeon, et al.
Published: (2025)
by: Pak, Gyuhyeon, et al.
Published: (2025)
PeLiCal: Targetless Extrinsic Calibration via Penetrating Lines for RGB-D Cameras with Limited Co-visibility
by: Shin, Jaeho, et al.
Published: (2024)
by: Shin, Jaeho, et al.
Published: (2024)
Viser: Imperative, Web-based 3D Visualization in Python
by: Yi, Brent, et al.
Published: (2025)
by: Yi, Brent, et al.
Published: (2025)
Similar Items
-
An intuitive multi-frequency feature representation for SO(3)-equivariant networks
by: Son, Dongwon, et al.
Published: (2024) -
NeuralSVCD for Efficient Swept Volume Collision Detection
by: Son, Dongwon, et al.
Published: (2025) -
FUSELOC: Fusing Global and Local Descriptors to Disambiguate 2D-3D Matching in Visual Localization
by: Nguyen, Son Tung, et al.
Published: (2024) -
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025) -
Multi-step manipulation task and motion planning guided by video demonstration
by: Zorina, Kateryna, et al.
Published: (2025)