Point Cloud Models Improve Visual Robustness in Robotic Learners
Fuente:
arXiv
Saved in:
| Main Authors: | Peri, Skand, Lee, Iain, Kim, Chanho, Fuxin, Li, Hermans, Tucker, Lee, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
by: Huang, Yixuan, et al.
Published: (2023)
by: Huang, Yixuan, et al.
Published: (2023)
Object Dynamics Modeling with Hierarchical Point Cloud-based Representations
by: Kim, Chanho, et al.
Published: (2024)
by: Kim, Chanho, et al.
Published: (2024)
Dense-depth map guided deep Lidar-Visual Odometry with Sparse Point Clouds and Images
by: Huang, JunYing, et al.
Published: (2025)
by: Huang, JunYing, et al.
Published: (2025)
GeomGS: LiDAR-Guided Geometry-Aware Gaussian Splatting for Robot Localization
by: Lee, Jaewon, et al.
Published: (2025)
by: Lee, Jaewon, et al.
Published: (2025)
Unsupervised Point Cloud Registration with Self-Distillation
by: Löwens, Christian, et al.
Published: (2024)
by: Löwens, Christian, et al.
Published: (2024)
Taming Transformers for Realistic Lidar Point Cloud Generation
by: Haghighi, Hamed, et al.
Published: (2024)
by: Haghighi, Hamed, et al.
Published: (2024)
DiffCloud: Real-to-Sim from Point Clouds with Differentiable Simulation and Rendering of Deformable Objects
by: Sundaresan, Priya, et al.
Published: (2022)
by: Sundaresan, Priya, et al.
Published: (2022)
Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning
by: Zhu, Haoyi, et al.
Published: (2024)
by: Zhu, Haoyi, et al.
Published: (2024)
Loss Distillation via Gradient Matching for Point Cloud Completion with Weighted Chamfer Distance
by: Lin, Fangzhou, et al.
Published: (2024)
by: Lin, Fangzhou, et al.
Published: (2024)
Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection
by: Khurana, Mehar, et al.
Published: (2024)
by: Khurana, Mehar, et al.
Published: (2024)
CANVAS: Commonsense-Aware Navigation System for Intuitive Human-Robot Interaction
by: Choi, Suhwan, et al.
Published: (2024)
by: Choi, Suhwan, et al.
Published: (2024)
Composable Part-Based Manipulation
by: Liu, Weiyu, et al.
Published: (2024)
by: Liu, Weiyu, et al.
Published: (2024)
Unsupervised Change Detection for Space Habitats Using 3D Point Clouds
by: Santos, Jamie, et al.
Published: (2023)
by: Santos, Jamie, et al.
Published: (2023)
Subtask-Aware Visual Reward Learning from Segmented Demonstrations
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
Audio-3DVG: Unified Audio -- Point Cloud Fusion for 3D Visual Grounding
by: Cao-Dinh, Duc, et al.
Published: (2025)
by: Cao-Dinh, Duc, et al.
Published: (2025)
Robustness Evaluation of Machine Learning Models for Robot Arm Action Recognition in Noisy Environments
by: Motamedi, Elaheh, et al.
Published: (2024)
by: Motamedi, Elaheh, et al.
Published: (2024)
SLNet: A Super-Lightweight Geometry-Adaptive Network for 3D Point Cloud Recognition
by: Saeid, Mohammad, et al.
Published: (2026)
by: Saeid, Mohammad, et al.
Published: (2026)
Test-Time Training for Visual Foresight Vision-Language-Action Models
by: Park, Sangwu, et al.
Published: (2026)
by: Park, Sangwu, et al.
Published: (2026)
Enhancing 3D Point Cloud Classification with ModelNet-R and Point-SkipNet
by: Saeid, Mohammad, et al.
Published: (2025)
by: Saeid, Mohammad, et al.
Published: (2025)
Deep Learning-Based Multi-Modal Fusion for Robust Robot Perception and Navigation
by: Lai, Delun, et al.
Published: (2025)
by: Lai, Delun, et al.
Published: (2025)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
by: Hwang, Dongyoon, et al.
Published: (2024)
by: Hwang, Dongyoon, et al.
Published: (2024)
Hierarchical Diffusion Policy for Kinematics-Aware Multi-Task Robotic Manipulation
by: Ma, Xiao, et al.
Published: (2024)
by: Ma, Xiao, et al.
Published: (2024)
HD Maps are Lane Detection Generalizers: A Novel Generative Framework for Single-Source Domain Generalization
by: Lee, Daeun, et al.
Published: (2023)
by: Lee, Daeun, et al.
Published: (2023)
Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
Leveraging Object Priors for Point Tracking
by: Boote, Bikram, et al.
Published: (2024)
by: Boote, Bikram, et al.
Published: (2024)
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention
by: Xiao, Lei, et al.
Published: (2025)
by: Xiao, Lei, et al.
Published: (2025)
NeurAll: Towards a Unified Visual Perception Model for Automated Driving
by: Sistu, Ganesh, et al.
Published: (2019)
by: Sistu, Ganesh, et al.
Published: (2019)
Nothing Stands Still: A Spatiotemporal Benchmark on 3D Point Cloud Registration Under Large Geometric and Temporal Change
by: Sun, Tao, et al.
Published: (2023)
by: Sun, Tao, et al.
Published: (2023)
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
by: Krauss, Henrik, et al.
Published: (2025)
by: Krauss, Henrik, et al.
Published: (2025)
Compressor-VLA: Instruction-Guided Visual Token Compression for Efficient Robotic Manipulation
by: Gao, Juntao, et al.
Published: (2025)
by: Gao, Juntao, et al.
Published: (2025)
Robot Synesthesia: In-Hand Manipulation with Visuotactile Sensing
by: Yuan, Ying, et al.
Published: (2023)
by: Yuan, Ying, et al.
Published: (2023)
BelHouse3D: A Benchmark Dataset for Assessing Occlusion Robustness in 3D Point Cloud Semantic Segmentation
by: Kumar, Umamaheswaran Raman, et al.
Published: (2024)
by: Kumar, Umamaheswaran Raman, et al.
Published: (2024)
4D Contrastive Superflows are Dense 3D Representation Learners
by: Xu, Xiang, et al.
Published: (2024)
by: Xu, Xiang, et al.
Published: (2024)
PGA: Personalizing Grasping Agents with Single Human-Robot Interaction
by: Kim, Junghyun, et al.
Published: (2023)
by: Kim, Junghyun, et al.
Published: (2023)
Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation
by: Zhang, Wenbo, et al.
Published: (2025)
by: Zhang, Wenbo, et al.
Published: (2025)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
by: Li, Chengmeng, et al.
Published: (2025)
by: Li, Chengmeng, et al.
Published: (2025)
AutoURDF: Unsupervised Robot Modeling from Point Cloud Frames Using Cluster Registration
by: Lin, Jiong, et al.
Published: (2024)
by: Lin, Jiong, et al.
Published: (2024)
GeRM: A Generalist Robotic Model with Mixture-of-experts for Quadruped Robot
by: Song, Wenxuan, et al.
Published: (2024)
by: Song, Wenxuan, et al.
Published: (2024)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
by: Lee, Seokmin, et al.
Published: (2026)
by: Lee, Seokmin, et al.
Published: (2026)
Similar Items
-
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
by: Huang, Yixuan, et al.
Published: (2023) -
Object Dynamics Modeling with Hierarchical Point Cloud-based Representations
by: Kim, Chanho, et al.
Published: (2024) -
Dense-depth map guided deep Lidar-Visual Odometry with Sparse Point Clouds and Images
by: Huang, JunYing, et al.
Published: (2025) -
GeomGS: LiDAR-Guided Geometry-Aware Gaussian Splatting for Robot Localization
by: Lee, Jaewon, et al.
Published: (2025) -
Unsupervised Point Cloud Registration with Self-Distillation
by: Löwens, Christian, et al.
Published: (2024)