MDHA: Multi-Scale Deformable Transformer with Hybrid Anchors for Multi-View 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Adeline, Michelle, Loo, Junn Yong, Baskaran, Vishnu Monn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conceptualizing Multi-scale Wavelet Attention and Ray-based Encoding for Human-Object Interaction Detection
by: Pay, Quan Bi, et al.
Published: (2025)
by: Pay, Quan Bi, et al.
Published: (2025)
Variational Potential Flow: A Novel Probabilistic Framework for Energy-Based Generative Modelling
by: Loo, Junn Yong, et al.
Published: (2024)
by: Loo, Junn Yong, et al.
Published: (2024)
SpaRTAN: Spatial Reinforcement Token-based Aggregation Network for Visual Recognition
by: Pay, Quan Bi, et al.
Published: (2025)
by: Pay, Quan Bi, et al.
Published: (2025)
Drive As You Like: Strategy-Level Motion Planning Based on A Multi-Head Diffusion Model
by: Ding, Fan, et al.
Published: (2025)
by: Ding, Fan, et al.
Published: (2025)
Cross-Domain Transfer Learning using Attention Latent Features for Multi-Agent Trajectory Prediction
by: Loh, Jia Quan, et al.
Published: (2024)
by: Loh, Jia Quan, et al.
Published: (2024)
Cross-domain Transfer Learning and State Inference for Soft Robots via a Semi-supervised Sequential Variational Bayes Framework
by: Sapai, Shageenderan, et al.
Published: (2023)
by: Sapai, Shageenderan, et al.
Published: (2023)
Energy-efficient Hybrid Model Predictive Trajectory Planning for Autonomous Electric Vehicles
by: Ding, Fan, et al.
Published: (2024)
by: Ding, Fan, et al.
Published: (2024)
DeformerNet: Learning Bimanual Manipulation of 3D Deformable Objects
by: Thach, Bao, et al.
Published: (2023)
by: Thach, Bao, et al.
Published: (2023)
Dynamic Manipulation of Deformable Objects in 3D: Simulation, Benchmark and Learning Strategy
by: Lan, Guanzhou, et al.
Published: (2025)
by: Lan, Guanzhou, et al.
Published: (2025)
RAPiD: Real-time Deterministic Trajectory Planning via Diffusion Behavior Priors for Safe and Efficient Autonomous Driving
by: Reddy, Ruturaj, et al.
Published: (2026)
by: Reddy, Ruturaj, et al.
Published: (2026)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
by: Banerjee, Prithviraj, et al.
Published: (2024)
by: Banerjee, Prithviraj, et al.
Published: (2024)
Multi-Object Tracking based on Imaging Radar 3D Object Detection
by: Palmer, Patrick, et al.
Published: (2024)
by: Palmer, Patrick, et al.
Published: (2024)
3D Multi-Object Tracking with Semi-Supervised GRU-Kalman Filter
by: Wang, Xiaoxiang, et al.
Published: (2024)
by: Wang, Xiaoxiang, et al.
Published: (2024)
GP3: A 3D Geometry-Aware Policy with Multi-View Images for Robotic Manipulation
by: Qian, Quanhao, et al.
Published: (2025)
by: Qian, Quanhao, et al.
Published: (2025)
PKRD-CoT: A Unified Chain-of-thought Prompting for Multi-Modal Large Language Models in Autonomous Driving
by: Luo, Xuewen, et al.
Published: (2024)
by: Luo, Xuewen, et al.
Published: (2024)
Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing
by: Hehli, Justin, et al.
Published: (2025)
by: Hehli, Justin, et al.
Published: (2025)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
by: Zhao, Ziyan, et al.
Published: (2025)
by: Zhao, Ziyan, et al.
Published: (2025)
AnchDrive: Bootstrapping Diffusion Policies with Hybrid Trajectory Anchors for End-to-End Driving
by: Chai, Jinhao, et al.
Published: (2025)
by: Chai, Jinhao, et al.
Published: (2025)
MA3DSG: Multi-Agent 3D Scene Graph Generation for Large-Scale Indoor Environments
by: Kim, Yirum, et al.
Published: (2026)
by: Kim, Yirum, et al.
Published: (2026)
Weakly Supervised Point Clouds Transformer for 3D Object Detection
by: Tang, Zuojin, et al.
Published: (2023)
by: Tang, Zuojin, et al.
Published: (2023)
LeHome: A Simulation Environment for Deformable Object Manipulation in Household Scenarios
by: Li, Zeyi, et al.
Published: (2026)
by: Li, Zeyi, et al.
Published: (2026)
AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models
by: Jia, Tingzheng, et al.
Published: (2026)
by: Jia, Tingzheng, et al.
Published: (2026)
Multi-Object Navigation in real environments using hybrid policies
by: Sadek, Assem, et al.
Published: (2024)
by: Sadek, Assem, et al.
Published: (2024)
ExoGait-MS: Learning Periodic Dynamics with Multi-Scale Graph Network for Exoskeleton Gait Recognition
by: Liu, Lijiang, et al.
Published: (2025)
by: Liu, Lijiang, et al.
Published: (2025)
Towards High-Consistency Embodied World Model with Multi-View Trajectory Videos
by: Su, Taiyi, et al.
Published: (2025)
by: Su, Taiyi, et al.
Published: (2025)
MV-UMI: A Scalable Multi-View Interface for Cross-Embodiment Learning
by: Rayyan, Omar, et al.
Published: (2025)
by: Rayyan, Omar, et al.
Published: (2025)
Greedy Perspectives: Multi-Drone View Planning for Collaborative Perception in Cluttered Environments
by: Suresh, Krishna, et al.
Published: (2023)
by: Suresh, Krishna, et al.
Published: (2023)
Single and bi-layered 2-D acoustic soft tactile skin (AST2)
by: Rajendran, Vishnu, et al.
Published: (2024)
by: Rajendran, Vishnu, et al.
Published: (2024)
Multi-Object Graph Affordance Network: Goal-Oriented Planning through Learned Compound Object Affordances
by: Girgin, Tuba, et al.
Published: (2023)
by: Girgin, Tuba, et al.
Published: (2023)
GenDOM: Generalizable One-shot Deformable Object Manipulation with Parameter-Aware Policy
by: Kuroki, So, et al.
Published: (2023)
by: Kuroki, So, et al.
Published: (2023)
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
by: Ortega-Peimbert, Jesús, et al.
Published: (2025)
by: Ortega-Peimbert, Jesús, et al.
Published: (2025)
A Multimodal Hybrid Late-Cascade Fusion Network for Enhanced 3D Object Detection
by: Sgaravatti, Carlo, et al.
Published: (2025)
by: Sgaravatti, Carlo, et al.
Published: (2025)
Sigma-point Kalman Filter with Nonlinear Unknown Input Estimation via Optimization and Data-driven Approach for Dynamic Systems
by: Loo, Junn Yong, et al.
Published: (2023)
by: Loo, Junn Yong, et al.
Published: (2023)
Understanding Physical Properties of Unseen Deformable Objects by Leveraging Large Language Models and Robot Actions
by: Park, Changmin, et al.
Published: (2025)
by: Park, Changmin, et al.
Published: (2025)
MoDeSuite: Robot Learning Task Suite for Benchmarking Mobile Manipulation with Deformable Objects
by: Zhang, Yuying, et al.
Published: (2025)
by: Zhang, Yuying, et al.
Published: (2025)
IndoorBEV: Joint Detection and Footprint Completion of Objects via Mask-based Prediction in Indoor Scenarios for Bird's-Eye View Perception
by: Li, Haichuan, et al.
Published: (2025)
by: Li, Haichuan, et al.
Published: (2025)
An Efficient LiDAR-Camera Fusion Network for Multi-Class 3D Dynamic Object Detection and Trajectory Prediction
by: He, Yushen, et al.
Published: (2025)
by: He, Yushen, et al.
Published: (2025)
Learning Sequential Kinematic Models from Demonstrations for Multi-Jointed Articulated Objects
by: Gupta, Anmol, et al.
Published: (2025)
by: Gupta, Anmol, et al.
Published: (2025)
AYDIV: Adaptable Yielding 3D Object Detection via Integrated Contextual Vision Transformer
by: Dam, Tanmoy, et al.
Published: (2024)
by: Dam, Tanmoy, et al.
Published: (2024)
LagMemo: Language 3D Gaussian Splatting Memory for Multi-modal Open-vocabulary Multi-goal Visual Navigation
by: Zhou, Haotian, et al.
Published: (2025)
by: Zhou, Haotian, et al.
Published: (2025)
Similar Items
-
Conceptualizing Multi-scale Wavelet Attention and Ray-based Encoding for Human-Object Interaction Detection
by: Pay, Quan Bi, et al.
Published: (2025) -
Variational Potential Flow: A Novel Probabilistic Framework for Energy-Based Generative Modelling
by: Loo, Junn Yong, et al.
Published: (2024) -
SpaRTAN: Spatial Reinforcement Token-based Aggregation Network for Visual Recognition
by: Pay, Quan Bi, et al.
Published: (2025) -
Drive As You Like: Strategy-Level Motion Planning Based on A Multi-Head Diffusion Model
by: Ding, Fan, et al.
Published: (2025) -
Cross-Domain Transfer Learning using Attention Latent Features for Multi-Agent Trajectory Prediction
by: Loh, Jia Quan, et al.
Published: (2024)