S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | He, Xuan, Yuan, Jin, Yang, Kailun, Zeng, Zhenchao, Li, Zhiyong |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MonoDETR: Depth-guided Transformer for Monocular 3D Object Detection
by: Zhang, Renrui, et al.
Published: (2022)
by: Zhang, Renrui, et al.
Published: (2022)
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026)
by: Rahman, Shantanu, et al.
Published: (2026)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
by: Jia, Wanjun, et al.
Published: (2025)
by: Jia, Wanjun, et al.
Published: (2025)
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
by: Zhang, Xu, et al.
Published: (2023)
by: Zhang, Xu, et al.
Published: (2023)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024)
by: Zeng, Kang, et al.
Published: (2024)
Language-Driven Dual Style Mixing for Single-Domain Generalized Object Detection
by: Qin, Hongda, et al.
Published: (2025)
by: Qin, Hongda, et al.
Published: (2025)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
by: Shi, Hao, et al.
Published: (2023)
by: Shi, Hao, et al.
Published: (2023)
NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving
by: Luo, Kai, et al.
Published: (2026)
by: Luo, Kai, et al.
Published: (2026)
LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras
by: Teng, Fei, et al.
Published: (2024)
by: Teng, Fei, et al.
Published: (2024)
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
by: Lin, Jiacheng, et al.
Published: (2024)
by: Lin, Jiacheng, et al.
Published: (2024)
HierDAMap: Towards Universal Domain Adaptive BEV Mapping via Hierarchical Perspective Priors
by: Li, Siyu, et al.
Published: (2025)
by: Li, Siyu, et al.
Published: (2025)
Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation
by: Qin, Yu, et al.
Published: (2026)
by: Qin, Yu, et al.
Published: (2026)
Towards Precise 3D Human Pose Estimation with Multi-Perspective Spatial-Temporal Relational Transformers
by: Jiao, Jianbin, et al.
Published: (2024)
by: Jiao, Jianbin, et al.
Published: (2024)
P2U-SLAM: A Monocular Wide-FoV SLAM System Based on Point Uncertainty and Pose Uncertainty
by: Zhang, Yufan, et al.
Published: (2024)
by: Zhang, Yufan, et al.
Published: (2024)
O3N: Omnidirectional Open-Vocabulary Occupancy Prediction
by: Duan, Mengfei, et al.
Published: (2026)
by: Duan, Mengfei, et al.
Published: (2026)
Towards Consistent Object Detection via LiDAR-Camera Synergy
by: Luo, Kai, et al.
Published: (2024)
by: Luo, Kai, et al.
Published: (2024)
TacShade A New 3D-printed Soft Optical Tactile Sensor Based on Light, Shadow and Greyscale for Shape Reconstruction
by: Lu, Zhenyu, et al.
Published: (2024)
by: Lu, Zhenyu, et al.
Published: (2024)
Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Dexterous Grasping
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
TS-CGNet: Temporal-Spatial Fusion Meets Centerline-Guided Diffusion for BEV Mapping
by: Hong, Xinying, et al.
Published: (2025)
by: Hong, Xinying, et al.
Published: (2025)
SDCM: Simulated Densifying and Compensatory Modeling Fusion for Radar-Vision 3-D Object Detection in Internet of Vehicles
by: Li, Shucong, et al.
Published: (2026)
by: Li, Shucong, et al.
Published: (2026)
EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras
by: Wang, Luming, et al.
Published: (2026)
by: Wang, Luming, et al.
Published: (2026)
DeProPose: Deficiency-Proof 3D Human Pose Estimation via Adaptive Multi-View Fusion
by: Jiao, Jianbin, et al.
Published: (2025)
by: Jiao, Jianbin, et al.
Published: (2025)
EdgeNavMamba: Mamba Optimized Object Detection for Energy Efficient Edge Devices
by: Aalishah, Romina, et al.
Published: (2025)
by: Aalishah, Romina, et al.
Published: (2025)
Event-guided 3D Gaussian Splatting for Dynamic Human and Scene Reconstruction
by: Yin, Xiaoting, et al.
Published: (2025)
by: Yin, Xiaoting, et al.
Published: (2025)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
by: Lin, Kaixin, et al.
Published: (2026)
by: Lin, Kaixin, et al.
Published: (2026)
GenMapping: Unleashing the Potential of Inverse Perspective Mapping for Robust Online HD Map Construction
by: Li, Siyu, et al.
Published: (2024)
by: Li, Siyu, et al.
Published: (2024)
PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments
by: Zhu, Guoliang, et al.
Published: (2026)
by: Zhu, Guoliang, et al.
Published: (2026)
CoBEVMoE: Heterogeneity-aware Feature Fusion with Dynamic Mixture-of-Experts for Collaborative Perception
by: Kong, Lingzhao, et al.
Published: (2025)
by: Kong, Lingzhao, et al.
Published: (2025)
NRSeg: Noise-Resilient Learning for BEV Semantic Segmentation via Driving World Models
by: Li, Siyu, et al.
Published: (2025)
by: Li, Siyu, et al.
Published: (2025)
CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction
by: Yang, Zhe, et al.
Published: (2026)
by: Yang, Zhe, et al.
Published: (2026)
DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking
by: Deng, Buyin, et al.
Published: (2025)
by: Deng, Buyin, et al.
Published: (2025)
Iterative PnP and its application in 3D-2D vascular image registration for robot navigation
by: Song, Jingwei, et al.
Published: (2023)
by: Song, Jingwei, et al.
Published: (2023)
LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection
by: Teng, Fei, et al.
Published: (2025)
by: Teng, Fei, et al.
Published: (2025)
A Late Collaborative Perception Framework for 3D Multi-Object and Multi-Source Association and Fusion
by: Fadili, Maryem, et al.
Published: (2025)
by: Fadili, Maryem, et al.
Published: (2025)
An iterative closest point algorithm for marker-free 3D shape registration of continuum robots
by: Hoffmann, Matthias K., et al.
Published: (2024)
by: Hoffmann, Matthias K., et al.
Published: (2024)
Omnidirectional Multi-Object Tracking
by: Luo, Kai, et al.
Published: (2025)
by: Luo, Kai, et al.
Published: (2025)
Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance
by: Zhao, Jiayi, et al.
Published: (2025)
by: Zhao, Jiayi, et al.
Published: (2025)
Exploring Event-based Human Pose Estimation with 3D Event Representations
by: Yin, Xiaoting, et al.
Published: (2023)
by: Yin, Xiaoting, et al.
Published: (2023)
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
Similar Items
-
MonoDETR: Depth-guided Transformer for Monocular 3D Object Detection
by: Zhang, Renrui, et al.
Published: (2022) -
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026) -
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
by: Jia, Wanjun, et al.
Published: (2025) -
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
by: Zhang, Xu, et al.
Published: (2023) -
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024)