Multimodal 3D Object Detection on Unseen Domains
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hegde, Deepti, Lohit, Suhas, Peng, Kuan-Chuan, Jones, Michael J., Patel, Vishal M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection
von: Hegde, Deepti, et al.
Veröffentlicht: (2024)
von: Hegde, Deepti, et al.
Veröffentlicht: (2024)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
Auto-Vocabulary 3D Object Detection
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
WISE: Weighted Iterative Society-of-Experts for Robust Multimodal Multi-Agent Debate
von: Cherian, Anoop, et al.
Veröffentlicht: (2025)
von: Cherian, Anoop, et al.
Veröffentlicht: (2025)
Multimodal Diffusion Bridge with Attention-Based SAR Fusion for Satellite Image Cloud Removal
von: Hu, Yuyang, et al.
Veröffentlicht: (2025)
von: Hu, Yuyang, et al.
Veröffentlicht: (2025)
Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
von: Cherian, Anoop, et al.
Veröffentlicht: (2024)
Joint Training of Image Generator and Detector for Road Defect Detection
von: Peng, Kuan-Chuan
Veröffentlicht: (2025)
von: Peng, Kuan-Chuan
Veröffentlicht: (2025)
An Empirical Study of the Generalization Ability of Lidar 3D Object Detectors to Unseen Domains
von: Eskandar, George, et al.
Veröffentlicht: (2024)
von: Eskandar, George, et al.
Veröffentlicht: (2024)
Towards Egocentric 3D Hand Pose Estimation in Unseen Domains
von: Mucha, Wiktor, et al.
Veröffentlicht: (2026)
von: Mucha, Wiktor, et al.
Veröffentlicht: (2026)
Programmatic Video Prediction Using Large Language Models
von: Tang, Hao, et al.
Veröffentlicht: (2025)
von: Tang, Hao, et al.
Veröffentlicht: (2025)
MatchU: Matching Unseen Objects for 6D Pose Estimation from RGB-D Images
von: Huang, Junwen, et al.
Veröffentlicht: (2024)
von: Huang, Junwen, et al.
Veröffentlicht: (2024)
UOD: Unseen Object Detection in 3D Point Cloud
von: Choi, Hyunjun, et al.
Veröffentlicht: (2024)
von: Choi, Hyunjun, et al.
Veröffentlicht: (2024)
Improving Open-World Object Localization by Discovering Background
von: Singh, Ashish, et al.
Veröffentlicht: (2025)
von: Singh, Ashish, et al.
Veröffentlicht: (2025)
Towards Zero-shot 3D Anomaly Localization
von: Wang, Yizhou, et al.
Veröffentlicht: (2024)
von: Wang, Yizhou, et al.
Veröffentlicht: (2024)
Investigating Domain Gaps for Indoor 3D Object Detection
von: Zhao, Zijing, et al.
Veröffentlicht: (2025)
von: Zhao, Zijing, et al.
Veröffentlicht: (2025)
FreBIS: Frequency-Based Stratification for Neural Implicit Surface Representations
von: Sawada, Naoko, et al.
Veröffentlicht: (2025)
von: Sawada, Naoko, et al.
Veröffentlicht: (2025)
Deployment Prior Injection for Run-time Calibratable Object Detection
von: Zhou, Mo, et al.
Veröffentlicht: (2024)
von: Zhou, Mo, et al.
Veröffentlicht: (2024)
Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
von: Li, Ruihuang, et al.
Veröffentlicht: (2024)
AssemblyBench: Physics-Aware Assembly of Complex Industrial Objects
von: Li, Danrui, et al.
Veröffentlicht: (2026)
von: Li, Danrui, et al.
Veröffentlicht: (2026)
Certainty and Uncertainty Guided Active Domain Adaptation
von: Safaei, Bardia, et al.
Veröffentlicht: (2025)
von: Safaei, Bardia, et al.
Veröffentlicht: (2025)
RemoteVAR: Autoregressive Visual Modeling for Remote Sensing Change Detection
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2026)
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2026)
FaceXBench: Evaluating Multimodal LLMs on Face Understanding
von: Narayan, Kartik, et al.
Veröffentlicht: (2025)
von: Narayan, Kartik, et al.
Veröffentlicht: (2025)
Generalizing Single-View 3D Shape Retrieval to Occlusions and Unseen Objects
von: Wu, Qirui, et al.
Veröffentlicht: (2023)
von: Wu, Qirui, et al.
Veröffentlicht: (2023)
PF3Det: A Prompted Foundation Feature Assisted Visual LiDAR 3D Detector
von: Li, Kaidong, et al.
Veröffentlicht: (2025)
von: Li, Kaidong, et al.
Veröffentlicht: (2025)
Long-Tailed Anomaly Detection with Learnable Class Names
von: Ho, Chih-Hui, et al.
Veröffentlicht: (2024)
von: Ho, Chih-Hui, et al.
Veröffentlicht: (2024)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
von: Gopal, Suhas, et al.
Veröffentlicht: (2025)
von: Gopal, Suhas, et al.
Veröffentlicht: (2025)
Generalizing to Unseen Domains in Diabetic Retinopathy with Disentangled Representations
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
Domain Generalization of 3D Object Detection by Density-Resampling
von: Li, Shuangzhi, et al.
Veröffentlicht: (2023)
von: Li, Shuangzhi, et al.
Veröffentlicht: (2023)
Toward Long-Tailed Online Anomaly Detection through Class-Agnostic Concepts
von: Yang, Chiao-An, et al.
Veröffentlicht: (2025)
von: Yang, Chiao-An, et al.
Veröffentlicht: (2025)
Multimodal Object Query Initialization for 3D Object Detection
von: van Geerenstein, Mathijs R., et al.
Veröffentlicht: (2023)
von: van Geerenstein, Mathijs R., et al.
Veröffentlicht: (2023)
Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
Distilling Multi-modal Large Language Models for Autonomous Driving
von: Hegde, Deepti, et al.
Veröffentlicht: (2025)
von: Hegde, Deepti, et al.
Veröffentlicht: (2025)
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2023)
Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2026)
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2026)
Time-Series U-Net with Recurrence for Noise-Robust Imaging Photoplethysmography
von: Shenoy, Vineet R., et al.
Veröffentlicht: (2025)
von: Shenoy, Vineet R., et al.
Veröffentlicht: (2025)
PoseStreamer: A Multi-modal Framework for 3D Tracking of Unseen Moving Objects
von: Yang, Huiming, et al.
Veröffentlicht: (2025)
von: Yang, Huiming, et al.
Veröffentlicht: (2025)
Temporal-Anchor3DLane: Enhanced 3D Lane Detection with Multi-Task Losses and LSTM Fusion
von: Suhas, D. Shainu, et al.
Veröffentlicht: (2025)
von: Suhas, D. Shainu, et al.
Veröffentlicht: (2025)
Object Pose Transformer: Unifying Unseen Object Pose Estimation
von: Li, Weihang, et al.
Veröffentlicht: (2026)
von: Li, Weihang, et al.
Veröffentlicht: (2026)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
Endless World: Real-Time 3D-Aware Long Video Generation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection
von: Hegde, Deepti, et al.
Veröffentlicht: (2024) -
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025) -
Auto-Vocabulary 3D Object Detection
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025) -
WISE: Weighted Iterative Society-of-Experts for Robust Multimodal Multi-Agent Debate
von: Cherian, Anoop, et al.
Veröffentlicht: (2025) -
Multimodal Diffusion Bridge with Attention-Based SAR Fusion for Satellite Image Cloud Removal
von: Hu, Yuyang, et al.
Veröffentlicht: (2025)