Scalable Object Detection in the Car Interior With Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Schmidt, Sebastian, Mészáros, Bálint, Firintepe, Ahmet, Günnemann, Stephan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Unified Approach Towards Active Learning and Out-of-Distribution Detection
by: Schmidt, Sebastian, et al.
Published: (2024)
by: Schmidt, Sebastian, et al.
Published: (2024)
A Machine Learning Perspective on Automated Driving Corner Cases
by: Schmidt, Sebastian, et al.
Published: (2025)
by: Schmidt, Sebastian, et al.
Published: (2025)
EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving
by: Schäfer, Finn Rasmus, et al.
Published: (2026)
by: Schäfer, Finn Rasmus, et al.
Published: (2026)
Joint Out-of-Distribution Filtering and Data Discovery Active Learning
by: Schmidt, Sebastian, et al.
Published: (2025)
by: Schmidt, Sebastian, et al.
Published: (2025)
Prior2Former -- Evidential Modeling of Mask Transformers for Assumption-Free Open-World Panoptic Segmentation
by: Schmidt, Sebastian, et al.
Published: (2025)
by: Schmidt, Sebastian, et al.
Published: (2025)
Amplified Patch-Level Differential Privacy for Free via Random Cropping
by: Durmaz, Kaan, et al.
Published: (2026)
by: Durmaz, Kaan, et al.
Published: (2026)
Towards Unbiased Source-Free Object Detection via Vision Foundation Models
by: Cai, Zhi, et al.
Published: (2026)
by: Cai, Zhi, et al.
Published: (2026)
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
by: Feng, Chen-Bin, et al.
Published: (2026)
by: Feng, Chen-Bin, et al.
Published: (2026)
Decoupled Prototype Matching with Vision Foundation Models for Few-Shot Industrial Object Detection
by: M., Hari Prasanth S., et al.
Published: (2026)
by: M., Hari Prasanth S., et al.
Published: (2026)
LipShiFT: A Certifiably Robust Shift-based Vision Transformer
by: Menon, Rohan, et al.
Published: (2025)
by: Menon, Rohan, et al.
Published: (2025)
Unexplored flaws in multiple-choice VQA evaluations
by: Rosenthal, Fabio, et al.
Published: (2025)
by: Rosenthal, Fabio, et al.
Published: (2025)
SPROUT: A Scalable Diffusion Foundation Model for Agricultural Vision
by: Xiang, Shuai, et al.
Published: (2026)
by: Xiang, Shuai, et al.
Published: (2026)
RT-DETRv4: Painlessly Furthering Real-Time Object Detection with Vision Foundation Models
by: Liao, Zijun, et al.
Published: (2025)
by: Liao, Zijun, et al.
Published: (2025)
Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection
by: Lundqvist, Lars, et al.
Published: (2026)
by: Lundqvist, Lars, et al.
Published: (2026)
Exploring Aleatoric Uncertainty in Object Detection via Vision Foundation Models
by: Cui, Peng, et al.
Published: (2024)
by: Cui, Peng, et al.
Published: (2024)
Vector-Quantized Vision Foundation Models for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2025)
by: Zhao, Rongzhen, et al.
Published: (2025)
Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy
by: Teuber, Carolin, et al.
Published: (2026)
by: Teuber, Carolin, et al.
Published: (2026)
SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
by: Dünkel, Olaf, et al.
Published: (2026)
by: Dünkel, Olaf, et al.
Published: (2026)
Not just Birds and Cars: Generic, Scalable and Explainable Models for Professional Visual Recognition
by: Wu, Junde, et al.
Published: (2024)
by: Wu, Junde, et al.
Published: (2024)
Enhancing Computer Vision Model Generalization in Warehouse Facilities: A Case Study on Anomaly Detection in Vertical Material Handling Systems
by: Liu, Ruiliang, et al.
Published: (2026)
by: Liu, Ruiliang, et al.
Published: (2026)
Training-Free Object-Agnostic Jam Detection in Fulfillment Centers
by: Liu, Ruiliang, et al.
Published: (2026)
by: Liu, Ruiliang, et al.
Published: (2026)
Test-Time Adaptive Object Detection with Foundation Model
by: Gao, Yingjie, et al.
Published: (2025)
by: Gao, Yingjie, et al.
Published: (2025)
Beyond Boundaries: Leveraging Vision Foundation Models for Source-Free Object Detection
by: Yao, Huizai, et al.
Published: (2025)
by: Yao, Huizai, et al.
Published: (2025)
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
by: Ding, Bonan, et al.
Published: (2024)
by: Ding, Bonan, et al.
Published: (2024)
Interior Object Geometry via Fitted Frames
by: Pizer, Stephen M., et al.
Published: (2024)
by: Pizer, Stephen M., et al.
Published: (2024)
FMG-Det: Foundation Model Guided Robust Object Detection
by: Hannan, Darryl, et al.
Published: (2025)
by: Hannan, Darryl, et al.
Published: (2025)
Template-based Object Detection Using a Foundation Model
by: Braeutigam, Valentin, et al.
Published: (2026)
by: Braeutigam, Valentin, et al.
Published: (2026)
MonoDINO-DETR: Depth-Enhanced Monocular 3D Object Detection Using a Vision Foundation Model
by: Kim, Jihyeok, et al.
Published: (2025)
by: Kim, Jihyeok, et al.
Published: (2025)
Looking Locally: Object-Centric Vision Transformers as Foundation Models for Efficient Segmentation
by: Traub, Manuel, et al.
Published: (2025)
by: Traub, Manuel, et al.
Published: (2025)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
by: Yang, Anqi Joyce, et al.
Published: (2026)
by: Yang, Anqi Joyce, et al.
Published: (2026)
RetFiner: A Vision-Language Refinement Scheme for Retinal Foundation Models
by: Fecso, Ronald, et al.
Published: (2025)
by: Fecso, Ronald, et al.
Published: (2025)
Foundation Model Priors Enhance Object Focus in Feature Space for Source-Free Object Detection
by: VCR, Sairam, et al.
Published: (2025)
by: VCR, Sairam, et al.
Published: (2025)
Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model
by: Le, Long, et al.
Published: (2024)
by: Le, Long, et al.
Published: (2024)
Detecting Car Speed using Object Detection and Depth Estimation: A Deep Learning Framework
by: Dasgupta, Subhasis, et al.
Published: (2024)
by: Dasgupta, Subhasis, et al.
Published: (2024)
DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather
by: Leitgeb, Christof, et al.
Published: (2026)
by: Leitgeb, Christof, et al.
Published: (2026)
Revisiting Few-Shot Object Detection with Vision-Language Models
by: Madan, Anish, et al.
Published: (2023)
by: Madan, Anish, et al.
Published: (2023)
Scalable Vision-Based 3D Object Detection and Monocular Depth Estimation for Autonomous Driving
by: Liu, Yuxuan
Published: (2024)
by: Liu, Yuxuan
Published: (2024)
DreamCar: Leveraging Car-specific Prior for in-the-wild 3D Car Reconstruction
by: Du, Xiaobiao, et al.
Published: (2024)
by: Du, Xiaobiao, et al.
Published: (2024)
Boosting Salient Object Detection with Knowledge Distillated from Large Foundation Models
by: He, Miaoyang, et al.
Published: (2025)
by: He, Miaoyang, et al.
Published: (2025)
Equipping Vision Foundation Model with Mixture of Experts for Out-of-Distribution Detection
by: Zhao, Shizhen, et al.
Published: (2025)
by: Zhao, Shizhen, et al.
Published: (2025)
Similar Items
-
A Unified Approach Towards Active Learning and Out-of-Distribution Detection
by: Schmidt, Sebastian, et al.
Published: (2024) -
A Machine Learning Perspective on Automated Driving Corner Cases
by: Schmidt, Sebastian, et al.
Published: (2025) -
EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving
by: Schäfer, Finn Rasmus, et al.
Published: (2026) -
Joint Out-of-Distribution Filtering and Data Discovery Active Learning
by: Schmidt, Sebastian, et al.
Published: (2025) -
Prior2Former -- Evidential Modeling of Mask Transformers for Assumption-Free Open-World Panoptic Segmentation
by: Schmidt, Sebastian, et al.
Published: (2025)