Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving
Fuente:
arXiv
Salvato in:
| Autori principali: | Theodoridis, Nikos, Mohandas, Reenu, Sistu, Ganesh, Scanlan, Anthony, Eising, Ciarán, Brophy, Tim |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating Small Vision-Language Models on Distance-Dependent Traffic Perception
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025)
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025)
Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025)
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025)
FIN: Fast Inference Network for Map Segmentation
di: Bispo, Ruan, et al.
Pubblicazione: (2025)
di: Bispo, Ruan, et al.
Pubblicazione: (2025)
Occluded nuScenes: A Multi-Sensor Dataset for Evaluating Perception Robustness in Automated Driving
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
Deformable Convolution Based Road Scene Semantic Segmentation of Fisheye Images in Autonomous Driving
di: Manzoor, Anam, et al.
Pubblicazione: (2024)
di: Manzoor, Anam, et al.
Pubblicazione: (2024)
SuperQuadricOcc: Real-Time Self-Supervised Semantic Occupancy Estimation with Superquadric Volume Rendering
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
Easy3D-Labels: Supervising Semantic Occupancy Estimation with 3D Pseudo-Labels for Automotive Perception
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
Optimizing Visual Question Answering Models for Driving: Bridging the Gap Between Human and Machine Attention Patterns
di: Rekanar, Kaavya, et al.
Pubblicazione: (2024)
di: Rekanar, Kaavya, et al.
Pubblicazione: (2024)
Revisiting Birds Eye View Perception Models with Frozen Foundation Models: DINOv2 and Metric3Dv2
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
di: Hayes, Seamie, et al.
Pubblicazione: (2025)
Evaluating the Impact of Weather-Induced Sensor Occlusion on BEVFusion for 3D Object Detection
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
MapsTP: HD Map Images Based Multimodal Trajectory Prediction for Automated Vehicles
di: Sharma, Sushil, et al.
Pubblicazione: (2024)
di: Sharma, Sushil, et al.
Pubblicazione: (2024)
Minimizing Occlusion Effect on Multi-View Camera Perception in BEV with Multi-Sensor Fusion
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
di: Kumar, Sanjay, et al.
Pubblicazione: (2025)
SS-SFR: Synthetic Scenes Spatial Frequency Response on Virtual KITTI and Degraded Automotive Simulations for Object Detection
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
Optimizing Ego Vehicle Trajectory Prediction: The Graph Enhancement Approach
di: Sharma, Sushil, et al.
Pubblicazione: (2023)
di: Sharma, Sushil, et al.
Pubblicazione: (2023)
NeurAll: Towards a Unified Visual Perception Model for Automated Driving
di: Sistu, Ganesh, et al.
Pubblicazione: (2019)
di: Sistu, Ganesh, et al.
Pubblicazione: (2019)
Reflective Teacher: Semi-Supervised Multimodal 3D Object Detection in Bird's-Eye-View via Uncertainty Measure
di: Hazra, Saheli, et al.
Pubblicazione: (2024)
di: Hazra, Saheli, et al.
Pubblicazione: (2024)
Fisheye Camera and Ultrasonic Sensor Fusion For Near-Field Obstacle Perception in Bird's-Eye-View
di: Das, Arindam, et al.
Pubblicazione: (2024)
di: Das, Arindam, et al.
Pubblicazione: (2024)
BEVMOSNet: Multimodal Fusion for BEV Moving Object Segmentation
di: Cong, Hiep Truong, et al.
Pubblicazione: (2025)
di: Cong, Hiep Truong, et al.
Pubblicazione: (2025)
FisheyeDetNet: 360° Surround view Fisheye Camera based Object Detection System for Autonomous Driving
di: Sistu, Ganesh, et al.
Pubblicazione: (2024)
di: Sistu, Ganesh, et al.
Pubblicazione: (2024)
PAN: Pillars-Attention-Based Network for 3D Object Detection
di: Bispo, Ruan, et al.
Pubblicazione: (2025)
di: Bispo, Ruan, et al.
Pubblicazione: (2025)
Velocity Driven Vision: Asynchronous Sensor Fusion Birds Eye View Models for Autonomous Vehicles
di: Hayes, Seamie, et al.
Pubblicazione: (2024)
di: Hayes, Seamie, et al.
Pubblicazione: (2024)
Measuring Natural Scenes SFR of Automotive Fisheye Cameras
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
Adapting Lightweight Vision Language Models for Radiological Visual Question Answering
di: Shourya, Aditya, et al.
Pubblicazione: (2025)
di: Shourya, Aditya, et al.
Pubblicazione: (2025)
Multi-Frame, Lightweight & Efficient Vision-Language Models for Question Answering in Autonomous Driving
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
di: Gopalkrishnan, Akshay, et al.
Pubblicazione: (2024)
FisheyeGaussianLift: BEV Feature Lifting for Surround-View Fisheye Camera Perception
di: Sonarghare, Shubham, et al.
Pubblicazione: (2025)
di: Sonarghare, Shubham, et al.
Pubblicazione: (2025)
Depth Estimation using Weighted-loss and Transfer Learning
di: Hafeez, Muhammad Adeel, et al.
Pubblicazione: (2024)
di: Hafeez, Muhammad Adeel, et al.
Pubblicazione: (2024)
CARE Drive A Framework for Evaluating Reason-Responsiveness of Vision Language Models in Automated Driving
di: Suryana, Lucas Elbert, et al.
Pubblicazione: (2026)
di: Suryana, Lucas Elbert, et al.
Pubblicazione: (2026)
Surround-View Fisheye Optics in Computer Vision and Simulation: Survey and Challenges
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
di: Jakab, Daniel, et al.
Pubblicazione: (2024)
Beyond the Known: Adversarial Autoencoders in Novelty Detection
di: Asad, Muhammad, et al.
Pubblicazione: (2024)
di: Asad, Muhammad, et al.
Pubblicazione: (2024)
Auto-Comp: An Automated Pipeline for Scalable Compositional Probing of Contrastive Vision-Language Models
di: Sbrolli, Cristian, et al.
Pubblicazione: (2026)
di: Sbrolli, Cristian, et al.
Pubblicazione: (2026)
Scalable and Efficient Hierarchical Visual Topological Mapping
di: Ramachandran, Saravanabalagi, et al.
Pubblicazione: (2024)
di: Ramachandran, Saravanabalagi, et al.
Pubblicazione: (2024)
Multi-Scale Spectral Attention Module-based Hyperspectral Segmentation in Autonomous Driving Scenarios
di: Shah, Imad Ali, et al.
Pubblicazione: (2025)
di: Shah, Imad Ali, et al.
Pubblicazione: (2025)
Perceiving Beyond Language Priors: Enhancing Visual Comprehension and Attention in Multimodal Models
di: Ghatkesar, Aarti, et al.
Pubblicazione: (2025)
di: Ghatkesar, Aarti, et al.
Pubblicazione: (2025)
WoodScape Motion Segmentation for Autonomous Driving -- CVPR 2023 OmniCV Workshop Challenge
di: Ramachandran, Saravanabalagi, et al.
Pubblicazione: (2023)
di: Ramachandran, Saravanabalagi, et al.
Pubblicazione: (2023)
LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model
di: Mei, Xiaodong, et al.
Pubblicazione: (2026)
di: Mei, Xiaodong, et al.
Pubblicazione: (2026)
A Vision-Language Foundation Model for Zero-shot Clinical Collaboration and Automated Concept Discovery in Dermatology
di: Yan, Siyuan, et al.
Pubblicazione: (2026)
di: Yan, Siyuan, et al.
Pubblicazione: (2026)
Language-Guided Invariance Probing of Vision-Language Models
di: Lee, Jae Joong
Pubblicazione: (2025)
di: Lee, Jae Joong
Pubblicazione: (2025)
Probing Vision-Language Understanding through the Visual Entailment Task: promises and pitfalls
di: Pitta, Elena, et al.
Pubblicazione: (2025)
di: Pitta, Elena, et al.
Pubblicazione: (2025)
Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?
di: He, Jingtao, et al.
Pubblicazione: (2026)
di: He, Jingtao, et al.
Pubblicazione: (2026)
Probing Perceptual Constancy in Large Vision-Language Models
di: Sun, Haoran, et al.
Pubblicazione: (2025)
di: Sun, Haoran, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Evaluating Small Vision-Language Models on Distance-Dependent Traffic Perception
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025) -
Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)
di: Theodoridis, Nikos, et al.
Pubblicazione: (2025) -
FIN: Fast Inference Network for Map Segmentation
di: Bispo, Ruan, et al.
Pubblicazione: (2025) -
Occluded nuScenes: A Multi-Sensor Dataset for Evaluating Perception Robustness in Automated Driving
di: Kumar, Sanjay, et al.
Pubblicazione: (2025) -
Deformable Convolution Based Road Scene Semantic Segmentation of Fisheye Images in Autonomous Driving
di: Manzoor, Anam, et al.
Pubblicazione: (2024)