Evaluating Small Vision-Language Models on Distance-Dependent Traffic Perception
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Theodoridis, Nikos, Brophy, Tim, Mohandas, Reenu, Sistu, Ganesh, Collins, Fiachra, Scanlan, Anthony, Eising, Ciaran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2025)
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2025)
Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2026)
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2026)
FIN: Fast Inference Network for Map Segmentation
von: Bispo, Ruan, et al.
Veröffentlicht: (2025)
von: Bispo, Ruan, et al.
Veröffentlicht: (2025)
Occluded nuScenes: A Multi-Sensor Dataset for Evaluating Perception Robustness in Automated Driving
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
Deformable Convolution Based Road Scene Semantic Segmentation of Fisheye Images in Autonomous Driving
von: Manzoor, Anam, et al.
Veröffentlicht: (2024)
von: Manzoor, Anam, et al.
Veröffentlicht: (2024)
SuperQuadricOcc: Real-Time Self-Supervised Semantic Occupancy Estimation with Superquadric Volume Rendering
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
Easy3D-Labels: Supervising Semantic Occupancy Estimation with 3D Pseudo-Labels for Automotive Perception
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
Revisiting Birds Eye View Perception Models with Frozen Foundation Models: DINOv2 and Metric3Dv2
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
von: Hayes, Seamie, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Weather-Induced Sensor Occlusion on BEVFusion for 3D Object Detection
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
Minimizing Occlusion Effect on Multi-View Camera Perception in BEV with Multi-Sensor Fusion
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025)
Optimizing Visual Question Answering Models for Driving: Bridging the Gap Between Human and Machine Attention Patterns
von: Rekanar, Kaavya, et al.
Veröffentlicht: (2024)
von: Rekanar, Kaavya, et al.
Veröffentlicht: (2024)
SS-SFR: Synthetic Scenes Spatial Frequency Response on Virtual KITTI and Degraded Automotive Simulations for Object Detection
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
MapsTP: HD Map Images Based Multimodal Trajectory Prediction for Automated Vehicles
von: Sharma, Sushil, et al.
Veröffentlicht: (2024)
von: Sharma, Sushil, et al.
Veröffentlicht: (2024)
Optimizing Ego Vehicle Trajectory Prediction: The Graph Enhancement Approach
von: Sharma, Sushil, et al.
Veröffentlicht: (2023)
von: Sharma, Sushil, et al.
Veröffentlicht: (2023)
Fisheye Camera and Ultrasonic Sensor Fusion For Near-Field Obstacle Perception in Bird's-Eye-View
von: Das, Arindam, et al.
Veröffentlicht: (2024)
von: Das, Arindam, et al.
Veröffentlicht: (2024)
SOLAS: Superpositioning an Optical Lens in Automotive Simulation
von: Jakab, Daniel, et al.
Veröffentlicht: (2025)
von: Jakab, Daniel, et al.
Veröffentlicht: (2025)
Reflective Teacher: Semi-Supervised Multimodal 3D Object Detection in Bird's-Eye-View via Uncertainty Measure
von: Hazra, Saheli, et al.
Veröffentlicht: (2024)
von: Hazra, Saheli, et al.
Veröffentlicht: (2024)
BEVMOSNet: Multimodal Fusion for BEV Moving Object Segmentation
von: Cong, Hiep Truong, et al.
Veröffentlicht: (2025)
von: Cong, Hiep Truong, et al.
Veröffentlicht: (2025)
FisheyeGaussianLift: BEV Feature Lifting for Surround-View Fisheye Camera Perception
von: Sonarghare, Shubham, et al.
Veröffentlicht: (2025)
von: Sonarghare, Shubham, et al.
Veröffentlicht: (2025)
PAN: Pillars-Attention-Based Network for 3D Object Detection
von: Bispo, Ruan, et al.
Veröffentlicht: (2025)
von: Bispo, Ruan, et al.
Veröffentlicht: (2025)
Velocity Driven Vision: Asynchronous Sensor Fusion Birds Eye View Models for Autonomous Vehicles
von: Hayes, Seamie, et al.
Veröffentlicht: (2024)
von: Hayes, Seamie, et al.
Veröffentlicht: (2024)
NeurAll: Towards a Unified Visual Perception Model for Automated Driving
von: Sistu, Ganesh, et al.
Veröffentlicht: (2019)
von: Sistu, Ganesh, et al.
Veröffentlicht: (2019)
Measuring Natural Scenes SFR of Automotive Fisheye Cameras
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
FisheyeDetNet: 360° Surround view Fisheye Camera based Object Detection System for Autonomous Driving
von: Sistu, Ganesh, et al.
Veröffentlicht: (2024)
von: Sistu, Ganesh, et al.
Veröffentlicht: (2024)
Surround-View Fisheye Optics in Computer Vision and Simulation: Survey and Challenges
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
von: Jakab, Daniel, et al.
Veröffentlicht: (2024)
Subgraph Clustering and Atom Learning for Improved Image Classification
von: Singh, Aryan, et al.
Veröffentlicht: (2024)
von: Singh, Aryan, et al.
Veröffentlicht: (2024)
Scalable and Efficient Hierarchical Visual Topological Mapping
von: Ramachandran, Saravanabalagi, et al.
Veröffentlicht: (2024)
von: Ramachandran, Saravanabalagi, et al.
Veröffentlicht: (2024)
WoodScape Motion Segmentation for Autonomous Driving -- CVPR 2023 OmniCV Workshop Challenge
von: Ramachandran, Saravanabalagi, et al.
Veröffentlicht: (2023)
von: Ramachandran, Saravanabalagi, et al.
Veröffentlicht: (2023)
Image Segmentation: Inducing graph-based learning
von: Singh, Aryan, et al.
Veröffentlicht: (2025)
von: Singh, Aryan, et al.
Veröffentlicht: (2025)
Time Series Anomaly Detection with CNN for Environmental Sensors in Healthcare-IoT
von: Khatun, Mirza Akhi, et al.
Veröffentlicht: (2024)
von: Khatun, Mirza Akhi, et al.
Veröffentlicht: (2024)
Depth Estimation using Weighted-loss and Transfer Learning
von: Hafeez, Muhammad Adeel, et al.
Veröffentlicht: (2024)
von: Hafeez, Muhammad Adeel, et al.
Veröffentlicht: (2024)
Beyond the Known: Adversarial Autoencoders in Novelty Detection
von: Asad, Muhammad, et al.
Veröffentlicht: (2024)
von: Asad, Muhammad, et al.
Veröffentlicht: (2024)
When Negation Is a Geometry Problem in Vision-Language Models
von: Sammani, Fawaz, et al.
Veröffentlicht: (2026)
von: Sammani, Fawaz, et al.
Veröffentlicht: (2026)
CSNR and JMIM Based Spectral Band Selection for Reducing Metamerism in Urban Driving
von: Li, Jiarong, et al.
Veröffentlicht: (2025)
von: Li, Jiarong, et al.
Veröffentlicht: (2025)
GPT-4V as Traffic Assistant: An In-depth Look at Vision Language Model on Complex Traffic Events
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
Same or Not? Enhancing Visual Perception in Vision-Language Models
von: Marsili, Damiano, et al.
Veröffentlicht: (2025)
von: Marsili, Damiano, et al.
Veröffentlicht: (2025)
HazardNet: A Small-Scale Vision Language Model for Real-Time Traffic Safety Detection at Edge Devices
von: Tami, Mohammad Abu, et al.
Veröffentlicht: (2025)
von: Tami, Mohammad Abu, et al.
Veröffentlicht: (2025)
Small Language Model Meets with Reinforced Vision Vocabulary
von: Wei, Haoran, et al.
Veröffentlicht: (2024)
von: Wei, Haoran, et al.
Veröffentlicht: (2024)
Contrastive Learning-Driven Traffic Sign Perception: Multi-Modal Fusion of Text and Vision
von: Lu, Qiang, et al.
Veröffentlicht: (2025)
von: Lu, Qiang, et al.
Veröffentlicht: (2025)
Hallucination Elimination and Semantic Enhancement Framework for Vision-Language Models in Traffic Scenarios
von: Fan, Jiaqi, et al.
Veröffentlicht: (2024)
von: Fan, Jiaqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2025) -
Probing Visual Concepts in Lightweight Vision-Language Models for Automated Driving
von: Theodoridis, Nikos, et al.
Veröffentlicht: (2026) -
FIN: Fast Inference Network for Map Segmentation
von: Bispo, Ruan, et al.
Veröffentlicht: (2025) -
Occluded nuScenes: A Multi-Sensor Dataset for Evaluating Perception Robustness in Automated Driving
von: Kumar, Sanjay, et al.
Veröffentlicht: (2025) -
Deformable Convolution Based Road Scene Semantic Segmentation of Fisheye Images in Autonomous Driving
von: Manzoor, Anam, et al.
Veröffentlicht: (2024)