Vision Technologies with Applications in Traffic Surveillance Systems: A Holistic Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Wei, Yang, Li, Zhao, Lei, Zhang, Runyu, Cui, Yifan, Huang, Hongpu, Qie, Kun, Wang, Chen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Physical Depth-aware Early Accident Anticipation: A Multi-dimensional Visual Feature Fusion Framework
von: Huang, Hongpu, et al.
Veröffentlicht: (2025)
von: Huang, Hongpu, et al.
Veröffentlicht: (2025)
OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models
von: Zhong, Yufeng, et al.
Veröffentlicht: (2026)
von: Zhong, Yufeng, et al.
Veröffentlicht: (2026)
TrafficLoc: Localizing Traffic Surveillance Cameras in 3D Scenes
von: Xia, Yan, et al.
Veröffentlicht: (2024)
von: Xia, Yan, et al.
Veröffentlicht: (2024)
Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video
von: Huang, Jiantang
Veröffentlicht: (2026)
von: Huang, Jiantang
Veröffentlicht: (2026)
ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)
Vision Transformer for Robust Occluded Person Reidentification in Complex Surveillance Scenes
von: Li, Bo, et al.
Veröffentlicht: (2025)
von: Li, Bo, et al.
Veröffentlicht: (2025)
Mobile Traffic Camera Calibration from Road Geometry for UAV-Based Traffic Surveillance
von: Popov, Alexey, et al.
Veröffentlicht: (2026)
von: Popov, Alexey, et al.
Veröffentlicht: (2026)
HoVLE: Unleashing the Power of Monolithic Vision-Language Models with Holistic Vision-Language Embedding
von: Tao, Chenxin, et al.
Veröffentlicht: (2024)
von: Tao, Chenxin, et al.
Veröffentlicht: (2024)
VHELM: A Holistic Evaluation of Vision Language Models
von: Lee, Tony, et al.
Veröffentlicht: (2024)
von: Lee, Tony, et al.
Veröffentlicht: (2024)
A Review of Vision-Based Assistive Systems for Visually Impaired People: Technologies, Applications, and Future Directions
von: Yao, Fulong, et al.
Veröffentlicht: (2025)
von: Yao, Fulong, et al.
Veröffentlicht: (2025)
A Survey of Medical Vision-and-Language Applications and Their Techniques
von: Chen, Qi, et al.
Veröffentlicht: (2024)
von: Chen, Qi, et al.
Veröffentlicht: (2024)
ProDisc-VAD: An Efficient System for Weakly-Supervised Anomaly Detection in Video Surveillance Applications
von: Zhu, Tao, et al.
Veröffentlicht: (2025)
von: Zhu, Tao, et al.
Veröffentlicht: (2025)
Improving Vision-language Models with Perception-centric Process Reward Models
von: Min, Yingqian, et al.
Veröffentlicht: (2026)
von: Min, Yingqian, et al.
Veröffentlicht: (2026)
Evaluation of Vision-LLMs in Surveillance Video
von: Benschop, Pascal, et al.
Veröffentlicht: (2025)
von: Benschop, Pascal, et al.
Veröffentlicht: (2025)
Hulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language Understanding
von: Jiang, Songtao, et al.
Veröffentlicht: (2025)
von: Jiang, Songtao, et al.
Veröffentlicht: (2025)
SIGMA: Bridging Structural and Distributional Gaps for Vision Foundation Model Adaptation
von: Xiong, Lingyu, et al.
Veröffentlicht: (2026)
von: Xiong, Lingyu, et al.
Veröffentlicht: (2026)
MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2025)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2025)
MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
von: Zhao, Kaikai, et al.
Veröffentlicht: (2025)
TSBOW -- Traffic Surveillance Benchmark for Occluded Vehicles Under Various Weather Conditions
von: Huynh, Ngoc Doan-Minh, et al.
Veröffentlicht: (2026)
von: Huynh, Ngoc Doan-Minh, et al.
Veröffentlicht: (2026)
CAS-ViT: Convolutional Additive Self-attention Vision Transformers for Efficient Mobile Applications
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
GPT-4V as Traffic Assistant: An In-depth Look at Vision Language Model on Complex Traffic Events
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
Vision-Language Models for Vision Tasks: A Survey
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
GPT4SGG: Synthesizing Scene Graphs from Holistic and Region-specific Narratives
von: Chen, Zuyao, et al.
Veröffentlicht: (2023)
von: Chen, Zuyao, et al.
Veröffentlicht: (2023)
Unified Language-Vision Pretraining in LLM with Dynamic Discrete Visual Tokenization
von: Jin, Yang, et al.
Veröffentlicht: (2023)
von: Jin, Yang, et al.
Veröffentlicht: (2023)
SurveillanceVQA-589K: A Benchmark for Comprehensive Surveillance Video-Language Understanding with Large Models
von: Liu, Bo, et al.
Veröffentlicht: (2025)
von: Liu, Bo, et al.
Veröffentlicht: (2025)
Wavelet-Enhanced Desnowing: A Novel Single Image Restoration Approach for Traffic Surveillance under Adverse Weather Conditions
von: Shen, Zihan, et al.
Veröffentlicht: (2025)
von: Shen, Zihan, et al.
Veröffentlicht: (2025)
A Survey on Remote Sensing Foundation Models: From Vision to Multimodality
von: Huang, Ziyue, et al.
Veröffentlicht: (2025)
von: Huang, Ziyue, et al.
Veröffentlicht: (2025)
A Unified Detection Pipeline for Robust Object Detection in Fisheye-Based Traffic Surveillance
von: Owor, Neema Jakisa, et al.
Veröffentlicht: (2025)
von: Owor, Neema Jakisa, et al.
Veröffentlicht: (2025)
Person Recognition in Aerial Surveillance: A Decade Survey
von: Nguyen, Kien, et al.
Veröffentlicht: (2025)
von: Nguyen, Kien, et al.
Veröffentlicht: (2025)
Composed Multi-modal Retrieval: A Survey of Approaches and Applications
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
Libra: Building Decoupled Vision System on Large Language Models
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
Physics-Based Adversarial Attack on Near-Infrared Human Detector for Nighttime Surveillance Camera Systems
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
Large Language Models for Video Surveillance Applications
von: De Silva, Ulindu, et al.
Veröffentlicht: (2025)
von: De Silva, Ulindu, et al.
Veröffentlicht: (2025)
Data-Centric Evolution in Autonomous Driving: A Comprehensive Survey of Big Data System, Data Mining, and Closed-Loop Technologies
von: Li, Lincan, et al.
Veröffentlicht: (2024)
von: Li, Lincan, et al.
Veröffentlicht: (2024)
Intelligent Traffic Surveillance for Real-Time Vehicle Detection, License Plate Recognition, and Speed Estimation
von: Mugizi, Bruce, et al.
Veröffentlicht: (2026)
von: Mugizi, Bruce, et al.
Veröffentlicht: (2026)
Towards Vision-Language Geo-Foundation Model: A Survey
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
THEMIS: Towards Holistic Evaluation of MLLMs for Scientific Paper Fraud Forensics
von: Ma, Tzu-Yen, et al.
Veröffentlicht: (2026)
von: Ma, Tzu-Yen, et al.
Veröffentlicht: (2026)
Deep Learning Technology for Face Forgery Detection: A Survey
von: Ma, Lixia, et al.
Veröffentlicht: (2024)
von: Ma, Lixia, et al.
Veröffentlicht: (2024)
Can 3D Vision-Language Models Truly Understand Natural Language?
von: Deng, Weipeng, et al.
Veröffentlicht: (2024)
von: Deng, Weipeng, et al.
Veröffentlicht: (2024)
Vision Mamba in Remote Sensing: A Comprehensive Survey of Techniques, Applications and Outlook
von: Bao, Muyi, et al.
Veröffentlicht: (2025)
von: Bao, Muyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Physical Depth-aware Early Accident Anticipation: A Multi-dimensional Visual Feature Fusion Framework
von: Huang, Hongpu, et al.
Veröffentlicht: (2025) -
OCRVerse: Towards Holistic OCR in End-to-End Vision-Language Models
von: Zhong, Yufeng, et al.
Veröffentlicht: (2026) -
TrafficLoc: Localizing Traffic Surveillance Cameras in 3D Scenes
von: Xia, Yan, et al.
Veröffentlicht: (2024) -
Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video
von: Huang, Jiantang
Veröffentlicht: (2026) -
ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation
von: Fu, Haoyu, et al.
Veröffentlicht: (2025)