Vision-Language Models can Identify Distracted Driver Behavior from Naturalistic Videos
Fuente:
arXiv
Guardado en:
| Autores principales: | Hasan, Md Zahid, Chen, Jiajing, Wang, Jiyang, Rahman, Mohammed Shaiqur, Joshi, Ameya, Velipasalar, Senem, Hegde, Chinmay, Sharma, Anuj, Sarkar, Soumik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Calibrated and Resource-Aware Super-Resolution for Reliable Driver Behavior Analysis
por: Shihab, Ibne Farabi, et al.
Publicado: (2025)
por: Shihab, Ibne Farabi, et al.
Publicado: (2025)
Block-As-Domain Adaptation for Workload Prediction from fNIRS Data
por: Wang, Jiyang, et al.
Publicado: (2024)
por: Wang, Jiyang, et al.
Publicado: (2024)
PRISM: Product Retrieval In Shopping Carts using Hybrid Matching
por: Kabadayi, Arda, et al.
Publicado: (2025)
por: Kabadayi, Arda, et al.
Publicado: (2025)
GaitPoint+: A Gait Recognition Network Incorporating Point Cloud Analysis and Recycling
por: Ren, Huantao, et al.
Publicado: (2024)
por: Ren, Huantao, et al.
Publicado: (2024)
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception
por: Waite, Joshua R., et al.
Publicado: (2025)
por: Waite, Joshua R., et al.
Publicado: (2025)
Driving as a Diagnostic Tool: Scenario-based Cognitive Assessment in Older Drivers from Driving Video
por: Hasan, Md Zahid, et al.
Publicado: (2025)
por: Hasan, Md Zahid, et al.
Publicado: (2025)
DeepLocalization: Using change point detection for Temporal Action Localization
por: Rahman, Mohammed Shaiqur, et al.
Publicado: (2024)
por: Rahman, Mohammed Shaiqur, et al.
Publicado: (2024)
Trans${^2}$-CBCT: A Dual-Transformer Framework for Sparse-View CBCT Reconstruction
por: Yang, Minmin, et al.
Publicado: (2025)
por: Yang, Minmin, et al.
Publicado: (2025)
DG-MVP: 3D Domain Generalization via Multiple Views of Point Clouds for Classification
por: Ren, Huantao, et al.
Publicado: (2025)
por: Ren, Huantao, et al.
Publicado: (2025)
3D-PointZshotS: Geometry-Aware 3D Point Cloud Zero-Shot Semantic Segmentation Narrowing the Visual-Semantic Gap
por: Yang, Minmin, et al.
Publicado: (2025)
por: Yang, Minmin, et al.
Publicado: (2025)
Predicting Mild Cognitive Impairment Using Naturalistic Driving and Trip Destination Modeling
por: Chattopadhyay, Souradeep, et al.
Publicado: (2025)
por: Chattopadhyay, Souradeep, et al.
Publicado: (2025)
Feature-based Federated Transfer Learning: Communication Efficiency, Robustness and Privacy
por: Wang, Feng, et al.
Publicado: (2024)
por: Wang, Feng, et al.
Publicado: (2024)
CLIP-BEVFormer: Enhancing Multi-View Image-Based BEV Detector with Ground Truth Flow
por: Pan, Chenbin, et al.
Publicado: (2024)
por: Pan, Chenbin, et al.
Publicado: (2024)
Balancing Utility and Privacy: Dynamically Private SGD with Random Projection
por: Jiang, Zhanhong, et al.
Publicado: (2025)
por: Jiang, Zhanhong, et al.
Publicado: (2025)
A Contextual Analysis of Driver-Facing and Dual-View Video Inputs for Distraction Detection in Naturalistic Driving Environments
por: Dontoh, Anthony, et al.
Publicado: (2025)
por: Dontoh, Anthony, et al.
Publicado: (2025)
Only My Model On My Data: A Privacy Preserving Approach Protecting one Model and Deceiving Unauthorized Black-Box Models
por: Chai, Weiheng, et al.
Publicado: (2024)
por: Chai, Weiheng, et al.
Publicado: (2024)
Geometry Matters: Benchmarking Scientific ML Approaches for Flow Prediction around Complex Geometries
por: Rabeh, Ali, et al.
Publicado: (2024)
por: Rabeh, Ali, et al.
Publicado: (2024)
FUSE: First-Order and Second-Order Unified SynthEsis in Stochastic Optimization
por: Jiang, Zhanhong, et al.
Publicado: (2025)
por: Jiang, Zhanhong, et al.
Publicado: (2025)
VLP: Vision Language Planning for Autonomous Driving
por: Pan, Chenbin, et al.
Publicado: (2024)
por: Pan, Chenbin, et al.
Publicado: (2024)
Zero‐shot insect detection via weak language supervision
por: Benjamin Feuer, et al.
Publicado: (2024)
por: Benjamin Feuer, et al.
Publicado: (2024)
Beyond the Dashboard: Investigating Distracted Driver Communication Preferences for ADAS
por: Hasan, Aamir, et al.
Publicado: (2024)
por: Hasan, Aamir, et al.
Publicado: (2024)
Fast Certification of Vision-Language Models Using Incremental Randomized Smoothing
por: Nirala, A K, et al.
Publicado: (2023)
por: Nirala, A K, et al.
Publicado: (2023)
ADKO: Agentic Decentralized Knowledge Optimization
por: Rillo, Lucas Nerone, et al.
Publicado: (2026)
por: Rillo, Lucas Nerone, et al.
Publicado: (2026)
DeCAF: Decentralized Consensus-And-Factorization for Low-Rank Adaptation of Foundation Models
por: Saadati, Nastaran, et al.
Publicado: (2025)
por: Saadati, Nastaran, et al.
Publicado: (2025)
STITCH: Surface reconstrucTion using Implicit neural representations with Topology Constraints and persistent Homology
por: Jignasu, Anushrut, et al.
Publicado: (2024)
por: Jignasu, Anushrut, et al.
Publicado: (2024)
GENESIS-RL: GEnerating Natural Edge-cases with Systematic Integration of Safety considerations and Reinforcement Learning
por: Yang, Hsin-Jung, et al.
Publicado: (2024)
por: Yang, Hsin-Jung, et al.
Publicado: (2024)
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
por: Doan, Gia-Bao, et al.
Publicado: (2026)
por: Doan, Gia-Bao, et al.
Publicado: (2026)
LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool
por: Ma, Yue, et al.
Publicado: (2024)
por: Ma, Yue, et al.
Publicado: (2024)
Search-contempt: a hybrid MCTS algorithm for training AlphaZero-like engines with better computational efficiency
por: Joshi, Ameya
Publicado: (2025)
por: Joshi, Ameya
Publicado: (2025)
WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training
por: Feuer, Benjamin, et al.
Publicado: (2025)
por: Feuer, Benjamin, et al.
Publicado: (2025)
Driver Age and Its Effect on Key Driving Metrics: Insights from Dynamic Vehicle Data
por: Joshi, Aparna, et al.
Publicado: (2025)
por: Joshi, Aparna, et al.
Publicado: (2025)
DIMAT: Decentralized Iterative Merging-And-Training for Deep Learning Models
por: Saadati, Nastaran, et al.
Publicado: (2024)
por: Saadati, Nastaran, et al.
Publicado: (2024)
EyeCue: Driver Cognitive Distraction Detection via Gaze-Empowered Egocentric Video Understanding
por: Zhang, Lang, et al.
Publicado: (2026)
por: Zhang, Lang, et al.
Publicado: (2026)
Leveraging Vision Language Models for Specialized Agricultural Tasks
por: Arshad, Muhammad Arbab, et al.
Publicado: (2024)
por: Arshad, Muhammad Arbab, et al.
Publicado: (2024)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
por: Ma, Yunsheng, et al.
Publicado: (2022)
por: Ma, Yunsheng, et al.
Publicado: (2022)
Zero-Shot Distracted Driver Detection via Vision Language Models with Double Decoupling
por: Miyata, Takamichi, et al.
Publicado: (2026)
por: Miyata, Takamichi, et al.
Publicado: (2026)
BioTrove: A Large Curated Image Dataset Enabling AI for Biodiversity
por: Yang, Chih-Hsuan, et al.
Publicado: (2024)
por: Yang, Chih-Hsuan, et al.
Publicado: (2024)
Towards Infusing Auxiliary Knowledge for Distracted Driver Detection
por: Balappanawar, Ishwar B, et al.
Publicado: (2024)
por: Balappanawar, Ishwar B, et al.
Publicado: (2024)
LLMs can be easily Confused by Instructional Distractions
por: Hwang, Yerin, et al.
Publicado: (2025)
por: Hwang, Yerin, et al.
Publicado: (2025)
Evaluating Parametric Car-Following Models in Naturalistic Congestion: Insights in Driver Behavior and Model Limitations
por: Hou, Huaidian, et al.
Publicado: (2025)
por: Hou, Huaidian, et al.
Publicado: (2025)
Ejemplares similares
-
Calibrated and Resource-Aware Super-Resolution for Reliable Driver Behavior Analysis
por: Shihab, Ibne Farabi, et al.
Publicado: (2025) -
Block-As-Domain Adaptation for Workload Prediction from fNIRS Data
por: Wang, Jiyang, et al.
Publicado: (2024) -
PRISM: Product Retrieval In Shopping Carts using Hybrid Matching
por: Kabadayi, Arda, et al.
Publicado: (2025) -
GaitPoint+: A Gait Recognition Network Incorporating Point Cloud Analysis and Recycling
por: Ren, Huantao, et al.
Publicado: (2024) -
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception
por: Waite, Joshua R., et al.
Publicado: (2025)