A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Doan, Gia-Bao, Huynh, Nam-Khoa, Ho, Minh-Nhat-Huy, Nguyen, Khanh-Thanh-Khoa, Le, Thanh-Hai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advanced Long-term Earth System Forecasting
von: Wu, Hao, et al.
Veröffentlicht: (2025)
von: Wu, Hao, et al.
Veröffentlicht: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
A Guide to Structureless Visual Localization
von: Panek, Vojtech, et al.
Veröffentlicht: (2025)
von: Panek, Vojtech, et al.
Veröffentlicht: (2025)
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
von: Panek, Vojtech, et al.
Veröffentlicht: (2024)
von: Panek, Vojtech, et al.
Veröffentlicht: (2024)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
von: Panek, Vojtech, et al.
Veröffentlicht: (2026)
von: Panek, Vojtech, et al.
Veröffentlicht: (2026)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
von: Tourani, Ali, et al.
Veröffentlicht: (2024)
von: Tourani, Ali, et al.
Veröffentlicht: (2024)
A Comprehensive Review of Fish Feeding Behavior Analysis in Aquaculture: Tasks, Techniques, and Applications
von: Zhang, Shulong, et al.
Veröffentlicht: (2025)
von: Zhang, Shulong, et al.
Veröffentlicht: (2025)
Positive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2025)
Accelerating Post-Tornado Disaster Assessment Using Advanced Deep Learning Models
von: Umeike, Robinson, et al.
Veröffentlicht: (2024)
von: Umeike, Robinson, et al.
Veröffentlicht: (2024)
Knee Osteoarthritis Severity Grading Using Optimized Deep Learning and LLM-Driven Intelligent AI on Computationally Limited Systems
von: Nadeem, Dayam, et al.
Veröffentlicht: (2026)
von: Nadeem, Dayam, et al.
Veröffentlicht: (2026)
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
von: Liu, Shanyuan, et al.
Veröffentlicht: (2023)
von: Liu, Shanyuan, et al.
Veröffentlicht: (2023)
Human-Centric Anomaly Detection in Surveillance Videos Using YOLO-World and Spatio-Temporal Deep Learning
von: Naeen, Mohammad Ali Etemadi, et al.
Veröffentlicht: (2025)
von: Naeen, Mohammad Ali Etemadi, et al.
Veröffentlicht: (2025)
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
A Light Perspective for 3D Object Detection
von: Pederiva, Marcelo Eduardo, et al.
Veröffentlicht: (2025)
von: Pederiva, Marcelo Eduardo, et al.
Veröffentlicht: (2025)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
von: Mahdian, Navid, et al.
Veröffentlicht: (2024)
von: Mahdian, Navid, et al.
Veröffentlicht: (2024)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
von: Seo, Huichan, et al.
Veröffentlicht: (2025)
von: Seo, Huichan, et al.
Veröffentlicht: (2025)
Reference Dataset and Benchmark for Reconstructing Laser Parameters from On-axis Video in Powder Bed Fusion of Bulk Stainless Steel
von: Blanc, Cyril, et al.
Veröffentlicht: (2024)
von: Blanc, Cyril, et al.
Veröffentlicht: (2024)
Facial Attribute Based Text Guided Face Anonymization
von: Muştu, Mustafa İzzet, et al.
Veröffentlicht: (2025)
von: Muştu, Mustafa İzzet, et al.
Veröffentlicht: (2025)
Cross-Domain Adversarial Augmentation: Stabilizing GANs for Medical and Handwriting Data Scarcity
von: Soad, Md. Sohanuzzaman, et al.
Veröffentlicht: (2026)
von: Soad, Md. Sohanuzzaman, et al.
Veröffentlicht: (2026)
Few TensoRF: Enhance the Few-shot on Tensorial Radiance Fields
von: Le, Thanh-Hai, et al.
Veröffentlicht: (2026)
von: Le, Thanh-Hai, et al.
Veröffentlicht: (2026)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
von: Dua, Karan, et al.
Veröffentlicht: (2025)
von: Dua, Karan, et al.
Veröffentlicht: (2025)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
von: Karam, Christophe, et al.
Veröffentlicht: (2024)
von: Karam, Christophe, et al.
Veröffentlicht: (2024)
EventFlow: Real-Time Neuromorphic Event-Driven Classification of Two-Phase Boiling Flow Regimes
von: Chang, Sanghyeon, et al.
Veröffentlicht: (2025)
von: Chang, Sanghyeon, et al.
Veröffentlicht: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization
von: He, Mengqi, et al.
Veröffentlicht: (2026)
von: He, Mengqi, et al.
Veröffentlicht: (2026)
Palmistry-Informed Feature Extraction and Analysis using Machine Learning
von: Patil, Shweta
Veröffentlicht: (2025)
von: Patil, Shweta
Veröffentlicht: (2025)
LightFFDNets: Lightweight Convolutional Neural Networks for Rapid Facial Forgery Detection
von: Jabbarlı, Günel, et al.
Veröffentlicht: (2024)
von: Jabbarlı, Günel, et al.
Veröffentlicht: (2024)
Dynamic Residual Encoding with Slide-Level Contrastive Learning for End-to-End Whole Slide Image Representation
von: Jin, Jing, et al.
Veröffentlicht: (2025)
von: Jin, Jing, et al.
Veröffentlicht: (2025)
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
von: Zhang, Yan, et al.
Veröffentlicht: (2026)
von: Zhang, Yan, et al.
Veröffentlicht: (2026)
Intelligent Vacuum Thermoforming Process
von: Kuswoyo, Andi, et al.
Veröffentlicht: (2025)
von: Kuswoyo, Andi, et al.
Veröffentlicht: (2025)
AVControl: Efficient Framework for Training Audio-Visual Controls
von: Ben-Yosef, Matan, et al.
Veröffentlicht: (2026)
von: Ben-Yosef, Matan, et al.
Veröffentlicht: (2026)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
Parameter-efficient fine-tuning (PEFT) of Vision Foundation Models for Atypical Mitotic Figure Classification
von: Ramchandani, Lavish, et al.
Veröffentlicht: (2025)
von: Ramchandani, Lavish, et al.
Veröffentlicht: (2025)
Intrinsic Image Fusion for Multi-View 3D Material Reconstruction
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
Revealing an Unattractivity Bias in Mental Reconstruction of Occluded Faces using Generative Image Models
von: Riedmann, Frederik, et al.
Veröffentlicht: (2024)
von: Riedmann, Frederik, et al.
Veröffentlicht: (2024)
IntrinsiX: High-Quality PBR Generation using Image Priors
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
von: Kocsis, Peter, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Advanced Long-term Earth System Forecasting
von: Wu, Hao, et al.
Veröffentlicht: (2025) -
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
von: Tourani, Ali, et al.
Veröffentlicht: (2023) -
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024) -
A Guide to Structureless Visual Localization
von: Panek, Vojtech, et al.
Veröffentlicht: (2025) -
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
von: Panek, Vojtech, et al.
Veröffentlicht: (2024)