Zero-Shot Distracted Driver Detection via Vision Language Models with Double Decoupling
Fuente:
arXiv
Salvato in:
| Autori principali: | Miyata, Takamichi, Miyata, Sumiko, Morris, Andrew |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
di: Javed, Sajid, et al.
Pubblicazione: (2024)
di: Javed, Sajid, et al.
Pubblicazione: (2024)
Lung-CADex: Fully automatic Zero-Shot Detection and Classification of Lung Nodules in Thoracic CT Images
di: Shaukat, Furqan, et al.
Pubblicazione: (2024)
di: Shaukat, Furqan, et al.
Pubblicazione: (2024)
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
di: Wang, Sicheng, et al.
Pubblicazione: (2024)
di: Wang, Sicheng, et al.
Pubblicazione: (2024)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
di: Guizilini, Vitor, et al.
Pubblicazione: (2024)
di: Guizilini, Vitor, et al.
Pubblicazione: (2024)
Robustly Optimized Deep Feature Decoupling Network for Fatty Liver Diseases Detection
di: Huang, Peng, et al.
Pubblicazione: (2024)
di: Huang, Peng, et al.
Pubblicazione: (2024)
Filter2Noise: A Framework for Interpretable and Zero-Shot Low-Dose CT Image Denoising
di: Sun, Yipeng, et al.
Pubblicazione: (2025)
di: Sun, Yipeng, et al.
Pubblicazione: (2025)
Zero-Splat TeleAssist: A Zero-Shot Pose Estimation Framework for Semantic Teleoperation
di: Dokania, Srijan, et al.
Pubblicazione: (2025)
di: Dokania, Srijan, et al.
Pubblicazione: (2025)
Blinking Beyond EAR: A Stable Eyelid Angle Metric for Driver Drowsiness Detection and Data Augmentation
di: Wolter, Mathis, et al.
Pubblicazione: (2025)
di: Wolter, Mathis, et al.
Pubblicazione: (2025)
VIP: Visual Information Protection through Adversarial Attacks on Vision-Language Models
di: Meftah, Hanene F. Z. Brachemi, et al.
Pubblicazione: (2025)
di: Meftah, Hanene F. Z. Brachemi, et al.
Pubblicazione: (2025)
Vision-Language Generative Model for View-Specific Chest X-ray Generation
di: Lee, Hyungyung, et al.
Pubblicazione: (2023)
di: Lee, Hyungyung, et al.
Pubblicazione: (2023)
HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet
di: Patro, Badri N., et al.
Pubblicazione: (2026)
di: Patro, Badri N., et al.
Pubblicazione: (2026)
Polyp SAM 2: Advancing Zero shot Polyp Segmentation in Colorectal Cancer Detection
di: Mansoori, Mobina, et al.
Pubblicazione: (2024)
di: Mansoori, Mobina, et al.
Pubblicazione: (2024)
Surgical Vision World Model
di: Koju, Saurabh, et al.
Pubblicazione: (2025)
di: Koju, Saurabh, et al.
Pubblicazione: (2025)
Zero-Shot Adaptation for Approximate Posterior Sampling of Diffusion Models in Inverse Problems
di: Alçalar, Yaşar Utku, et al.
Pubblicazione: (2024)
di: Alçalar, Yaşar Utku, et al.
Pubblicazione: (2024)
One-Shot Image Restoration
di: Pereg, Deborah
Pubblicazione: (2024)
di: Pereg, Deborah
Pubblicazione: (2024)
AutoRad-Lung: A Radiomic-Guided Prompting Autoregressive Vision-Language Model for Lung Nodule Malignancy Prediction
di: Khademi, Sadaf, et al.
Pubblicazione: (2025)
di: Khademi, Sadaf, et al.
Pubblicazione: (2025)
Few-Shot Continual Learning for 3D Brain MRI with Frozen Foundation Models
di: Chen, Chi-Sheng, et al.
Pubblicazione: (2026)
di: Chen, Chi-Sheng, et al.
Pubblicazione: (2026)
SAM3D: Zero-Shot 3D Object Detection via Segment Anything Model
di: Zhang, Dingyuan, et al.
Pubblicazione: (2023)
di: Zhang, Dingyuan, et al.
Pubblicazione: (2023)
MSEG-VCUQ: Multimodal SEGmentation with Enhanced Vision Foundation Models, Convolutional Neural Networks, and Uncertainty Quantification for High-Speed Video Phase Detection Data
di: Maduabuchi, Chika, et al.
Pubblicazione: (2024)
di: Maduabuchi, Chika, et al.
Pubblicazione: (2024)
An Intrinsically Explainable Approach to Detecting Vertebral Compression Fractures in CT Scans via Neurosymbolic Modeling
di: Inigo, Blanca, et al.
Pubblicazione: (2024)
di: Inigo, Blanca, et al.
Pubblicazione: (2024)
Zero-TPrune: Zero-Shot Token Pruning through Leveraging of the Attention Graph in Pre-Trained Transformers
di: Wang, Hongjie, et al.
Pubblicazione: (2023)
di: Wang, Hongjie, et al.
Pubblicazione: (2023)
Multi-Modal Zero-Shot Prediction of Color Trajectories in Food Drying
di: Li, Shichen, et al.
Pubblicazione: (2025)
di: Li, Shichen, et al.
Pubblicazione: (2025)
Polyp and Surgical Instrument Segmentation with Double Encoder-Decoder Networks
di: Galdran, Adrian
Pubblicazione: (2024)
di: Galdran, Adrian
Pubblicazione: (2024)
Self-supervised Vision Transformer are Scalable Generative Models for Domain Generalization
di: Doerrich, Sebastian, et al.
Pubblicazione: (2024)
di: Doerrich, Sebastian, et al.
Pubblicazione: (2024)
BenchCloudVision: A Benchmark Analysis of Deep Learning Approaches for Cloud Detection and Segmentation in Remote Sensing Imagery
di: Fabio, Loddo, et al.
Pubblicazione: (2024)
di: Fabio, Loddo, et al.
Pubblicazione: (2024)
Towards a General-Purpose Zero-Shot Synthetic Low-Light Image and Video Pipeline
di: Lin, Joanne, et al.
Pubblicazione: (2025)
di: Lin, Joanne, et al.
Pubblicazione: (2025)
Separated Inter/Intra-Modal Fusion Prompts for Compositional Zero-Shot Learning
di: Jung, Sua
Pubblicazione: (2025)
di: Jung, Sua
Pubblicazione: (2025)
A Hybrid Wavelet-Fourier Method for Next-Generation Conditional Diffusion Models
di: Kiruluta, Andrew, et al.
Pubblicazione: (2025)
di: Kiruluta, Andrew, et al.
Pubblicazione: (2025)
SAMDA: Leveraging SAM on Few-Shot Domain Adaptation for Electronic Microscopy Segmentation
di: Wang, Yiran, et al.
Pubblicazione: (2024)
di: Wang, Yiran, et al.
Pubblicazione: (2024)
Multilabel Classification for Lung Disease Detection: Integrating Deep Learning and Natural Language Processing
di: Efimovich, Maria, et al.
Pubblicazione: (2024)
di: Efimovich, Maria, et al.
Pubblicazione: (2024)
Diffusion Models with Implicit Guidance for Medical Anomaly Detection
di: Bercea, Cosmin I., et al.
Pubblicazione: (2024)
di: Bercea, Cosmin I., et al.
Pubblicazione: (2024)
Computer Vision for Increased Operative Efficiency via Identification of Instruments in the Neurosurgical Operating Room: A Proof-of-Concept Study
di: Zachem, Tanner J., et al.
Pubblicazione: (2023)
di: Zachem, Tanner J., et al.
Pubblicazione: (2023)
GCS-M3VLT: Guided Context Self-Attention based Multi-modal Medical Vision Language Transformer for Retinal Image Captioning
di: Cherukuri, Teja Krishna, et al.
Pubblicazione: (2024)
di: Cherukuri, Teja Krishna, et al.
Pubblicazione: (2024)
Detection of Malaria Vector Breeding Habitats using Topographic Models
di: Jadhav, Aishwarya
Pubblicazione: (2020)
di: Jadhav, Aishwarya
Pubblicazione: (2020)
CLIP-Driven Universal Model for Organ Segmentation and Tumor Detection
di: Liu, Jie, et al.
Pubblicazione: (2023)
di: Liu, Jie, et al.
Pubblicazione: (2023)
Low-Trace Adaptation of Zero-shot Self-supervised Blind Image Denoising
di: Hu, Jintong, et al.
Pubblicazione: (2024)
di: Hu, Jintong, et al.
Pubblicazione: (2024)
Semantic Meta-Split Learning: A TinyML Scheme for Few-Shot Wireless Image Classification
di: Eldeeb, Eslam, et al.
Pubblicazione: (2024)
di: Eldeeb, Eslam, et al.
Pubblicazione: (2024)
Concrete Surface Crack Detection with Convolutional-based Deep Learning Models
di: Zadeh, Sara Shomal, et al.
Pubblicazione: (2024)
di: Zadeh, Sara Shomal, et al.
Pubblicazione: (2024)
Diffusion based Zero-shot Medical Image-to-Image Translation for Cross Modality Segmentation
di: Wang, Zihao, et al.
Pubblicazione: (2024)
di: Wang, Zihao, et al.
Pubblicazione: (2024)
GlaLSTM: A Concurrent LSTM Stream Framework for Glaucoma Detection via Biomarker Mining
di: Huang, Cheng, et al.
Pubblicazione: (2024)
di: Huang, Cheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
di: Javed, Sajid, et al.
Pubblicazione: (2024) -
Lung-CADex: Fully automatic Zero-Shot Detection and Classification of Lung Nodules in Thoracic CT Images
di: Shaukat, Furqan, et al.
Pubblicazione: (2024) -
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
di: Wang, Sicheng, et al.
Pubblicazione: (2024) -
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
di: Guizilini, Vitor, et al.
Pubblicazione: (2024) -
Robustly Optimized Deep Feature Decoupling Network for Fatty Liver Diseases Detection
di: Huang, Peng, et al.
Pubblicazione: (2024)