ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yunsheng, Wang, Ziran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026)
von: Wang, Feng, et al.
Veröffentlicht: (2026)
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
von: Ma, Xiaochen, et al.
Veröffentlicht: (2023)
von: Ma, Xiaochen, et al.
Veröffentlicht: (2023)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
Spiking-DD: Neuromorphic Event Camera based Driver Distraction Detection with Spiking Neural Network
von: Shariff, Waseem, et al.
Veröffentlicht: (2024)
von: Shariff, Waseem, et al.
Veröffentlicht: (2024)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
von: Ibtehaz, Nabil, et al.
Veröffentlicht: (2024)
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions
von: Xia, Chunlong, et al.
Veröffentlicht: (2024)
von: Xia, Chunlong, et al.
Veröffentlicht: (2024)
ViT-Explainer: An Interactive Walkthrough of the Vision Transformer Pipeline
von: Hernandez, Juan Manuel, et al.
Veröffentlicht: (2026)
von: Hernandez, Juan Manuel, et al.
Veröffentlicht: (2026)
ViT-FIQA: Assessing Face Image Quality using Vision Transformers
von: Atzori, Andrea, et al.
Veröffentlicht: (2025)
von: Atzori, Andrea, et al.
Veröffentlicht: (2025)
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer
von: Tai, Yu-Shan, et al.
Veröffentlicht: (2024)
von: Tai, Yu-Shan, et al.
Veröffentlicht: (2024)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
ViT-1.58b: Mobile Vision Transformers in the 1-bit Era
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2024)
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2024)
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024)
von: Yao, Ting, et al.
Veröffentlicht: (2024)
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
ADFQ-ViT: Activation-Distribution-Friendly Post-Training Quantization for Vision Transformers
von: Jiang, Yanfeng, et al.
Veröffentlicht: (2024)
von: Jiang, Yanfeng, et al.
Veröffentlicht: (2024)
Hyb-KAN ViT: Hybrid Kolmogorov-Arnold Networks Augmented Vision Transformer
von: Dey, Sainath, et al.
Veröffentlicht: (2025)
von: Dey, Sainath, et al.
Veröffentlicht: (2025)
SVD-ViT: Does SVD Make Vision Transformers Attend More to the Foreground?
von: Murata, Haruhiko, et al.
Veröffentlicht: (2026)
von: Murata, Haruhiko, et al.
Veröffentlicht: (2026)
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
Which Direction to Choose? An Analysis on the Representation Power of Self-Supervised ViTs in Downstream Tasks
von: Kaltampanidis, Yannis, et al.
Veröffentlicht: (2025)
von: Kaltampanidis, Yannis, et al.
Veröffentlicht: (2025)
When CNN Meet with ViT: Towards Semi-Supervised Learning for Multi-Class Medical Image Semantic Segmentation
von: Wang, Ziyang, et al.
Veröffentlicht: (2022)
von: Wang, Ziyang, et al.
Veröffentlicht: (2022)
APHQ-ViT: Post-Training Quantization with Average Perturbation Hessian Based Reconstruction for Vision Transformers
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
CAS-ViT: Convolutional Additive Self-attention Vision Transformers for Efficient Mobile Applications
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer
von: Elharrouss, Omar, et al.
Veröffentlicht: (2025)
von: Elharrouss, Omar, et al.
Veröffentlicht: (2025)
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
von: Ancey, Pierre, et al.
Veröffentlicht: (2025)
von: Ancey, Pierre, et al.
Veröffentlicht: (2025)
ViT-VS: On the Applicability of Pretrained Vision Transformer Features for Generalizable Visual Servoing
von: Scherl, Alessandro, et al.
Veröffentlicht: (2025)
von: Scherl, Alessandro, et al.
Veröffentlicht: (2025)
SAC-ViT: Semantic-Aware Clustering Vision Transformer with Early Exit
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
GTP-ViT: Efficient Vision Transformers via Graph-based Token Propagation
von: Xu, Xuwei, et al.
Veröffentlicht: (2023)
von: Xu, Xuwei, et al.
Veröffentlicht: (2023)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
DeNAS-ViT: Data Efficient NAS-Optimized Vision Transformer for Ultrasound Image Segmentation
von: Chen, Renqi, et al.
Veröffentlicht: (2024)
von: Chen, Renqi, et al.
Veröffentlicht: (2024)
PaW-ViT: A Patch-based Warping Vision Transformer for Robust Ear Verification
von: Arun, Deeksha, et al.
Veröffentlicht: (2026)
von: Arun, Deeksha, et al.
Veröffentlicht: (2026)
ViT-EnsembleAttack: Augmenting Ensemble Models for Stronger Adversarial Transferability in Vision Transformers
von: Cao, Hanwen, et al.
Veröffentlicht: (2025)
von: Cao, Hanwen, et al.
Veröffentlicht: (2025)
Trio-ViT: Post-Training Quantization and Acceleration for Softmax-Free Efficient Vision Transformer
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
DFQ-ViT: Data-Free Quantization for Vision Transformers without Fine-tuning
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026) -
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025) -
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025) -
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
von: Ma, Xiaochen, et al.
Veröffentlicht: (2023) -
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
von: Li, Wenhao, et al.
Veröffentlicht: (2025)