A Weighted Vision Transformer-Based Multi-Task Learning Framework for Predicting ADAS-Cog Scores
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hamid, Nur Amirah Abd, Shapiai, Mohd Ibrahim, Lai, Daphne Teck Ching |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Interpretability in Alzheimer's Prediction via Joint Learning of ADAS-Cog Scores
von: Hamid, Nur Amirah Abd, et al.
Veröffentlicht: (2025)
von: Hamid, Nur Amirah Abd, et al.
Veröffentlicht: (2025)
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning
von: Zhong, Hanwen, et al.
Veröffentlicht: (2025)
von: Zhong, Hanwen, et al.
Veröffentlicht: (2025)
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
von: Gwak, Chanyoung, et al.
Veröffentlicht: (2026)
von: Gwak, Chanyoung, et al.
Veröffentlicht: (2026)
An Investigation on The Position Encoding in Vision-Based Dynamics Prediction
von: Zhu, Jiageng, et al.
Veröffentlicht: (2024)
von: Zhu, Jiageng, et al.
Veröffentlicht: (2024)
SatelliteCalculator: A Multi-Task Vision Foundation Model for Quantitative Remote Sensing Inversion
von: Yu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Yu, Zhenyu, et al.
Veröffentlicht: (2025)
Road Segmentation for ADAS/AD Applications
von: Ramasamy, Mathanesh Vellingiri, et al.
Veröffentlicht: (2025)
von: Ramasamy, Mathanesh Vellingiri, et al.
Veröffentlicht: (2025)
HPE-CogVLM: Advancing Vision Language Models with a Head Pose Grounding Task
von: Tian, Yu, et al.
Veröffentlicht: (2024)
von: Tian, Yu, et al.
Veröffentlicht: (2024)
An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2024)
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2024)
ADAS-TO: A Large-Scale Multimodal Naturalistic Dataset and Empirical Characterization of Human Takeovers during ADAS Engagement
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Reasoning in Computer Vision: Taxonomy, Models, Tasks, and Methodologies
von: Sarkar, Ayushman, et al.
Veröffentlicht: (2025)
von: Sarkar, Ayushman, et al.
Veröffentlicht: (2025)
Inadequate contrast ratio of road markings as an indicator for ADAS failure
von: Certad, Novel, et al.
Veröffentlicht: (2024)
von: Certad, Novel, et al.
Veröffentlicht: (2024)
Analytical Uncertainty-Based Loss Weighting in Multi-Task Learning
von: Kirchdorfer, Lukas, et al.
Veröffentlicht: (2024)
von: Kirchdorfer, Lukas, et al.
Veröffentlicht: (2024)
Progressive Pretext Task Learning for Human Trajectory Prediction
von: Lin, Xiaotong, et al.
Veröffentlicht: (2024)
von: Lin, Xiaotong, et al.
Veröffentlicht: (2024)
SGW-based Multi-Task Learning in Vision Tasks
von: Zhang, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyuan, et al.
Veröffentlicht: (2024)
Towards SAR Automatic Target Recognition MultiCategory SAR Image Classification Based on Light Weight Vision Transformer
von: Zhao, Guibin, et al.
Veröffentlicht: (2024)
von: Zhao, Guibin, et al.
Veröffentlicht: (2024)
Generative Pre-training for Subjective Tasks: A Diffusion Transformer-Based Framework for Facial Beauty Prediction
von: Boukhari, Djamel Eddine, et al.
Veröffentlicht: (2025)
von: Boukhari, Djamel Eddine, et al.
Veröffentlicht: (2025)
Robust ADAS: Enhancing Robustness of Machine Learning-based Advanced Driver Assistance Systems for Adverse Weather
von: Shahzad, Muhammad Zaeem, et al.
Veröffentlicht: (2024)
von: Shahzad, Muhammad Zaeem, et al.
Veröffentlicht: (2024)
TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning
von: Baek, Seungmin, et al.
Veröffentlicht: (2025)
von: Baek, Seungmin, et al.
Veröffentlicht: (2025)
Vision Transformers for Preoperative CT-Based Prediction of Histopathologic Chemotherapy Response Score in High-Grade Serous Ovarian Carcinoma
von: Fati, Francesca, et al.
Veröffentlicht: (2026)
von: Fati, Francesca, et al.
Veröffentlicht: (2026)
Adaptive Multi Scale Document Binarisation Using Vision Mamba
von: Azfar, Mohd., et al.
Veröffentlicht: (2024)
von: Azfar, Mohd., et al.
Veröffentlicht: (2024)
Efficient Domain-Adaptive Multi-Task Dense Prediction with Vision Foundation Models
von: Kang, Beomseok, et al.
Veröffentlicht: (2025)
von: Kang, Beomseok, et al.
Veröffentlicht: (2025)
Task Indicating Transformer for Task-conditional Dense Predictions
von: Lu, Yuxiang, et al.
Veröffentlicht: (2024)
von: Lu, Yuxiang, et al.
Veröffentlicht: (2024)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
von: Ma, Yunsheng, et al.
Veröffentlicht: (2022)
von: Ma, Yunsheng, et al.
Veröffentlicht: (2022)
VLMDiff: Leveraging Vision-Language Models for Multi-Class Anomaly Detection with Diffusion
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2025)
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2025)
Patch Pruning Strategy Based on Robust Statistical Measures of Attention Weight Diversity in Vision Transformers
von: Igaue, Yuki, et al.
Veröffentlicht: (2025)
von: Igaue, Yuki, et al.
Veröffentlicht: (2025)
Efficient Cross-Country Data Acquisition Strategy for ADAS via Street-View Imagery
von: Wu, Yin, et al.
Veröffentlicht: (2026)
von: Wu, Yin, et al.
Veröffentlicht: (2026)
Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
von: Singh, Raghavendra
Veröffentlicht: (2024)
von: Singh, Raghavendra
Veröffentlicht: (2024)
CogDoc: Towards Unified thinking in Documents
von: Xu, Qixin, et al.
Veröffentlicht: (2025)
von: Xu, Qixin, et al.
Veröffentlicht: (2025)
Multidimensional Task Learning: A Unified Tensor Framework for Computer Vision Tasks
von: Ichi, Alaa El, et al.
Veröffentlicht: (2026)
von: Ichi, Alaa El, et al.
Veröffentlicht: (2026)
Proximal Vision Transformer: Enhancing Feature Representation through Two-Stage Manifold Geometry
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
Deep Learning Based Wildfire Detection for Peatland Fires Using Transfer Learning
von: Hamdan, Emadeldeen, et al.
Veröffentlicht: (2026)
von: Hamdan, Emadeldeen, et al.
Veröffentlicht: (2026)
Safety-Critical Camera Reliability Monitoring for ADAS via Degradation-Aware Uncertainty Pattern Analysis
von: Aher, Shiva
Veröffentlicht: (2026)
von: Aher, Shiva
Veröffentlicht: (2026)
Feature Learning with Multi-Stage Vision Transformers on Inter-Modality HER2 Status Scoring and Tumor Classification on Whole Slides
von: Oyelade, Olaide N., et al.
Veröffentlicht: (2025)
von: Oyelade, Olaide N., et al.
Veröffentlicht: (2025)
CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
On Convolutional Vision Transformers for Yield Prediction
von: Inderka, Alvin, et al.
Veröffentlicht: (2024)
von: Inderka, Alvin, et al.
Veröffentlicht: (2024)
CogVLM: Visual Expert for Pretrained Language Models
von: Wang, Weihan, et al.
Veröffentlicht: (2023)
von: Wang, Weihan, et al.
Veröffentlicht: (2023)
EULER-ADAS: Energy-Efficient & SIMD-Unified Logarithmic-Posit Engine for Precision-Reconfigurable Approximate ADAS Acceleration
von: Lokhande, Mukul, et al.
Veröffentlicht: (2026)
von: Lokhande, Mukul, et al.
Veröffentlicht: (2026)
A Vanilla Multi-Task Framework for Dense Visual Prediction Solution to 1st VCL Challenge -- Multi-Task Robustness Track
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Improving Interpretability in Alzheimer's Prediction via Joint Learning of ADAS-Cog Scores
von: Hamid, Nur Amirah Abd, et al.
Veröffentlicht: (2025) -
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning
von: Zhong, Hanwen, et al.
Veröffentlicht: (2025) -
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
von: Gwak, Chanyoung, et al.
Veröffentlicht: (2026) -
An Investigation on The Position Encoding in Vision-Based Dynamics Prediction
von: Zhu, Jiageng, et al.
Veröffentlicht: (2024) -
SatelliteCalculator: A Multi-Task Vision Foundation Model for Quantitative Remote Sensing Inversion
von: Yu, Zhenyu, et al.
Veröffentlicht: (2025)