Efficient Inter-Task Attention for Multitask Transformer Models
Fuente:
arXiv
Saved in:
| Main Authors: | Bohn, Christian, Kurbiel, Thomas, Friedrichs, Klaus, Tercan, Hasan, Meisen, Tobias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Faster Training, Fewer Labels: Self-Supervised Pretraining for Fine-Grained BEV Segmentation
by: Busch, Daniel, et al.
Published: (2026)
by: Busch, Daniel, et al.
Published: (2026)
Rethinking Backbone Design for Lightweight 3D Object Detection in LiDAR
by: Chandorkar, Adwait, et al.
Published: (2025)
by: Chandorkar, Adwait, et al.
Published: (2025)
Graph Query Networks for Object Detection with Automotive Radar
by: Saini, Loveneet, et al.
Published: (2025)
by: Saini, Loveneet, et al.
Published: (2025)
AttentiveGRU: Recurrent Spatio-Temporal Modeling for Advanced Radar-Based BEV Object Detection
by: Saini, Loveneet, et al.
Published: (2025)
by: Saini, Loveneet, et al.
Published: (2025)
Detection Transformers Under the Knife: A Neuroscience-Inspired Approach to Ablations
by: Hütten, Nils, et al.
Published: (2025)
by: Hütten, Nils, et al.
Published: (2025)
Out-of-Distribution Object Detection in Street Scenes via Synthetic Outlier Exposure and Transfer Learning
by: Ilyas, Sadia, et al.
Published: (2026)
by: Ilyas, Sadia, et al.
Published: (2026)
Task Weighting through Gradient Projection for Multitask Learning
by: Bohn, Christian, et al.
Published: (2024)
by: Bohn, Christian, et al.
Published: (2024)
CASPFormer: Trajectory Prediction from BEV Images with Deformable Attention
by: Yadav, Harsh, et al.
Published: (2024)
by: Yadav, Harsh, et al.
Published: (2024)
Principal Component Clustering for Semantic Segmentation in Synthetic Data Generation
by: Stillger, Felix, et al.
Published: (2024)
by: Stillger, Felix, et al.
Published: (2024)
LMFormer: Lane based Motion Prediction Transformer
by: Yadav, Harsh, et al.
Published: (2025)
by: Yadav, Harsh, et al.
Published: (2025)
Improved Single Camera BEV Perception Using Multi-Camera Training
by: Busch, Daniel, et al.
Published: (2024)
by: Busch, Daniel, et al.
Published: (2024)
InCaRPose: In-Cabin Relative Camera Pose Estimation Model and Dataset
by: Stillger, Felix, et al.
Published: (2026)
by: Stillger, Felix, et al.
Published: (2026)
Inferring Driving Maps by Deep Learning-based Trail Map Extraction
by: Hubbertz, Michael, et al.
Published: (2025)
by: Hubbertz, Michael, et al.
Published: (2025)
Video Super-Resolution Transformer with Masked Inter&Intra-Frame Attention
by: Zhou, Xingyu, et al.
Published: (2024)
by: Zhou, Xingyu, et al.
Published: (2024)
Group Diffusion Transformers are Unsupervised Multitask Learners
by: Huang, Lianghua, et al.
Published: (2024)
by: Huang, Lianghua, et al.
Published: (2024)
Failure Modes for Deep Learning-Based Online Mapping: How to Measure and Address Them
by: Hubbertz, Michael, et al.
Published: (2026)
by: Hubbertz, Michael, et al.
Published: (2026)
Efficient Multitask Dense Predictor via Binarization
by: Shang, Yuzhang, et al.
Published: (2024)
by: Shang, Yuzhang, et al.
Published: (2024)
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
by: Huang, You, et al.
Published: (2025)
by: Huang, You, et al.
Published: (2025)
Transformer-based Multimodal Change Detection with Multitask Consistency Constraints
by: Liu, Biyuan, et al.
Published: (2023)
by: Liu, Biyuan, et al.
Published: (2023)
Fine-Tuning Video Transformers for Word-Level Bangla Sign Language: A Comparative Analysis for Classification Tasks
by: Shawon, Jubayer Ahmed Bhuiyan, et al.
Published: (2025)
by: Shawon, Jubayer Ahmed Bhuiyan, et al.
Published: (2025)
Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers
by: Wang, Chung-Shien Brian, et al.
Published: (2025)
by: Wang, Chung-Shien Brian, et al.
Published: (2025)
Beyond Attention Magnitude: Leveraging Inter-layer Rank Consistency for Efficient Vision-Language-Action Models
by: Liu, Peiju, et al.
Published: (2026)
by: Liu, Peiju, et al.
Published: (2026)
InterACT: Inter-dependency Aware Action Chunking with Hierarchical Attention Transformers for Bimanual Manipulation
by: Lee, Andrew, et al.
Published: (2024)
by: Lee, Andrew, et al.
Published: (2024)
Transformer based Multitask Learning for Image Captioning and Object Detection
by: Basak, Debolena, et al.
Published: (2024)
by: Basak, Debolena, et al.
Published: (2024)
Attention Is not Everything: Efficient Alternatives for Vision
by: Kazi, Nur Mohammad, et al.
Published: (2026)
by: Kazi, Nur Mohammad, et al.
Published: (2026)
TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning
by: Baek, Seungmin, et al.
Published: (2025)
by: Baek, Seungmin, et al.
Published: (2025)
Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile
by: Ding, Hangliang, et al.
Published: (2025)
by: Ding, Hangliang, et al.
Published: (2025)
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning
by: Zhong, Hanwen, et al.
Published: (2025)
by: Zhong, Hanwen, et al.
Published: (2025)
Cross-Task Affinity Learning for Multitask Dense Scene Predictions
by: Sinodinos, Dimitrios, et al.
Published: (2024)
by: Sinodinos, Dimitrios, et al.
Published: (2024)
Stack Transformer Based Spatial-Temporal Attention Model for Dynamic Sign Language and Fingerspelling Recognition
by: Hirooka, Koki, et al.
Published: (2025)
by: Hirooka, Koki, et al.
Published: (2025)
Exploring the Integration of Key-Value Attention Into Pure and Hybrid Transformers for Semantic Segmentation
by: Hwa, DeShin, et al.
Published: (2025)
by: Hwa, DeShin, et al.
Published: (2025)
Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers
by: Zhou, Yifan, et al.
Published: (2025)
by: Zhou, Yifan, et al.
Published: (2025)
Artifacts and Attention Sinks: Structured Approximations for Efficient Vision Transformers
by: Lu, Andrew, et al.
Published: (2025)
by: Lu, Andrew, et al.
Published: (2025)
S2AFormer: Strip Self-Attention for Efficient Vision Transformer
by: Xu, Guoan, et al.
Published: (2025)
by: Xu, Guoan, et al.
Published: (2025)
ToSA: Token Selective Attention for Efficient Vision Transformers
by: Singh, Manish Kumar, et al.
Published: (2024)
by: Singh, Manish Kumar, et al.
Published: (2024)
Efficient Diffusion Transformer with Step-wise Dynamic Attention Mediators
by: Pu, Yifan, et al.
Published: (2024)
by: Pu, Yifan, et al.
Published: (2024)
PISA: Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers
by: Li, Haopeng, et al.
Published: (2026)
by: Li, Haopeng, et al.
Published: (2026)
ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention
by: He, Chenhang, et al.
Published: (2024)
by: He, Chenhang, et al.
Published: (2024)
RadVLM: A Multitask Conversational Vision-Language Model for Radiology
by: Deperrois, Nicolas, et al.
Published: (2025)
by: Deperrois, Nicolas, et al.
Published: (2025)
Edge Detection based on Channel Attention and Inter-region Independence Test
by: Yan, Ru-yu, et al.
Published: (2025)
by: Yan, Ru-yu, et al.
Published: (2025)
Similar Items
-
Faster Training, Fewer Labels: Self-Supervised Pretraining for Fine-Grained BEV Segmentation
by: Busch, Daniel, et al.
Published: (2026) -
Rethinking Backbone Design for Lightweight 3D Object Detection in LiDAR
by: Chandorkar, Adwait, et al.
Published: (2025) -
Graph Query Networks for Object Detection with Automotive Radar
by: Saini, Loveneet, et al.
Published: (2025) -
AttentiveGRU: Recurrent Spatio-Temporal Modeling for Advanced Radar-Based BEV Object Detection
by: Saini, Loveneet, et al.
Published: (2025) -
Detection Transformers Under the Knife: A Neuroscience-Inspired Approach to Ablations
by: Hütten, Nils, et al.
Published: (2025)