DenseMTL: Cross-task Attention Mechanism for Dense Multi-task Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Lopes, Ivan, Vu, Tuan-Hung, de Charette, Raoul |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets
by: Cao, Anh-Quan, et al.
Published: (2025)
by: Cao, Anh-Quan, et al.
Published: (2025)
AeroLite-MDNet: Lightweight Multi-task Deviation Detection Network for UAV Landing
by: Yang, Haiping, et al.
Published: (2025)
by: Yang, Haiping, et al.
Published: (2025)
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
by: Cai, Xiongyi, et al.
Published: (2025)
by: Cai, Xiongyi, et al.
Published: (2025)
CMI-MTL: Cross-Mamba interaction based multi-task learning for medical visual question answering
by: Jin, Qiangguo, et al.
Published: (2025)
by: Jin, Qiangguo, et al.
Published: (2025)
SGS-SLAM: Semantic Gaussian Splatting For Neural Dense SLAM
by: Li, Mingrui, et al.
Published: (2024)
by: Li, Mingrui, et al.
Published: (2024)
MANSION: Multi-floor lANguage-to-3D Scene generatIOn for loNg-horizon tasks
by: Che, Lirong, et al.
Published: (2026)
by: Che, Lirong, et al.
Published: (2026)
IDD-X: A Multi-View Dataset for Ego-relative Important Object Localization and Explanation in Dense and Unstructured Traffic
by: Parikh, Chirag, et al.
Published: (2024)
by: Parikh, Chirag, et al.
Published: (2024)
Cycle-Correspondence Loss: Learning Dense View-Invariant Visual Features from Unlabeled and Unordered RGB Images
by: Adrian, David B., et al.
Published: (2024)
by: Adrian, David B., et al.
Published: (2024)
DePT3R: Joint Dense Point Tracking and 3D Reconstruction of Dynamic Scenes in a Single Forward Pass
by: Alumootil, Vivek, et al.
Published: (2025)
by: Alumootil, Vivek, et al.
Published: (2025)
SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM
by: Keetha, Nikhil, et al.
Published: (2023)
by: Keetha, Nikhil, et al.
Published: (2023)
StruMPL: Multi-task Dense Regression under Disjoint Partial Supervision and MNAR Labels
by: Asiyabi, Reza M., et al.
Published: (2026)
by: Asiyabi, Reza M., et al.
Published: (2026)
Dense Depth from Event Focal Stack
by: Horikawa, Kenta, et al.
Published: (2024)
by: Horikawa, Kenta, et al.
Published: (2024)
FGS-SLAM: Fourier-based Gaussian Splatting for Real-time SLAM with Sparse and Dense Map Fusion
by: Xu, Yansong, et al.
Published: (2025)
by: Xu, Yansong, et al.
Published: (2025)
LoD-Loc v3: Generalized Aerial Localization in Dense Cities using Instance Silhouette Alignment
by: Peng, Shuaibang, et al.
Published: (2026)
by: Peng, Shuaibang, et al.
Published: (2026)
PaSCo: Urban 3D Panoptic Scene Completion with Uncertainty Awareness
by: Cao, Anh-Quan, et al.
Published: (2023)
by: Cao, Anh-Quan, et al.
Published: (2023)
Non-rigid Relative Placement through 3D Dense Diffusion
by: Cai, Eric, et al.
Published: (2024)
by: Cai, Eric, et al.
Published: (2024)
Privacy-Preserving Multi-Stage Fall Detection Framework with Semi-supervised Federated Learning and Robotic Vision Confirmation
by: Azghadi, Seyed Alireza Rahimi, et al.
Published: (2025)
by: Azghadi, Seyed Alireza Rahimi, et al.
Published: (2025)
Frequency-Dynamic Attention Modulation for Dense Prediction
by: Chen, Linwei, et al.
Published: (2025)
by: Chen, Linwei, et al.
Published: (2025)
LiDPM: Rethinking Point Diffusion for Lidar Scene Completion
by: Martyniuk, Tetiana, et al.
Published: (2025)
by: Martyniuk, Tetiana, et al.
Published: (2025)
CLIP's Visual Embedding Projector is a Few-shot Cornucopia
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
A Simple Recipe for Language-guided Domain Generalized Segmentation
by: Fahes, Mohammad, et al.
Published: (2023)
by: Fahes, Mohammad, et al.
Published: (2023)
AffordTissue: Dense Affordance Prediction for Tool-Action Specific Tissue Interaction
by: Maksutova, Aiza, et al.
Published: (2026)
by: Maksutova, Aiza, et al.
Published: (2026)
UMBRAE: Unified Multimodal Brain Decoding
by: Xia, Weihao, et al.
Published: (2024)
by: Xia, Weihao, et al.
Published: (2024)
V3D-SLAM: Robust RGB-D SLAM in Dynamic Environments with 3D Semantic Geometry Voting
by: Dang, Tuan, et al.
Published: (2024)
by: Dang, Tuan, et al.
Published: (2024)
VaViM and VaVAM: Autonomous Driving through Video Generative Modeling
by: Bartoccioni, Florent, et al.
Published: (2025)
by: Bartoccioni, Florent, et al.
Published: (2025)
Context-aware Multi-task Learning for Pedestrian Intent and Trajectory Prediction
by: Munir, Farzeen, et al.
Published: (2024)
by: Munir, Farzeen, et al.
Published: (2024)
CSAOT: Cooperative Multi-Agent System for Active Object Tracking
by: Nguyen, Hy, et al.
Published: (2025)
by: Nguyen, Hy, et al.
Published: (2025)
Multi-task Cross-modal Learning for Chest X-ray Image Retrieval
by: Liang, Zhaohui, et al.
Published: (2026)
by: Liang, Zhaohui, et al.
Published: (2026)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
M2DA: Multi-Modal Fusion Transformer Incorporating Driver Attention for Autonomous Driving
by: Xu, Dongyang, et al.
Published: (2024)
by: Xu, Dongyang, et al.
Published: (2024)
VISTA: A Vision and Intent-Aware Social Attention Framework for Multi-Agent Trajectory Prediction
by: Martins, Stephane Da Silva, et al.
Published: (2025)
by: Martins, Stephane Da Silva, et al.
Published: (2025)
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
by: Zhou, Jiaming, et al.
Published: (2025)
by: Zhou, Jiaming, et al.
Published: (2025)
Parameter Aware Mamba Model for Multi-task Dense Prediction
by: Yu, Xinzhuo, et al.
Published: (2025)
by: Yu, Xinzhuo, et al.
Published: (2025)
Bridging Cross-task Protocol Inconsistency for Distillation in Dense Object Detection
by: Yang, Longrong, et al.
Published: (2023)
by: Yang, Longrong, et al.
Published: (2023)
DualCross: Cross-Modality Cross-Domain Adaptation for Monocular BEV Perception
by: Man, Yunze, et al.
Published: (2023)
by: Man, Yunze, et al.
Published: (2023)
FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2025)
by: Benigmim, Yasser, et al.
Published: (2025)
Domain Adaptation with a Single Vision-Language Embedding
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
DenseFormer: Learning Dense Depth Map from Sparse Depth and Image via Conditional Diffusion Model
by: Yuan, Ming, et al.
Published: (2025)
by: Yuan, Ming, et al.
Published: (2025)
Dense Connector for MLLMs
by: Yao, Huanjin, et al.
Published: (2024)
by: Yao, Huanjin, et al.
Published: (2024)
Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints
by: Dai, Ming, et al.
Published: (2025)
by: Dai, Ming, et al.
Published: (2025)
Similar Items
-
StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets
by: Cao, Anh-Quan, et al.
Published: (2025) -
AeroLite-MDNet: Lightweight Multi-task Deviation Detection Network for UAV Landing
by: Yang, Haiping, et al.
Published: (2025) -
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
by: Cai, Xiongyi, et al.
Published: (2025) -
CMI-MTL: Cross-Mamba interaction based multi-task learning for medical visual question answering
by: Jin, Qiangguo, et al.
Published: (2025) -
SGS-SLAM: Semantic Gaussian Splatting For Neural Dense SLAM
by: Li, Mingrui, et al.
Published: (2024)