Transformer-based Multimodal Change Detection with Multitask Consistency Constraints
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Biyuan, Chen, Huaixin, Li, Kun, Yang, Michael Ying |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Change-Aware Siamese Network for Surface Defects Segmentation under Complex Background
di: Liu, Biyuan, et al.
Pubblicazione: (2024)
di: Liu, Biyuan, et al.
Pubblicazione: (2024)
Transformer based Multitask Learning for Image Captioning and Object Detection
di: Basak, Debolena, et al.
Pubblicazione: (2024)
di: Basak, Debolena, et al.
Pubblicazione: (2024)
Multimodal Rationales for Explainable Visual Question Answering
di: Li, Kun, et al.
Pubblicazione: (2024)
di: Li, Kun, et al.
Pubblicazione: (2024)
Group Diffusion Transformers are Unsupervised Multitask Learners
di: Huang, Lianghua, et al.
Pubblicazione: (2024)
di: Huang, Lianghua, et al.
Pubblicazione: (2024)
EgoM2P: Egocentric Multimodal Multitask Pretraining
di: Li, Gen, et al.
Pubblicazione: (2025)
di: Li, Gen, et al.
Pubblicazione: (2025)
Multimodal Fake News Detection: MFND Dataset and Shallow-Deep Multitask Learning
di: Zhu, Ye, et al.
Pubblicazione: (2025)
di: Zhu, Ye, et al.
Pubblicazione: (2025)
MMSF: Multitask and Multimodal Supervised Framework for WSI Classification and Survival Analysis
di: She, Chengying, et al.
Pubblicazione: (2026)
di: She, Chengying, et al.
Pubblicazione: (2026)
M$^3$GPT: An Advanced Multimodal, Multitask Framework for Motion Comprehension and Generation
di: Luo, Mingshuang, et al.
Pubblicazione: (2024)
di: Luo, Mingshuang, et al.
Pubblicazione: (2024)
OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning
di: Srivastava, Siddharth, et al.
Pubblicazione: (2025)
di: Srivastava, Siddharth, et al.
Pubblicazione: (2025)
A Survey on fMRI-based Brain Decoding for Reconstructing Multimodal Stimuli
di: Liu, Pengyu, et al.
Pubblicazione: (2025)
di: Liu, Pengyu, et al.
Pubblicazione: (2025)
Consistency Change Detection Framework for Unsupervised Remote Sensing Change Detection
di: Liu, Yating, et al.
Pubblicazione: (2025)
di: Liu, Yating, et al.
Pubblicazione: (2025)
MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
di: Ying, Kaining, et al.
Pubblicazione: (2024)
di: Ying, Kaining, et al.
Pubblicazione: (2024)
Efficient Inter-Task Attention for Multitask Transformer Models
di: Bohn, Christian, et al.
Pubblicazione: (2025)
di: Bohn, Christian, et al.
Pubblicazione: (2025)
CloudMatch: Weak-to-Strong Consistency Learning for Semi-Supervised Cloud Detection
di: Zhao, Jiayi, et al.
Pubblicazione: (2026)
di: Zhao, Jiayi, et al.
Pubblicazione: (2026)
TaCo: Capturing Spatio-Temporal Semantic Consistency in Remote Sensing Change Detection
di: Guo, Han, et al.
Pubblicazione: (2025)
di: Guo, Han, et al.
Pubblicazione: (2025)
Weakly Supervised Multimodal Temporal Forgery Localization via Multitask Learning
di: Xu, Wenbo, et al.
Pubblicazione: (2025)
di: Xu, Wenbo, et al.
Pubblicazione: (2025)
CoRegOVCD: Consistency-Regularized Open-Vocabulary Change Detection
di: Tang, Weidong, et al.
Pubblicazione: (2026)
di: Tang, Weidong, et al.
Pubblicazione: (2026)
SChanger: Change Detection from a Semantic Change and Spatial Consistency Perspective
di: Zhou, Ziyu, et al.
Pubblicazione: (2025)
di: Zhou, Ziyu, et al.
Pubblicazione: (2025)
Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks
di: Lee, Jusung, et al.
Pubblicazione: (2024)
di: Lee, Jusung, et al.
Pubblicazione: (2024)
Prior-guided Fusion of Multimodal Features for Change Detection from Optical-SAR Images
di: Liu, Xuanguang, et al.
Pubblicazione: (2026)
di: Liu, Xuanguang, et al.
Pubblicazione: (2026)
Cross Branch Feature Fusion Decoder for Consistency Regularization-based Semi-Supervised Change Detection
di: Xing, Yan, et al.
Pubblicazione: (2024)
di: Xing, Yan, et al.
Pubblicazione: (2024)
UniChange: Unifying Change Detection with Multimodal Large Language Model
di: Zhang, Xu, et al.
Pubblicazione: (2025)
di: Zhang, Xu, et al.
Pubblicazione: (2025)
HSACNet: Hierarchical Scale-Aware Consistency Regularized Semi-Supervised Change Detection
di: Xu, Qi'ao, et al.
Pubblicazione: (2025)
di: Xu, Qi'ao, et al.
Pubblicazione: (2025)
Learning Streaming Video Representation via Multitask Training
di: Yan, Yibin, et al.
Pubblicazione: (2025)
di: Yan, Yibin, et al.
Pubblicazione: (2025)
Document Image Rectification Bases on Self-Adaptive Multitask Fusion
di: Li, Heng, et al.
Pubblicazione: (2025)
di: Li, Heng, et al.
Pubblicazione: (2025)
Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation
di: Li, Yiheng, et al.
Pubblicazione: (2025)
di: Li, Yiheng, et al.
Pubblicazione: (2025)
GTPC-SSCD: Gate-guided Two-level Perturbation Consistency-based Semi-Supervised Change Detection
di: Xing, Yan, et al.
Pubblicazione: (2024)
di: Xing, Yan, et al.
Pubblicazione: (2024)
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
di: Jiang, Feibo, et al.
Pubblicazione: (2026)
di: Jiang, Feibo, et al.
Pubblicazione: (2026)
RoadscapesQA: A Multitask, Multimodal Dataset for Visual Question Answering on Indian Roads
di: Iyer, Vijayasri, et al.
Pubblicazione: (2026)
di: Iyer, Vijayasri, et al.
Pubblicazione: (2026)
Scale-wise Bidirectional Alignment Network for Referring Remote Sensing Image Segmentation
di: Li, Kun, et al.
Pubblicazione: (2025)
di: Li, Kun, et al.
Pubblicazione: (2025)
ChangeViT: Unleashing Plain Vision Transformers for Change Detection
di: Zhu, Duowang, et al.
Pubblicazione: (2024)
di: Zhu, Duowang, et al.
Pubblicazione: (2024)
Test-Time Intensity Consistency Adaptation for Shadow Detection
di: Zhu, Leyi, et al.
Pubblicazione: (2024)
di: Zhu, Leyi, et al.
Pubblicazione: (2024)
Multitask Learning for SAR Ship Detection with Gaussian-Mask Joint Segmentation
di: Zhao, Ming, et al.
Pubblicazione: (2024)
di: Zhao, Ming, et al.
Pubblicazione: (2024)
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling
di: Wang, Yu, et al.
Pubblicazione: (2022)
di: Wang, Yu, et al.
Pubblicazione: (2022)
VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection
di: Han, Jianhong, et al.
Pubblicazione: (2025)
di: Han, Jianhong, et al.
Pubblicazione: (2025)
Efficient Multitask Dense Predictor via Binarization
di: Shang, Yuzhang, et al.
Pubblicazione: (2024)
di: Shang, Yuzhang, et al.
Pubblicazione: (2024)
ConsistCompose: Unified Multimodal Layout Control for Image Composition
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer
di: Yu, Ning, et al.
Pubblicazione: (2022)
di: Yu, Ning, et al.
Pubblicazione: (2022)
Training-Free Multimodal Deepfake Detection via Graph Reasoning
di: Liu, Yuxin, et al.
Pubblicazione: (2025)
di: Liu, Yuxin, et al.
Pubblicazione: (2025)
Query-Guided Spatial-Temporal-Frequency Interaction for Music Audio-Visual Question Answering
di: Li, Kun, et al.
Pubblicazione: (2026)
di: Li, Kun, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Change-Aware Siamese Network for Surface Defects Segmentation under Complex Background
di: Liu, Biyuan, et al.
Pubblicazione: (2024) -
Transformer based Multitask Learning for Image Captioning and Object Detection
di: Basak, Debolena, et al.
Pubblicazione: (2024) -
Multimodal Rationales for Explainable Visual Question Answering
di: Li, Kun, et al.
Pubblicazione: (2024) -
Group Diffusion Transformers are Unsupervised Multitask Learners
di: Huang, Lianghua, et al.
Pubblicazione: (2024) -
EgoM2P: Egocentric Multimodal Multitask Pretraining
di: Li, Gen, et al.
Pubblicazione: (2025)