Task-Generalized Adaptive Cross-Domain Learning for Multimodal Image Fusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Mengyu, Liu, Zhenyu, Li, Kun, Wang, Yu, Wang, Yuwei, Wei, Yanyan, Wang, Fei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Exploiting Ensemble Learning for Cross-View Isolated Sign Language Recognition
von: Wang, Fei, et al.
Veröffentlicht: (2025)
von: Wang, Fei, et al.
Veröffentlicht: (2025)
Towards Generalizable Deepfake Detection with Spatial-Frequency Collaborative Learning and Hierarchical Cross-Modal Fusion
von: Qiao, Mengyu, et al.
Veröffentlicht: (2025)
von: Qiao, Mengyu, et al.
Veröffentlicht: (2025)
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
von: Liu, Yuxin, et al.
Veröffentlicht: (2026)
von: Liu, Yuxin, et al.
Veröffentlicht: (2026)
GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Cross-Sample Relational Fusion: Unifying Domain Generalization and Class-Incremental Learning
von: Xie, Zhen-Hao, et al.
Veröffentlicht: (2026)
von: Xie, Zhen-Hao, et al.
Veröffentlicht: (2026)
TRACE: Task-Adaptive Reasoning and Representation Learning for Universal Multimodal Retrieval
von: Hao, Xiangzhao, et al.
Veröffentlicht: (2026)
von: Hao, Xiangzhao, et al.
Veröffentlicht: (2026)
Enhancing Adaptive Deep Networks for Image Classification via Uncertainty-aware Decision Fusion
von: Zhang, Xu, et al.
Veröffentlicht: (2024)
von: Zhang, Xu, et al.
Veröffentlicht: (2024)
MoVL:Exploring Fusion Strategies for the Domain-Adaptive Application of Pretrained Models in Medical Imaging Tasks
von: Tian, Haijiang, et al.
Veröffentlicht: (2024)
von: Tian, Haijiang, et al.
Veröffentlicht: (2024)
Entity-Guided Multi-Task Learning for Infrared and Visible Image Fusion
von: Shao, Wenyu, et al.
Veröffentlicht: (2026)
von: Shao, Wenyu, et al.
Veröffentlicht: (2026)
Exploring Task-Solving Paradigm for Generalized Cross-Domain Face Anti-Spoofing via Reinforcement Fine-Tuning
von: Jiang, Fangling, et al.
Veröffentlicht: (2025)
von: Jiang, Fangling, et al.
Veröffentlicht: (2025)
Adaptive Domain Shift in Diffusion Models for Cross-Modality Image Translation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational Adaptation
von: Wang, Fei, et al.
Veröffentlicht: (2026)
von: Wang, Fei, et al.
Veröffentlicht: (2026)
Revisiting Cross-Attention Mechanisms: Leveraging Beneficial Noise for Domain-Adaptive Learning
von: Zang, Zelin, et al.
Veröffentlicht: (2026)
von: Zang, Zelin, et al.
Veröffentlicht: (2026)
CSFMamba: Cross State Fusion Mamba Operator for Multimodal Remote Sensing Image Classification
von: Wang, Qingyu, et al.
Veröffentlicht: (2025)
von: Wang, Qingyu, et al.
Veröffentlicht: (2025)
Training-Free Multimodal Deepfake Detection via Graph Reasoning
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
von: Liu, Yuxin, et al.
Veröffentlicht: (2025)
MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks
von: Xu, Yushen, et al.
Veröffentlicht: (2025)
von: Xu, Yushen, et al.
Veröffentlicht: (2025)
DreamFuse: Adaptive Image Fusion with Diffusion Transformer
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
TACFN: Transformer-based Adaptive Cross-modal Fusion Network for Multimodal Emotion Recognition
von: Liu, Feng, et al.
Veröffentlicht: (2025)
von: Liu, Feng, et al.
Veröffentlicht: (2025)
Multimodal Fusion via Self-Consistent Task-Gradient Fields
von: Xiong, Jiayu, et al.
Veröffentlicht: (2024)
von: Xiong, Jiayu, et al.
Veröffentlicht: (2024)
CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object Detection
von: Wu, Yuchen, et al.
Veröffentlicht: (2026)
von: Wu, Yuchen, et al.
Veröffentlicht: (2026)
3D Smoke Scene Reconstruction Guided by Vision Priors from Multimodal Large Language Models
von: Zheng, Xinye, et al.
Veröffentlicht: (2026)
von: Zheng, Xinye, et al.
Veröffentlicht: (2026)
Rethinking Where to Edit: Task-Aware Localization for Instruction-Based Image Editing
von: He, Jingxuan, et al.
Veröffentlicht: (2026)
von: He, Jingxuan, et al.
Veröffentlicht: (2026)
Double Helix Diffusion for Cross-Domain Anomaly Image Generation
von: Wu, Linchun, et al.
Veröffentlicht: (2025)
von: Wu, Linchun, et al.
Veröffentlicht: (2025)
Unsupervised Cross-Domain Image Retrieval via Prototypical Optimal Transport
von: Li, Bin, et al.
Veröffentlicht: (2024)
von: Li, Bin, et al.
Veröffentlicht: (2024)
MAUGIF: Mechanism-Aware Unsupervised General Image Fusion via Dual Cross-Image Autoencoders
von: Yang, Kunjing, et al.
Veröffentlicht: (2025)
von: Yang, Kunjing, et al.
Veröffentlicht: (2025)
Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
von: Huang, Hai, et al.
Veröffentlicht: (2025)
von: Huang, Hai, et al.
Veröffentlicht: (2025)
Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action Recognition
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning
von: Liu, Shih-Wen, et al.
Veröffentlicht: (2025)
von: Liu, Shih-Wen, et al.
Veröffentlicht: (2025)
Pulling Target to Source: A New Perspective on Domain Adaptive Semantic Segmentation
von: Wang, Haochen, et al.
Veröffentlicht: (2023)
von: Wang, Haochen, et al.
Veröffentlicht: (2023)
Interactive Multimodal Fusion with Temporal Modeling
von: Yu, Jun, et al.
Veröffentlicht: (2025)
von: Yu, Jun, et al.
Veröffentlicht: (2025)
Customized Fusion: A Closed-Loop Dynamic Network for Adaptive Multi-Task-Aware Infrared-Visible Image Fusion
von: Yang, Zengyi, et al.
Veröffentlicht: (2026)
von: Yang, Zengyi, et al.
Veröffentlicht: (2026)
LLM-Enhanced Multimodal Fusion for Cross-Domain Sequential Recommendation
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
von: Wu, Wangyu, et al.
Veröffentlicht: (2025)
Beyond Global Scanning: Adaptive Visual State Space Modeling for Salient Object Detection in Optical Remote Sensing Images
von: Ren, Mengyu, et al.
Veröffentlicht: (2025)
von: Ren, Mengyu, et al.
Veröffentlicht: (2025)
Infrared and Visible Image Fusion: From Data Compatibility to Task Adaption
von: Liu, Jinyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jinyuan, et al.
Veröffentlicht: (2025)
Let Synthetic Data Shine: Domain Reassembly and Soft-Fusion for Single Domain Generalization
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Balancing Task-invariant Interaction and Task-specific Adaptation for Unified Image Fusion
von: Hu, Xingyu, et al.
Veröffentlicht: (2025)
von: Hu, Xingyu, et al.
Veröffentlicht: (2025)
Meta-Exploiting Frequency Prior for Cross-Domain Few-Shot Learning
von: Zhou, Fei, et al.
Veröffentlicht: (2024)
von: Zhou, Fei, et al.
Veröffentlicht: (2024)
Hybrid Fusion: One-Minute Efficient Training for Zero-Shot Cross-Domain Image Fusion
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
von: Gu, Jihao, et al.
Veröffentlicht: (2025) -
Exploiting Ensemble Learning for Cross-View Isolated Sign Language Recognition
von: Wang, Fei, et al.
Veröffentlicht: (2025) -
Towards Generalizable Deepfake Detection with Spatial-Frequency Collaborative Learning and Hierarchical Cross-Modal Fusion
von: Qiao, Mengyu, et al.
Veröffentlicht: (2025) -
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
von: Liu, Yuxin, et al.
Veröffentlicht: (2026) -
GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)