One Model for ALL: Low-Level Task Interaction Is a Key to Task-Agnostic Image Fusion
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Chunyang, Xu, Tianyang, Feng, Zhenhua, Wu, Xiaojun, ZhangyongTang, Li, Hui, Zhang, Zeyang, Atito, Sara, Awais, Muhammad, Kittler, Josef |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Information theoretic underpinning of self-supervised learning by clustering
by: Kittler, Josef, et al.
Published: (2026)
by: Kittler, Josef, et al.
Published: (2026)
DailyMAE: Towards Pretraining Masked Autoencoders in One Day
by: Wu, Jiantao, et al.
Published: (2024)
by: Wu, Jiantao, et al.
Published: (2024)
TextFusion: Unveiling the Power of Textual Semantics for Controllable Image Fusion
by: Cheng, Chunyang, et al.
Published: (2023)
by: Cheng, Chunyang, et al.
Published: (2023)
Investigating Self-Supervised Methods for Label-Efficient Learning
by: Nandam, Srinivasa Rao, et al.
Published: (2024)
by: Nandam, Srinivasa Rao, et al.
Published: (2024)
Pseudo Labelling for Enhanced Masked Autoencoders
by: Nandam, Srinivasa Rao, et al.
Published: (2024)
by: Nandam, Srinivasa Rao, et al.
Published: (2024)
EvaNet: Towards More Efficient and Consistent Infrared and Visible Image Fusion Assessment
by: Cheng, Chunyang, et al.
Published: (2026)
by: Cheng, Chunyang, et al.
Published: (2026)
Probabilistically Aligned View-unaligned Clustering with Adaptive Template Selection
by: Dong, Wenhua, et al.
Published: (2024)
by: Dong, Wenhua, et al.
Published: (2024)
C2C: Component-to-Composition Learning for Zero-Shot Compositional Action Recognition
by: Li, Rongchang, et al.
Published: (2024)
by: Li, Rongchang, et al.
Published: (2024)
Serial Over Parallel: Learning Continual Unification for Multi-Modal Visual Object Tracking and Benchmarking
by: Tang, Zhangyong, et al.
Published: (2025)
by: Tang, Zhangyong, et al.
Published: (2025)
ASiT: Local-Global Audio Spectrogram vIsion Transformer for Event Classification
by: Atito, Sara, et al.
Published: (2022)
by: Atito, Sara, et al.
Published: (2022)
Revisiting RGBT Tracking Benchmarks from the Perspective of Modality Validity: A New Benchmark, Problem, and Solution
by: Tang, Zhangyong, et al.
Published: (2024)
by: Tang, Zhangyong, et al.
Published: (2024)
BusReF: Infrared-Visible images registration and fusion focus on reconstructible area using one set of features
by: Zhang, Zeyang, et al.
Published: (2023)
by: Zhang, Zeyang, et al.
Published: (2023)
FusionBooster: A Unified Image Fusion Boosting Paradigm
by: Cheng, Chunyang, et al.
Published: (2023)
by: Cheng, Chunyang, et al.
Published: (2023)
Omni Survey for Multimodality Analysis in Visual Object Tracking
by: Tang, Zhangyong, et al.
Published: (2025)
by: Tang, Zhangyong, et al.
Published: (2025)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
by: Nazarieh, Fatemeh, et al.
Published: (2024)
by: Nazarieh, Fatemeh, et al.
Published: (2024)
TENet: Targetness Entanglement Incorporating with Multi-Scale Pooling and Mutually-Guided Fusion for RGB-E Object Tracking
by: Shao, Pengcheng, et al.
Published: (2024)
by: Shao, Pengcheng, et al.
Published: (2024)
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
MMDRFuse: Distilled Mini-Model with Dynamic Refresh for Multi-Modality Image Fusion
by: Deng, Yanglin, et al.
Published: (2024)
by: Deng, Yanglin, et al.
Published: (2024)
Beyond Strict Pairing: Arbitrarily Paired Training for High-Performance Infrared and Visible Image Fusion
by: Deng, Yanglin, et al.
Published: (2026)
by: Deng, Yanglin, et al.
Published: (2026)
Learning Progressive Adaptation for Multi-Modal Tracking
by: Wang, He, et al.
Published: (2026)
by: Wang, He, et al.
Published: (2026)
UASTrack: A Unified Adaptive Selection Framework with Modality-Customization in Single Object Tracking
by: Wang, He, et al.
Published: (2025)
by: Wang, He, et al.
Published: (2025)
Rethinking Positive Pairs in Contrastive Learning
by: Wu, Jiantao, et al.
Published: (2024)
by: Wu, Jiantao, et al.
Published: (2024)
DC-ViT: Modulating Spatial and Channel Interactions for Multi-Channel Images
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
SMLNet: A SPD Manifold Learning Network for Infrared and Visible Image Fusion
by: Kang, Huan, et al.
Published: (2024)
by: Kang, Huan, et al.
Published: (2024)
GrFormer: A Novel Transformer on Grassmann Manifold for Infrared and Visible Image Fusion
by: Kang, Huan, et al.
Published: (2025)
by: Kang, Huan, et al.
Published: (2025)
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
by: Nazarieh, Fatemeh, et al.
Published: (2025)
by: Nazarieh, Fatemeh, et al.
Published: (2025)
Channel-Aware Probing for Multi-Channel Imaging
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
Hybrid Batch Normalisation: Resolving the Dilemma of Batch Normalisation in Federated Learning
by: Chen, Hongyao, et al.
Published: (2025)
by: Chen, Hongyao, et al.
Published: (2025)
C3R: Channel Conditioned Cell Representations for unified evaluation in microscopy imaging
by: Marikkar, Umar, et al.
Published: (2025)
by: Marikkar, Umar, et al.
Published: (2025)
Adaptive Hyper-Graph Convolution Network for Skeleton-based Human Action Recognition with Virtual Connections
by: Zhou, Youwei, et al.
Published: (2024)
by: Zhou, Youwei, et al.
Published: (2024)
Dynamic Subframe Splitting and Spatio-Temporal Motion Entangled Sparse Attention for RGB-E Tracking
by: Shao, Pengcheng, et al.
Published: (2024)
by: Shao, Pengcheng, et al.
Published: (2024)
CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks
by: Suharitdamrong, Wish, et al.
Published: (2026)
by: Suharitdamrong, Wish, et al.
Published: (2026)
One Latent Space to Rule All Degradations: Unifying Restoration Knowledge for Image Fusion
by: Ma, Haolong, et al.
Published: (2025)
by: Ma, Haolong, et al.
Published: (2025)
An Improved Graph Pooling Network for Skeleton-Based Action Recognition
by: Wu, Cong, et al.
Published: (2024)
by: Wu, Cong, et al.
Published: (2024)
Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs
by: Liang, Zhangyong
Published: (2026)
by: Liang, Zhangyong
Published: (2026)
DeepChest: Dynamic Gradient-Free Task Weighting for Effective Multi-Task Learning in Chest X-ray Classification
by: Mohamed, Youssef, et al.
Published: (2025)
by: Mohamed, Youssef, et al.
Published: (2025)
Towards Highly Transferable Vision-Language Attack via Semantic-Augmented Dynamic Contrastive Interaction
by: Li, Yuanbo, et al.
Published: (2026)
by: Li, Yuanbo, et al.
Published: (2026)
S4Fusion: Saliency-aware Selective State Space Model for Infrared Visible Image Fusion
by: Ma, Haolong, et al.
Published: (2024)
by: Ma, Haolong, et al.
Published: (2024)
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
CoMoFusion: Fast and High-quality Fusion of Infrared and Visible Image with Consistency Model
by: Meng, Zhiming, et al.
Published: (2024)
by: Meng, Zhiming, et al.
Published: (2024)
Similar Items
-
Information theoretic underpinning of self-supervised learning by clustering
by: Kittler, Josef, et al.
Published: (2026) -
DailyMAE: Towards Pretraining Masked Autoencoders in One Day
by: Wu, Jiantao, et al.
Published: (2024) -
TextFusion: Unveiling the Power of Textual Semantics for Controllable Image Fusion
by: Cheng, Chunyang, et al.
Published: (2023) -
Investigating Self-Supervised Methods for Label-Efficient Learning
by: Nandam, Srinivasa Rao, et al.
Published: (2024) -
Pseudo Labelling for Enhanced Masked Autoencoders
by: Nandam, Srinivasa Rao, et al.
Published: (2024)