Asymmetric Reinforcing against Multi-modal Representation Bias
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Xiyuan, Cao, Bing, Zhu, Pengfei, Wang, Nannan, Hu, Qinghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dynamic Brightness Adaptation for Robust Multi-modal Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2024)
von: Sun, Yiming, et al.
Veröffentlicht: (2024)
RGBX-R1: Visual Modality Chain-of-Thought Guided Reinforcement Learning for Multimodal Grounding
von: Wu, Jiahe, et al.
Veröffentlicht: (2026)
von: Wu, Jiahe, et al.
Veröffentlicht: (2026)
Visible and Clear: Finding Tiny Objects in Difference Map
von: Cao, Bing, et al.
Veröffentlicht: (2024)
von: Cao, Bing, et al.
Veröffentlicht: (2024)
Task-Customized Mixture of Adapters for General Image Fusion
von: Zhu, Pengfei, et al.
Veröffentlicht: (2024)
von: Zhu, Pengfei, et al.
Veröffentlicht: (2024)
Exploring Diverse Representations for Open Set Recognition
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
Dream-IF: Dynamic Relative EnhAnceMent for Image Fusion
von: Xu, Xingxin, et al.
Veröffentlicht: (2025)
von: Xu, Xingxin, et al.
Veröffentlicht: (2025)
Bi-directional Self-Registration for Misaligned Infrared-Visible Image Fusion
von: Li, Timing, et al.
Veröffentlicht: (2025)
von: Li, Timing, et al.
Veröffentlicht: (2025)
Efficient Masked AutoEncoder for Video Object Counting and A Large-Scale Benchmark
von: Cao, Bing, et al.
Veröffentlicht: (2024)
von: Cao, Bing, et al.
Veröffentlicht: (2024)
Reversible Efficient Diffusion for Image Fusion
von: Xu, Xingxin, et al.
Veröffentlicht: (2026)
von: Xu, Xingxin, et al.
Veröffentlicht: (2026)
DIFF-MF: A Difference-Driven Channel-Spatial State Space Model for Multi-Modal Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
AMU-Tuning: Effective Logit Bias for CLIP-based Few-shot Learning
von: Tang, Yuwei, et al.
Veröffentlicht: (2024)
von: Tang, Yuwei, et al.
Veröffentlicht: (2024)
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
von: Wang, Qing, et al.
Veröffentlicht: (2025)
von: Wang, Qing, et al.
Veröffentlicht: (2025)
Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and Evaluations
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
CSHNet: A Novel Information Asymmetric Image Translation Method
von: Yang, Xi, et al.
Veröffentlicht: (2025)
von: Yang, Xi, et al.
Veröffentlicht: (2025)
$S^3$: Synonymous Semantic Space for Improving Zero-Shot Generalization of Vision-Language Models
von: Yin, Xiaojie, et al.
Veröffentlicht: (2024)
von: Yin, Xiaojie, et al.
Veröffentlicht: (2024)
Conditional Controllable Image Fusion
von: Cao, Bing, et al.
Veröffentlicht: (2024)
von: Cao, Bing, et al.
Veröffentlicht: (2024)
Decoupled Multi-Predictor Optimization for Inference-Efficient Model Tuning
von: Luo, Liwei, et al.
Veröffentlicht: (2025)
von: Luo, Liwei, et al.
Veröffentlicht: (2025)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
CKD: Contrastive Knowledge Distillation from A Sample-wise Perspective
von: Zhu, Wencheng, et al.
Veröffentlicht: (2024)
von: Zhu, Wencheng, et al.
Veröffentlicht: (2024)
CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
Generalized Few-Shot Out-of-Distribution Detection
von: Li, Pinxuan, et al.
Veröffentlicht: (2025)
von: Li, Pinxuan, et al.
Veröffentlicht: (2025)
Federated Face Forgery Detection Learning with Personalized Representation
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
von: Liu, Decheng, et al.
Veröffentlicht: (2024)
Dynamic Sub-graph Distillation for Robust Semi-supervised Continual Learning
von: Fan, Yan, et al.
Veröffentlicht: (2023)
von: Fan, Yan, et al.
Veröffentlicht: (2023)
Multi-view Deep Subspace Clustering Networks
von: Zhu, Pengfei, et al.
Veröffentlicht: (2019)
von: Zhu, Pengfei, et al.
Veröffentlicht: (2019)
Hyperbolic Cycle Alignment for Infrared-Visible Image Fusion
von: Li, Timing, et al.
Veröffentlicht: (2025)
von: Li, Timing, et al.
Veröffentlicht: (2025)
BackMix: Regularizing Open Set Recognition by Removing Underlying Fore-Background Priors
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
Test-Time Dynamic Image Fusion
von: Cao, Bing, et al.
Veröffentlicht: (2024)
von: Cao, Bing, et al.
Veröffentlicht: (2024)
Unleashing Degradation-Carrying Features in Symmetric U-Net: Simpler and Stronger Baselines for All-in-One Image Restoration
von: Jiao, Wenlong, et al.
Veröffentlicht: (2025)
von: Jiao, Wenlong, et al.
Veröffentlicht: (2025)
Dig2DIG: Dig into Diffusion Information Gains for Image Fusion
von: Cao, Bing, et al.
Veröffentlicht: (2025)
von: Cao, Bing, et al.
Veröffentlicht: (2025)
Asymmetrical Siamese Network for Point Clouds Normal Estimation
von: Jin, Wei, et al.
Veröffentlicht: (2024)
von: Jin, Wei, et al.
Veröffentlicht: (2024)
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition
von: Wang, Teng, et al.
Veröffentlicht: (2024)
von: Wang, Teng, et al.
Veröffentlicht: (2024)
Object Affordance Recognition and Grounding via Multi-scale Cross-modal Representation Learning
von: Wan, Xinhang, et al.
Veröffentlicht: (2025)
von: Wan, Xinhang, et al.
Veröffentlicht: (2025)
Towards Generalized Proactive Defense against Face Swapping with Contour-Hybrid Watermark
von: Xia, Ruiyang, et al.
Veröffentlicht: (2025)
von: Xia, Ruiyang, et al.
Veröffentlicht: (2025)
MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection
von: Wang, Hanshi, et al.
Veröffentlicht: (2025)
von: Wang, Hanshi, et al.
Veröffentlicht: (2025)
HYDRA: Unifying Multi-modal Generation and Understanding via Representation-Harmonized Tokenization
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
CrackSegDiff: Diffusion Probability Model-based Multi-modal Crack Segmentation
von: Jiang, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Jiang, Xiaoyan, et al.
Veröffentlicht: (2024)
OTTER: Open-Tagging via Text-Image Representation for Multi-modal Understanding
von: Ouyang, Jieer, et al.
Veröffentlicht: (2025)
von: Ouyang, Jieer, et al.
Veröffentlicht: (2025)
Multi Attribute Bias Mitigation via Representation Learning
von: Dwivedi, Rajeev Ranjan, et al.
Veröffentlicht: (2025)
von: Dwivedi, Rajeev Ranjan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dynamic Brightness Adaptation for Robust Multi-modal Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2024) -
RGBX-R1: Visual Modality Chain-of-Thought Guided Reinforcement Learning for Multimodal Grounding
von: Wu, Jiahe, et al.
Veröffentlicht: (2026) -
Visible and Clear: Finding Tiny Objects in Difference Map
von: Cao, Bing, et al.
Veröffentlicht: (2024) -
Task-Customized Mixture of Adapters for General Image Fusion
von: Zhu, Pengfei, et al.
Veröffentlicht: (2024) -
Exploring Diverse Representations for Open Set Recognition
von: Wang, Yu, et al.
Veröffentlicht: (2024)