CrossFlowDG: Bridging the Modality Gap with Cross-modal Flow Matching for Domain Generalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kritikos, Antonios, Spanos, Nikolaos, Voulodimos, Athanasios |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Explaining Vision GNNs: A Semantic and Visual Analysis of Graph-based Image Classification
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
Complex Style Image Transformations for Domain Generalization in Medical Images
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2024)
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2024)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025)
Bridging Modality Gap for Visual Grounding with Effecitve Cross-modal Distillation
von: Wang, Jiaxi, et al.
Veröffentlicht: (2023)
von: Wang, Jiaxi, et al.
Veröffentlicht: (2023)
Caption-Matching: A Multimodal Approach for Cross-Domain Image Retrieval
von: Iijima, Lucas, et al.
Veröffentlicht: (2024)
von: Iijima, Lucas, et al.
Veröffentlicht: (2024)
Exploring Cross-Modal Flows for Few-Shot Learning
von: Jiang, Ziqi, et al.
Veröffentlicht: (2025)
von: Jiang, Ziqi, et al.
Veröffentlicht: (2025)
Analyze-Prompt-Reason: A Collaborative Agent-Based Framework for Multi-Image Vision-Language Reasoning
von: Vlachos, Angelos, et al.
Veröffentlicht: (2025)
von: Vlachos, Angelos, et al.
Veröffentlicht: (2025)
Multi-modal Generation via Cross-Modal In-Context Learning
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
Reasoning or Pattern Matching? Probing Large Vision-Language Models with Visual Puzzles
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2026)
von: Lymperaiou, Maria, et al.
Veröffentlicht: (2026)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
von: Zhang, Zelin, et al.
Veröffentlicht: (2026)
von: Zhang, Zelin, et al.
Veröffentlicht: (2026)
$x^2$-Fusion: Cross-Modality and Cross-Dimension Flow Estimation in Event Edge Space
von: Guo, Ruishan, et al.
Veröffentlicht: (2026)
von: Guo, Ruishan, et al.
Veröffentlicht: (2026)
MER-DG: Modality-Entropy Regularization for Multimodal Domain Generalization
von: Yarici, Yavuz, et al.
Veröffentlicht: (2026)
von: Yarici, Yavuz, et al.
Veröffentlicht: (2026)
RareFlow: Physics-Aware Flow-Matching for Cross-Sensor Super-Resolution of Rare-Earth Features
von: Fallah, Forouzan, et al.
Veröffentlicht: (2025)
von: Fallah, Forouzan, et al.
Veröffentlicht: (2025)
Bridging the Inter-Domain Gap through Low-Level Features for Cross-Modal Medical Image Segmentation
von: Lyu, Pengfei, et al.
Veröffentlicht: (2025)
von: Lyu, Pengfei, et al.
Veröffentlicht: (2025)
FSDA-DG: Improving Cross-Domain Generalizability of Medical Image Segmentation with Few Source Domain Annotations
von: Ye, Zanting, et al.
Veröffentlicht: (2023)
von: Ye, Zanting, et al.
Veröffentlicht: (2023)
XoFTR: Cross-modal Feature Matching Transformer
von: Tuzcuoğlu, Önder, et al.
Veröffentlicht: (2024)
von: Tuzcuoğlu, Önder, et al.
Veröffentlicht: (2024)
Dino-Diffusion Modular Designs Bridge the Cross-Domain Gap in Autonomous Parking
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection
von: Bai, Yuhu, et al.
Veröffentlicht: (2025)
von: Bai, Yuhu, et al.
Veröffentlicht: (2025)
Mind the Modality Gap: Towards a Remote Sensing Vision-Language Model via Cross-modal Alignment
von: Zavras, Angelos, et al.
Veröffentlicht: (2024)
von: Zavras, Angelos, et al.
Veröffentlicht: (2024)
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2026)
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2026)
Non-target Divergence Hypothesis: Toward Understanding Domain Gaps in Cross-Modal Knowledge Distillation
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
AsyncBEV: Cross-modal Flow Alignment in Asynchronous 3D Object Detection
von: Wang, Shiming, et al.
Veröffentlicht: (2026)
von: Wang, Shiming, et al.
Veröffentlicht: (2026)
Bridging the Gap: Multi-Level Cross-Modality Joint Alignment for Visible-Infrared Person Re-Identification
von: Liang, Tengfei, et al.
Veröffentlicht: (2023)
von: Liang, Tengfei, et al.
Veröffentlicht: (2023)
LayoutFlow: Flow Matching for Layout Generation
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
Cross-modal Information Flow in Multimodal Large Language Models
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhi, et al.
Veröffentlicht: (2024)
FGML-DG: Feynman-Inspired Cognitive Science Paradigm for Cross-Domain Medical Image Segmentation
von: Song, Yucheng, et al.
Veröffentlicht: (2026)
von: Song, Yucheng, et al.
Veröffentlicht: (2026)
TrajFlow: Multi-modal Motion Prediction via Flow Matching
von: Yan, Qi, et al.
Veröffentlicht: (2025)
von: Yan, Qi, et al.
Veröffentlicht: (2025)
CM-Bench: A Comprehensive Cross-Modal Feature Matching Benchmark Bridging Visible and Infrared Images
von: Sun, Liangzheng, et al.
Veröffentlicht: (2026)
von: Sun, Liangzheng, et al.
Veröffentlicht: (2026)
Fast Post-Hoc Confidence Fusion for 3-Class Open-Set Aerial Object Detection
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
von: Loukovitis, Spyridon, et al.
Veröffentlicht: (2025)
Flow Matching for Medical Image Synthesis: Bridging the Gap Between Speed and Quality
von: Yazdani, Milad, et al.
Veröffentlicht: (2025)
von: Yazdani, Milad, et al.
Veröffentlicht: (2025)
Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency Constraint
von: Zhang, Runmin, et al.
Veröffentlicht: (2025)
von: Zhang, Runmin, et al.
Veröffentlicht: (2025)
Cross-Modal-Domain Generalization Through Semantically Aligned Discrete Representations
von: Sen, Souptik, et al.
Veröffentlicht: (2026)
von: Sen, Souptik, et al.
Veröffentlicht: (2026)
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
von: Mistretta, Marco, et al.
Veröffentlicht: (2025)
von: Mistretta, Marco, et al.
Veröffentlicht: (2025)
Cross-Modal Mapping: Mitigating the Modality Gap for Few-Shot Image Classification
von: Yang, Xi, et al.
Veröffentlicht: (2024)
von: Yang, Xi, et al.
Veröffentlicht: (2024)
Multi-Modal LLM based Image Captioning in ICT: Bridging the Gap Between General and Industry Domain
von: Chao, Lianying, et al.
Veröffentlicht: (2026)
von: Chao, Lianying, et al.
Veröffentlicht: (2026)
Weakly Supervised Cross-Modal Learning for 4D Radar Scene Flow Estimation
von: Fu, Jingyun, et al.
Veröffentlicht: (2026)
von: Fu, Jingyun, et al.
Veröffentlicht: (2026)
Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality Evolution
von: Liu, Qihao, et al.
Veröffentlicht: (2024)
von: Liu, Qihao, et al.
Veröffentlicht: (2024)
Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis
von: Chefer, Hila, et al.
Veröffentlicht: (2026)
von: Chefer, Hila, et al.
Veröffentlicht: (2026)
Flow Matching Posterior Sampling: A Training-free Conditional Generation for Flow Matching
von: Song, Kaiyu, et al.
Veröffentlicht: (2024)
von: Song, Kaiyu, et al.
Veröffentlicht: (2024)
Blockwise Flow Matching: Improving Flow Matching Models For Efficient High-Quality Generation
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Explaining Vision GNNs: A Semantic and Visual Analysis of Graph-based Image Classification
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025) -
Complex Style Image Transformations for Domain Generalization in Medical Images
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2024) -
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
von: Spanos, Nikolaos, et al.
Veröffentlicht: (2025) -
Bridging Modality Gap for Visual Grounding with Effecitve Cross-modal Distillation
von: Wang, Jiaxi, et al.
Veröffentlicht: (2023) -
Caption-Matching: A Multimodal Approach for Cross-Domain Image Retrieval
von: Iijima, Lucas, et al.
Veröffentlicht: (2024)