Bridge: Basis-Driven Causal Inference Marries VFMs for Domain Generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Hong, Mingbo, Liu, Feng, Gevaert, Caroline, Vosselman, George, Cheng, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
ACPV-Net: All-Class Polygonal Vectorization for Seamless Vector Map Generation from Aerial Imagery
by: Jiao, Weiqin, et al.
Published: (2026)
by: Jiao, Weiqin, et al.
Published: (2026)
LDPoly: Latent Diffusion for Polygonal Road Outline Extraction in Large-Scale Topographic Mapping
by: Jiao, Weiqin, et al.
Published: (2025)
by: Jiao, Weiqin, et al.
Published: (2025)
RoIPoly: Vectorized Building Outline Extraction Using Vertex and Logit Embeddings
by: Jiao, Weiqin, et al.
Published: (2024)
by: Jiao, Weiqin, et al.
Published: (2024)
Learning from Exemplars for Interactive Image Segmentation
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
A Comparison of Multi-View Stereo Methods for Photogrammetric 3D Reconstruction: From Traditional to Learning-Based Approaches
by: Li, Yawen, et al.
Published: (2026)
by: Li, Yawen, et al.
Published: (2026)
PolyR-CNN: R-CNN for end-to-end polygonal building outline extraction
by: Jiao, Weiqin, et al.
Published: (2024)
by: Jiao, Weiqin, et al.
Published: (2024)
Multimodal Rationales for Explainable Visual Question Answering
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
Scale-wise Bidirectional Alignment Network for Referring Remote Sensing Image Segmentation
by: Li, Kun, et al.
Published: (2025)
by: Li, Kun, et al.
Published: (2025)
Incremental Semantics-Aided Meshing from LiDAR-Inertial Odometry and RGB Direct Label Transfer
by: Affan, Muhammad, et al.
Published: (2026)
by: Affan, Muhammad, et al.
Published: (2026)
CausalSR: Structural Causal Model-Driven Super-Resolution with Counterfactual Inference
by: Lu, Zhengyang, et al.
Published: (2025)
by: Lu, Zhengyang, et al.
Published: (2025)
MFH: Marrying Frequency Domain with Handwritten Mathematical Expression Recognition
by: Yang, Huanxin, et al.
Published: (2025)
by: Yang, Huanxin, et al.
Published: (2025)
M2H-MX: Multi-Task Semantic and Geometric Perception for Real-Time Monocular 3D Scene Graph Construction
by: Udugama, U. V. B. L., et al.
Published: (2026)
by: Udugama, U. V. B. L., et al.
Published: (2026)
DVLO4D: Deep Visual-Lidar Odometry with Sparse Spatial-temporal Fusion
by: Liu, Mengmeng, et al.
Published: (2025)
by: Liu, Mengmeng, et al.
Published: (2025)
Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
by: Udugama, U. V. B. L., et al.
Published: (2026)
by: Udugama, U. V. B. L., et al.
Published: (2026)
Exploiting Domain Properties in Language-Driven Domain Generalization for Semantic Segmentation
by: Jeon, Seogkyu, et al.
Published: (2025)
by: Jeon, Seogkyu, et al.
Published: (2025)
TK-Mamba: Marrying KAN With Mamba for Text-Driven 3D Medical Image Segmentation
by: Yang, Haoyu, et al.
Published: (2025)
by: Yang, Haoyu, et al.
Published: (2025)
Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition
by: Kuang, Jidong, et al.
Published: (2026)
by: Kuang, Jidong, et al.
Published: (2026)
SCHNet: SAM Marries CLIP for Human Parsing
by: Liu, Kunliang, et al.
Published: (2025)
by: Liu, Kunliang, et al.
Published: (2025)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
by: Liu, Shilong, et al.
Published: (2023)
by: Liu, Shilong, et al.
Published: (2023)
Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
by: Zhang, David Junhao, et al.
Published: (2023)
by: Zhang, David Junhao, et al.
Published: (2023)
You Only Look Around: Learning Illumination Invariant Feature for Low-light Object Detection
by: Hong, Mingbo, et al.
Published: (2024)
by: Hong, Mingbo, et al.
Published: (2024)
Casual Inference via Style Bias Deconfounding for Domain Generalization
by: Li, Jiaxi, et al.
Published: (2025)
by: Li, Jiaxi, et al.
Published: (2025)
Feasibility of Indoor Frame-Wise Lidar Semantic Segmentation via Distillation from Visual Foundation Model
by: Wu, Haiyang, et al.
Published: (2026)
by: Wu, Haiyang, et al.
Published: (2026)
Marrying Autoregressive Transformer and Diffusion with Multi-Reference Autoregression
by: Zhen, Dingcheng, et al.
Published: (2025)
by: Zhen, Dingcheng, et al.
Published: (2025)
Spectral Property-Driven Data Augmentation for Hyperspectral Single-Source Domain Generalization
by: Chen, Taiqin, et al.
Published: (2026)
by: Chen, Taiqin, et al.
Published: (2026)
Optimization-Driven Statistical Models of Anatomies using Radial Basis Function Shape Representation
by: Xu, Hong, et al.
Published: (2024)
by: Xu, Hong, et al.
Published: (2024)
Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation
by: Ye, Maoyuan, et al.
Published: (2024)
by: Ye, Maoyuan, et al.
Published: (2024)
Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
by: Huang, Hai, et al.
Published: (2025)
by: Huang, Hai, et al.
Published: (2025)
Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection
by: Wang, Haoxuan, et al.
Published: (2024)
by: Wang, Haoxuan, et al.
Published: (2024)
Planner3D: LLM-enhanced graph prior meets 3D indoor scene explicit regularization
by: Wei, Yao, et al.
Published: (2024)
by: Wei, Yao, et al.
Published: (2024)
AI-Derived Structural Building Intelligence for Urban Resilience: An Application in Saint Vincent and the Grenadines
by: Tingzon, Isabelle, et al.
Published: (2025)
by: Tingzon, Isabelle, et al.
Published: (2025)
TriC-Motion: Tri-Domain Causal Modeling Grounded Text-to-Motion Generation
by: Cao, Yiyang, et al.
Published: (2026)
by: Cao, Yiyang, et al.
Published: (2026)
Neural Spectral Decomposition for Dataset Distillation
by: Yang, Shaolei, et al.
Published: (2024)
by: Yang, Shaolei, et al.
Published: (2024)
Deconfounding Causal Inference through Two-Branch Framework with Early-Forking for Sensor-Based Cross-Domain Activity Recognition
by: Xiong, Di, et al.
Published: (2025)
by: Xiong, Di, et al.
Published: (2025)
LFSamba: Marry SAM with Mamba for Light Field Salient Object Detection
by: Liu, Zhengyi, et al.
Published: (2024)
by: Liu, Zhengyi, et al.
Published: (2024)
FlexEdit: Marrying Free-Shape Masks to VLLM for Flexible Image Editing
by: Yuan, Tianshuo, et al.
Published: (2024)
by: Yuan, Tianshuo, et al.
Published: (2024)
A Causal Inspired Early-Branching Structure for Domain Generalization
by: Chen, Liang, et al.
Published: (2024)
by: Chen, Liang, et al.
Published: (2024)
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
by: Udugama, U. V. B. L, et al.
Published: (2025)
by: Udugama, U. V. B. L, et al.
Published: (2025)
Causality-inspired Discriminative Feature Learning in Triple Domains for Gait Recognition
by: Xiong, Haijun, et al.
Published: (2024)
by: Xiong, Haijun, et al.
Published: (2024)
Similar Items
-
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026) -
ACPV-Net: All-Class Polygonal Vectorization for Seamless Vector Map Generation from Aerial Imagery
by: Jiao, Weiqin, et al.
Published: (2026) -
LDPoly: Latent Diffusion for Polygonal Road Outline Extraction in Large-Scale Topographic Mapping
by: Jiao, Weiqin, et al.
Published: (2025) -
RoIPoly: Vectorized Building Outline Extraction Using Vertex and Logit Embeddings
by: Jiao, Weiqin, et al.
Published: (2024) -
Learning from Exemplars for Interactive Image Segmentation
by: Li, Kun, et al.
Published: (2024)