Learning Invariant Causal Mechanism from Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Song, Zeen, Zhao, Siyu, Zhang, Xingyu, Li, Jiangmeng, Zheng, Changwen, Qiang, Wenwen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Generalization and Causal Explanation in Self-Supervised Learning
di: Qiang, Wenwen, et al.
Pubblicazione: (2024)
di: Qiang, Wenwen, et al.
Pubblicazione: (2024)
On the Discriminability of Self-Supervised Representation Learning
di: Song, Zeen, et al.
Pubblicazione: (2024)
di: Song, Zeen, et al.
Pubblicazione: (2024)
Self-Supervised Video Representation Learning in a Heuristic Decoupled Perspective
di: Song, Zeen, et al.
Pubblicazione: (2024)
di: Song, Zeen, et al.
Pubblicazione: (2024)
Rethinking Misalignment in Vision-Language Model Adaptation from a Causal Perspective
di: Zhang, Yanan, et al.
Pubblicazione: (2024)
di: Zhang, Yanan, et al.
Pubblicazione: (2024)
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
di: Li, Jiangmeng, et al.
Pubblicazione: (2022)
di: Li, Jiangmeng, et al.
Pubblicazione: (2022)
Causal Prompt Calibration Guided Segment Anything Model for Open-Vocabulary Multi-Entity Segmentation
di: Wang, Jingyao, et al.
Pubblicazione: (2025)
di: Wang, Jingyao, et al.
Pubblicazione: (2025)
AmPLe: Supporting Vision-Language Models via Adaptive-Debiased Ensemble Multi-Prompt Learning
di: Song, Fei, et al.
Pubblicazione: (2025)
di: Song, Fei, et al.
Pubblicazione: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
di: Song, Fei, et al.
Pubblicazione: (2025)
di: Song, Fei, et al.
Pubblicazione: (2025)
On the Transferability and Discriminability of Repersentation Learning in Unsupervised Domain Adaptation
di: Qiang, Wenwen, et al.
Pubblicazione: (2025)
di: Qiang, Wenwen, et al.
Pubblicazione: (2025)
Rethinking Generalizability and Discriminability of Self-Supervised Learning from Evolutionary Game Theory Perspective
di: Li, Jiangmeng, et al.
Pubblicazione: (2024)
di: Li, Jiangmeng, et al.
Pubblicazione: (2024)
Rethinking Meta-Learning from a Learning Lens
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
Unbiased Image Synthesis via Manifold Guidance in Diffusion Models
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
Spatio-Temporal Fuzzy-oriented Multi-Modal Meta-Learning for Fine-grained Emotion Recognition
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
Meta-Auxiliary Learning for Micro-Expression Recognition
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
Towards Task Sampler Learning for Meta-Learning
di: Wang, Jingyao, et al.
Pubblicazione: (2023)
di: Wang, Jingyao, et al.
Pubblicazione: (2023)
Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models
di: Zang, Zehua, et al.
Pubblicazione: (2026)
di: Zang, Zehua, et al.
Pubblicazione: (2026)
Exploring Transferability of Self-Supervised Learning by Task Conflict Calibration
di: Guo, Huijie, et al.
Pubblicazione: (2025)
di: Guo, Huijie, et al.
Pubblicazione: (2025)
AwesomeMeta+: A Mixed-Prototyping Meta-Learning System Supporting AI Application Design Anywhere
di: Wang, Jingyao, et al.
Pubblicazione: (2023)
di: Wang, Jingyao, et al.
Pubblicazione: (2023)
Vision-Language Attribute Disentanglement and Reinforcement for Lifelong Person Re-Identification
di: Xu, Kunlun, et al.
Pubblicazione: (2026)
di: Xu, Kunlun, et al.
Pubblicazione: (2026)
Beyond All-to-All: Causal-Aligned Transformer with Dynamic Structure Learning for Multivariate Time Series Forecasting
di: Zhang, Xingyu, et al.
Pubblicazione: (2025)
di: Zhang, Xingyu, et al.
Pubblicazione: (2025)
Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
di: Wei, Yana, et al.
Pubblicazione: (2025)
di: Wei, Yana, et al.
Pubblicazione: (2025)
On the Out-of-Distribution Generalization of Self-Supervised Learning
di: Qiang, Wenwen, et al.
Pubblicazione: (2025)
di: Qiang, Wenwen, et al.
Pubblicazione: (2025)
Dynamic Multimodal Prototype Learning in Vision-Language Models
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
Intriguing Property and Counterfactual Explanation of GAN for Remote Sensing Image Generation
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
Manifold Constraint Regularization for Remote Sensing Image Generation
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
Multimodal Interpretation of Remote Sensing Images: Dynamic Resolution Input Strategy and Multi-scale Vision-Language Alignment Mechanism
di: Zhang, Siyu, et al.
Pubblicazione: (2025)
di: Zhang, Siyu, et al.
Pubblicazione: (2025)
BayesTTA: Continual-Temporal Test-Time Adaptation for Vision-Language Models via Gaussian Discriminant Analysis
di: Cui, Shuang, et al.
Pubblicazione: (2025)
di: Cui, Shuang, et al.
Pubblicazione: (2025)
All-in-One Image Restoration via Causal-Deconfounding Wavelet-Disentangled Prompt Network
di: Wang, Bingnan, et al.
Pubblicazione: (2026)
di: Wang, Bingnan, et al.
Pubblicazione: (2026)
Domain-Invariant Prompt Learning for Vision-Language Models
di: Khoee, Arsham Gholamzadeh, et al.
Pubblicazione: (2026)
di: Khoee, Arsham Gholamzadeh, et al.
Pubblicazione: (2026)
Causal-Tune: Mining Causal Factors from Vision Foundation Models for Domain Generalized Semantic Segmentation
di: Zhang, Yin, et al.
Pubblicazione: (2025)
di: Zhang, Yin, et al.
Pubblicazione: (2025)
CELLO: Causal Evaluation of Large Vision-Language Models
di: Chen, Meiqi, et al.
Pubblicazione: (2024)
di: Chen, Meiqi, et al.
Pubblicazione: (2024)
Cascade Prompt Learning for Vision-Language Model Adaptation
di: Wu, Ge, et al.
Pubblicazione: (2024)
di: Wu, Ge, et al.
Pubblicazione: (2024)
Image-based Freeform Handwriting Authentication with Energy-oriented Self-Supervised Learning
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
di: Wang, Jingyao, et al.
Pubblicazione: (2024)
MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
Making Large Vision Language Models to be Good Few-shot Learners
di: Liu, Fan, et al.
Pubblicazione: (2024)
di: Liu, Fan, et al.
Pubblicazione: (2024)
Principled Steering via Null-space Projection for Jailbreak Defense in Vision-Language Models
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
di: Zhu, Xingyu, et al.
Pubblicazione: (2026)
Sparse but not Simpler: A Multi-Level Interpretability Analysis of Vision Transformers
di: Zhang, Siyu
Pubblicazione: (2026)
di: Zhang, Siyu
Pubblicazione: (2026)
Enhanced Continual Learning of Vision-Language Models with Model Fusion
di: Gao, Haoyuan, et al.
Pubblicazione: (2025)
di: Gao, Haoyuan, et al.
Pubblicazione: (2025)
MAO: Efficient Model-Agnostic Optimization of Prompt Tuning for Vision-Language Models
di: Li, Haoyang, et al.
Pubblicazione: (2025)
di: Li, Haoyang, et al.
Pubblicazione: (2025)
Learning an Adaptive and View-Invariant Vision Transformer for Real-Time UAV Tracking
di: Wu, You, et al.
Pubblicazione: (2024)
di: Wu, You, et al.
Pubblicazione: (2024)
Documenti analoghi
-
On the Generalization and Causal Explanation in Self-Supervised Learning
di: Qiang, Wenwen, et al.
Pubblicazione: (2024) -
On the Discriminability of Self-Supervised Representation Learning
di: Song, Zeen, et al.
Pubblicazione: (2024) -
Self-Supervised Video Representation Learning in a Heuristic Decoupled Perspective
di: Song, Zeen, et al.
Pubblicazione: (2024) -
Rethinking Misalignment in Vision-Language Model Adaptation from a Causal Perspective
di: Zhang, Yanan, et al.
Pubblicazione: (2024) -
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
di: Li, Jiangmeng, et al.
Pubblicazione: (2022)