PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection
Fuente:
arXiv
Saved in:
| Main Authors: | Molahasani, Mahdiyar, Motamedi, Azadeh, Greenspan, Michael, Kim, Il-Min, Etemad, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On The Relationship Between Continual Learning and Long-Tailed Recognition
by: Molahasani, Mahdiyar, et al.
Published: (2023)
by: Molahasani, Mahdiyar, et al.
Published: (2023)
Federated Unsupervised Domain Generalization using Global and Local Alignment of Gradients
by: Pourpanah, Farhad, et al.
Published: (2024)
by: Pourpanah, Farhad, et al.
Published: (2024)
Federated Domain Generalization with Label Smoothing and Balanced Decentralized Training
by: Soltany, Milad, et al.
Published: (2024)
by: Soltany, Milad, et al.
Published: (2024)
Diffusion Models with Deterministic Normalizing Flow Priors
by: Zand, Mohsen, et al.
Published: (2023)
by: Zand, Mohsen, et al.
Published: (2023)
Socially-Informed Reconstruction for Pedestrian Trajectory Forecasting
by: Damirchi, Haleh, et al.
Published: (2024)
by: Damirchi, Haleh, et al.
Published: (2024)
CollideNet: Hierarchical Multi-scale Video Representation Learning with Disentanglement for Time-To-Collision Forecasting
by: Desai, Nishq Poorav, et al.
Published: (2026)
by: Desai, Nishq Poorav, et al.
Published: (2026)
CycleCrash: A Dataset of Bicycle Collision Videos for Collision Prediction and Analysis
by: Desai, Nishq Poorav, et al.
Published: (2024)
by: Desai, Nishq Poorav, et al.
Published: (2024)
Consistency-guided Prompt Learning for Vision-Language Models
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
SelfPrompt: Confidence-Aware Semi-Supervised Tuning for Robust Vision-Language Model Adaptation
by: Roy, Shuvendu, et al.
Published: (2025)
by: Roy, Shuvendu, et al.
Published: (2025)
Self-alignment of Large Video Language Models with Refined Regularized Preference Optimization
by: Sarkar, Pritam, et al.
Published: (2025)
by: Sarkar, Pritam, et al.
Published: (2025)
VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models
by: Sarkar, Pritam, et al.
Published: (2025)
by: Sarkar, Pritam, et al.
Published: (2025)
Identifying Implicit Social Biases in Vision-Language Models
by: Hamidieh, Kimia, et al.
Published: (2024)
by: Hamidieh, Kimia, et al.
Published: (2024)
GLOV: Guided Large Language Models as Implicit Optimizers for Vision Language Models
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
Vision Language Models are Biased
by: Vo, An, et al.
Published: (2025)
by: Vo, An, et al.
Published: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
by: Huang, Jen-tse, et al.
Published: (2025)
by: Huang, Jen-tse, et al.
Published: (2025)
Consistency-Guided Asynchronous Contrastive Tuning for Few-Shot Class-Incremental Tuning of Foundation Models
by: Roy, Shuvendu, et al.
Published: (2024)
by: Roy, Shuvendu, et al.
Published: (2024)
DLTPose: 6DoF Pose Estimation From Accurate Dense Surface Point Estimates
by: Jadhav, Akash, et al.
Published: (2025)
by: Jadhav, Akash, et al.
Published: (2025)
Pseudo-keypoint RKHS Learning for Self-supervised 6DoF Pose Estimation
by: Wu, Yangzheng, et al.
Published: (2023)
by: Wu, Yangzheng, et al.
Published: (2023)
Impact of Strategic Sampling and Supervision Policies on Semi-supervised Learning
by: Roy, Shuvendu, et al.
Published: (2022)
by: Roy, Shuvendu, et al.
Published: (2022)
Exploring the Boundaries of Semi-Supervised Facial Expression Recognition using In-Distribution, Out-of-Distribution, and Unconstrained Data
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Partial Label Learning for Emotion Recognition from EEG
by: Zhang, Guangyi, et al.
Published: (2023)
by: Zhang, Guangyi, et al.
Published: (2023)
Reducing Spurious Correlation for Federated Domain Generalization
by: Ma, Shuran, et al.
Published: (2024)
by: Ma, Shuran, et al.
Published: (2024)
Ego: Embedding-Guided Personalization of Vision-Language Models
by: Seifi, Soroush, et al.
Published: (2026)
by: Seifi, Soroush, et al.
Published: (2026)
iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models
by: Cui, Xuezhi, et al.
Published: (2026)
by: Cui, Xuezhi, et al.
Published: (2026)
Are Large Vision-Language Models Ready to Guide Blind and Low-Vision Individuals?
by: Kim, Eunki, et al.
Published: (2025)
by: Kim, Eunki, et al.
Published: (2025)
Scaling Up Semi-supervised Learning with Unconstrained Unlabelled Data
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Human Pose Estimation from Ambiguous Pressure Recordings with Spatio-temporal Masked Transformers
by: Davoodnia, Vandad, et al.
Published: (2023)
by: Davoodnia, Vandad, et al.
Published: (2023)
Identifying Spurious Biases Early in Training through the Lens of Simplicity Bias
by: Yang, Yu, et al.
Published: (2023)
by: Yang, Yu, et al.
Published: (2023)
Anatomical Token Uncertainty for Transformer-Guided Active MRI Acquisition
by: Ayzenberg, Lev, et al.
Published: (2026)
by: Ayzenberg, Lev, et al.
Published: (2026)
Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning
by: Roy, Shuvendu, et al.
Published: (2024)
by: Roy, Shuvendu, et al.
Published: (2024)
How Reasoning Influences Intersectional Biases in Vision Language Models
by: Desai, Adit, et al.
Published: (2025)
by: Desai, Adit, et al.
Published: (2025)
ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models
by: Yang, Cheng, et al.
Published: (2026)
by: Yang, Cheng, et al.
Published: (2026)
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
by: Varma, Maya, et al.
Published: (2024)
by: Varma, Maya, et al.
Published: (2024)
MM-SpuBench: Towards Better Understanding of Spurious Biases in Multimodal LLMs
by: Ye, Wenqian, et al.
Published: (2024)
by: Ye, Wenqian, et al.
Published: (2024)
Training Unbiased Diffusion Models From Biased Dataset
by: Kim, Yeongmin, et al.
Published: (2024)
by: Kim, Yeongmin, et al.
Published: (2024)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
by: Lee, Dong-Jae, et al.
Published: (2026)
by: Lee, Dong-Jae, et al.
Published: (2026)
PRISM: Progressive Reasoning through Iterative Slot Memory for Vision
by: Wang, Ziyu, et al.
Published: (2026)
by: Wang, Ziyu, et al.
Published: (2026)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
by: Lee, Donghwna, et al.
Published: (2024)
by: Lee, Donghwna, et al.
Published: (2024)
ETTA: Efficient Test-Time Adaptation for Vision-Language Models through Dynamic Embedding Updates
by: Dastmalchi, Hamidreza, et al.
Published: (2025)
by: Dastmalchi, Hamidreza, et al.
Published: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
by: Ma, Huan, et al.
Published: (2024)
by: Ma, Huan, et al.
Published: (2024)
Similar Items
-
On The Relationship Between Continual Learning and Long-Tailed Recognition
by: Molahasani, Mahdiyar, et al.
Published: (2023) -
Federated Unsupervised Domain Generalization using Global and Local Alignment of Gradients
by: Pourpanah, Farhad, et al.
Published: (2024) -
Federated Domain Generalization with Label Smoothing and Balanced Decentralized Training
by: Soltany, Milad, et al.
Published: (2024) -
Diffusion Models with Deterministic Normalizing Flow Priors
by: Zand, Mohsen, et al.
Published: (2023) -
Socially-Informed Reconstruction for Pedestrian Trajectory Forecasting
by: Damirchi, Haleh, et al.
Published: (2024)