DiSa: Directional Saliency-Aware Prompt Learning for Generalizable Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Talemi, Niloufar Alipour, Kashiani, Hossein, Nowdeh, Hossein R., Afghah, Fatemeh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
FreqDebias: Towards Generalizable Deepfake Detection via Consistency-Driven Frequency Debiasing
by: Kashiani, Hossein, et al.
Published: (2025)
by: Kashiani, Hossein, et al.
Published: (2025)
ROADS: Robust Prompt-driven Multi-Class Anomaly Detection under Domain Shift
by: Kashiani, Hossein, et al.
Published: (2024)
by: Kashiani, Hossein, et al.
Published: (2024)
CATFace: Cross-Attribute-Guided Transformer with Self-Attention Distillation for Low-Quality Face Recognition
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
Agentic AI in Remote Sensing: Foundations, Taxonomy, and Emerging Systems
by: Talemi, Niloufar Alipour, et al.
Published: (2026)
by: Talemi, Niloufar Alipour, et al.
Published: (2026)
PromptMAD: Cross-Modal Prompting for Multi-Class Visual Anomaly Localization
by: McCain, Duncan, et al.
Published: (2026)
by: McCain, Duncan, et al.
Published: (2026)
Modality-Aware SAM: Sharpness-Aware-Minimization Driven Gradient Modulation for Harmonized Multimodal Learning
by: Nowdeh, Hossein R., et al.
Published: (2025)
by: Nowdeh, Hossein R., et al.
Published: (2025)
DiSa: Saliency-Aware Foreground-Background Disentangled Framework for Open-Vocabulary Semantic Segmentation
by: Yao, Zhen, et al.
Published: (2026)
by: Yao, Zhen, et al.
Published: (2026)
WildFireVQA: A Large-Scale Radiometric Thermal VQA Benchmark for Aerial Wildfire Monitoring
by: Habibpour, Mobin, et al.
Published: (2026)
by: Habibpour, Mobin, et al.
Published: (2026)
FlameFinder: Illuminating Obscured Fire through Smoke with Attentive Deep Metric Learning
by: Rajoli, Hossein, et al.
Published: (2024)
by: Rajoli, Hossein, et al.
Published: (2024)
Thermal Image Calibration and Correction using Unpaired Cycle-Consistent Adversarial Networks
by: Rajoli, Hossein, et al.
Published: (2024)
by: Rajoli, Hossein, et al.
Published: (2024)
Layout-Independent License Plate Recognition via Integrated Vision and Language Models
by: Shabaninia, Elham, et al.
Published: (2025)
by: Shabaninia, Elham, et al.
Published: (2025)
Generalizable Prompt Tuning for Vision-Language Models
by: Zhang, Qian
Published: (2024)
by: Zhang, Qian
Published: (2024)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
by: Basu, Abhishek, et al.
Published: (2025)
by: Basu, Abhishek, et al.
Published: (2025)
Diversity Covariance-Aware Prompt Learning for Vision-Language Models
by: Dong, Songlin, et al.
Published: (2025)
by: Dong, Songlin, et al.
Published: (2025)
TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning
by: Chen, Lihong, et al.
Published: (2025)
by: Chen, Lihong, et al.
Published: (2025)
Weak Distribution Detectors Lead to Stronger Generalizability of Vision-Language Prompt Tuning
by: Ding, Kun, et al.
Published: (2024)
by: Ding, Kun, et al.
Published: (2024)
3D Aware Region Prompted Vision Language Model
by: Cheng, An-Chieh, et al.
Published: (2025)
by: Cheng, An-Chieh, et al.
Published: (2025)
LOD1 3D City Model from LiDAR: The Impact of Segmentation Accuracy on Quality of Urban 3D Modeling and Morphology Extraction
by: Chajaei, Fatemeh, et al.
Published: (2025)
by: Chajaei, Fatemeh, et al.
Published: (2025)
IAP: Improving Continual Learning of Vision-Language Models via Instance-Aware Prompting
by: Fu, Hao, et al.
Published: (2025)
by: Fu, Hao, et al.
Published: (2025)
Cluster-Aware Prompt Ensemble Learning for Few-Shot Vision-Language Model Adaptation
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
Dude: Dual Distribution-Aware Context Prompt Learning For Large Vision-Language Model
by: Nguyen, Duy M. H., et al.
Published: (2024)
by: Nguyen, Duy M. H., et al.
Published: (2024)
Hardware Acceleration for Real-Time Wildfire Detection Onboard Drone Networks
by: Briley, Austin, et al.
Published: (2024)
by: Briley, Austin, et al.
Published: (2024)
Active Prompt Learning in Vision Language Models
by: Bang, Jihwan, et al.
Published: (2023)
by: Bang, Jihwan, et al.
Published: (2023)
In the Era of Prompt Learning with Vision-Language Models
by: Jha, Ankit
Published: (2024)
by: Jha, Ankit
Published: (2024)
Mixture of Prompt Learning for Vision Language Models
by: Du, Yu, et al.
Published: (2024)
by: Du, Yu, et al.
Published: (2024)
GLAD: Generalizable Tuning for Vision-Language Models
by: Peng, Yuqi, et al.
Published: (2025)
by: Peng, Yuqi, et al.
Published: (2025)
Foveated Reasoning: Stateful, Action-based Visual Focusing for Vision-Language Models
by: Min, Juhong, et al.
Published: (2026)
by: Min, Juhong, et al.
Published: (2026)
Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM
by: Kang, Jingxuan, et al.
Published: (2026)
by: Kang, Jingxuan, et al.
Published: (2026)
Saliency-Aware Model Merging
by: Park, Jungin, et al.
Published: (2026)
by: Park, Jungin, et al.
Published: (2026)
Active Prompt Learning with Vision-Language Model Priors
by: Kim, Hoyoung, et al.
Published: (2024)
by: Kim, Hoyoung, et al.
Published: (2024)
Integrated Structural Prompt Learning for Vision-Language Models
by: Wang, Jiahui, et al.
Published: (2025)
by: Wang, Jiahui, et al.
Published: (2025)
Modular Prompt Learning Improves Vision-Language Models
by: Huang, Zhenhan, et al.
Published: (2025)
by: Huang, Zhenhan, et al.
Published: (2025)
Consistency-guided Prompt Learning for Vision-Language Models
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Cascade Prompt Learning for Vision-Language Model Adaptation
by: Wu, Ge, et al.
Published: (2024)
by: Wu, Ge, et al.
Published: (2024)
On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
ConsensusDrop: Fusing Visual and Cross-Modal Saliency for Efficient Vision Language Models
by: Parikh, Dhruv, et al.
Published: (2026)
by: Parikh, Dhruv, et al.
Published: (2026)
MIRACLE3D: Memory-efficient Integrated Robust Approach for Continual Learning on Point Clouds via Shape Model Construction
by: Resani, Hossein, et al.
Published: (2024)
by: Resani, Hossein, et al.
Published: (2024)
Recent Advances of Continual Learning in Computer Vision: An Overview
by: Qu, Haoxuan, et al.
Published: (2021)
by: Qu, Haoxuan, et al.
Published: (2021)
Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models
by: Lai, Yuxiang, et al.
Published: (2025)
by: Lai, Yuxiang, et al.
Published: (2025)
Similar Items
-
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
by: Talemi, Niloufar Alipour, et al.
Published: (2024) -
FreqDebias: Towards Generalizable Deepfake Detection via Consistency-Driven Frequency Debiasing
by: Kashiani, Hossein, et al.
Published: (2025) -
ROADS: Robust Prompt-driven Multi-Class Anomaly Detection under Domain Shift
by: Kashiani, Hossein, et al.
Published: (2024) -
CATFace: Cross-Attribute-Guided Transformer with Self-Attention Distillation for Low-Quality Face Recognition
by: Talemi, Niloufar Alipour, et al.
Published: (2024) -
Agentic AI in Remote Sensing: Foundations, Taxonomy, and Emerging Systems
by: Talemi, Niloufar Alipour, et al.
Published: (2026)