O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Sharifdeen, Ashshak, Munir, Muhammad Akhtar, Baliah, Sanoojan, Khan, Salman, Khan, Muhammad Haris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Calibrating Prompt Tuning of Vision-Language Models
por: Sharifdeen, Ashshak, et al.
Publicado: (2026)
por: Sharifdeen, Ashshak, et al.
Publicado: (2026)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
por: Basu, Abhishek, et al.
Publicado: (2025)
por: Basu, Abhishek, et al.
Publicado: (2025)
A-TPT: Angular Diversity Calibration Properties for Test-Time Prompt Tuning of Vision-Language Models
por: Ahamed, Shihab Aaqil, et al.
Publicado: (2025)
por: Ahamed, Shihab Aaqil, et al.
Publicado: (2025)
Towards Generalizing to Unseen Domains with Few Labels
por: Galappaththige, Chamuditha Jayanga, et al.
Publicado: (2024)
por: Galappaththige, Chamuditha Jayanga, et al.
Publicado: (2024)
Synergistic Neural Forecasting of Air Pollution with Stochastic Sampling
por: Abeysinghe, Yohan, et al.
Publicado: (2025)
por: Abeysinghe, Yohan, et al.
Publicado: (2025)
Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models
por: Baliah, Sanoojan, et al.
Publicado: (2024)
por: Baliah, Sanoojan, et al.
Publicado: (2024)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
por: Baliah, Sanoojan, et al.
Publicado: (2026)
por: Baliah, Sanoojan, et al.
Publicado: (2026)
MetaTPT: Meta Test-time Prompt Tuning for Vision-Language Models
por: Lei, Yuqing, et al.
Publicado: (2025)
por: Lei, Yuqing, et al.
Publicado: (2025)
D-TPT: Dimensional Entropy Maximization for Calibrating Test-Time Prompt Tuning in Vision-Language Models
por: Han, Jisu, et al.
Publicado: (2025)
por: Han, Jisu, et al.
Publicado: (2025)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
por: Danish, Muhammad Sohail, et al.
Publicado: (2025)
por: Danish, Muhammad Sohail, et al.
Publicado: (2025)
Agentic AI for Remote Sensing: Technical Challenges and Research Directions
por: Munir, Muhammad Akhtar, et al.
Publicado: (2026)
por: Munir, Muhammad Akhtar, et al.
Publicado: (2026)
GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks
por: Danish, Muhammad Sohail, et al.
Publicado: (2024)
por: Danish, Muhammad Sohail, et al.
Publicado: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
por: Ali, Eman, et al.
Publicado: (2023)
por: Ali, Eman, et al.
Publicado: (2023)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
por: Danish, Muhammad Sohail, et al.
Publicado: (2024)
por: Danish, Muhammad Sohail, et al.
Publicado: (2024)
ThinkGeo: Evaluating Tool-Augmented Agents for Remote Sensing Tasks
por: Shabbir, Akashah, et al.
Publicado: (2025)
por: Shabbir, Akashah, et al.
Publicado: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
por: Ali, Eman, et al.
Publicado: (2024)
por: Ali, Eman, et al.
Publicado: (2024)
C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion
por: Yoon, Hee Suk, et al.
Publicado: (2024)
por: Yoon, Hee Suk, et al.
Publicado: (2024)
OpenEarthAgent: A Unified Framework for Tool-Augmented Geospatial Agents
por: Shabbir, Akashah, et al.
Publicado: (2026)
por: Shabbir, Akashah, et al.
Publicado: (2026)
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
por: Maaz, Muhammad, et al.
Publicado: (2023)
por: Maaz, Muhammad, et al.
Publicado: (2023)
Robust and Label-Efficient Deep Waste Detection
por: Abid, Hassan, et al.
Publicado: (2025)
por: Abid, Hassan, et al.
Publicado: (2025)
R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning
por: Sheng, Lijun, et al.
Publicado: (2025)
por: Sheng, Lijun, et al.
Publicado: (2025)
CountZES: Counting via Zero-Shot Exemplar Selection
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
por: Siddiqui, Muhammad Ibraheem, et al.
Publicado: (2025)
TLAC: Two-stage LMM Augmented CLIP for Zero-Shot Classification
por: Munir, Ans, et al.
Publicado: (2025)
por: Munir, Ans, et al.
Publicado: (2025)
Compositional Zero-Shot Learning: A Survey
por: Munir, Ans, et al.
Publicado: (2025)
por: Munir, Ans, et al.
Publicado: (2025)
Align Your Prompts: Test-Time Prompting with Distribution Alignment for Zero-Shot Generalization
por: Hassan, Jameel, et al.
Publicado: (2023)
por: Hassan, Jameel, et al.
Publicado: (2023)
Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models
por: Maaz, Muhammad, et al.
Publicado: (2025)
por: Maaz, Muhammad, et al.
Publicado: (2025)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
por: Ali, Muhammad, et al.
Publicado: (2024)
por: Ali, Muhammad, et al.
Publicado: (2024)
Underwater Object Detection Enhancement via Channel Stabilization
por: Ali, Muhammad, et al.
Publicado: (2024)
por: Ali, Muhammad, et al.
Publicado: (2024)
SoC: Semantic Orthogonal Calibration for Test-Time Prompt Tuning
por: Fillioux, Leo, et al.
Publicado: (2026)
por: Fillioux, Leo, et al.
Publicado: (2026)
Attention Based Simple Primitives for Open World Compositional Zero-Shot Learning
por: Munir, Ans, et al.
Publicado: (2024)
por: Munir, Ans, et al.
Publicado: (2024)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
por: Khan, Adnan, et al.
Publicado: (2024)
por: Khan, Adnan, et al.
Publicado: (2024)
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
por: Hanif, Asif, et al.
Publicado: (2024)
por: Hanif, Asif, et al.
Publicado: (2024)
DEFT: Decompositional Efficient Fine-Tuning for Text-to-Image Models
por: Kumar, Komal, et al.
Publicado: (2025)
por: Kumar, Komal, et al.
Publicado: (2025)
Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments
por: Ali, Muhammad, et al.
Publicado: (2025)
por: Ali, Muhammad, et al.
Publicado: (2025)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
por: Liu, Yu, et al.
Publicado: (2024)
por: Liu, Yu, et al.
Publicado: (2024)
Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model
por: Chen, Shiming, et al.
Publicado: (2025)
por: Chen, Shiming, et al.
Publicado: (2025)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
por: Chokuwa, Sharon, et al.
Publicado: (2024)
por: Chokuwa, Sharon, et al.
Publicado: (2024)
VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
por: Maaz, Muhammad, et al.
Publicado: (2024)
por: Maaz, Muhammad, et al.
Publicado: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
por: Gunduboina, Hariseetharam, et al.
Publicado: (2025)
por: Gunduboina, Hariseetharam, et al.
Publicado: (2025)
AirCast: Improving Air Pollution Forecasting Through Multi-Variable Data Alignment
por: Nedungadi, Vishal, et al.
Publicado: (2025)
por: Nedungadi, Vishal, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Calibrating Prompt Tuning of Vision-Language Models
por: Sharifdeen, Ashshak, et al.
Publicado: (2026) -
Calibration-Aware Prompt Learning for Medical Vision-Language Models
por: Basu, Abhishek, et al.
Publicado: (2025) -
A-TPT: Angular Diversity Calibration Properties for Test-Time Prompt Tuning of Vision-Language Models
por: Ahamed, Shihab Aaqil, et al.
Publicado: (2025) -
Towards Generalizing to Unseen Domains with Few Labels
por: Galappaththige, Chamuditha Jayanga, et al.
Publicado: (2024) -
Synergistic Neural Forecasting of Air Pollution with Stochastic Sampling
por: Abeysinghe, Yohan, et al.
Publicado: (2025)