Visual Modality Prompt for Adapting Vision-Language Object Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Medeiros, Heitor R., Belal, Atif, Muralidharan, Srikanth, Granger, Eric, Pedersoli, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors
von: Belal, Atif, et al.
Veröffentlicht: (2025)
von: Belal, Atif, et al.
Veröffentlicht: (2025)
WiSE-OD: Benchmarking Robustness in Infrared Object Detection
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2025)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2025)
Multi-Source Domain Adaptation for Object Detection with Prototype-based Mean-teacher
von: Belal, Atif, et al.
Veröffentlicht: (2023)
von: Belal, Atif, et al.
Veröffentlicht: (2023)
High-Rate Mixout: Revisiting Mixout for Robust Domain Generalization
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)
Attention-based Class-Conditioned Alignment for Multi-Source Domain Adaptation of Object Detectors
von: Belal, Atif, et al.
Veröffentlicht: (2024)
von: Belal, Atif, et al.
Veröffentlicht: (2024)
Modality Translation for Object Detection Adaptation Without Forgetting Prior Knowledge
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2024)
MiPa: Mixed Patch Infrared-Visible Modality Agnostic Object Detection
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
Source-Free Domain Adaptation for YOLO Object Detection
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
HalluciDet: Hallucinating RGB Modality for Person Detection Through Privileged Information
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
Revisiting Mixout: An Overlooked Path to Robust Finetuning
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
Low-Rank Expert Merging for Multi-Source Domain Adaptation in Person Re-Identification
von: Nehdi, Taha Mustapha, et al.
Veröffentlicht: (2025)
von: Nehdi, Taha Mustapha, et al.
Veröffentlicht: (2025)
Progressive Multi-Source Domain Adaptation for Personalized Facial Expression Recognition
von: Zeeshan, Muhammad Osama, et al.
Veröffentlicht: (2025)
von: Zeeshan, Muhammad Osama, et al.
Veröffentlicht: (2025)
A Realistic Protocol for Evaluation of Weakly Supervised Object Localization
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
Domain Generalization by Rejecting Extreme Augmentations
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2023)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2023)
TeD-Loc: Text Distillation for Weakly Supervised Object Localization
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2025)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2025)
Jailbreaking Vision-Language Models Through the Visual Modality
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
Distilling Privileged Multimodal Information for Expression Recognition using Optimal Transport
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
von: Aslam, Muhammad Haseeb, et al.
Veröffentlicht: (2024)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
von: Li, Xu, et al.
Veröffentlicht: (2025)
von: Li, Xu, et al.
Veröffentlicht: (2025)
Adapting Vision-Language Models for Evaluating World Models
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2025)
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2025)
Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
Visual Language Models as Zero-Shot Deepfake Detectors
von: Pirogov, Viacheslav
Veröffentlicht: (2025)
von: Pirogov, Viacheslav
Veröffentlicht: (2025)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
von: Lee, Jihoon, et al.
Veröffentlicht: (2025)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
Multi-Modal Adapter for Vision-Language Models
von: Seputis, Dominykas, et al.
Veröffentlicht: (2024)
von: Seputis, Dominykas, et al.
Veröffentlicht: (2024)
Adapting Lightweight Vision Language Models for Radiological Visual Question Answering
von: Shourya, Aditya, et al.
Veröffentlicht: (2025)
von: Shourya, Aditya, et al.
Veröffentlicht: (2025)
VIAssist: Adapting Multi-modal Large Language Models for Users with Visual Impairments
von: Yang, Bufang, et al.
Veröffentlicht: (2024)
von: Yang, Bufang, et al.
Veröffentlicht: (2024)
LT-Soups: Bridging Head and Tail Classes via Subsampled Model Soups
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Personalized Feature Translation for Expression Recognition: An Efficient Source-Free Domain Adaptation Method
von: Sharafi, Masoumeh, et al.
Veröffentlicht: (2025)
von: Sharafi, Masoumeh, et al.
Veröffentlicht: (2025)
Adversarially Trained Object Detector for Unsupervised Domain Adaptation
von: Fujii, Kazuma, et al.
Veröffentlicht: (2021)
von: Fujii, Kazuma, et al.
Veröffentlicht: (2021)
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
Biomed-DPT: Dual Modality Prompt Tuning for Biomedical Vision-Language Models
von: Peng, Wei, et al.
Veröffentlicht: (2025)
von: Peng, Wei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors
von: Belal, Atif, et al.
Veröffentlicht: (2025) -
WiSE-OD: Benchmarking Robustness in Infrared Object Detection
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2025) -
Multi-Source Domain Adaptation for Object Detection with Prototype-based Mean-teacher
von: Belal, Atif, et al.
Veröffentlicht: (2023) -
High-Rate Mixout: Revisiting Mixout for Robust Domain Generalization
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025) -
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
von: Muralidharan, Srikanth, et al.
Veröffentlicht: (2025)