Koo-Fu CLIP: Closed-Form Adaptation of Vision-Language Models via Fukunaga-Koontz Linear Discriminant Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Suchanek, Matej, Janouskova, Klara, Vasatko, Ondrej, Matas, Jiri |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025)
by: Volkov, Illia, et al.
Published: (2025)
Single Image Test-Time Adaptation for Segmentation
by: Janouskova, Klara, et al.
Published: (2023)
by: Janouskova, Klara, et al.
Published: (2023)
Multimodal Large Language Models as Image Classifiers
by: Kisel, Nikita, et al.
Published: (2026)
by: Kisel, Nikita, et al.
Published: (2026)
Bringing the Context Back into Object Recognition, Robustly
by: Janouskova, Klara, et al.
Published: (2024)
by: Janouskova, Klara, et al.
Published: (2024)
Robust Context-Aware Object Recognition
by: Janouskova, Klara, et al.
Published: (2025)
by: Janouskova, Klara, et al.
Published: (2025)
Human Pose-Constrained UV Map Estimation
by: Suchanek, Matej, et al.
Published: (2025)
by: Suchanek, Matej, et al.
Published: (2025)
Flaws of ImageNet, Computer Vision's Favourite Dataset
by: Kisel, Nikita, et al.
Published: (2024)
by: Kisel, Nikita, et al.
Published: (2024)
FungiTastic: A multi-modal dataset and benchmark for image categorization
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
SAM2RL: Towards Reinforcement Learning Memory Control in Segment Anything Model 2
by: Adamyan, Alen, et al.
Published: (2025)
by: Adamyan, Alen, et al.
Published: (2025)
Unlocking the Hidden Potential of CLIP in Generalizable Deepfake Detection
by: Yermakov, Andrii, et al.
Published: (2025)
by: Yermakov, Andrii, et al.
Published: (2025)
Detection, Pose Estimation and Segmentation for Multiple Bodies: Closing the Virtuous Circle
by: Purkrabek, Miroslav, et al.
Published: (2024)
by: Purkrabek, Miroslav, et al.
Published: (2024)
A New Dataset and a Distractor-Aware Architecture for Transparent Object Tracking
by: Lukezic, Alan, et al.
Published: (2024)
by: Lukezic, Alan, et al.
Published: (2024)
BayesTTA: Continual-Temporal Test-Time Adaptation for Vision-Language Models via Gaussian Discriminant Analysis
by: Cui, Shuang, et al.
Published: (2025)
by: Cui, Shuang, et al.
Published: (2025)
Video shutter angle estimation using optical flow and linear blur
by: Korcak, David, et al.
Published: (2023)
by: Korcak, David, et al.
Published: (2023)
Accurate Planar Tracking With Robust Re-Detection
by: Serych, Jonas, et al.
Published: (2026)
by: Serych, Jonas, et al.
Published: (2026)
Improving 2D Human Pose Estimation in Rare Camera Views with Synthetic Data
by: Purkrabek, Miroslav, et al.
Published: (2023)
by: Purkrabek, Miroslav, et al.
Published: (2023)
ProbPose: A Probabilistic Approach to 2D Human Pose Estimation
by: Purkrabek, Miroslav, et al.
Published: (2024)
by: Purkrabek, Miroslav, et al.
Published: (2024)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
by: Yang, Kaicheng, et al.
Published: (2024)
by: Yang, Kaicheng, et al.
Published: (2024)
Animal Identification with Independent Foreground and Background Modeling
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
by: Koleilat, Taha, et al.
Published: (2025)
by: Koleilat, Taha, et al.
Published: (2025)
Test-Time Adaptation with CLIP Reward for Zero-Shot Generalization in Vision-Language Models
by: Zhao, Shuai, et al.
Published: (2023)
by: Zhao, Shuai, et al.
Published: (2023)
Closed-Form Linear-Probe Dataset Distillation for Pre-trained Vision Models
by: Peng, Bincheng, et al.
Published: (2026)
by: Peng, Bincheng, et al.
Published: (2026)
NEARL-CLIP: Interacted Query Adaptation with Orthogonal Regularization for Medical Vision-Language Understanding
by: Peng, Zelin, et al.
Published: (2025)
by: Peng, Zelin, et al.
Published: (2025)
Closing the Confusion Loop: CLIP-Guided Alignment for Source-Free Domain Adaptation
by: Wang, Shanshan, et al.
Published: (2026)
by: Wang, Shanshan, et al.
Published: (2026)
SAM-pose2seg: Pose-Guided Human Instance Segmentation in Crowds
by: Kolomiiets, Constantin, et al.
Published: (2026)
by: Kolomiiets, Constantin, et al.
Published: (2026)
Dense Matchers for Dense Tracking
by: Jelínek, Tomáš, et al.
Published: (2024)
by: Jelínek, Tomáš, et al.
Published: (2024)
MFTIQ: Multi-Flow Tracker with Independent Matching Quality Estimation
by: Serych, Jonas, et al.
Published: (2024)
by: Serych, Jonas, et al.
Published: (2024)
PixOOD: Pixel-Level Out-of-Distribution Detection
by: Vojíř, Tomáš, et al.
Published: (2024)
by: Vojíř, Tomáš, et al.
Published: (2024)
BBoxMaskPose v2: Expanding Mutual Conditioning to 3D
by: Purkrabek, Miroslav, et al.
Published: (2026)
by: Purkrabek, Miroslav, et al.
Published: (2026)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
by: Alyami, Sarah, et al.
Published: (2025)
by: Alyami, Sarah, et al.
Published: (2025)
Q-CLIP: Unleashing the Power of Vision-Language Models for Video Quality Assessment through Unified Cross-Modal Adaptation
by: Mi, Yachun, et al.
Published: (2025)
by: Mi, Yachun, et al.
Published: (2025)
MedProbCLIP: Probabilistic Adaptation of Vision-Language Foundation Model for Reliable Radiograph-Report Retrieval
by: Elallaf, Ahmad, et al.
Published: (2026)
by: Elallaf, Ahmad, et al.
Published: (2026)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
by: Lan, Mengcheng, et al.
Published: (2024)
by: Lan, Mengcheng, et al.
Published: (2024)
A Closed-Form Solution for Debiasing Vision-Language Models with Utility Guarantees Across Modalities and Tasks
by: Lian, Tangzheng, et al.
Published: (2026)
by: Lian, Tangzheng, et al.
Published: (2026)
TripletCLIP: Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives
by: Patel, Maitreya, et al.
Published: (2024)
by: Patel, Maitreya, et al.
Published: (2024)
CLIP the Divergence: Language-guided Unsupervised Domain Adaptation
by: Zhu, Jinjing, et al.
Published: (2024)
by: Zhu, Jinjing, et al.
Published: (2024)
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
by: Alawode, Basit, et al.
Published: (2025)
by: Alawode, Basit, et al.
Published: (2025)
AF-CLIP: Zero-Shot Anomaly Detection via Anomaly-Focused CLIP Adaptation
by: Fang, Qingqing, et al.
Published: (2025)
by: Fang, Qingqing, et al.
Published: (2025)
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels
by: Singh, Darshan, et al.
Published: (2024)
by: Singh, Darshan, et al.
Published: (2024)
Three Things to Know about Deep Metric Learning
by: Patel, Yash, et al.
Published: (2024)
by: Patel, Yash, et al.
Published: (2024)
Similar Items
-
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025) -
Single Image Test-Time Adaptation for Segmentation
by: Janouskova, Klara, et al.
Published: (2023) -
Multimodal Large Language Models as Image Classifiers
by: Kisel, Nikita, et al.
Published: (2026) -
Bringing the Context Back into Object Recognition, Robustly
by: Janouskova, Klara, et al.
Published: (2024) -
Robust Context-Aware Object Recognition
by: Janouskova, Klara, et al.
Published: (2025)