Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baron, Ethan, Tankel, Idan, Tu, Peter, Ben-Yosef, Guy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT
von: Tankel, Idan, et al.
Veröffentlicht: (2025)
von: Tankel, Idan, et al.
Veröffentlicht: (2025)
Common Data Properties Limit Object-Attribute Binding in CLIP
von: Gurung, Bijay, et al.
Veröffentlicht: (2025)
von: Gurung, Bijay, et al.
Veröffentlicht: (2025)
DesCLIP: Robust Continual Learning via General Attribute Descriptions for VLM-Based Visual Recognition
von: He, Chiyuan, et al.
Veröffentlicht: (2025)
von: He, Chiyuan, et al.
Veröffentlicht: (2025)
EchoNarrator: Generating natural text explanations for ejection fraction predictions
von: Thomas, Sarina, et al.
Veröffentlicht: (2024)
von: Thomas, Sarina, et al.
Veröffentlicht: (2024)
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
The Double-Ellipsoid Geometry of CLIP
von: Levi, Meir Yossef, et al.
Veröffentlicht: (2024)
von: Levi, Meir Yossef, et al.
Veröffentlicht: (2024)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
Adaptive Prototype Model for Attribute-based Multi-label Few-shot Action Recognition
von: Xiao, Juefeng, et al.
Veröffentlicht: (2025)
von: Xiao, Juefeng, et al.
Veröffentlicht: (2025)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
von: Kim, Yehna, et al.
Veröffentlicht: (2025)
von: Kim, Yehna, et al.
Veröffentlicht: (2025)
CLIP-FTI: Fine-Grained Face Template Inversion via CLIP-Driven Attribute Conditioning
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
Multimodal Multilabel Classification by CLIP
von: Guo, Yanming
Veröffentlicht: (2024)
von: Guo, Yanming
Veröffentlicht: (2024)
Emotion Recognition with CLIP and Sequential Learning
von: Zhou, Weiwei, et al.
Veröffentlicht: (2025)
von: Zhou, Weiwei, et al.
Veröffentlicht: (2025)
MoCLIP-Lite: Efficient Video Recognition by Fusing CLIP with Motion Vectors
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
von: Alyami, Sarah, et al.
Veröffentlicht: (2025)
von: Alyami, Sarah, et al.
Veröffentlicht: (2025)
SANER: Annotation-free Societal Attribute Neutralizer for Debiasing CLIP
von: Hirota, Yusuke, et al.
Veröffentlicht: (2024)
von: Hirota, Yusuke, et al.
Veröffentlicht: (2024)
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
von: Vardi, Ben, et al.
Veröffentlicht: (2025)
von: Vardi, Ben, et al.
Veröffentlicht: (2025)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
von: Song, Dan, et al.
Veröffentlicht: (2023)
von: Song, Dan, et al.
Veröffentlicht: (2023)
CLIP-FSAC++: Few-Shot Anomaly Classification with Anomaly Descriptor Based on CLIP
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification
von: Ma, Yiming, et al.
Veröffentlicht: (2024)
von: Ma, Yiming, et al.
Veröffentlicht: (2024)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
von: Song, Yehun, et al.
Veröffentlicht: (2025)
von: Song, Yehun, et al.
Veröffentlicht: (2025)
Toward a Holistic Evaluation of Robustness in CLIP Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
LaVPR: Benchmarking Language and Vision for Place Recognition
von: Idan, Ofer, et al.
Veröffentlicht: (2026)
von: Idan, Ofer, et al.
Veröffentlicht: (2026)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
PSA-MIL: A Probabilistic Spatial Attention-Based Multiple Instance Learning for Whole Slide Image Classification
von: Peled, Sharon, et al.
Veröffentlicht: (2025)
von: Peled, Sharon, et al.
Veröffentlicht: (2025)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
von: Ali, Muhammad, et al.
Veröffentlicht: (2024)
von: Ali, Muhammad, et al.
Veröffentlicht: (2024)
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
Are CLIP features all you need for Universal Synthetic Image Origin Attribution?
von: Cioni, Dario, et al.
Veröffentlicht: (2024)
von: Cioni, Dario, et al.
Veröffentlicht: (2024)
BlendCLIP: Bridging Synthetic and Real Domains for Zero-Shot 3D Object Classification with Multimodal Pretraining
von: Khoche, Ajinkya, et al.
Veröffentlicht: (2025)
von: Khoche, Ajinkya, et al.
Veröffentlicht: (2025)
Informative Sample Selection Model for Skeleton-based Action Recognition with Limited Training Samples
von: Tu, Zhigang, et al.
Veröffentlicht: (2025)
von: Tu, Zhigang, et al.
Veröffentlicht: (2025)
CLIP3D-AD: Extending CLIP for 3D Few-Shot Anomaly Detection with Multi-View Images Generation
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
von: Zuo, Zuo, et al.
Veröffentlicht: (2024)
Exploring Weak-to-Strong Generalization for CLIP-based Classification
von: Li, Jinhao, et al.
Veröffentlicht: (2025)
von: Li, Jinhao, et al.
Veröffentlicht: (2025)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
von: Pang, Li, et al.
Veröffentlicht: (2025)
von: Pang, Li, et al.
Veröffentlicht: (2025)
FairCLIP: Social Bias Elimination based on Attribute Prototype Learning and Representation Neutralization
von: Wang, Junyang, et al.
Veröffentlicht: (2022)
von: Wang, Junyang, et al.
Veröffentlicht: (2022)
Robust Light-Weight Facial Affective Behavior Recognition with CLIP
von: Lin, Li, et al.
Veröffentlicht: (2024)
von: Lin, Li, et al.
Veröffentlicht: (2024)
Exploring Open-Vocabulary Object Recognition in Images using CLIP
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
von: Wang, Xiang, et al.
Veröffentlicht: (2023)
von: Wang, Xiang, et al.
Veröffentlicht: (2023)
Tell Me What You See: Text-Guided Real-World Image Denoising
von: Yosef, Erez, et al.
Veröffentlicht: (2023)
von: Yosef, Erez, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT
von: Tankel, Idan, et al.
Veröffentlicht: (2025) -
Common Data Properties Limit Object-Attribute Binding in CLIP
von: Gurung, Bijay, et al.
Veröffentlicht: (2025) -
DesCLIP: Robust Continual Learning via General Attribute Descriptions for VLM-Based Visual Recognition
von: He, Chiyuan, et al.
Veröffentlicht: (2025) -
EchoNarrator: Generating natural text explanations for ejection fraction predictions
von: Thomas, Sarina, et al.
Veröffentlicht: (2024) -
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)