Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Baron, Ethan, Tankel, Idan, Tu, Peter, Ben-Yosef, Guy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT
di: Tankel, Idan, et al.
Pubblicazione: (2025)
di: Tankel, Idan, et al.
Pubblicazione: (2025)
Common Data Properties Limit Object-Attribute Binding in CLIP
di: Gurung, Bijay, et al.
Pubblicazione: (2025)
di: Gurung, Bijay, et al.
Pubblicazione: (2025)
DesCLIP: Robust Continual Learning via General Attribute Descriptions for VLM-Based Visual Recognition
di: He, Chiyuan, et al.
Pubblicazione: (2025)
di: He, Chiyuan, et al.
Pubblicazione: (2025)
EchoNarrator: Generating natural text explanations for ejection fraction predictions
di: Thomas, Sarina, et al.
Pubblicazione: (2024)
di: Thomas, Sarina, et al.
Pubblicazione: (2024)
SuperCLIP: CLIP with Simple Classification Supervision
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
The Double-Ellipsoid Geometry of CLIP
di: Levi, Meir Yossef, et al.
Pubblicazione: (2024)
di: Levi, Meir Yossef, et al.
Pubblicazione: (2024)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
di: Wang, Xiao, et al.
Pubblicazione: (2023)
di: Wang, Xiao, et al.
Pubblicazione: (2023)
Adaptive Prototype Model for Attribute-based Multi-label Few-shot Action Recognition
di: Xiao, Juefeng, et al.
Pubblicazione: (2025)
di: Xiao, Juefeng, et al.
Pubblicazione: (2025)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
di: Kim, Yehna, et al.
Pubblicazione: (2025)
di: Kim, Yehna, et al.
Pubblicazione: (2025)
CLIP-FTI: Fine-Grained Face Template Inversion via CLIP-Driven Attribute Conditioning
di: Dai, Longchen, et al.
Pubblicazione: (2025)
di: Dai, Longchen, et al.
Pubblicazione: (2025)
Multimodal Multilabel Classification by CLIP
di: Guo, Yanming
Pubblicazione: (2024)
di: Guo, Yanming
Pubblicazione: (2024)
Emotion Recognition with CLIP and Sequential Learning
di: Zhou, Weiwei, et al.
Pubblicazione: (2025)
di: Zhou, Weiwei, et al.
Pubblicazione: (2025)
MoCLIP-Lite: Efficient Video Recognition by Fusing CLIP with Motion Vectors
di: Huang, Binhua, et al.
Pubblicazione: (2025)
di: Huang, Binhua, et al.
Pubblicazione: (2025)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
SANER: Annotation-free Societal Attribute Neutralizer for Debiasing CLIP
di: Hirota, Yusuke, et al.
Pubblicazione: (2024)
di: Hirota, Yusuke, et al.
Pubblicazione: (2024)
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
di: Vardi, Ben, et al.
Pubblicazione: (2025)
di: Vardi, Ben, et al.
Pubblicazione: (2025)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
di: Song, Dan, et al.
Pubblicazione: (2023)
di: Song, Dan, et al.
Pubblicazione: (2023)
CLIP-FSAC++: Few-Shot Anomaly Classification with Anomaly Descriptor Based on CLIP
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification
di: Ma, Yiming, et al.
Pubblicazione: (2024)
di: Ma, Yiming, et al.
Pubblicazione: (2024)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
di: Song, Yehun, et al.
Pubblicazione: (2025)
di: Song, Yehun, et al.
Pubblicazione: (2025)
Toward a Holistic Evaluation of Robustness in CLIP Models
di: Tu, Weijie, et al.
Pubblicazione: (2024)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
LaVPR: Benchmarking Language and Vision for Place Recognition
di: Idan, Ofer, et al.
Pubblicazione: (2026)
di: Idan, Ofer, et al.
Pubblicazione: (2026)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
PSA-MIL: A Probabilistic Spatial Attention-Based Multiple Instance Learning for Whole Slide Image Classification
di: Peled, Sharon, et al.
Pubblicazione: (2025)
di: Peled, Sharon, et al.
Pubblicazione: (2025)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
di: Liu, Mushui, et al.
Pubblicazione: (2024)
di: Liu, Mushui, et al.
Pubblicazione: (2024)
CLIP-Decoder : ZeroShot Multilabel Classification using Multimodal CLIP Aligned Representation
di: Ali, Muhammad, et al.
Pubblicazione: (2024)
di: Ali, Muhammad, et al.
Pubblicazione: (2024)
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
di: Wang, Jiapeng, et al.
Pubblicazione: (2024)
di: Wang, Jiapeng, et al.
Pubblicazione: (2024)
Are CLIP features all you need for Universal Synthetic Image Origin Attribution?
di: Cioni, Dario, et al.
Pubblicazione: (2024)
di: Cioni, Dario, et al.
Pubblicazione: (2024)
BlendCLIP: Bridging Synthetic and Real Domains for Zero-Shot 3D Object Classification with Multimodal Pretraining
di: Khoche, Ajinkya, et al.
Pubblicazione: (2025)
di: Khoche, Ajinkya, et al.
Pubblicazione: (2025)
Informative Sample Selection Model for Skeleton-based Action Recognition with Limited Training Samples
di: Tu, Zhigang, et al.
Pubblicazione: (2025)
di: Tu, Zhigang, et al.
Pubblicazione: (2025)
CLIP3D-AD: Extending CLIP for 3D Few-Shot Anomaly Detection with Multi-View Images Generation
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
Exploring Weak-to-Strong Generalization for CLIP-based Classification
di: Li, Jinhao, et al.
Pubblicazione: (2025)
di: Li, Jinhao, et al.
Pubblicazione: (2025)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
di: Pang, Li, et al.
Pubblicazione: (2025)
di: Pang, Li, et al.
Pubblicazione: (2025)
FairCLIP: Social Bias Elimination based on Attribute Prototype Learning and Representation Neutralization
di: Wang, Junyang, et al.
Pubblicazione: (2022)
di: Wang, Junyang, et al.
Pubblicazione: (2022)
Robust Light-Weight Facial Affective Behavior Recognition with CLIP
di: Lin, Li, et al.
Pubblicazione: (2024)
di: Lin, Li, et al.
Pubblicazione: (2024)
Exploring Open-Vocabulary Object Recognition in Images using CLIP
di: Chen, Wei Yu, et al.
Pubblicazione: (2026)
di: Chen, Wei Yu, et al.
Pubblicazione: (2026)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
di: Wang, Xiang, et al.
Pubblicazione: (2023)
di: Wang, Xiang, et al.
Pubblicazione: (2023)
Tell Me What You See: Text-Guided Real-World Image Denoising
di: Yosef, Erez, et al.
Pubblicazione: (2023)
di: Yosef, Erez, et al.
Pubblicazione: (2023)
Documenti analoghi
-
INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT
di: Tankel, Idan, et al.
Pubblicazione: (2025) -
Common Data Properties Limit Object-Attribute Binding in CLIP
di: Gurung, Bijay, et al.
Pubblicazione: (2025) -
DesCLIP: Robust Continual Learning via General Attribute Descriptions for VLM-Based Visual Recognition
di: He, Chiyuan, et al.
Pubblicazione: (2025) -
EchoNarrator: Generating natural text explanations for ejection fraction predictions
di: Thomas, Sarina, et al.
Pubblicazione: (2024) -
SuperCLIP: CLIP with Simple Classification Supervision
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)