MulCLIP: A Multi-level Alignment Framework for Enhancing Fine-grained Long-context CLIP
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Truong, Chau, Quang, Hieu Ta, Le, Dung D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
von: Nguyen, Nghia, et al.
Veröffentlicht: (2024)
von: Nguyen, Nghia, et al.
Veröffentlicht: (2024)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
von: Zhang, Beichen, et al.
Veröffentlicht: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
FG-CLIP: Fine-Grained Visual and Textual Alignment
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
Taming CLIP for Fine-grained and Structured Visual Understanding of Museum Exhibits
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
MARBLE: Material Recomposition and Blending in CLIP-Space
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
von: Chen, Haodong, et al.
Veröffentlicht: (2024)
von: Chen, Haodong, et al.
Veröffentlicht: (2024)
GeoAlignCLIP: Enhancing Fine-Grained Vision-Language Alignment in Remote Sensing via Multi-Granular Consistency Learning
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
von: Maldonado, Gabriel, et al.
Veröffentlicht: (2025)
von: Maldonado, Gabriel, et al.
Veröffentlicht: (2025)
MedSAD-CLIP: Supervised CLIP with Token-Patch Cross-Attention for Medical Anomaly Detection and Segmentation
von: Tran, Thuy Truong, et al.
Veröffentlicht: (2026)
von: Tran, Thuy Truong, et al.
Veröffentlicht: (2026)
Breaking the Limits of Open-Weight CLIP: An Optimization Framework for Self-supervised Fine-tuning of CLIP
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
PixCLIP: Achieving Fine-grained Visual Language Understanding via Any-granularity Pixel-Text Alignment Learning
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
AgriCLIP: Adapting CLIP for Agriculture and Livestock via Domain-Specialized Cross-Model Alignment
von: Nawaz, Umair, et al.
Veröffentlicht: (2024)
von: Nawaz, Umair, et al.
Veröffentlicht: (2024)
CLIP-FTI: Fine-Grained Face Template Inversion via CLIP-Driven Attribute Conditioning
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
von: Ali, Eman, et al.
Veröffentlicht: (2025)
von: Ali, Eman, et al.
Veröffentlicht: (2025)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
von: Liu, Yanqing, et al.
Veröffentlicht: (2024)
von: Liu, Yanqing, et al.
Veröffentlicht: (2024)
Enhancing CLIP Robustness via Cross-Modality Alignment
von: Zhu, Xingyu, et al.
Veröffentlicht: (2025)
von: Zhu, Xingyu, et al.
Veröffentlicht: (2025)
Omni-NegCLIP: Enhancing CLIP with Front-Layer Contrastive Fine-Tuning for Comprehensive Negation Understanding
von: Xu, Jingqi
Veröffentlicht: (2026)
von: Xu, Jingqi
Veröffentlicht: (2026)
CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification
von: Ma, Yiming, et al.
Veröffentlicht: (2024)
von: Ma, Yiming, et al.
Veröffentlicht: (2024)
Semantic-aware Adversarial Fine-tuning for CLIP
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2026)
CLIP Under the Microscope: A Fine-Grained Analysis of Multi-Object Representation
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
von: Abbasi, Reza, et al.
Veröffentlicht: (2025)
VITRIX-CLIPIN: Enhancing Fine-Grained Visual Understanding in CLIP via Instruction Editing Data and Long Captions
von: Wang, Ziteng, et al.
Veröffentlicht: (2025)
von: Wang, Ziteng, et al.
Veröffentlicht: (2025)
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
PowerCLIP: Powerset Alignment for Contrastive Pre-Training
von: Kawamura, Masaki, et al.
Veröffentlicht: (2025)
von: Kawamura, Masaki, et al.
Veröffentlicht: (2025)
DeCLIP: Decoupled Prompting for CLIP-based Multi-Label Class-Incremental Learning
von: Du, Kaile, et al.
Veröffentlicht: (2025)
von: Du, Kaile, et al.
Veröffentlicht: (2025)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
von: Song, Dan, et al.
Veröffentlicht: (2023)
von: Song, Dan, et al.
Veröffentlicht: (2023)
Is CLIP the main roadblock for fine-grained open-world perception?
von: Bianchi, Lorenzo, et al.
Veröffentlicht: (2024)
von: Bianchi, Lorenzo, et al.
Veröffentlicht: (2024)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
CleanerCLIP: Fine-grained Counterfactual Semantic Augmentation for Backdoor Defense in Contrastive Learning
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
DiCLIP: Diffusion Model Enhances CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
von: Yang, Zhiwei, et al.
Veröffentlicht: (2026)
von: Yang, Zhiwei, et al.
Veröffentlicht: (2026)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
von: Karimian, Banafsheh, et al.
Veröffentlicht: (2025)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
von: Silva, Sathira, et al.
Veröffentlicht: (2025)
von: Silva, Sathira, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
von: Nguyen, Nghia, et al.
Veröffentlicht: (2024) -
Long-CLIP: Unlocking the Long-Text Capability of CLIP
von: Zhang, Beichen, et al.
Veröffentlicht: (2024) -
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024) -
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024) -
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)