CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Javed, Sajid, Mahmood, Arif, Ganapathi, Iyyakutti Iyappan, Dharejo, Fayaz Ali, Werghi, Naoufel, Bennamoun, Mohammed |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
par: Alawode, Basit, et autres
Publié: (2025)
par: Alawode, Basit, et autres
Publié: (2025)
Advancing Histopathology with Deep Learning Under Data Scarcity: A Decade in Review
par: Obeid, Ahmad, et autres
Publié: (2024)
par: Obeid, Ahmad, et autres
Publié: (2024)
Multi-Resolution Pathology-Language Pre-training Model with Text-Guided Visual Representation
par: Albastaki, Shahad, et autres
Publié: (2025)
par: Albastaki, Shahad, et autres
Publié: (2025)
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
par: Xu, Kepeng, et autres
Publié: (2024)
par: Xu, Kepeng, et autres
Publié: (2024)
Multi-Modal Attention Networks for Enhanced Segmentation and Depth Estimation of Subsurface Defects in Pulse Thermography
par: Salah, Mohammed, et autres
Publié: (2025)
par: Salah, Mohammed, et autres
Publié: (2025)
Off-the-shelf Vision Models Benefit Image Manipulation Localization
par: Zhang, Zhengxuan, et autres
Publié: (2026)
par: Zhang, Zhengxuan, et autres
Publié: (2026)
NIC-RobustBench: A Comprehensive Open-Source Toolkit for Neural Image Compression and Robustness Analysis
par: Bychkov, Georgii, et autres
Publié: (2025)
par: Bychkov, Georgii, et autres
Publié: (2025)
Predicting the Best of N Visual Trackers
par: Alawode, Basit, et autres
Publié: (2024)
par: Alawode, Basit, et autres
Publié: (2024)
A Near-Raw Talking-Head Video Dataset for Various Computer Vision Tasks
par: Naderi, Babak, et autres
Publié: (2026)
par: Naderi, Babak, et autres
Publié: (2026)
ARIQA-3DS: A Stereoscopic Image Quality Assessment Dataset for Realistic Augmented Reality
par: Sekhri, Aymen, et autres
Publié: (2026)
par: Sekhri, Aymen, et autres
Publié: (2026)
DIGITWISE: Digital Twin-based Modeling of Adaptive Video Streaming Engagement
par: Artioli, Emanuele, et autres
Publié: (2025)
par: Artioli, Emanuele, et autres
Publié: (2025)
Low-complexity 8-point DCT Approximation Based on Angle Similarity for Image and Video Coding
par: Oliveira, R. S., et autres
Publié: (2018)
par: Oliveira, R. S., et autres
Publié: (2018)
Subjective assessment of the impact of a content adaptive optimiser for compressing 4K HDR content with AV1
par: Vibhoothi, et autres
Publié: (2023)
par: Vibhoothi, et autres
Publié: (2023)
Out-Of-Distribution Detection for Audio-visual Generalized Zero-Shot Learning: A General Framework
par: Wen, Liuyuan
Publié: (2024)
par: Wen, Liuyuan
Publié: (2024)
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes
par: Zhai, Jiajun, et autres
Publié: (2026)
par: Zhai, Jiajun, et autres
Publié: (2026)
STING-BEE: Towards Vision-Language Model for Real-World X-ray Baggage Security Inspection
par: Velayudhan, Divya, et autres
Publié: (2025)
par: Velayudhan, Divya, et autres
Publié: (2025)
SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation
par: Lu, Zhenyu, et autres
Publié: (2026)
par: Lu, Zhenyu, et autres
Publié: (2026)
CLDTracker: A Comprehensive Language Description for Visual Tracking
par: Alansari, Mohamad, et autres
Publié: (2025)
par: Alansari, Mohamad, et autres
Publié: (2025)
Deep Video Codec Control for Vision Models
par: Reich, Christoph, et autres
Publié: (2023)
par: Reich, Christoph, et autres
Publié: (2023)
TIACam: Text-Anchored Invariant Feature Learning with Auto-Augmentation for Camera-Robust Zero-Watermarking
par: Tanvir, Abdullah All, et autres
Publié: (2026)
par: Tanvir, Abdullah All, et autres
Publié: (2026)
Scaling Up Single Image Dehazing Algorithm by Cross-Data Vision Alignment for Richer Representation Learning and Beyond
par: Shi, Yukai, et autres
Publié: (2024)
par: Shi, Yukai, et autres
Publié: (2024)
From Darkness to Detail: Frequency-Aware SSMs for Low-Light Vision
par: Adhikarla, Eashan, et autres
Publié: (2024)
par: Adhikarla, Eashan, et autres
Publié: (2024)
In-Loop Filtering Using Learned Look-Up Tables for Video Coding
par: Li, Zhuoyuan, et autres
Publié: (2025)
par: Li, Zhuoyuan, et autres
Publié: (2025)
T2IW: Joint Text to Image & Watermark Generation
par: Liu, An-An, et autres
Publié: (2023)
par: Liu, An-An, et autres
Publié: (2023)
Resolution limit of the eye: how many pixels can we see?
par: Ashraf, Maliha, et autres
Publié: (2024)
par: Ashraf, Maliha, et autres
Publié: (2024)
JND-Guided Light-Weight Neural Pre-Filter for Perceptual Image Coding
par: He, Chenlong, et autres
Publié: (2025)
par: He, Chenlong, et autres
Publié: (2025)
UPDA: Unsupervised Progressive Domain Adaptation for No-Reference Point Cloud Quality Assessment
par: Xie, Bingxu, et autres
Publié: (2026)
par: Xie, Bingxu, et autres
Publié: (2026)
Surveillance Facial Image Quality Assessment: A Multi-dimensional Dataset and Lightweight Model
par: Jiang, Yanwei, et autres
Publié: (2026)
par: Jiang, Yanwei, et autres
Publié: (2026)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
par: Chen, Tung-I, et autres
Publié: (2026)
par: Chen, Tung-I, et autres
Publié: (2026)
Hierarchical Prior-based Super Resolution for Point Cloud Geometry Compression
par: Li, Dingquan, et autres
Publié: (2024)
par: Li, Dingquan, et autres
Publié: (2024)
Perceptual Learned Image Compression via End-to-End JND-Based Optimization
par: Pakdaman, Farhad, et autres
Publié: (2024)
par: Pakdaman, Farhad, et autres
Publié: (2024)
Joint End-to-End Image Compression and Denoising: Leveraging Contrastive Learning and Multi-Scale Self-ONNs
par: Xie, Yuxin, et autres
Publié: (2024)
par: Xie, Yuxin, et autres
Publié: (2024)
EchoSR: Efficient Context Harnessing for Lightweight Image Super-Resolution
par: Zhao, Hanli, et autres
Publié: (2026)
par: Zhao, Hanli, et autres
Publié: (2026)
Automated Retinal Image Analysis and Medical Report Generation through Deep Learning
par: Huang, Jia-Hong
Publié: (2024)
par: Huang, Jia-Hong
Publié: (2024)
Perceptual Depth Quality Assessment of Stereoscopic Omnidirectional Images
par: Zhou, Wei, et autres
Publié: (2024)
par: Zhou, Wei, et autres
Publié: (2024)
Exploiting Frequency Correlation for Hyperspectral Image Reconstruction
par: Yan, Muge, et autres
Publié: (2024)
par: Yan, Muge, et autres
Publié: (2024)
Analysis of Video Quality Datasets via Design of Minimalistic Video Quality Models
par: Sun, Wei, et autres
Publié: (2023)
par: Sun, Wei, et autres
Publié: (2023)
Temporal Inconsistency Guidance for Super-resolution Video Quality Assessment
par: Li, Yixiao, et autres
Publié: (2024)
par: Li, Yixiao, et autres
Publié: (2024)
Brain-Grasp: Graph-based Saliency Priors for Improved fMRI-based Visual Brain Decoding
par: Moradi, Mohammad, et autres
Publié: (2026)
par: Moradi, Mohammad, et autres
Publié: (2026)
A Dynamic Prognostic Prediction Method for Colorectal Cancer Liver Metastasis
par: Yang, Wei, et autres
Publié: (2025)
par: Yang, Wei, et autres
Publié: (2025)
Documents similaires
-
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
par: Alawode, Basit, et autres
Publié: (2025) -
Advancing Histopathology with Deep Learning Under Data Scarcity: A Decade in Review
par: Obeid, Ahmad, et autres
Publié: (2024) -
Multi-Resolution Pathology-Language Pre-training Model with Text-Guided Visual Representation
par: Albastaki, Shahad, et autres
Publié: (2025) -
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
par: Xu, Kepeng, et autres
Publié: (2024) -
Multi-Modal Attention Networks for Enhanced Segmentation and Depth Estimation of Subsurface Defects in Pulse Thermography
par: Salah, Mohammed, et autres
Publié: (2025)