CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Wahab, Haroon, Mehmood, Irfan, Ugail, Hassan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DFA-CON: A Contrastive Learning Approach for Detecting Copyright Infringement in DeepFake Art
by: Wahab, Haroon, et al.
Published: (2025)
by: Wahab, Haroon, et al.
Published: (2025)
Ensemble-Based Deepfake Detection using State-of-the-Art Models with Robust Cross-Dataset Generalisation
by: Wahab, Haroon, et al.
Published: (2025)
by: Wahab, Haroon, et al.
Published: (2025)
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification
by: Alawode, Basit, et al.
Published: (2024)
by: Alawode, Basit, et al.
Published: (2024)
Capsule Vision 2024 Challenge: Multi-Class Abnormality Classification for Video Capsule Endoscopy
by: Handa, Palak, et al.
Published: (2024)
by: Handa, Palak, et al.
Published: (2024)
Capsule Vision Challenge 2024: Multi-Class Abnormality Classification for Video Capsule Endoscopy
by: Bansal, Aakarsh, et al.
Published: (2024)
by: Bansal, Aakarsh, et al.
Published: (2024)
MoralCLIP: Contrastive Alignment of Vision-and-Language Representations with Moral Foundations Theory
by: Condez, Ana Carolina, et al.
Published: (2025)
by: Condez, Ana Carolina, et al.
Published: (2025)
Automated Bleeding Detection and Classification in Wireless Capsule Endoscopy with YOLOv8-X
by: Shekar, Pavan C, et al.
Published: (2024)
by: Shekar, Pavan C, et al.
Published: (2024)
A Highlight Removal Method for Capsule Endoscopy Images
by: Zhang, Shaojie, et al.
Published: (2024)
by: Zhang, Shaojie, et al.
Published: (2024)
CAVE-Net: Classifying Abnormalities in Video Capsule Endoscopy
by: Harish, Ishita, et al.
Published: (2024)
by: Harish, Ishita, et al.
Published: (2024)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
by: Lan, Mengcheng, et al.
Published: (2024)
by: Lan, Mengcheng, et al.
Published: (2024)
Integrating Visual and X-Ray Machine Learning Features in the Study of Paintings by Goya
by: Ugail, Hassan, et al.
Published: (2025)
by: Ugail, Hassan, et al.
Published: (2025)
Handcrafted Feature-Assisted One-Class Learning for Artist Authentication in Historical Drawings
by: Ugail, Hassan, et al.
Published: (2026)
by: Ugail, Hassan, et al.
Published: (2026)
PrunedCaps: A Case For Primary Capsules Discrimination
by: Sharifi, Ramin, et al.
Published: (2025)
by: Sharifi, Ramin, et al.
Published: (2025)
Capsule Endoscopy Image Enhancement for Small Intestinal Villi Clarity
by: Zhang, Shaojie, et al.
Published: (2024)
by: Zhang, Shaojie, et al.
Published: (2024)
RWKV-CLIP: A Robust Vision-Language Representation Learner
by: Gu, Tiancheng, et al.
Published: (2024)
by: Gu, Tiancheng, et al.
Published: (2024)
LE-CapsNet: A Light and Enhanced Capsule Network
by: Shiri, Pouya, et al.
Published: (2025)
by: Shiri, Pouya, et al.
Published: (2025)
DL-CapsNet: A Deep and Light Capsule Network
by: Shiri, Pouya, et al.
Published: (2025)
by: Shiri, Pouya, et al.
Published: (2025)
Enhanced Anomaly Detection for Capsule Endoscopy Using Ensemble Learning Strategies
by: Werner, Julia, et al.
Published: (2025)
by: Werner, Julia, et al.
Published: (2025)
Seeing More with Less: Video Capsule Endoscopy with Multi-Task Learning
by: Werner, Julia, et al.
Published: (2025)
by: Werner, Julia, et al.
Published: (2025)
Learning to Adapt Foundation Model DINOv2 for Capsule Endoscopy Diagnosis
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Convolutional Fully-Connected Capsule Network (CFC-CapsNet): A Novel and Fast Capsule Network
by: Shiri, Pouya, et al.
Published: (2025)
by: Shiri, Pouya, et al.
Published: (2025)
Clinical Evaluation of Medical Image Synthesis: A Case Study in Wireless Capsule Endoscopy
by: Gatoula, Panagiota, et al.
Published: (2024)
by: Gatoula, Panagiota, et al.
Published: (2024)
A Robust Pipeline for Classification and Detection of Bleeding Frames in Wireless Capsule Endoscopy using Swin Transformer and RT-DETR
by: Alavala, Sasidhar, et al.
Published: (2024)
by: Alavala, Sasidhar, et al.
Published: (2024)
Reliable Mislabel Detection for Video Capsule Endoscopy Data
by: Werner, Julia, et al.
Published: (2026)
by: Werner, Julia, et al.
Published: (2026)
Quick-CapsNet (QCN): A fast alternative to Capsule Networks
by: Shiri, Pouya, et al.
Published: (2025)
by: Shiri, Pouya, et al.
Published: (2025)
EndoOOD: Uncertainty-aware Out-of-distribution Detection in Capsule Endoscopy Diagnosis
by: Tan, Qiaozhi, et al.
Published: (2024)
by: Tan, Qiaozhi, et al.
Published: (2024)
V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder
by: Hu, Xiaoxing, et al.
Published: (2025)
by: Hu, Xiaoxing, et al.
Published: (2025)
Using Multi-Instance Learning to Identify Unique Polyps in Colon Capsule Endoscopy Images
by: Sharma, Puneet, et al.
Published: (2026)
by: Sharma, Puneet, et al.
Published: (2026)
ProtoCaps: A Fast and Non-Iterative Capsule Network Routing Method
by: Everett, Miles, et al.
Published: (2023)
by: Everett, Miles, et al.
Published: (2023)
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
by: Koleilat, Taha, et al.
Published: (2025)
by: Koleilat, Taha, et al.
Published: (2025)
ParseCaps: An Interpretable Parsing Capsule Network for Medical Image Diagnosis
by: Geng, Xinyu, et al.
Published: (2024)
by: Geng, Xinyu, et al.
Published: (2024)
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
by: Li, Yunheng, et al.
Published: (2024)
by: Li, Yunheng, et al.
Published: (2024)
$β$-CLIP: Text-Conditioned Contrastive Learning for Multi-Granular Vision-Language Alignment
by: Zohra, Fatimah, et al.
Published: (2025)
by: Zohra, Fatimah, et al.
Published: (2025)
CLIP-GS: Unifying Vision-Language Representation with 3D Gaussian Splatting
by: Jiao, Siyu, et al.
Published: (2024)
by: Jiao, Siyu, et al.
Published: (2024)
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
by: Wu, Ruijia, et al.
Published: (2025)
by: Wu, Ruijia, et al.
Published: (2025)
CapsoNet: A CNN-Transformer Ensemble for Multi-Class Abnormality Detection in Video Capsule Endoscopy
by: Samal, Arnav, et al.
Published: (2024)
by: Samal, Arnav, et al.
Published: (2024)
Image Compression with Bubble-Aware Frame Rate Adaptation for Energy-Efficient Video Capsule Endoscopy
by: Bause, Oliver, et al.
Published: (2026)
by: Bause, Oliver, et al.
Published: (2026)
EquiCaps: Predictor-Free Pose-Aware Pre-Trained Capsule Networks
by: Konstantinou, Athinoulla, et al.
Published: (2025)
by: Konstantinou, Athinoulla, et al.
Published: (2025)
CapStARE: Capsule-based Spatiotemporal Architecture for Robust and Efficient Gaze Estimation
by: Samaniego, Miren, et al.
Published: (2025)
by: Samaniego, Miren, et al.
Published: (2025)
Similar Items
-
DFA-CON: A Contrastive Learning Approach for Detecting Copyright Infringement in DeepFake Art
by: Wahab, Haroon, et al.
Published: (2025) -
Ensemble-Based Deepfake Detection using State-of-the-Art Models with Robust Cross-Dataset Generalisation
by: Wahab, Haroon, et al.
Published: (2025) -
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification
by: Alawode, Basit, et al.
Published: (2024) -
Capsule Vision 2024 Challenge: Multi-Class Abnormality Classification for Video Capsule Endoscopy
by: Handa, Palak, et al.
Published: (2024) -
Capsule Vision Challenge 2024: Multi-Class Abnormality Classification for Video Capsule Endoscopy
by: Bansal, Aakarsh, et al.
Published: (2024)