Multi-Resolution Pathology-Language Pre-training Model with Text-Guided Visual Representation
Fuente:
arXiv
Saved in:
| Main Authors: | Albastaki, Shahad, Sohail, Anabia, Ganapathi, Iyyakutti Iyappan, Alawode, Basit, Khan, Asim, Javed, Sajid, Werghi, Naoufel, Bennamoun, Mohammed, Mahmood, Arif |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
by: Alawode, Basit, et al.
Published: (2025)
by: Alawode, Basit, et al.
Published: (2025)
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
by: Javed, Sajid, et al.
Published: (2024)
by: Javed, Sajid, et al.
Published: (2024)
MLLM-HWSI: A Multimodal Large Language Model for Hierarchical Whole Slide Image Understanding
by: Alawode, Basit, et al.
Published: (2026)
by: Alawode, Basit, et al.
Published: (2026)
Advancing Histopathology with Deep Learning Under Data Scarcity: A Decade in Review
by: Obeid, Ahmad, et al.
Published: (2024)
by: Obeid, Ahmad, et al.
Published: (2024)
Predicting the Best of N Visual Trackers
by: Alawode, Basit, et al.
Published: (2024)
by: Alawode, Basit, et al.
Published: (2024)
CLDTracker: A Comprehensive Language Description for Visual Tracking
by: Alansari, Mohamad, et al.
Published: (2025)
by: Alansari, Mohamad, et al.
Published: (2025)
DyCON: Dynamic Uncertainty-aware Consistency and Contrastive Learning for Semi-supervised Medical Image Segmentation
by: Assefa, Maregu, et al.
Published: (2025)
by: Assefa, Maregu, et al.
Published: (2025)
Rethinking Memory Design in SAM-Based Visual Object Tracking
by: Alansari, Mohamad, et al.
Published: (2025)
by: Alansari, Mohamad, et al.
Published: (2025)
Video Anomaly Detection in 10 Years: A Survey and Outlook
by: Abdalla, Moshira, et al.
Published: (2024)
by: Abdalla, Moshira, et al.
Published: (2024)
Cytoplasmic Strings Analysis in Human Embryo Time-Lapse Videos using Deep Learning Framework
by: Sohail, Anabia, et al.
Published: (2025)
by: Sohail, Anabia, et al.
Published: (2025)
Implicit to Explicit Entropy Regularization: Benchmarking ViT Fine-tuning under Noisy Labels
by: Marrium, Maria, et al.
Published: (2024)
by: Marrium, Maria, et al.
Published: (2024)
SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs
by: Alansari, Mohamad, et al.
Published: (2026)
by: Alansari, Mohamad, et al.
Published: (2026)
SHREC'25 Track on Multiple Relief Patterns: Report and Analysis
by: Paolini, Gabriele, et al.
Published: (2025)
by: Paolini, Gabriele, et al.
Published: (2025)
Dense-Sparse Deep Convolutional Neural Networks Training for Image Denoising
by: Alawode, Basit O., et al.
Published: (2021)
by: Alawode, Basit O., et al.
Published: (2021)
Multi-Modal Attention Networks for Enhanced Segmentation and Depth Estimation of Subsurface Defects in Pulse Thermography
by: Salah, Mohammed, et al.
Published: (2025)
by: Salah, Mohammed, et al.
Published: (2025)
AdaRD-key: Adaptive Relevance-Diversity Keyframe Sampling for Long-form Video understanding
by: Zhang, Xian, et al.
Published: (2025)
by: Zhang, Xian, et al.
Published: (2025)
BENet: A Cross-domain Robust Network for Detecting Face Forgeries via Bias Expansion and Latent-space Attention
by: Liu, Weihua, et al.
Published: (2024)
by: Liu, Weihua, et al.
Published: (2024)
Tomato Maturity Recognition with Convolutional Transformers
by: Khan, Asim, et al.
Published: (2023)
by: Khan, Asim, et al.
Published: (2023)
Transformer-Based Wireless Capsule Endoscopy Bleeding Tissue Detection and Classification
by: Alawode, Basit, et al.
Published: (2024)
by: Alawode, Basit, et al.
Published: (2024)
Accurate and Efficient Urban Street Tree Inventory with Deep Learning on Mobile Phone Imagery
by: Khan, Asim, et al.
Published: (2024)
by: Khan, Asim, et al.
Published: (2024)
RING‐Box E3 Ligase Target N‐Terminal Lysine 55 to Regulate Turnover of Sp7 Protein
by: Abeera Sikandar, et al.
Published: (2025)
by: Abeera Sikandar, et al.
Published: (2025)
Multiple Cases‐Based Learning in Oral Pathology
by: Shahad A. Waheed
Published: (2025)
by: Shahad A. Waheed
Published: (2025)
RobMOT: Robust 3D Multi-Object Tracking by Observational Noise and State Estimation Drift Mitigation on LiDAR PointCloud
by: Nagy, Mohamed, et al.
Published: (2024)
by: Nagy, Mohamed, et al.
Published: (2024)
Towards Accurate State Estimation: Kalman Filter Incorporating Motion Dynamics for 3D Multi-Object Tracking
by: Nagy, Mohamed, et al.
Published: (2025)
by: Nagy, Mohamed, et al.
Published: (2025)
A survey of the Vision Transformers and their CNN-Transformer based Variants
by: Khan, Asifullah, et al.
Published: (2023)
by: Khan, Asifullah, et al.
Published: (2023)
GeoPixel: Pixel Grounding Large Multimodal Model in Remote Sensing
by: Shabbir, Akashah, et al.
Published: (2025)
by: Shabbir, Akashah, et al.
Published: (2025)
QA-HFL: Quality-Aware Hierarchical Federated Learning for Resource-Constrained Mobile Devices with Heterogeneous Image Quality
by: Hussain, Sajid, et al.
Published: (2025)
by: Hussain, Sajid, et al.
Published: (2025)
SEMFED: Semantic-Aware Resource-Efficient Federated Learning for Heterogeneous NLP Tasks
by: Hussain, Sajid, et al.
Published: (2025)
by: Hussain, Sajid, et al.
Published: (2025)
A Robust Adversary Detection-Deactivation Method for Metaverse-oriented Collaborative Deep Learning
by: Li, Pengfei, et al.
Published: (2023)
by: Li, Pengfei, et al.
Published: (2023)
STING-BEE: Towards Vision-Language Model for Real-World X-ray Baggage Security Inspection
by: Velayudhan, Divya, et al.
Published: (2025)
by: Velayudhan, Divya, et al.
Published: (2025)
NeuGen: Amplifying the 'Neural' in Neural Radiance Fields for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
Improving Generative Adversarial Network Generalization for Facial Expression Synthesis
by: Akram, Arbish, et al.
Published: (2026)
by: Akram, Arbish, et al.
Published: (2026)
An Improved Quantum Software Challenges Classification Approach using Transfer Learning and Explainable AI
by: Khan, Nek Dil, et al.
Published: (2025)
by: Khan, Nek Dil, et al.
Published: (2025)
Pathology Image Compression with Pre-trained Autoencoders
by: Yellapragada, Srikar, et al.
Published: (2025)
by: Yellapragada, Srikar, et al.
Published: (2025)
Machine Learning-Driven Insights into Excitonic Effects in 2D Materials
by: Javed, Ahsan, et al.
Published: (2025)
by: Javed, Ahsan, et al.
Published: (2025)
Linear and Nonlinear Energy Harvesting in Concurrent Cellular and D2D Communication
by: Md. Fazlul Kader, et al.
Published: (2025)
by: Md. Fazlul Kader, et al.
Published: (2025)
Reconstructing Deep Neural Networks: Unleashing the Optimization Potential of Natural Gradient Descent
by: Liu, Weihua, et al.
Published: (2024)
by: Liu, Weihua, et al.
Published: (2024)
MUOT_3M: A 3 Million Frame Multimodal Underwater Benchmark and the MUTrack Tracking Method
by: Bakht, Ahsan Baidar, et al.
Published: (2026)
by: Bakht, Ahsan Baidar, et al.
Published: (2026)
Depth Attention for Robust RGB Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Similar Items
-
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
by: Alawode, Basit, et al.
Published: (2025) -
CPLIP: Zero-Shot Learning for Histopathology with Comprehensive Vision-Language Alignment
by: Javed, Sajid, et al.
Published: (2024) -
MLLM-HWSI: A Multimodal Large Language Model for Hierarchical Whole Slide Image Understanding
by: Alawode, Basit, et al.
Published: (2026) -
Advancing Histopathology with Deep Learning Under Data Scarcity: A Decade in Review
by: Obeid, Ahmad, et al.
Published: (2024) -
Predicting the Best of N Visual Trackers
by: Alawode, Basit, et al.
Published: (2024)