Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roy, Shuvendu, Parhizkar, Yasaman, Ogidi, Franklin, Khazaie, Vahid Reza, Colacci, Michael, Etemad, Ali, Dolatabadi, Elham, Afkanpour, Arash |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Medical Representation Learning Through High-Quality Data
von: Baghbanzadeh, Negin, et al.
Veröffentlicht: (2025)
von: Baghbanzadeh, Negin, et al.
Veröffentlicht: (2025)
A Shared Encoder Approach to Multimodal Representation Learning
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
Consistency-Guided Asynchronous Contrastive Tuning for Few-Shot Class-Incremental Tuning of Foundation Models
von: Roy, Shuvendu, et al.
Veröffentlicht: (2024)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2024)
Can Generative Models Improve Self-Supervised Representation Learning?
von: Ayromlou, Sana, et al.
Veröffentlicht: (2024)
von: Ayromlou, Sana, et al.
Veröffentlicht: (2024)
Consistency-guided Prompt Learning for Vision-Language Models
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
SelfPrompt: Confidence-Aware Semi-Supervised Tuning for Robust Vision-Language Model Adaptation
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
Open-PMC-18M: A High-Fidelity Large Scale Medical Dataset for Multimodal Representation Learning
von: Baghbanzadeh, Negin, et al.
Veröffentlicht: (2025)
von: Baghbanzadeh, Negin, et al.
Veröffentlicht: (2025)
Impact of Strategic Sampling and Supervision Policies on Semi-supervised Learning
von: Roy, Shuvendu, et al.
Veröffentlicht: (2022)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2022)
Scaling Up Semi-supervised Learning with Unconstrained Unlabelled Data
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
Exploring the Boundaries of Semi-Supervised Facial Expression Recognition using In-Distribution, Out-of-Distribution, and Unconstrained Data
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)
Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models
von: Islam, Mohammed Saidul, et al.
Veröffentlicht: (2026)
von: Islam, Mohammed Saidul, et al.
Veröffentlicht: (2026)
Automated Capability Evaluation of Foundation Models
von: Afkanpour, Arash, et al.
Veröffentlicht: (2025)
von: Afkanpour, Arash, et al.
Veröffentlicht: (2025)
A Bag of Tricks for Few-Shot Class-Incremental Learning
von: Roy, Shuvendu, et al.
Veröffentlicht: (2024)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2024)
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
von: Narayanan, Aravind, et al.
Veröffentlicht: (2025)
von: Narayanan, Aravind, et al.
Veröffentlicht: (2025)
LinguaMark: Do Multimodal Models Speak Fairly? A Benchmark-Based Evaluation
von: Raval, Ananya, et al.
Veröffentlicht: (2025)
von: Raval, Ananya, et al.
Veröffentlicht: (2025)
Self-Supervised Human Activity Recognition with Localized Time-Frequency Contrastive Representation Learning
von: Taghanaki, Setareh Rahimi, et al.
Veröffentlicht: (2022)
von: Taghanaki, Setareh Rahimi, et al.
Veröffentlicht: (2022)
Exploring Bias and Prediction Metrics to Characterise the Fairness of Machine Learning for Equity-Centered Public Health Decision-Making: A Narrative Review
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
von: Liskavets, Barys, et al.
Veröffentlicht: (2024)
von: Liskavets, Barys, et al.
Veröffentlicht: (2024)
Task-agnostic Prompt Compression with Context-aware Sentence Embedding and Reward-guided Task Descriptor
von: Liskavets, Barys, et al.
Veröffentlicht: (2025)
von: Liskavets, Barys, et al.
Veröffentlicht: (2025)
Feminismos y comunicación: Pilares sustantivos de la Extensión Crítica
von: Romina Colacci
Veröffentlicht: (2023)
von: Romina Colacci
Veröffentlicht: (2023)
A Flexible Fairness Framework with Surrogate Loss Reweighting for Addressing Sociodemographic Disparities
von: Xu, Wen, et al.
Veröffentlicht: (2025)
von: Xu, Wen, et al.
Veröffentlicht: (2025)
VLDBench Evaluating Multimodal Disinformation with Regulatory Alignment
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
Developing and validating perceived intercultural communication anxiety/apprehension scale
von: Ali Roohani, et al.
Veröffentlicht: (2024)
von: Ali Roohani, et al.
Veröffentlicht: (2024)
Signal Processing in the Retina: Interpretable Graph Classifier to Predict Ganglion Cell Responses
von: Parhizkar, Yasaman, et al.
Veröffentlicht: (2024)
von: Parhizkar, Yasaman, et al.
Veröffentlicht: (2024)
ViLBias: Detecting and Reasoning about Bias in Multimodal Content
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
von: Raza, Shaina, et al.
Veröffentlicht: (2024)
When Does RL Help Medical VLMs? Disentangling Vision, SFT, and RL Gains
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2026)
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2026)
CollideNet: Hierarchical Multi-scale Video Representation Learning with Disentanglement for Time-To-Collision Forecasting
von: Desai, Nishq Poorav, et al.
Veröffentlicht: (2026)
von: Desai, Nishq Poorav, et al.
Veröffentlicht: (2026)
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
Learning Time-Series Representations by Hierarchical Uniformity-Tolerance Latent Balancing
von: Jalali, Amin, et al.
Veröffentlicht: (2025)
von: Jalali, Amin, et al.
Veröffentlicht: (2025)
Fruit preservation with bioethanol obtained from the fermentation of brewer's spent grain with Saccharomyces carlsbergensis
von: Clement Olusola Ogidi
Veröffentlicht: (2020)
von: Clement Olusola Ogidi
Veröffentlicht: (2020)
In-Distribution and Out-of-Distribution Self-supervised ECG Representation Learning for Arrhythmia Detection
von: Soltanieh, Sahar, et al.
Veröffentlicht: (2023)
von: Soltanieh, Sahar, et al.
Veröffentlicht: (2023)
Enhancing Anomaly Detection Generalization through Knowledge Exposure: The Dual Effects of Augmentation
von: Anvari, Mohammad Akhavan, et al.
Veröffentlicht: (2024)
von: Anvari, Mohammad Akhavan, et al.
Veröffentlicht: (2024)
Naturally reductive homogeneous $(α,β)$-metric spaces
von: Parhizkar, Mojtaba, et al.
Veröffentlicht: (2013)
von: Parhizkar, Mojtaba, et al.
Veröffentlicht: (2013)
Self-alignment of Large Video Language Models with Refined Regularized Preference Optimization
von: Sarkar, Pritam, et al.
Veröffentlicht: (2025)
von: Sarkar, Pritam, et al.
Veröffentlicht: (2025)
The June 2025 Israeli War: Iran's Assessment and Regional Consequences
von: Ali Bagheri Dolatabadi
Veröffentlicht: (2025)
von: Ali Bagheri Dolatabadi
Veröffentlicht: (2025)
VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models
von: Sarkar, Pritam, et al.
Veröffentlicht: (2025)
von: Sarkar, Pritam, et al.
Veröffentlicht: (2025)
Evaluating LLM-driven User-Intent Formalization for Verification-Aware Languages
von: Lahiri, Shuvendu K.
Veröffentlicht: (2024)
von: Lahiri, Shuvendu K.
Veröffentlicht: (2024)
Partial Label Learning for Emotion Recognition from EEG
von: Zhang, Guangyi, et al.
Veröffentlicht: (2023)
von: Zhang, Guangyi, et al.
Veröffentlicht: (2023)
A Multi-Modal Foundational Model for Wireless Communication and Sensing
von: Yazdnian, Vahid, et al.
Veröffentlicht: (2026)
von: Yazdnian, Vahid, et al.
Veröffentlicht: (2026)
On The Relationship Between Continual Learning and Long-Tailed Recognition
von: Molahasani, Mahdiyar, et al.
Veröffentlicht: (2023)
von: Molahasani, Mahdiyar, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Advancing Medical Representation Learning Through High-Quality Data
von: Baghbanzadeh, Negin, et al.
Veröffentlicht: (2025) -
A Shared Encoder Approach to Multimodal Representation Learning
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025) -
Consistency-Guided Asynchronous Contrastive Tuning for Few-Shot Class-Incremental Tuning of Foundation Models
von: Roy, Shuvendu, et al.
Veröffentlicht: (2024) -
Can Generative Models Improve Self-Supervised Representation Learning?
von: Ayromlou, Sana, et al.
Veröffentlicht: (2024) -
Consistency-guided Prompt Learning for Vision-Language Models
von: Roy, Shuvendu, et al.
Veröffentlicht: (2023)