Contrastive vision-language learning with paraphrasing and negation
Fuente:
arXiv
Saved in:
| Main Authors: | Ngan, Kwun Ho, Afgeh, Saman Sadeghi, Townsend, Joe, Garcez, Artur d'Avila |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
by: Sikar, Daniel, et al.
Published: (2025)
by: Sikar, Daniel, et al.
Published: (2025)
What do vision-language models see in the context? Investigating multimodal in-context learning
by: Santos, Gabriel O. dos, et al.
Published: (2025)
by: Santos, Gabriel O. dos, et al.
Published: (2025)
Video Annotator: A framework for efficiently building video classifiers using vision-language models and active learning
by: Ziai, Amir, et al.
Published: (2024)
by: Ziai, Amir, et al.
Published: (2024)
Non-negative Contrastive Learning
by: Wang, Yifei, et al.
Published: (2024)
by: Wang, Yifei, et al.
Published: (2024)
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024)
by: Venkatraman, Siddarth, et al.
Published: (2024)
Visual hallucination detection in large vision-language models via evidential conflict
by: Huang, Tao, et al.
Published: (2025)
by: Huang, Tao, et al.
Published: (2025)
ViSTa Dataset: Do vision-language models understand sequential tasks?
by: Wybitul, Evžen, et al.
Published: (2024)
by: Wybitul, Evžen, et al.
Published: (2024)
FlySearch: Exploring how vision-language models explore
by: Pardyl, Adam, et al.
Published: (2025)
by: Pardyl, Adam, et al.
Published: (2025)
Unsupervised Continual Anomaly Detection with Contrastively-learned Prompt
by: Liu, Jiaqi, et al.
Published: (2024)
by: Liu, Jiaqi, et al.
Published: (2024)
Aligning Visual Contrastive learning models via Preference Optimization
by: Afzali, Amirabbas, et al.
Published: (2024)
by: Afzali, Amirabbas, et al.
Published: (2024)
The in-context inductive biases of vision-language models differ across modalities
by: Allen, Kelsey, et al.
Published: (2025)
by: Allen, Kelsey, et al.
Published: (2025)
BRAVE: Broadening the visual encoding of vision-language models
by: Kar, Oğuzhan Fatih, et al.
Published: (2024)
by: Kar, Oğuzhan Fatih, et al.
Published: (2024)
Linking heterogeneous microstructure informatics with expert characterization knowledge through customized and hybrid vision-language representations for industrial qualification
by: Safdar, Mutahar, et al.
Published: (2025)
by: Safdar, Mutahar, et al.
Published: (2025)
Bridging visual saliency and large language models for explainable deep learning in medical imaging
by: Nguezet, Paul Valery, et al.
Published: (2026)
by: Nguezet, Paul Valery, et al.
Published: (2026)
Investigating the Effectiveness of Cross-Attention to Unlock Zero-Shot Editing of Text-to-Video Diffusion Models
by: Motamed, Saman, et al.
Published: (2024)
by: Motamed, Saman, et al.
Published: (2024)
A preliminary study on continual learning in computer vision using Kolmogorov-Arnold Networks
by: Cacciatore, Alessandro, et al.
Published: (2024)
by: Cacciatore, Alessandro, et al.
Published: (2024)
A Contrastive Learning-Guided Confident Meta-learning for Zero Shot Anomaly Detection
by: Aqeel, Muhammad, et al.
Published: (2025)
by: Aqeel, Muhammad, et al.
Published: (2025)
Forward-Backward Knowledge Distillation for Continual Clustering
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
FocusLearn: Fully-Interpretable, High-Performance Modular Neural Networks for Time Series
by: Su, Qiqi, et al.
Published: (2023)
by: Su, Qiqi, et al.
Published: (2023)
GP-VLS: A general-purpose vision language model for surgery
by: Schmidgall, Samuel, et al.
Published: (2024)
by: Schmidgall, Samuel, et al.
Published: (2024)
Distributed solar generation forecasting using attention-based deep neural networks for cloud movement prediction
by: Perera, Maneesha, et al.
Published: (2024)
by: Perera, Maneesha, et al.
Published: (2024)
CommonForms: A Large, Diverse Dataset for Form Field Detection
by: Barrow, Joe
Published: (2025)
by: Barrow, Joe
Published: (2025)
Outlier detection by ensembling uncertainty with negative objectness
by: Delić, Anja, et al.
Published: (2024)
by: Delić, Anja, et al.
Published: (2024)
Cross-modal linkage risk in clinical vision-language models
by: Arasteh, Soroosh Tayebi, et al.
Published: (2026)
by: Arasteh, Soroosh Tayebi, et al.
Published: (2026)
Reproducible scaling laws for contrastive language-image learning
by: Cherti, Mehdi, et al.
Published: (2022)
by: Cherti, Mehdi, et al.
Published: (2022)
VISion On Request: Enhanced VLLM efficiency with sparse, dynamically selected, vision-language interactions
by: Bulat, Adrian, et al.
Published: (2026)
by: Bulat, Adrian, et al.
Published: (2026)
Contrastive Semantic Projection: Faithful Neuron Labeling with Contrastive Examples
by: Bouanani, Oussama, et al.
Published: (2026)
by: Bouanani, Oussama, et al.
Published: (2026)
Eyes on the Image: Gaze Supervised Multimodal Learning for Chest X-ray Diagnosis and Report Generation
by: Riju, Tanjim Islam, et al.
Published: (2025)
by: Riju, Tanjim Islam, et al.
Published: (2025)
Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models
by: Motamed, Saman, et al.
Published: (2023)
by: Motamed, Saman, et al.
Published: (2023)
RealKIE: Five Novel Datasets for Enterprise Key Information Extraction
by: Townsend, Benjamin, et al.
Published: (2024)
by: Townsend, Benjamin, et al.
Published: (2024)
ContrastCAD: Contrastive Learning-based Representation Learning for Computer-Aided Design Models
by: Jung, Minseop, et al.
Published: (2024)
by: Jung, Minseop, et al.
Published: (2024)
$\mathbb{X}$-Sample Contrastive Loss: Improving Contrastive Learning with Sample Similarity Graphs
by: Sobal, Vlad, et al.
Published: (2024)
by: Sobal, Vlad, et al.
Published: (2024)
Advancing vision-language models in front-end development via data synthesis
by: Ge, Tong, et al.
Published: (2025)
by: Ge, Tong, et al.
Published: (2025)
PathAlign: A vision-language model for whole slide images in histopathology
by: Ahmed, Faruk, et al.
Published: (2024)
by: Ahmed, Faruk, et al.
Published: (2024)
Variational Supervised Contrastive Learning
by: Wang, Ziwen, et al.
Published: (2025)
by: Wang, Ziwen, et al.
Published: (2025)
Utility-Fairness Trade-Offs and How to Find Them
by: Dehdashtian, Sepehr, et al.
Published: (2024)
by: Dehdashtian, Sepehr, et al.
Published: (2024)
Rethinking Positive Pairs in Contrastive Learning
by: Wu, Jiantao, et al.
Published: (2024)
by: Wu, Jiantao, et al.
Published: (2024)
A Mathematical Perspective On Contrastive Learning
by: Baptista, Ricardo, et al.
Published: (2025)
by: Baptista, Ricardo, et al.
Published: (2025)
Contrastive Learning for Regression on Hyperspectral Data
by: Dhaini, Mohamad, et al.
Published: (2024)
by: Dhaini, Mohamad, et al.
Published: (2024)
Your CLIP has 164 dimensions of noise: Exploring the embeddings covariance eigenspectrum of contrastively pretrained vision-language transformers
by: Grzywaczewski, Jakub, et al.
Published: (2026)
by: Grzywaczewski, Jakub, et al.
Published: (2026)
Similar Items
-
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
by: Sikar, Daniel, et al.
Published: (2025) -
What do vision-language models see in the context? Investigating multimodal in-context learning
by: Santos, Gabriel O. dos, et al.
Published: (2025) -
Video Annotator: A framework for efficiently building video classifiers using vision-language models and active learning
by: Ziai, Amir, et al.
Published: (2024) -
Non-negative Contrastive Learning
by: Wang, Yifei, et al.
Published: (2024) -
Amortizing intractable inference in diffusion models for vision, language, and control
by: Venkatraman, Siddarth, et al.
Published: (2024)