CLIPTime: Time-Aware Multimodal Representation Learning from Images and Text
Fuente:
arXiv
Saved in:
| Main Authors: | Rani, Anju, Ortiz-Arroyo, Daniel, Durdevic, Petar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Imagery Dataset for Remaining Useful Life Estimation of Synthetic Fibre Ropes
by: Rani, Anju, et al.
Published: (2026)
by: Rani, Anju, et al.
Published: (2026)
Advancements in Point Cloud-Based 3D Defect Detection and Classification for Industrial Systems: A Comprehensive Survey
by: Rani, Anju, et al.
Published: (2024)
by: Rani, Anju, et al.
Published: (2024)
FungalZSL: Zero-Shot Fungal Classification with Image Captioning Using a Synthetic Data Approach
by: Rani, Anju, et al.
Published: (2025)
by: Rani, Anju, et al.
Published: (2025)
DART: A Vision-Language Foundation Model for Comprehensive Rope Condition Monitoring
by: Rani, Anju, et al.
Published: (2026)
by: Rani, Anju, et al.
Published: (2026)
Defect Detection in Synthetic Fibre Ropes using Detectron2 Framework
by: Rani, Anju, et al.
Published: (2023)
by: Rani, Anju, et al.
Published: (2023)
Synthetic Fungi Datasets: A Time-Aligned Approach
by: Rani, A., et al.
Published: (2025)
by: Rani, A., et al.
Published: (2025)
Hyperbolic Image-Text Representations
by: Desai, Karan, et al.
Published: (2023)
by: Desai, Karan, et al.
Published: (2023)
FairJudge: Abstention-Aware Multimodal Judges for Fairness and Alignment Evaluation in Text-to-Image Models
by: Sahili, Zahraa Al, et al.
Published: (2025)
by: Sahili, Zahraa Al, et al.
Published: (2025)
Improved Probabilistic Image-Text Representations
by: Chun, Sanghyuk
Published: (2023)
by: Chun, Sanghyuk
Published: (2023)
Fill the Gap: Quantifying and Reducing the Modality Gap in Image-Text Representation Learning
by: Role, François, et al.
Published: (2025)
by: Role, François, et al.
Published: (2025)
ARGENT: Adaptive Hierarchical Image-Text Representations
by: Huynh, Chuong, et al.
Published: (2026)
by: Huynh, Chuong, et al.
Published: (2026)
A Survey on Self-supervised Contrastive Learning for Multimodal Text-Image Analysis
by: Khan, Asifullah, et al.
Published: (2025)
by: Khan, Asifullah, et al.
Published: (2025)
Advanced Multimodal Deep Learning Architecture for Image-Text Matching
by: Wang, Jinyin, et al.
Published: (2024)
by: Wang, Jinyin, et al.
Published: (2024)
Expert Knowledge-Aware Image Difference Graph Representation Learning for Difference-Aware Medical Visual Question Answering
by: Hu, Xinyue, et al.
Published: (2023)
by: Hu, Xinyue, et al.
Published: (2023)
Cost-Aware Routing for Efficient Text-To-Image Generation
by: Li, Qinchan, et al.
Published: (2025)
by: Li, Qinchan, et al.
Published: (2025)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
by: Arazi, Alan, et al.
Published: (2026)
by: Arazi, Alan, et al.
Published: (2026)
Multimodal Representation Learning by Alternating Unimodal Adaptation
by: Zhang, Xiaohui, et al.
Published: (2023)
by: Zhang, Xiaohui, et al.
Published: (2023)
Principled Multimodal Representation Learning
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
Semimage: HSV-Based Semantic Image Encoding for Disentangled Text Representation
by: Zare, Mohammad
Published: (2025)
by: Zare, Mohammad
Published: (2025)
The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models
by: Serra, Alessandro Pietro, et al.
Published: (2024)
by: Serra, Alessandro Pietro, et al.
Published: (2024)
ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval
by: Xing, Eric, et al.
Published: (2025)
by: Xing, Eric, et al.
Published: (2025)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025)
by: Cai, Yichao, et al.
Published: (2025)
Hierarchy-Guided Multimodal Representation Learning for Taxonomic Inference
by: Ahmed, Sk Miraj, et al.
Published: (2026)
by: Ahmed, Sk Miraj, et al.
Published: (2026)
Toward Unified Multimodal Representation Learning for Autonomous Driving
by: Tao, Ximeng, et al.
Published: (2026)
by: Tao, Ximeng, et al.
Published: (2026)
DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models
by: Ryu, Hyogon, et al.
Published: (2025)
by: Ryu, Hyogon, et al.
Published: (2025)
VTBench: A Multimodal Framework for Time-Series Classification with Chart-Based Representations
by: Venkatesan, Madhumitha, et al.
Published: (2026)
by: Venkatesan, Madhumitha, et al.
Published: (2026)
Multimodal ML: Quantifying the Improvement of Calorie Estimation Through Image-Text Pairs
by: Narang, Arya
Published: (2025)
by: Narang, Arya
Published: (2025)
Text-Conditional JEPA for Learning Semantically Rich Visual Representations
by: Huang, Chen, et al.
Published: (2026)
by: Huang, Chen, et al.
Published: (2026)
Context-Aware Multimodal Representation Learning for Spatio-Temporally Explicit Environmental Modelling
by: Peters, Julia, et al.
Published: (2025)
by: Peters, Julia, et al.
Published: (2025)
TerraFlow: Multimodal, Multitemporal Representation Learning for Earth Observation
by: Puriy, Nazar, et al.
Published: (2026)
by: Puriy, Nazar, et al.
Published: (2026)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
by: Lv, Song-Lin, et al.
Published: (2025)
by: Lv, Song-Lin, et al.
Published: (2025)
Text-to-Image GAN with Pretrained Representations
by: You, Xiaozhou, et al.
Published: (2024)
by: You, Xiaozhou, et al.
Published: (2024)
Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics
by: Chen, Jiayuan, et al.
Published: (2026)
by: Chen, Jiayuan, et al.
Published: (2026)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
by: Chatterjee, Abhiroop, et al.
Published: (2025)
by: Chatterjee, Abhiroop, et al.
Published: (2025)
Reward Incremental Learning in Text-to-Image Generation
by: Wang, Maorong, et al.
Published: (2024)
by: Wang, Maorong, et al.
Published: (2024)
Learned Image Compression with Text Quality Enhancement
by: Lai, Chih-Yu, et al.
Published: (2024)
by: Lai, Chih-Yu, et al.
Published: (2024)
COMPRER: A Multimodal Multi-Objective Pretraining Framework for Enhanced Medical Image Representation
by: Lutsker, Guy, et al.
Published: (2024)
by: Lutsker, Guy, et al.
Published: (2024)
Learning Topological Representations for Deep Image Understanding
by: Hu, Xiaoling
Published: (2024)
by: Hu, Xiaoling
Published: (2024)
CLARGA: Multimodal Graph Representation Learning over Arbitrary Sets of Modalities
by: Patapati, Santosh
Published: (2025)
by: Patapati, Santosh
Published: (2025)
Multimodal Adaptive Retrieval Augmented Generation through Internal Representation Learning
by: Du, Ruoshuang, et al.
Published: (2026)
by: Du, Ruoshuang, et al.
Published: (2026)
Similar Items
-
Imagery Dataset for Remaining Useful Life Estimation of Synthetic Fibre Ropes
by: Rani, Anju, et al.
Published: (2026) -
Advancements in Point Cloud-Based 3D Defect Detection and Classification for Industrial Systems: A Comprehensive Survey
by: Rani, Anju, et al.
Published: (2024) -
FungalZSL: Zero-Shot Fungal Classification with Image Captioning Using a Synthetic Data Approach
by: Rani, Anju, et al.
Published: (2025) -
DART: A Vision-Language Foundation Model for Comprehensive Rope Condition Monitoring
by: Rani, Anju, et al.
Published: (2026) -
Defect Detection in Synthetic Fibre Ropes using Detectron2 Framework
by: Rani, Anju, et al.
Published: (2023)