CLIMP: Contrastive Language-Image Mamba Pretraining
Fuente:
arXiv
Saved in:
| Main Authors: | Shabtay, Nimrod, Zimerman, Itamar, Schwartz, Eli, Giryes, Raja |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PIP: Positional-encoding Image Prior
by: Shabtay, Nimrod, et al.
Published: (2022)
by: Shabtay, Nimrod, et al.
Published: (2022)
Deep Phase Coded Image Prior
by: Shabtay, Nimrod, et al.
Published: (2024)
by: Shabtay, Nimrod, et al.
Published: (2024)
CARES: Context-Aware Resolution Selector for VLMs
by: Kimhi, Moshe, et al.
Published: (2025)
by: Kimhi, Moshe, et al.
Published: (2025)
Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
by: Shabtay, Nimrod, et al.
Published: (2026)
by: Shabtay, Nimrod, et al.
Published: (2026)
Teaching VLMs to Localize Specific Objects from In-context Examples
by: Doveh, Sivan, et al.
Published: (2024)
by: Doveh, Sivan, et al.
Published: (2024)
DIP-GS: Deep Image Prior For Gaussian Splatting Sparse View Recovery
by: Khatib, Rajaei, et al.
Published: (2025)
by: Khatib, Rajaei, et al.
Published: (2025)
MAEDAY: MAE for few and zero shot AnomalY-Detection
by: Schwartz, Eli, et al.
Published: (2022)
by: Schwartz, Eli, et al.
Published: (2022)
TriNeRFLet: A Wavelet Based Triplane NeRF Representation
by: Khatib, Rajaei, et al.
Published: (2024)
by: Khatib, Rajaei, et al.
Published: (2024)
3VL: Using Trees to Improve Vision-Language Models' Interpretability
by: Yellinek, Nir, et al.
Published: (2023)
by: Yellinek, Nir, et al.
Published: (2023)
Tell Me What You See: Text-Guided Real-World Image Denoising
by: Yosef, Erez, et al.
Published: (2023)
by: Yosef, Erez, et al.
Published: (2023)
LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content
by: Shabtay, Nimrod, et al.
Published: (2024)
by: Shabtay, Nimrod, et al.
Published: (2024)
ProtoSAM: One-Shot Medical Image Segmentation With Foundational Models
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
Pruning at Initialization -- A Sketching Perspective
by: Bar, Noga, et al.
Published: (2023)
by: Bar, Noga, et al.
Published: (2023)
DINOv2 based Self Supervised Learning For Few Shot Medical Image Segmentation
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
ADIR: Adaptive Diffusion for Image Reconstruction
by: Abu-Hussein, Shady, et al.
Published: (2022)
by: Abu-Hussein, Shady, et al.
Published: (2022)
Guided Lensless Polarization Imaging
by: Kraicer, Noa, et al.
Published: (2026)
by: Kraicer, Noa, et al.
Published: (2026)
Diverse Subset Selection via Norm-Based Sampling and Orthogonality
by: Bar, Noga, et al.
Published: (2024)
by: Bar, Noga, et al.
Published: (2024)
Group Orthogonalization Regularization For Vision Models Adaptation and Robustness
by: Kurtz, Yoav, et al.
Published: (2023)
by: Kurtz, Yoav, et al.
Published: (2023)
DifuzCam: Replacing Camera Lens with a Mask and a Diffusion Model
by: Yosef, Erez, et al.
Published: (2024)
by: Yosef, Erez, et al.
Published: (2024)
ICC: Quantifying Image Caption Concreteness for Multimodal Dataset Curation
by: Yanuka, Moran, et al.
Published: (2024)
by: Yanuka, Moran, et al.
Published: (2024)
Autoregressive Pretraining with Mamba in Vision
by: Ren, Sucheng, et al.
Published: (2024)
by: Ren, Sucheng, et al.
Published: (2024)
Image-Adaptive GAN based Reconstruction
by: Hussein, Shady Abu, et al.
Published: (2019)
by: Hussein, Shady Abu, et al.
Published: (2019)
Anatomical Token Uncertainty for Transformer-Guided Active MRI Acquisition
by: Ayzenberg, Lev, et al.
Published: (2026)
by: Ayzenberg, Lev, et al.
Published: (2026)
Jacobian-aware Posterior Sampling for Inverse Problems
by: Hen, Liav, et al.
Published: (2025)
by: Hen, Liav, et al.
Published: (2025)
Detecting Backdoor Samples in Contrastive Language Image Pretraining
by: Huang, Hanxun, et al.
Published: (2025)
by: Huang, Hanxun, et al.
Published: (2025)
Object-centric Binding in Contrastive Language-Image Pretraining
by: Assouel, Rim, et al.
Published: (2025)
by: Assouel, Rim, et al.
Published: (2025)
CLIP-Mamba: CLIP Pretrained Mamba Models with OOD and Hessian Evaluation
by: Huang, Weiquan, et al.
Published: (2024)
by: Huang, Weiquan, et al.
Published: (2024)
UDPM: Upsampling Diffusion Probabilistic Models
by: Abu-Hussein, Shady, et al.
Published: (2023)
by: Abu-Hussein, Shady, et al.
Published: (2023)
Time-Contrastive Pretraining for In-Context Image and Video Segmentation
by: Wahd, Assefa, et al.
Published: (2025)
by: Wahd, Assefa, et al.
Published: (2025)
Contrastive Heliophysical Image Pretraining for Solar Dynamics Observatory Records
by: Shen, Shiyu, et al.
Published: (2025)
by: Shen, Shiyu, et al.
Published: (2025)
Video Analysis and Generation via a Semantic Progress Function
by: Metzer, Gal, et al.
Published: (2026)
by: Metzer, Gal, et al.
Published: (2026)
ConMamba: Contrastive Vision Mamba for Plant Disease Detection
by: Mamun, Abdullah Al, et al.
Published: (2025)
by: Mamun, Abdullah Al, et al.
Published: (2025)
Neural Brain Fields: A NeRF-Inspired Approach for Generating Nonexistent EEG Electrodes
by: Kedem, Shahar Ain, et al.
Published: (2025)
by: Kedem, Shahar Ain, et al.
Published: (2025)
X-ray2CTPA: Leveraging Diffusion Models to Enhance Pulmonary Embolism Classification
by: Cahan, Noa, et al.
Published: (2024)
by: Cahan, Noa, et al.
Published: (2024)
GS-CLIP: Gaussian Splatting for Contrastive Language-Image-3D Pretraining from Real-World Data
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation
by: Sincan, Ozge Mercanoglu, et al.
Published: (2025)
by: Sincan, Ozge Mercanoglu, et al.
Published: (2025)
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
by: Joshi, Siddharth, et al.
Published: (2024)
by: Joshi, Siddharth, et al.
Published: (2024)
ID-LoRA: Identity-Driven Audio-Video Personalization with In-Context LoRA
by: Dahan, Aviad, et al.
Published: (2026)
by: Dahan, Aviad, et al.
Published: (2026)
VideoMAP: Toward Scalable Mamba-based Video Autoregressive Pretraining
by: Liu, Yunze, et al.
Published: (2025)
by: Liu, Yunze, et al.
Published: (2025)
Separators in Enhancing Autoregressive Pretraining for Vision Mamba
by: Liu, Hanpeng, et al.
Published: (2026)
by: Liu, Hanpeng, et al.
Published: (2026)
Similar Items
-
PIP: Positional-encoding Image Prior
by: Shabtay, Nimrod, et al.
Published: (2022) -
Deep Phase Coded Image Prior
by: Shabtay, Nimrod, et al.
Published: (2024) -
CARES: Context-Aware Resolution Selector for VLMs
by: Kimhi, Moshe, et al.
Published: (2025) -
Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
by: Shabtay, Nimrod, et al.
Published: (2026) -
Teaching VLMs to Localize Specific Objects from In-context Examples
by: Doveh, Sivan, et al.
Published: (2024)