Llama Nemoretriever Colembed: Top-Performing Text-Image Retrieval Model
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Mengyao, Moreira, Gabriel, Ak, Ronay, Osmulski, Radek, Babakhin, Yauhen, Yu, Zhiding, Schifferer, Benedikt, Oldridge, Even |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Llama-Embed-Nemotron-8B: A Universal Text Embedding Model for Multilingual and Cross-Lingual Tasks
by: Babakhin, Yauhen, et al.
Published: (2025)
by: Babakhin, Yauhen, et al.
Published: (2025)
Omni-Embed-Nemotron: A Unified Multimodal Retrieval Model for Text, Image, Audio, and Video
by: Xu, Mengyao, et al.
Published: (2025)
by: Xu, Mengyao, et al.
Published: (2025)
MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark
by: Osmulski, Radek, et al.
Published: (2025)
by: Osmulski, Radek, et al.
Published: (2025)
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
Nemotron ColEmbed V2: Top-Performing Late Interaction Embedding Models for Visual Document Retrieval
by: Moreira, Gabriel de Souza P., et al.
Published: (2026)
by: Moreira, Gabriel de Souza P., et al.
Published: (2026)
NV-Retriever: Improving text embedding models with effective hard-negative mining
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
by: Moreira, Gabriel de Souza P., et al.
Published: (2024)
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
by: Xiao, J., et al.
Published: (2025)
by: Xiao, J., et al.
Published: (2025)
H2O-Danube3 Technical Report
by: Pfeiffer, Pascal, et al.
Published: (2024)
by: Pfeiffer, Pascal, et al.
Published: (2024)
LlamaSeg: Image Segmentation via Autoregressive Mask Generation
by: Deng, Jiru, et al.
Published: (2025)
by: Deng, Jiru, et al.
Published: (2025)
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
by: Sun, Peize, et al.
Published: (2024)
by: Sun, Peize, et al.
Published: (2024)
Towards Text-Image Interleaved Retrieval
by: Zhang, Xin, et al.
Published: (2025)
by: Zhang, Xin, et al.
Published: (2025)
Multi-Spectral Remote Sensing Image Retrieval Using Geospatial Foundation Models
by: Blumenstiel, Benedikt, et al.
Published: (2024)
by: Blumenstiel, Benedikt, et al.
Published: (2024)
Vezetők, testületek, felelősség a felsőoktatási intézményekben, különös tekintettel az állami egyetemekre
by: Rónay, Zoltán
Published: (2025)
by: Rónay, Zoltán
Published: (2025)
De la evagación a la evolución?
by: Alicia Ronay
Published: (2004)
by: Alicia Ronay
Published: (2004)
Anatomy-Aware Conditional Image-Text Retrieval
by: Zheng, Meng, et al.
Published: (2025)
by: Zheng, Meng, et al.
Published: (2025)
Zero-shot Composed Text-Image Retrieval
by: Liu, Yikun, et al.
Published: (2023)
by: Liu, Yikun, et al.
Published: (2023)
Brain-Inspired Multimodal Spiking Neural Network for Image-Text Retrieval
by: Zong, Xintao, et al.
Published: (2026)
by: Zong, Xintao, et al.
Published: (2026)
Knowledge-aware Text-Image Retrieval for Remote Sensing Images
by: Mi, Li, et al.
Published: (2024)
by: Mi, Li, et al.
Published: (2024)
H2O-Danube-1.8B Technical Report
by: Singer, Philipp, et al.
Published: (2024)
by: Singer, Philipp, et al.
Published: (2024)
Multi-path Exploration and Feedback Adjustment for Text-to-Image Person Retrieval
by: Kang, Bin, et al.
Published: (2024)
by: Kang, Bin, et al.
Published: (2024)
Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated Images
by: Xu, Shicheng, et al.
Published: (2023)
by: Xu, Shicheng, et al.
Published: (2023)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
by: Yuan, Chao, et al.
Published: (2026)
by: Yuan, Chao, et al.
Published: (2026)
Compositional Image-Text Matching and Retrieval by Grounding Entities
by: Vongala, Madhukar Reddy, et al.
Published: (2025)
by: Vongala, Madhukar Reddy, et al.
Published: (2025)
INQUIRE: A Natural World Text-to-Image Retrieval Benchmark
by: Vendrow, Edward, et al.
Published: (2024)
by: Vendrow, Edward, et al.
Published: (2024)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
by: Liu, Delong, et al.
Published: (2023)
by: Liu, Delong, et al.
Published: (2023)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
by: Lee, Saehyung, et al.
Published: (2024)
by: Lee, Saehyung, et al.
Published: (2024)
Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations
by: Chi, Jianfeng, et al.
Published: (2024)
by: Chi, Jianfeng, et al.
Published: (2024)
The Llama 3 Herd of Models
by: Grattafiori, Aaron, et al.
Published: (2024)
by: Grattafiori, Aaron, et al.
Published: (2024)
CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval
by: Wang, Haoran, et al.
Published: (2022)
by: Wang, Haoran, et al.
Published: (2022)
TEMA: Anchor the Image, Follow the Text for Multi-Modification Composed Image Retrieval
by: Li, Zixu, et al.
Published: (2026)
by: Li, Zixu, et al.
Published: (2026)
The Unmet Promise of Synthetic Training Images: Using Retrieved Real Images Performs Better
by: Geng, Scott, et al.
Published: (2024)
by: Geng, Scott, et al.
Published: (2024)
Robust Remote Sensing Image-Text Retrieval with Noisy Correspondence
by: Song, Qiya, et al.
Published: (2026)
by: Song, Qiya, et al.
Published: (2026)
DIR-TIR: Dialog-Iterative Refinement for Text-to-Image Retrieval
by: Zhen, Zongwei, et al.
Published: (2025)
by: Zhen, Zongwei, et al.
Published: (2025)
EFSA: Episodic Few-Shot Adaptation for Text-to-Image Retrieval
by: Huzaifa, Muhammad, et al.
Published: (2024)
by: Huzaifa, Muhammad, et al.
Published: (2024)
MIRAGE: Retrieval and Generation of Multimodal Images and Texts for Medical Education
by: Benito, Miguel Diaz, et al.
Published: (2026)
by: Benito, Miguel Diaz, et al.
Published: (2026)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
by: Lim, Youngsun, et al.
Published: (2024)
by: Lim, Youngsun, et al.
Published: (2024)
ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval
by: Xing, Eric, et al.
Published: (2025)
by: Xing, Eric, et al.
Published: (2025)
ADaFuSE: Adaptive Diffusion-generated Image and Text Fusion for Interactive Text-to-Image Retrieval
by: Zhang, Zhuocheng, et al.
Published: (2026)
by: Zhang, Zhuocheng, et al.
Published: (2026)
QLIP: Text-Aligned Visual Tokenization Unifies Auto-Regressive Multimodal Understanding and Generation
by: Zhao, Yue, et al.
Published: (2025)
by: Zhao, Yue, et al.
Published: (2025)
TIGER: Text-Instructed 3D Gaussian Retrieval and Coherent Editing
by: Xu, Teng, et al.
Published: (2024)
by: Xu, Teng, et al.
Published: (2024)
Similar Items
-
Llama-Embed-Nemotron-8B: A Universal Text Embedding Model for Multilingual and Cross-Lingual Tasks
by: Babakhin, Yauhen, et al.
Published: (2025) -
Omni-Embed-Nemotron: A Unified Multimodal Retrieval Model for Text, Image, Audio, and Video
by: Xu, Mengyao, et al.
Published: (2025) -
MIRACL-VISION: A Large, multilingual, visual document retrieval benchmark
by: Osmulski, Radek, et al.
Published: (2025) -
Enhancing Q&A Text Retrieval with Ranking Models: Benchmarking, fine-tuning and deploying Rerankers for RAG
by: Moreira, Gabriel de Souza P., et al.
Published: (2024) -
Nemotron ColEmbed V2: Top-Performing Late Interaction Embedding Models for Visual Document Retrieval
by: Moreira, Gabriel de Souza P., et al.
Published: (2026)