xVLM2Vec: Adapting LVLM-based embedding models to multilinguality using Self-Knowledge Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Musacchio, Elio, Siciliani, Lucia, Basile, Pierpaolo, Semeraro, Giovanni |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring the Word Sense Disambiguation Capabilities of Large Language Models
by: Basile, Pierpaolo, et al.
Published: (2025)
by: Basile, Pierpaolo, et al.
Published: (2025)
Mimir: Large-scale Multilingual Concept Modeling
by: Musacchio, Elio, et al.
Published: (2026)
by: Musacchio, Elio, et al.
Published: (2026)
Understanding and Mitigating the Threat of Vec2Text to Dense Retrieval Systems
by: Zhuang, Shengyao, et al.
Published: (2024)
by: Zhuang, Shengyao, et al.
Published: (2024)
Does Knowledge Distillation Matter for Large Language Model based Bundle Generation?
by: Feng, Kaidong, et al.
Published: (2025)
by: Feng, Kaidong, et al.
Published: (2025)
Routing Distilled Knowledge via Mixture of LoRA Experts for Large Language Model based Bundle Generation
by: Feng, Kaidong, et al.
Published: (2025)
by: Feng, Kaidong, et al.
Published: (2025)
Scaling Knowledge Graph Construction through Synthetic Data Generation and Distillation
by: Choubey, Prafulla Kumar, et al.
Published: (2024)
by: Choubey, Prafulla Kumar, et al.
Published: (2024)
DiSCo: LLM Knowledge Distillation for Efficient Sparse Retrieval in Conversational Search
by: Lupart, Simon, et al.
Published: (2024)
by: Lupart, Simon, et al.
Published: (2024)
RDRec: Rationale Distillation for LLM-based Recommendation
by: Wang, Xinfeng, et al.
Published: (2024)
by: Wang, Xinfeng, et al.
Published: (2024)
Interact2Vec -- An efficient neural network-based model for simultaneously learning users and items embeddings in recommender systems
by: Pires, Pedro R., et al.
Published: (2025)
by: Pires, Pedro R., et al.
Published: (2025)
Adapting General-Purpose Embedding Models to Private Datasets Using Keyword-based Retrieval
by: Wei, Yubai, et al.
Published: (2025)
by: Wei, Yubai, et al.
Published: (2025)
Advanced Natural-based interaction for the ITAlian language: LLaMAntino-3-ANITA
by: Polignano, Marco, et al.
Published: (2024)
by: Polignano, Marco, et al.
Published: (2024)
PairDistill: Pairwise Relevance Distillation for Dense Retrieval
by: Huang, Chao-Wei, et al.
Published: (2024)
by: Huang, Chao-Wei, et al.
Published: (2024)
Optimization of embeddings storage for RAG systems using quantization and dimensionality reduction techniques
by: Huerga-Pérez, Naamán, et al.
Published: (2025)
by: Huerga-Pérez, Naamán, et al.
Published: (2025)
Improving Neural Topic Models with Wasserstein Knowledge Distillation
by: Adhya, Suman, et al.
Published: (2023)
by: Adhya, Suman, et al.
Published: (2023)
A comparative analysis of embedding models for patent similarity
by: Ascione, Grazia Sveva, et al.
Published: (2024)
by: Ascione, Grazia Sveva, et al.
Published: (2024)
Translate-Distill: Learning Cross-Language Dense Retrieval by Translation and Distillation
by: Yang, Eugene, et al.
Published: (2024)
by: Yang, Eugene, et al.
Published: (2024)
From Tokens to Concepts: Leveraging SAE for SPLADE
by: Zong, Yuxuan, et al.
Published: (2026)
by: Zong, Yuxuan, et al.
Published: (2026)
Distilling Reasoning Without Knowledge: A Framework for Reliable LLMs
by: Kietkajornrit, Auksarapak, et al.
Published: (2026)
by: Kietkajornrit, Auksarapak, et al.
Published: (2026)
Structured Distillation for Personalized Agent Memory: 11x Token Reduction with Retrieval Preservation
by: Lewis, Sydney
Published: (2026)
by: Lewis, Sydney
Published: (2026)
VLM2GeoVec: Toward Universal Multimodal Embeddings for Remote Sensing
by: Aimar, Emanuel Sánchez, et al.
Published: (2025)
by: Aimar, Emanuel Sánchez, et al.
Published: (2025)
SAKE: Self-aware Knowledge Exploitation-Exploration for Grounded Multimodal Named Entity Recognition
by: Tang, Jielong, et al.
Published: (2026)
by: Tang, Jielong, et al.
Published: (2026)
Distillation for Multilingual Information Retrieval
by: Yang, Eugene, et al.
Published: (2024)
by: Yang, Eugene, et al.
Published: (2024)
Health Misinformation Detection in Web Content via Web2Vec: A Structural-, Content-based, and Context-aware Approach based on Web2Vec
by: Upadhyay, Rishabh, et al.
Published: (2024)
by: Upadhyay, Rishabh, et al.
Published: (2024)
Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning
by: Liang, Zihan, et al.
Published: (2026)
by: Liang, Zihan, et al.
Published: (2026)
Training the Knowledge Base through Evidence Distillation and Write-Back Enrichment
by: Lu, Yuxing, et al.
Published: (2026)
by: Lu, Yuxing, et al.
Published: (2026)
LEAF: Knowledge Distillation of Text Embedding Models with Teacher-Aligned Representations
by: Vujanic, Robin, et al.
Published: (2025)
by: Vujanic, Robin, et al.
Published: (2025)
Cross-modal Retrieval for Knowledge-based Visual Question Answering
by: Lerner, Paul, et al.
Published: (2024)
by: Lerner, Paul, et al.
Published: (2024)
ChatRetriever: Adapting Large Language Models for Generalized and Robust Conversational Dense Retrieval
by: Mao, Kelong, et al.
Published: (2024)
by: Mao, Kelong, et al.
Published: (2024)
TaxoAdapt: Aligning LLM-Based Multidimensional Taxonomy Construction to Evolving Research Corpora
by: Kargupta, Priyanka, et al.
Published: (2025)
by: Kargupta, Priyanka, et al.
Published: (2025)
Empowering Large Language Models to Set up a Knowledge Retrieval Indexer via Self-Learning
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
Optimizing Knowledge Integration in Retrieval-Augmented Generation with Self-Selection
by: Weng, Yan, et al.
Published: (2025)
by: Weng, Yan, et al.
Published: (2025)
Hypothetical Documents or Knowledge Leakage? Rethinking LLM-based Query Expansion
by: Yoon, Yejun, et al.
Published: (2025)
by: Yoon, Yejun, et al.
Published: (2025)
Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge
by: Wang, Mengyu, et al.
Published: (2026)
by: Wang, Mengyu, et al.
Published: (2026)
SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning
by: Ma, Yufei, et al.
Published: (2026)
by: Ma, Yufei, et al.
Published: (2026)
Less is More: Adapting Text Embeddings for Low-Resource Languages with Small Scale Noisy Synthetic Data
by: Navasardyan, Zaruhi, et al.
Published: (2026)
by: Navasardyan, Zaruhi, et al.
Published: (2026)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
by: Abdullah, Abdullah, et al.
Published: (2025)
by: Abdullah, Abdullah, et al.
Published: (2025)
Skill matching at scale: freelancer-project alignment for efficient multilingual candidate retrieval
by: Jouanneau, Warren, et al.
Published: (2024)
by: Jouanneau, Warren, et al.
Published: (2024)
Distillation versus Contrastive Learning: How to Train Your Rerankers
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
Unveiling the Magic: Investigating Attention Distillation in Retrieval-augmented Generation
by: Li, Zizhong, et al.
Published: (2024)
by: Li, Zizhong, et al.
Published: (2024)
Retrieval Augmented Generation using Engineering Design Knowledge
by: Siddharth, L., et al.
Published: (2023)
by: Siddharth, L., et al.
Published: (2023)
Similar Items
-
Exploring the Word Sense Disambiguation Capabilities of Large Language Models
by: Basile, Pierpaolo, et al.
Published: (2025) -
Mimir: Large-scale Multilingual Concept Modeling
by: Musacchio, Elio, et al.
Published: (2026) -
Understanding and Mitigating the Threat of Vec2Text to Dense Retrieval Systems
by: Zhuang, Shengyao, et al.
Published: (2024) -
Does Knowledge Distillation Matter for Large Language Model based Bundle Generation?
by: Feng, Kaidong, et al.
Published: (2025) -
Routing Distilled Knowledge via Mixture of LoRA Experts for Large Language Model based Bundle Generation
by: Feng, Kaidong, et al.
Published: (2025)