Let Multimodal Embedders Learn When to Augment Query via Adaptive Query Augmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Wongyu, Lee, Hochang, Lee, Sanghak, Kim, Yoonsung, Park, Jaehyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
by: Ju, Yeong-Joon, et al.
Published: (2024)
by: Ju, Yeong-Joon, et al.
Published: (2024)
Automating Steering for Safe Multimodal Large Language Models
by: Wu, Lyucheng, et al.
Published: (2025)
by: Wu, Lyucheng, et al.
Published: (2025)
Unified Hallucination Detection for Multimodal Large Language Models
by: Chen, Xiang, et al.
Published: (2024)
by: Chen, Xiang, et al.
Published: (2024)
ChatDiet: Empowering Personalized Nutrition-Oriented Food Recommender Chatbots through an LLM-Augmented Framework
by: Yang, Zhongqi, et al.
Published: (2024)
by: Yang, Zhongqi, et al.
Published: (2024)
Multimodal Music Recommendation System using LLMs
by: Kandagatla, Srikar Prabhas, et al.
Published: (2026)
by: Kandagatla, Srikar Prabhas, et al.
Published: (2026)
Knowledge-Augmented Large Language Models for Personalized Contextual Query Suggestion
by: Baek, Jinheon, et al.
Published: (2023)
by: Baek, Jinheon, et al.
Published: (2023)
A Multimodal Single-Branch Embedding Network for Recommendation in Cold-Start and Missing Modality Scenarios
by: Ganhör, Christian, et al.
Published: (2024)
by: Ganhör, Christian, et al.
Published: (2024)
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
by: Jo, Hwiyeol, et al.
Published: (2024)
by: Jo, Hwiyeol, et al.
Published: (2024)
LightThinker: Thinking Step-by-Step Compression
by: Zhang, Jintian, et al.
Published: (2025)
by: Zhang, Jintian, et al.
Published: (2025)
LightThinker++: From Reasoning Compression to Memory Management
by: Zhu, Yuqi, et al.
Published: (2026)
by: Zhu, Yuqi, et al.
Published: (2026)
General Item Representation Learning for Cold-start Content Recommendations
by: Kim, Jooeun, et al.
Published: (2024)
by: Kim, Jooeun, et al.
Published: (2024)
TelcoAI: Advancing 3GPP Technical Specification Search through Agentic Multi-Modal Retrieval-Augmented Generation
by: Ghosh, Rahul, et al.
Published: (2025)
by: Ghosh, Rahul, et al.
Published: (2025)
Unifying Ranking and Generation in Query Auto-Completion via Retrieval-Augmented Generation and Multi-Objective Alignment
by: Yuan, Kai, et al.
Published: (2026)
by: Yuan, Kai, et al.
Published: (2026)
From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model
by: Ju, Yeong-Joon, et al.
Published: (2025)
by: Ju, Yeong-Joon, et al.
Published: (2025)
Database-Augmented Query Representation for Information Retrieval
by: Jeong, Soyeong, et al.
Published: (2024)
by: Jeong, Soyeong, et al.
Published: (2024)
VDCook:DIY video data cook your MLLMs
by: Wu, Chengwei
Published: (2026)
by: Wu, Chengwei
Published: (2026)
MCA: Modality Composition Awareness for Robust Composed Multimodal Retrieval
by: Wu, Qiyu, et al.
Published: (2025)
by: Wu, Qiyu, et al.
Published: (2025)
Learning Audio-Visual Embeddings with Inferred Latent Interaction Graphs
by: Zeng, Donghuo, et al.
Published: (2026)
by: Zeng, Donghuo, et al.
Published: (2026)
Composite Sketch+Text Queries for Retrieving Objects with Elusive Names and Complex Interactions
by: Gatti, Prajwal, et al.
Published: (2025)
by: Gatti, Prajwal, et al.
Published: (2025)
QuickGrasp: Responsive Video-Language Querying Service via Accelerated Tokenization and Edge-Augmented Inference
by: Zhang, Miao, et al.
Published: (2026)
by: Zhang, Miao, et al.
Published: (2026)
Semantic In-Domain Product Identification for Search Queries
by: Sharma, Sanat, et al.
Published: (2024)
by: Sharma, Sanat, et al.
Published: (2024)
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
by: Luo, Pengfei, et al.
Published: (2025)
by: Luo, Pengfei, et al.
Published: (2025)
Learning Smooth Distance Functions via Queries
by: Kumar, Akash, et al.
Published: (2024)
by: Kumar, Akash, et al.
Published: (2024)
Retrieval Augmented Structured Generation: Business Document Information Extraction As Tool Use
by: Cesista, Franz Louis, et al.
Published: (2024)
by: Cesista, Franz Louis, et al.
Published: (2024)
Enhancing Retrieval Processes for Language Generation with Augmented Queries
by: Ghali, Julien Pierre Edmond, et al.
Published: (2024)
by: Ghali, Julien Pierre Edmond, et al.
Published: (2024)
Query Augmentation by Decoding Semantics from Brain Signals
by: Ye, Ziyi, et al.
Published: (2024)
by: Ye, Ziyi, et al.
Published: (2024)
Segment First, Retrieve Better: Realistic Legal Search via Rhetorical Role-Based Queries
by: Nigam, Shubham Kumar, et al.
Published: (2025)
by: Nigam, Shubham Kumar, et al.
Published: (2025)
Multi-Facet Blending for Faceted Query-by-Example Retrieval
by: Do, Heejin, et al.
Published: (2024)
by: Do, Heejin, et al.
Published: (2024)
Semantic Item Graph Enhancement for Multimodal Recommendation
by: Zhang, Xiaoxiong, et al.
Published: (2025)
by: Zhang, Xiaoxiong, et al.
Published: (2025)
Personalized Image Generation with Large Multimodal Models
by: Xu, Yiyan, et al.
Published: (2024)
by: Xu, Yiyan, et al.
Published: (2024)
EventCast: Hybrid Demand Forecasting in E-Commerce with LLM-Based Event Knowledge
by: Hu, Congcong, et al.
Published: (2026)
by: Hu, Congcong, et al.
Published: (2026)
Towards Unified Multi-Modal Personalization: Large Vision-Language Models for Generative Recommendation and Beyond
by: Wei, Tianxin, et al.
Published: (2024)
by: Wei, Tianxin, et al.
Published: (2024)
Mitigating the Impact of Reference Quality on Evaluation of Summarization Systems with Reference-Free Metrics
by: Gigant, Théo, et al.
Published: (2024)
by: Gigant, Théo, et al.
Published: (2024)
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
by: McCallum, Matthew C., et al.
Published: (2024)
by: McCallum, Matthew C., et al.
Published: (2024)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
by: Wu, Jiaxin, et al.
Published: (2025)
by: Wu, Jiaxin, et al.
Published: (2025)
Synthetic Prefixes to Mitigate Bias in Real-Time Neural Query Autocomplete
by: Rajan, Adithya, et al.
Published: (2025)
by: Rajan, Adithya, et al.
Published: (2025)
Self-Augmented In-Context Learning for Unsupervised Word Translation
by: Li, Yaoyiran, et al.
Published: (2024)
by: Li, Yaoyiran, et al.
Published: (2024)
Generating Query-Relevant Document Summaries via Reinforcement Learning
by: Yadav, Nitin, et al.
Published: (2025)
by: Yadav, Nitin, et al.
Published: (2025)
SPARC-RAG: Adaptive Sequential-Parallel Scaling with Context Management for Retrieval-Augmented Generation
by: Yang, Yuxin, et al.
Published: (2026)
by: Yang, Yuxin, et al.
Published: (2026)
REMOTE: A Unified Multimodal Relation Extraction Framework with Multilevel Optimal Transport and Mixture-of-Experts
by: Lin, Xinkui, et al.
Published: (2025)
by: Lin, Xinkui, et al.
Published: (2025)
Similar Items
-
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
by: Ju, Yeong-Joon, et al.
Published: (2024) -
Automating Steering for Safe Multimodal Large Language Models
by: Wu, Lyucheng, et al.
Published: (2025) -
Unified Hallucination Detection for Multimodal Large Language Models
by: Chen, Xiang, et al.
Published: (2024) -
ChatDiet: Empowering Personalized Nutrition-Oriented Food Recommender Chatbots through an LLM-Augmented Framework
by: Yang, Zhongqi, et al.
Published: (2024) -
Multimodal Music Recommendation System using LLMs
by: Kandagatla, Srikar Prabhas, et al.
Published: (2026)