VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Giahi, Ramin, Yao, Kehui, Kollipara, Sriram, Zhao, Kai, Mirjalili, Vahid, Xu, Jianpeng, Biswas, Topojoy, Korpeoglu, Evren, Achan, Kannan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2025)
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2025)
Decoding Style: Efficient Fine-Tuning of LLMs for Image-Guided Outfit Recommendation with Preference
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024)
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024)
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
von: Mirjalili, Vahid, et al.
Veröffentlicht: (2025)
von: Mirjalili, Vahid, et al.
Veröffentlicht: (2025)
CARTS: Collaborative Agents for Recommendation Textual Summarization
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024)
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024)
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation
von: Cao, Yanan, et al.
Veröffentlicht: (2026)
von: Cao, Yanan, et al.
Veröffentlicht: (2026)
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction
von: Fang, Chenhao, et al.
Veröffentlicht: (2024)
von: Fang, Chenhao, et al.
Veröffentlicht: (2024)
Grocery to General Merchandise: A Cross-Pollination Recommender using LLMs and Real-Time Cart Context
von: Kekuda, Akshay, et al.
Veröffentlicht: (2025)
von: Kekuda, Akshay, et al.
Veröffentlicht: (2025)
Latent Customer Segmentation and Value-Based Recommendation Leveraging a Two-Stage Model with Missing Labels
von: Gopalakrishnan, Keerthi, et al.
Veröffentlicht: (2026)
von: Gopalakrishnan, Keerthi, et al.
Veröffentlicht: (2026)
Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking
von: Che, Yiming, et al.
Veröffentlicht: (2026)
von: Che, Yiming, et al.
Veröffentlicht: (2026)
GRACE: Generative Recommendation via Journey-Aware Sparse Attention on Chain-of-Thought Tokenization
von: Ma, Luyi, et al.
Veröffentlicht: (2025)
von: Ma, Luyi, et al.
Veröffentlicht: (2025)
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
von: Fan, Zezhong, et al.
Veröffentlicht: (2025)
Triple Modality Fusion: Aligning Visual, Textual, and Graph Data with Large Language Models for Multi-Behavior Recommendations
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
Event-based Product Carousel Recommendation with Query-Click Graph
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
S2SRec2: Set-to-Set Recommendation for Basket Completion with Recipe
von: Cao, Yanan, et al.
Veröffentlicht: (2025)
von: Cao, Yanan, et al.
Veröffentlicht: (2025)
Leveraging User-Generated Reviews for Recommender Systems with Dynamic Headers
von: Vashishtha, Shanu, et al.
Veröffentlicht: (2024)
von: Vashishtha, Shanu, et al.
Veröffentlicht: (2024)
Dynamic Decision Making in Engineering System Design: A Deep Q-Learning Approach
von: Giahi, Ramin, et al.
Veröffentlicht: (2023)
von: Giahi, Ramin, et al.
Veröffentlicht: (2023)
MetaSynth: Multi-Agent Metadata Generation from Implicit Feedback in Black-Box Systems
von: Srirangamsridharan, Shreeranjani, et al.
Veröffentlicht: (2025)
von: Srirangamsridharan, Shreeranjani, et al.
Veröffentlicht: (2025)
CRAB: Codebook Rebalancing for Bias Mitigation in Generative Recommendation
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
NeuroCLIP: Brain-Inspired Prompt Tuning for EEG-to-Image Multimodal Contrastive Learning
von: Wang, Jiyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiyuan, et al.
Veröffentlicht: (2025)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks
von: Ma, Luyi, et al.
Veröffentlicht: (2026)
von: Ma, Luyi, et al.
Veröffentlicht: (2026)
Segment and Matte Anything in a Unified Model
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
von: Fan, Zezhong, et al.
Veröffentlicht: (2026)
On the Brittleness of CLIP Text Encoders
von: Tran, Allie, et al.
Veröffentlicht: (2025)
von: Tran, Allie, et al.
Veröffentlicht: (2025)
Integrating Visual and Textual Inputs for Searching Large-Scale Map Collections with CLIP
von: Mahowald, Jamie, et al.
Veröffentlicht: (2024)
von: Mahowald, Jamie, et al.
Veröffentlicht: (2024)
ARAG: Agentic Retrieval Augmented Generation for Personalized Recommendation
von: Maragheh, Reza Yousefi, et al.
Veröffentlicht: (2025)
von: Maragheh, Reza Yousefi, et al.
Veröffentlicht: (2025)
Improving Sequential Recommender Systems with Online and In-store User Behavior
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
von: Ma, Luyi, et al.
Veröffentlicht: (2024)
CLIP-Branches: Interactive Fine-Tuning for Text-Image Retrieval
von: Lülf, Christian, et al.
Veröffentlicht: (2024)
von: Lülf, Christian, et al.
Veröffentlicht: (2024)
Serendipitous Recommendation with Multimodal LLM
von: Wang, Haoting, et al.
Veröffentlicht: (2025)
von: Wang, Haoting, et al.
Veröffentlicht: (2025)
Stealthy LLM-Driven Data Poisoning Attacks Against Embedding-Based Retrieval-Augmented Recommender Systems
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2025)
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2025)
ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2025)
von: Nazary, Fatemeh, et al.
Veröffentlicht: (2025)
TutorLLM: Customizing Learning Recommendations with Knowledge Tracing and Retrieval-Augmented Generation
von: Li, Zhaoxing, et al.
Veröffentlicht: (2025)
von: Li, Zhaoxing, et al.
Veröffentlicht: (2025)
ACE: Anisotropy-Controllable Embedding for LLM-enhanced Sequential Recommendation
von: Lee, Dongcheol, et al.
Veröffentlicht: (2026)
von: Lee, Dongcheol, et al.
Veröffentlicht: (2026)
MOSAIC: Multimodal Multistakeholder-aware Visual Art Recommendation
von: Yilma, Bereket A., et al.
Veröffentlicht: (2024)
von: Yilma, Bereket A., et al.
Veröffentlicht: (2024)
EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models
von: Meng, GuangHao, et al.
Veröffentlicht: (2025)
von: Meng, GuangHao, et al.
Veröffentlicht: (2025)
Knowledge Graph Retrieval-Augmented Generation for LLM-based Recommendation
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
Jina CLIP: Your CLIP Model Is Also Your Text Retriever
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
CoLLM: Integrating Collaborative Embeddings into Large Language Models for Recommendation
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
von: Zhang, Yang, et al.
Veröffentlicht: (2023)
Continuous Input Embedding Size Search For Recommender Systems
von: Qu, Yunke, et al.
Veröffentlicht: (2023)
von: Qu, Yunke, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
CAL-RAG: Retrieval-Augmented Multi-Agent Generation for Content-Aware Layout Design
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2025) -
Decoding Style: Efficient Fine-Tuning of LLMs for Image-Guided Outfit Recommendation with Preference
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024) -
Spatial Reasoning in Foundation Models: Benchmarking Object-Centric Spatial Understanding
von: Mirjalili, Vahid, et al.
Veröffentlicht: (2025) -
CARTS: Collaborative Agents for Recommendation Textual Summarization
von: Chen, Jiao, et al.
Veröffentlicht: (2025) -
Character-based Outfit Generation with Vision-augmented Style Extraction via LLMs
von: Forouzandehmehr, Najmeh, et al.
Veröffentlicht: (2024)