CART: A Generative Cross-Modal Retrieval Framework with Coarse-To-Fine Semantic Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Minghui, Ji, Shengpeng, Zuo, Jialong, Huang, Hai, Xia, Yan, Zhu, Jieming, Cheng, Xize, Yang, Xiaoda, Liu, Wenrui, Wang, Gang, Dong, Zhenhua, Zhao, Zhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
von: Hu, Ruofan, et al.
Veröffentlicht: (2025)
von: Hu, Ruofan, et al.
Veröffentlicht: (2025)
EAGER-LLM: Enhancing Large Language Models as Recommenders through Exogenous Behavior-Semantic Integration
von: Hong, Minjie, et al.
Veröffentlicht: (2025)
von: Hong, Minjie, et al.
Veröffentlicht: (2025)
UNGER: Generative Recommendation with A Unified Code via Semantic and Collaborative Integration
von: Xiao, Longtao, et al.
Veröffentlicht: (2025)
von: Xiao, Longtao, et al.
Veröffentlicht: (2025)
MLLM-Driven Semantic Identifier Generation for Generative Cross-Modal Retrieval
von: Li, Tianyuan, et al.
Veröffentlicht: (2025)
von: Li, Tianyuan, et al.
Veröffentlicht: (2025)
CoST: Contrastive Quantization based Semantic Tokenization for Generative Recommendation
von: Zhu, Jieming, et al.
Veröffentlicht: (2024)
von: Zhu, Jieming, et al.
Veröffentlicht: (2024)
Enhancing Multimodal Unified Representations for Cross Modal Generalization
von: Huang, Hai, et al.
Veröffentlicht: (2024)
von: Huang, Hai, et al.
Veröffentlicht: (2024)
Multimodal Pretraining and Generation for Recommendation: A Tutorial
von: Zhu, Jieming, et al.
Veröffentlicht: (2024)
von: Zhu, Jieming, et al.
Veröffentlicht: (2024)
Entropy-based Coarse and Compressed Semantic Speech Representation Learning
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
Recall-Augmented Ranking: Enhancing Click-Through Rate Prediction Accuracy with Cross-Stage Data
von: Huang, Junjie, et al.
Veröffentlicht: (2024)
von: Huang, Junjie, et al.
Veröffentlicht: (2024)
Rhythm Controllable and Efficient Zero-Shot Voice Conversion via Shortcut Flow Matching
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
Counteracting Duration Bias in Video Recommendation via Counterfactual Watch Time
von: Zhao, Haiyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Haiyuan, et al.
Veröffentlicht: (2024)
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
EAGER: Two-Stream Generative Recommender with Behavior-Semantic Collaboration
von: Wang, Ye, et al.
Veröffentlicht: (2024)
von: Wang, Ye, et al.
Veröffentlicht: (2024)
A Unified Optimal Transport Framework for Cross-Modal Retrieval with Noisy Labels
von: Han, Haochen, et al.
Veröffentlicht: (2024)
von: Han, Haochen, et al.
Veröffentlicht: (2024)
FunnelRAG: A Coarse-to-Fine Progressive Retrieval Paradigm for RAG
von: Zhao, Xinping, et al.
Veröffentlicht: (2024)
von: Zhao, Xinping, et al.
Veröffentlicht: (2024)
Towards Cross-Modal Text-Molecule Retrieval with Better Modality Alignment
von: Song, Jia, et al.
Veröffentlicht: (2024)
von: Song, Jia, et al.
Veröffentlicht: (2024)
RREH: Reconstruction Relations Embedded Hashing for Semi-Paired Cross-Modal Retrieval
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
CORONA: A Coarse-to-Fine Framework for Graph-based Recommendation with Large Language Models
von: Chen, Junze, et al.
Veröffentlicht: (2025)
von: Chen, Junze, et al.
Veröffentlicht: (2025)
XR: Cross-Modal Agents for Composed Image Retrieval
von: Yang, Zhongyu, et al.
Veröffentlicht: (2026)
von: Yang, Zhongyu, et al.
Veröffentlicht: (2026)
RecBase: Generative Foundation Model Pretraining for Zero-Shot Recommendation
von: Zhou, Sashuai, et al.
Veröffentlicht: (2025)
von: Zhou, Sashuai, et al.
Veröffentlicht: (2025)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
Learning Multi-Aspect Item Palette: A Semantic Tokenization Framework for Generative Recommendation
von: Liu, Qijiong, et al.
Veröffentlicht: (2024)
von: Liu, Qijiong, et al.
Veröffentlicht: (2024)
Semantic-enhanced Modality-asymmetric Retrieval for Online E-commerce Search
von: Zhou, Zhigong, et al.
Veröffentlicht: (2025)
von: Zhou, Zhigong, et al.
Veröffentlicht: (2025)
TayFCS: Towards Light Feature Combination Selection for Deep Recommender Systems
von: Wang, Xianquan, et al.
Veröffentlicht: (2025)
von: Wang, Xianquan, et al.
Veröffentlicht: (2025)
FairFS: Addressing Deep Feature Selection Biases for Recommender System
von: Wang, Xianquan, et al.
Veröffentlicht: (2026)
von: Wang, Xianquan, et al.
Veröffentlicht: (2026)
MSAM: Multi-Semantic Adaptive Mining for Cross-Modal Drone Video-Text Retrieval
von: Huang, Jinghao, et al.
Veröffentlicht: (2025)
von: Huang, Jinghao, et al.
Veröffentlicht: (2025)
OmniSep: Unified Omni-Modality Sound Separation with Query-Mixup
von: Cheng, Xize, et al.
Veröffentlicht: (2024)
von: Cheng, Xize, et al.
Veröffentlicht: (2024)
Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark
von: Liu, Qijiong, et al.
Veröffentlicht: (2025)
von: Liu, Qijiong, et al.
Veröffentlicht: (2025)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
von: Dai, Sunhao, et al.
Veröffentlicht: (2024)
Rebalancing Contrastive Alignment with Bottlenecked Semantic Increments in Text-Video Retrieval
von: Xiao, Jian, et al.
Veröffentlicht: (2025)
von: Xiao, Jian, et al.
Veröffentlicht: (2025)
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
Retrieval Augmented Cross-Modal Tag Recommendation in Software Q&A Sites
von: Lu, Sijin, et al.
Veröffentlicht: (2024)
von: Lu, Sijin, et al.
Veröffentlicht: (2024)
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
Improving Semantic Proximity in Information Retrieval through Cross-Lingual Alignment
von: Hong, Seongtae, et al.
Veröffentlicht: (2026)
von: Hong, Seongtae, et al.
Veröffentlicht: (2026)
Cross-Modal Attention Network with Dual Graph Learning in Multimodal Recommendation
von: Dai, Ji, et al.
Veröffentlicht: (2026)
von: Dai, Ji, et al.
Veröffentlicht: (2026)
HASH-RAG: Bridging Deep Hashing with Retriever for Efficient, Fine Retrieval and Augmented Generation
von: Guo, Jinyu, et al.
Veröffentlicht: (2025)
von: Guo, Jinyu, et al.
Veröffentlicht: (2025)
Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity Documents
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Cross-Lingual Retrieval
von: Zuo, Longfei, et al.
Veröffentlicht: (2025)
von: Zuo, Longfei, et al.
Veröffentlicht: (2025)
FAIR: Focused Attention Is All You Need for Generative Recommendation
von: Xiao, Longtao, et al.
Veröffentlicht: (2025)
von: Xiao, Longtao, et al.
Veröffentlicht: (2025)
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
von: Hu, Ruofan, et al.
Veröffentlicht: (2025) -
EAGER-LLM: Enhancing Large Language Models as Recommenders through Exogenous Behavior-Semantic Integration
von: Hong, Minjie, et al.
Veröffentlicht: (2025) -
UNGER: Generative Recommendation with A Unified Code via Semantic and Collaborative Integration
von: Xiao, Longtao, et al.
Veröffentlicht: (2025) -
MLLM-Driven Semantic Identifier Generation for Generative Cross-Modal Retrieval
von: Li, Tianyuan, et al.
Veröffentlicht: (2025) -
CoST: Contrastive Quantization based Semantic Tokenization for Generative Recommendation
von: Zhu, Jieming, et al.
Veröffentlicht: (2024)