LD4MRec: Simplifying and Powering Diffusion Model for Multimedia Recommendation

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Zhu, Jiarui, Hou, Jun, Yu, Penghang, Tan, Zhiyi, Bao, Bing-Kun
Natura: Preprint
Pubblicazione: 2023
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910909384884224
author Zhu, Jiarui
Hou, Jun
Yu, Penghang
Tan, Zhiyi
Bao, Bing-Kun
author_facet Zhu, Jiarui
Hou, Jun
Yu, Penghang
Tan, Zhiyi
Bao, Bing-Kun
contents Multimedia recommendation aims to predict users' future behaviors based on observed behaviors and item content information. However, the inherent noise contained in observed behaviors easily leads to suboptimal recommendation performance. Recently, the diffusion model's ability to generate information from noise presents a promising solution to this issue, prompting us to explore its application in multimedia recommendation. Nonetheless, several challenges must be addressed: 1) The diffusion model requires simplification to meet the efficiency requirements of real-time recommender systems, 2) The generated behaviors must align with user preference. To address these challenges, we propose a Light Diffusion model for Multimedia Recommendation (LD4MRec). LD4MRec largely reduces computational complexity by employing a forward-free inference strategy, which directly predicts future behaviors from observed noisy behaviors. Meanwhile, to ensure the alignment between generated behaviors and user preference, we propose a novel Conditional neural Network (C-Net). C-Net achieves guided generation by leveraging two key signals, collaborative signals and personalized modality preference signals, thereby improving the semantic consistency between generated behaviors and user preference. Experiments conducted on three real-world datasets demonstrate the effectiveness of LD4MRec.
format Preprint
id arxiv_https___arxiv_org_abs_2309_15363
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle LD4MRec: Simplifying and Powering Diffusion Model for Multimedia Recommendation
Zhu, Jiarui
Hou, Jun
Yu, Penghang
Tan, Zhiyi
Bao, Bing-Kun
Information Retrieval
Multimedia recommendation aims to predict users' future behaviors based on observed behaviors and item content information. However, the inherent noise contained in observed behaviors easily leads to suboptimal recommendation performance. Recently, the diffusion model's ability to generate information from noise presents a promising solution to this issue, prompting us to explore its application in multimedia recommendation. Nonetheless, several challenges must be addressed: 1) The diffusion model requires simplification to meet the efficiency requirements of real-time recommender systems, 2) The generated behaviors must align with user preference. To address these challenges, we propose a Light Diffusion model for Multimedia Recommendation (LD4MRec). LD4MRec largely reduces computational complexity by employing a forward-free inference strategy, which directly predicts future behaviors from observed noisy behaviors. Meanwhile, to ensure the alignment between generated behaviors and user preference, we propose a novel Conditional neural Network (C-Net). C-Net achieves guided generation by leveraging two key signals, collaborative signals and personalized modality preference signals, thereby improving the semantic consistency between generated behaviors and user preference. Experiments conducted on three real-world datasets demonstrate the effectiveness of LD4MRec.
title LD4MRec: Simplifying and Powering Diffusion Model for Multimedia Recommendation
topic Information Retrieval
url https://arxiv.org/abs/2309.15363