StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Bi, Yuxi, Gao, Yunfan, Wang, Haofen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FashionDPO:Fine-tune Fashion Outfit Generation Model using Direct Preference Optimization
por: Yu, Mingzhe, et al.
Publicado: (2025)
por: Yu, Mingzhe, et al.
Publicado: (2025)
Breaking the Curse of Knowledge: Towards Effective Multimodal Recommendation using Knowledge Soft Integration
por: Ouyang, Kai, et al.
Publicado: (2023)
por: Ouyang, Kai, et al.
Publicado: (2023)
RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation
por: Tourani, Ali, et al.
Publicado: (2025)
por: Tourani, Ali, et al.
Publicado: (2025)
Synergizing RAG and Reasoning: A Systematic Review
por: Gao, Yunfan, et al.
Publicado: (2025)
por: Gao, Yunfan, et al.
Publicado: (2025)
MHier-RAG: Multi-Modal RAG for Visual-Rich Document Question-Answering via Hierarchical and Multi-Granularity Reasoning
por: Gong, Ziyu, et al.
Publicado: (2025)
por: Gong, Ziyu, et al.
Publicado: (2025)
U-Sticker: A Large-Scale Multi-Domain User Sticker Dataset for Retrieval and Personalization
por: Chee, Heng Er Metilda, et al.
Publicado: (2025)
por: Chee, Heng Er Metilda, et al.
Publicado: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
por: Mo, Xian, et al.
Publicado: (2025)
por: Mo, Xian, et al.
Publicado: (2025)
Multimodal Graph Neural Network for Recommendation with Dynamic De-redundancy and Modality-Guided Feature De-noisy
por: Mo, Feng, et al.
Publicado: (2024)
por: Mo, Feng, et al.
Publicado: (2024)
Uni-Retrieval: A Multi-Style Retrieval Framework for STEM's Education
por: Jia, Yanhao, et al.
Publicado: (2025)
por: Jia, Yanhao, et al.
Publicado: (2025)
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
por: Xu, Jinfeng, et al.
Publicado: (2026)
por: Xu, Jinfeng, et al.
Publicado: (2026)
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
por: Luo, Pengfei, et al.
Publicado: (2025)
por: Luo, Pengfei, et al.
Publicado: (2025)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
por: Wu, Jiaxin, et al.
Publicado: (2025)
por: Wu, Jiaxin, et al.
Publicado: (2025)
Towards Unified Multi-Modal Personalization: Large Vision-Language Models for Generative Recommendation and Beyond
por: Wei, Tianxin, et al.
Publicado: (2024)
por: Wei, Tianxin, et al.
Publicado: (2024)
MDF: A Dynamic Fusion Model for Multi-modal Fake News Detection
por: Lv, Hongzhen, et al.
Publicado: (2024)
por: Lv, Hongzhen, et al.
Publicado: (2024)
OpenLifelogQA: An Open-Ended Multi-Modal Lifelog Question-Answering Dataset
por: Tran, Quang-Linh, et al.
Publicado: (2025)
por: Tran, Quang-Linh, et al.
Publicado: (2025)
Agentic Mixed-Source Multi-Modal Misinformation Detection with Adaptive Test-Time Scaling
por: Jiang, Wei, et al.
Publicado: (2026)
por: Jiang, Wei, et al.
Publicado: (2026)
Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning
por: Zhang, Lingzi, et al.
Publicado: (2023)
por: Zhang, Lingzi, et al.
Publicado: (2023)
Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative Learning
por: Bin, Yi, et al.
Publicado: (2024)
por: Bin, Yi, et al.
Publicado: (2024)
Personalized Image Generation with Large Multimodal Models
por: Xu, Yiyan, et al.
Publicado: (2024)
por: Xu, Yiyan, et al.
Publicado: (2024)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
por: Wang, Haoyu, et al.
Publicado: (2025)
por: Wang, Haoyu, et al.
Publicado: (2025)
Don't Lose Yourself: Boosting Multimodal Recommendation via Reducing Node-neighbor Discrepancy in Graph Convolutional Network
por: Chen, Zheyu, et al.
Publicado: (2024)
por: Chen, Zheyu, et al.
Publicado: (2024)
Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale
por: Pan, Yongsen, et al.
Publicado: (2026)
por: Pan, Yongsen, et al.
Publicado: (2026)
CM$^3$: Calibrating Multimodal Recommendation
por: Zhou, Xin, et al.
Publicado: (2025)
por: Zhou, Xin, et al.
Publicado: (2025)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
por: Jaspal, Amit, et al.
Publicado: (2025)
por: Jaspal, Amit, et al.
Publicado: (2025)
Enhancing Image-Text Matching with Adaptive Feature Aggregation
por: Wang, Zuhui, et al.
Publicado: (2024)
por: Wang, Zuhui, et al.
Publicado: (2024)
CLEAR: Null-Space Projection for Cross-Modal De-Redundancy in Multimodal Recommendation
por: Zhan, Hao, et al.
Publicado: (2026)
por: Zhan, Hao, et al.
Publicado: (2026)
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
por: Li, Yang, et al.
Publicado: (2025)
por: Li, Yang, et al.
Publicado: (2025)
Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation
por: Ma, Hongjian, et al.
Publicado: (2026)
por: Ma, Hongjian, et al.
Publicado: (2026)
Small Stickers, Big Meanings: A Multilingual Sticker Semantic Understanding Dataset with a Gamified Approach
por: Chee, Heng Er Metilda, et al.
Publicado: (2025)
por: Chee, Heng Er Metilda, et al.
Publicado: (2025)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
por: Sun, Huatuan, et al.
Publicado: (2025)
por: Sun, Huatuan, et al.
Publicado: (2025)
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
por: Wang, Tianshi, et al.
Publicado: (2023)
por: Wang, Tianshi, et al.
Publicado: (2023)
A Survey on Multimodal Recommender Systems: Recent Advances and Future Directions
por: Xu, Jinfeng, et al.
Publicado: (2025)
por: Xu, Jinfeng, et al.
Publicado: (2025)
The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval
por: Fu, Junchen, et al.
Publicado: (2026)
por: Fu, Junchen, et al.
Publicado: (2026)
Jamendo-MT-QA: A Benchmark for Multi-Track Comparative Music Question Answering
por: Koh, Junyoung, et al.
Publicado: (2026)
por: Koh, Junyoung, et al.
Publicado: (2026)
A Comprehensive Survey of Knowledge-Based Vision Question Answering Systems: The Lifecycle of Knowledge in Visual Reasoning Task
por: Deng, Jiaqi, et al.
Publicado: (2025)
por: Deng, Jiaqi, et al.
Publicado: (2025)
Enhancing Automatic Chord Recognition via Pseudo-Labeling and Knowledge Distillation
por: Phan, Nghia, et al.
Publicado: (2026)
por: Phan, Nghia, et al.
Publicado: (2026)
Learning Item Representations Directly from Multimodal Features for Effective Recommendation
por: Zhou, Xin, et al.
Publicado: (2025)
por: Zhou, Xin, et al.
Publicado: (2025)
Does Multimodality Improve Recommender Systems as Expected? A Critical Analysis and Future Directions
por: Zhou, Hongyu, et al.
Publicado: (2025)
por: Zhou, Hongyu, et al.
Publicado: (2025)
lifeXplore at the Lifelog Search Challenge 2021
por: Leibetseder, Andreas, et al.
Publicado: (2025)
por: Leibetseder, Andreas, et al.
Publicado: (2025)
Robust Relevance Feedback for Interactive Known-Item Video Search
por: Ma, Zhixin, et al.
Publicado: (2025)
por: Ma, Zhixin, et al.
Publicado: (2025)
Ejemplares similares
-
FashionDPO:Fine-tune Fashion Outfit Generation Model using Direct Preference Optimization
por: Yu, Mingzhe, et al.
Publicado: (2025) -
Breaking the Curse of Knowledge: Towards Effective Multimodal Recommendation using Knowledge Soft Integration
por: Ouyang, Kai, et al.
Publicado: (2023) -
RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation
por: Tourani, Ali, et al.
Publicado: (2025) -
Synergizing RAG and Reasoning: A Systematic Review
por: Gao, Yunfan, et al.
Publicado: (2025) -
MHier-RAG: Multi-Modal RAG for Visual-Rich Document Question-Answering via Hierarchical and Multi-Granularity Reasoning
por: Gong, Ziyu, et al.
Publicado: (2025)