Gespeichert in:
| Hauptverfasser: | Rossetto, Luca, Schoeffmann, Klaus, Gurrin, Cathal, Lokoč, Jakub, Bailer, Werner |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2509.12000 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Results of the 2024 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
von: Tran, Allie, et al.
Veröffentlicht: (2025)
von: Tran, Allie, et al.
Veröffentlicht: (2025)
diveXplore at the Video Browser Showdown 2024
von: Schoeffmann, Klaus, et al.
Veröffentlicht: (2025)
von: Schoeffmann, Klaus, et al.
Veröffentlicht: (2025)
lifeXplore at the Lifelog Search Challenge 2021
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
OpenLifelogQA: An Open-Ended Multi-Modal Lifelog Question-Answering Dataset
von: Tran, Quang-Linh, et al.
Veröffentlicht: (2025)
von: Tran, Quang-Linh, et al.
Veröffentlicht: (2025)
On the Brittleness of CLIP Text Encoders
von: Tran, Allie, et al.
Veröffentlicht: (2025)
von: Tran, Allie, et al.
Veröffentlicht: (2025)
Performance Evaluation in Multimedia Retrieval
von: Sauter, Loris, et al.
Veröffentlicht: (2024)
von: Sauter, Loris, et al.
Veröffentlicht: (2024)
Evaluating Keyframe Layouts for Visual Known-Item Search in Homogeneous Collections
von: Jäckl, Bastian, et al.
Veröffentlicht: (2025)
von: Jäckl, Bastian, et al.
Veröffentlicht: (2025)
VCR: Video representation for Contextual Retrieval
von: Nir, Oron, et al.
Veröffentlicht: (2024)
von: Nir, Oron, et al.
Veröffentlicht: (2024)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
Robust Relevance Feedback for Interactive Known-Item Video Search
von: Ma, Zhixin, et al.
Veröffentlicht: (2025)
von: Ma, Zhixin, et al.
Veröffentlicht: (2025)
Leveraging User-Generated Metadata of Online Videos for Cover Song Identification
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
von: Wu, Jiaxin, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxin, et al.
Veröffentlicht: (2025)
Learning Item Representations Directly from Multimodal Features for Effective Recommendation
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning
von: Bi, Yuxi, et al.
Veröffentlicht: (2025)
von: Bi, Yuxi, et al.
Veröffentlicht: (2025)
Does Multimodality Improve Recommender Systems as Expected? A Critical Analysis and Future Directions
von: Zhou, Hongyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyu, et al.
Veröffentlicht: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
von: Mo, Xian, et al.
Veröffentlicht: (2025)
von: Mo, Xian, et al.
Veröffentlicht: (2025)
U-Sticker: A Large-Scale Multi-Domain User Sticker Dataset for Retrieval and Personalization
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
CM$^3$: Calibrating Multimodal Recommendation
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
MHier-RAG: Multi-Modal RAG for Visual-Rich Document Question-Answering via Hierarchical and Multi-Granularity Reasoning
von: Gong, Ziyu, et al.
Veröffentlicht: (2025)
von: Gong, Ziyu, et al.
Veröffentlicht: (2025)
Small Stickers, Big Meanings: A Multilingual Sticker Semantic Understanding Dataset with a Gamified Approach
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
HistLLM: A Unified Framework for LLM-Based Multimodal Recommendation with User History Encoding and Compression
von: Zhang, Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Chen, et al.
Veröffentlicht: (2025)
RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
From ID-based to ID-free: Rethinking ID Effectiveness in Multimodal Collaborative Filtering Recommendation
von: Li, Guohao, et al.
Veröffentlicht: (2025)
von: Li, Guohao, et al.
Veröffentlicht: (2025)
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
FashionDPO:Fine-tune Fashion Outfit Generation Model using Direct Preference Optimization
von: Yu, Mingzhe, et al.
Veröffentlicht: (2025)
von: Yu, Mingzhe, et al.
Veröffentlicht: (2025)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
A Survey on Multimodal Recommender Systems: Recent Advances and Future Directions
von: Xu, Jinfeng, et al.
Veröffentlicht: (2025)
von: Xu, Jinfeng, et al.
Veröffentlicht: (2025)
Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative Learning
von: Bin, Yi, et al.
Veröffentlicht: (2024)
von: Bin, Yi, et al.
Veröffentlicht: (2024)
Dual-Diffusional Generative Fashion Recommendation
von: Yu, Mingzhe, et al.
Veröffentlicht: (2026)
von: Yu, Mingzhe, et al.
Veröffentlicht: (2026)
Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale
von: Pan, Yongsen, et al.
Veröffentlicht: (2026)
von: Pan, Yongsen, et al.
Veröffentlicht: (2026)
Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
CLEAR: Null-Space Projection for Cross-Modal De-Redundancy in Multimodal Recommendation
von: Zhan, Hao, et al.
Veröffentlicht: (2026)
von: Zhan, Hao, et al.
Veröffentlicht: (2026)
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
von: Xu, Jinfeng, et al.
Veröffentlicht: (2026)
von: Xu, Jinfeng, et al.
Veröffentlicht: (2026)
Multimodal Graph Neural Network for Recommendation with Dynamic De-redundancy and Modality-Guided Feature De-noisy
von: Mo, Feng, et al.
Veröffentlicht: (2024)
von: Mo, Feng, et al.
Veröffentlicht: (2024)
The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval
von: Fu, Junchen, et al.
Veröffentlicht: (2026)
von: Fu, Junchen, et al.
Veröffentlicht: (2026)
Enhancing Image-Text Matching with Adaptive Feature Aggregation
von: Wang, Zuhui, et al.
Veröffentlicht: (2024)
von: Wang, Zuhui, et al.
Veröffentlicht: (2024)
Multimodal Learned Sparse Retrieval for Image Suggestion
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Results of the 2024 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2024) -
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
von: Tran, Allie, et al.
Veröffentlicht: (2025) -
diveXplore at the Video Browser Showdown 2024
von: Schoeffmann, Klaus, et al.
Veröffentlicht: (2025) -
lifeXplore at the Lifelog Search Challenge 2021
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025) -
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)