Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jun, Lou, Xuhang, Wang, Jinpeng, Wang, Yuting, Wang, Yaowei, Xia, Shu-Tao, Chen, Bin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval
von: Li, Jun, et al.
Veröffentlicht: (2026)
von: Li, Jun, et al.
Veröffentlicht: (2026)
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval
von: Wang, Yuting, et al.
Veröffentlicht: (2023)
von: Wang, Yuting, et al.
Veröffentlicht: (2023)
AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing
von: Lian, Niu, et al.
Veröffentlicht: (2025)
von: Lian, Niu, et al.
Veröffentlicht: (2025)
Efficient Self-Supervised Video Hashing with Selective State Spaces
von: Wang, Jinpeng, et al.
Veröffentlicht: (2024)
von: Wang, Jinpeng, et al.
Veröffentlicht: (2024)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
von: Lian, Niu, et al.
Veröffentlicht: (2026)
von: Lian, Niu, et al.
Veröffentlicht: (2026)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
von: Jaspal, Amit, et al.
Veröffentlicht: (2025)
Love Me, Love My Label: Rethinking the Role of Labels in Prompt Retrieval for Visual In-Context Learning
von: Luo, Tianci, et al.
Veröffentlicht: (2026)
von: Luo, Tianci, et al.
Veröffentlicht: (2026)
Robust Relevance Feedback for Interactive Known-Item Video Search
von: Ma, Zhixin, et al.
Veröffentlicht: (2025)
von: Ma, Zhixin, et al.
Veröffentlicht: (2025)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
VCR: Video representation for Contextual Retrieval
von: Nir, Oron, et al.
Veröffentlicht: (2024)
von: Nir, Oron, et al.
Veröffentlicht: (2024)
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale
von: Pan, Yongsen, et al.
Veröffentlicht: (2026)
von: Pan, Yongsen, et al.
Veröffentlicht: (2026)
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
von: Mo, Xian, et al.
Veröffentlicht: (2025)
von: Mo, Xian, et al.
Veröffentlicht: (2025)
StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning
von: Bi, Yuxi, et al.
Veröffentlicht: (2025)
von: Bi, Yuxi, et al.
Veröffentlicht: (2025)
Enhancing Image-Text Matching with Adaptive Feature Aggregation
von: Wang, Zuhui, et al.
Veröffentlicht: (2024)
von: Wang, Zuhui, et al.
Veröffentlicht: (2024)
RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
U-Sticker: A Large-Scale Multi-Domain User Sticker Dataset for Retrieval and Personalization
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval
von: Fu, Junchen, et al.
Veröffentlicht: (2026)
von: Fu, Junchen, et al.
Veröffentlicht: (2026)
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
von: Xu, Jinfeng, et al.
Veröffentlicht: (2026)
von: Xu, Jinfeng, et al.
Veröffentlicht: (2026)
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
von: Luo, Pengfei, et al.
Veröffentlicht: (2025)
von: Luo, Pengfei, et al.
Veröffentlicht: (2025)
Performance Evaluation in Multimedia Retrieval
von: Sauter, Loris, et al.
Veröffentlicht: (2024)
von: Sauter, Loris, et al.
Veröffentlicht: (2024)
Multimodal Learned Sparse Retrieval for Image Suggestion
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
von: Wu, Jiaxin, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxin, et al.
Veröffentlicht: (2025)
Results of the 2025 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
Results of the 2024 Video Browser Showdown
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
von: Rossetto, Luca, et al.
Veröffentlicht: (2024)
Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative Learning
von: Bin, Yi, et al.
Veröffentlicht: (2024)
von: Bin, Yi, et al.
Veröffentlicht: (2024)
From Swath to Full-Disc: Advancing Precipitation Retrieval with Multimodal Knowledge Expansion
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
CM$^3$: Calibrating Multimodal Recommendation
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
von: Zhou, Xin, et al.
Veröffentlicht: (2025)
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
von: Tran, Allie, et al.
Veröffentlicht: (2025)
von: Tran, Allie, et al.
Veröffentlicht: (2025)
Leveraging User-Generated Metadata of Online Videos for Cover Song Identification
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
CLEAR: Null-Space Projection for Cross-Modal De-Redundancy in Multimodal Recommendation
von: Zhan, Hao, et al.
Veröffentlicht: (2026)
von: Zhan, Hao, et al.
Veröffentlicht: (2026)
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval
von: Ning, Hailong, et al.
Veröffentlicht: (2025)
von: Ning, Hailong, et al.
Veröffentlicht: (2025)
Multimodal Graph Neural Network for Recommendation with Dynamic De-redundancy and Modality-Guided Feature De-noisy
von: Mo, Feng, et al.
Veröffentlicht: (2024)
von: Mo, Feng, et al.
Veröffentlicht: (2024)
Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
Small Stickers, Big Meanings: A Multilingual Sticker Semantic Understanding Dataset with a Gamified Approach
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
von: Chee, Heng Er Metilda, et al.
Veröffentlicht: (2025)
FashionDPO:Fine-tune Fashion Outfit Generation Model using Direct Preference Optimization
von: Yu, Mingzhe, et al.
Veröffentlicht: (2025)
von: Yu, Mingzhe, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval
von: Li, Jun, et al.
Veröffentlicht: (2026) -
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
von: Li, Jun, et al.
Veröffentlicht: (2025) -
GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval
von: Wang, Yuting, et al.
Veröffentlicht: (2023) -
AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing
von: Lian, Niu, et al.
Veröffentlicht: (2025) -
Efficient Self-Supervised Video Hashing with Selective State Spaces
von: Wang, Jinpeng, et al.
Veröffentlicht: (2024)