Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Jiacheng, Xiao, Mingyuan, Wang, Weijian, Du, Yuxin, Wu, Zhengze, Hua, Cheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M3TR: Temporal Retrieval Enhanced Multi-Modal Micro-video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web
von: Heo, Ryang, et al.
Veröffentlicht: (2026)
von: Heo, Ryang, et al.
Veröffentlicht: (2026)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
von: Wang, Dali, et al.
Veröffentlicht: (2026)
von: Wang, Dali, et al.
Veröffentlicht: (2026)
FinCall-Surprise: A Large Scale Multi-modal Benchmark for Earning Surprise Prediction
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
Compression Metadata-assisted RoI Extraction and Adaptive Inference for Efficient Video Analytics
von: Wang, Chengzhi, et al.
Veröffentlicht: (2025)
von: Wang, Chengzhi, et al.
Veröffentlicht: (2025)
PopSim: Social Network Simulation for Social Media Popularity Prediction
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
Anchoring Trends: Mitigating Social Media Popularity Prediction Drift via Feature Clustering and Expansion
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
von: Lee, Chia-Ming, et al.
Veröffentlicht: (2025)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
von: Sun, Huatuan, et al.
Veröffentlicht: (2025)
High-level Codes and Fine-grained Weights for Online Multi-modal Hashing Retrieval
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2024)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2024)
Characterizing Multimedia Information Environment through Multi-modal Clustering of YouTube Videos
von: Yousefi, Niloofar, et al.
Veröffentlicht: (2024)
von: Yousefi, Niloofar, et al.
Veröffentlicht: (2024)
Clinical Multi-modal Fusion with Heterogeneous Graph and Disease Correlation Learning for Multi-Disease Prediction
von: Jiang, Yueheng, et al.
Veröffentlicht: (2025)
von: Jiang, Yueheng, et al.
Veröffentlicht: (2025)
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2024)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
von: Wang, Sen, et al.
Veröffentlicht: (2024)
von: Wang, Sen, et al.
Veröffentlicht: (2024)
Routing Experts: Learning to Route Dynamic Experts in Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
Towards Structure-aware Model for Multi-modal Knowledge Graph Completion
von: Li, Linyu, et al.
Veröffentlicht: (2025)
von: Li, Linyu, et al.
Veröffentlicht: (2025)
Deep Mamba Multi-modal Learning
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
Leveraging User-Generated Metadata of Online Videos for Cover Song Identification
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
Revisiting Vision-Language Features Adaptation and Inconsistency for Social Media Popularity Prediction
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
When Top-ranked Recommendations Fail: Modeling Multi-Granular Negative Feedback for Explainable and Robust Video Recommendation
von: Chen, Siran, et al.
Veröffentlicht: (2025)
von: Chen, Siran, et al.
Veröffentlicht: (2025)
Multi-view Hypergraph-based Contrastive Learning Model for Cold-Start Micro-video Recommendation
von: Lyu, Sisuo, et al.
Veröffentlicht: (2024)
von: Lyu, Sisuo, et al.
Veröffentlicht: (2024)
EEmo-Bench: A Benchmark for Multi-modal Large Language Models on Image Evoked Emotion Assessment
von: Gao, Lancheng, et al.
Veröffentlicht: (2025)
von: Gao, Lancheng, et al.
Veröffentlicht: (2025)
Audio Matters Too! Enhancing Markerless Motion Capture with Audio Signals for String Performance Capture
von: Jin, Yitong, et al.
Veröffentlicht: (2024)
von: Jin, Yitong, et al.
Veröffentlicht: (2024)
Accelerating Multi-Condition T2I Generation via Adaptive Condition Offloading and Pruning
von: Kong, Yuxin, et al.
Veröffentlicht: (2026)
von: Kong, Yuxin, et al.
Veröffentlicht: (2026)
Short-Form Video Viewing Behavior Analysis and Multi-Step Viewing Time Prediction
von: Yen, Vu Thi Hai, et al.
Veröffentlicht: (2026)
von: Yen, Vu Thi Hai, et al.
Veröffentlicht: (2026)
Copycat vs. Original: Multi-modal Pretraining and Variable Importance in Box-office Prediction
von: Chao, Qin, et al.
Veröffentlicht: (2025)
von: Chao, Qin, et al.
Veröffentlicht: (2025)
M3SD: Multi-modal, Multi-scenario and Multi-language Speaker Diarization Dataset
von: Wu, Shilong
Veröffentlicht: (2025)
von: Wu, Shilong
Veröffentlicht: (2025)
CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
von: He, Zheqi, et al.
Veröffentlicht: (2024)
von: He, Zheqi, et al.
Veröffentlicht: (2024)
Challenging Dataset and Multi-modal Gated Mixture of Experts Model for Remote Sensing Copy-Move Forgery Understanding
von: Zhang, Ze, et al.
Veröffentlicht: (2025)
von: Zhang, Ze, et al.
Veröffentlicht: (2025)
Robust Multi-modal Task-oriented Communications with Redundancy-aware Representations
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2023)
Multi-modal Segment Assemblage Network for Ad Video Editing with Importance-Coherence Reward
von: Tang, Yolo Yunlong, et al.
Veröffentlicht: (2022)
von: Tang, Yolo Yunlong, et al.
Veröffentlicht: (2022)
A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Not All Attention is Needed: Parameter and Computation Efficient Transfer Learning for Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
HyperFusion: Hierarchical Multimodal Ensemble Learning for Social Media Popularity Prediction
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
The Future is Meta: Metadata, Formats and Perspectives towards Interactive and Personalized AV Content
von: Weller, Alexander, et al.
Veröffentlicht: (2024)
von: Weller, Alexander, et al.
Veröffentlicht: (2024)
Tri-Ergon: Fine-grained Video-to-Audio Generation with Multi-modal Conditions and LUFS Control
von: Li, Bingliang, et al.
Veröffentlicht: (2024)
von: Li, Bingliang, et al.
Veröffentlicht: (2024)
AIM: Let Any Multi-modal Large Language Models Embrace Efficient In-Context Learning
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
Voices, Faces, and Feelings: Multi-modal Emotion-Cognition Captioning for Mental Health Understanding
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
M3TR: Temporal Retrieval Enhanced Multi-Modal Micro-video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024) -
Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web
von: Heo, Ryang, et al.
Veröffentlicht: (2026) -
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
von: Wang, Dali, et al.
Veröffentlicht: (2026) -
FinCall-Surprise: A Large Scale Multi-modal Benchmark for Earning Surprise Prediction
von: Shu, Dong, et al.
Veröffentlicht: (2025) -
Compression Metadata-assisted RoI Extraction and Adaptive Inference for Efficient Video Analytics
von: Wang, Chengzhi, et al.
Veröffentlicht: (2025)