MultiMediate'24: Multi-Domain Engagement Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Müller, Philipp, Balazia, Michal, Baur, Tobias, Dietz, Michael, Heimerl, Alexander, Penzkofer, Anna, Schiller, Dominik, Brémond, François, Alexandersson, Jan, André, Elisabeth, Bulling, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Identifying Surgical Instruments in Pedagogical Cataract Surgery Videos through an Optimized Aggregation Network
by: Sinha, Sanya, et al.
Published: (2025)
by: Sinha, Sanya, et al.
Published: (2025)
Transforming Video Subjective Testing with Training, Engagement, and Real-Time Feedback
by: Rahul, Kumar, et al.
Published: (2026)
by: Rahul, Kumar, et al.
Published: (2026)
Clinical Multi-modal Fusion with Heterogeneous Graph and Disease Correlation Learning for Multi-Disease Prediction
by: Jiang, Yueheng, et al.
Published: (2025)
by: Jiang, Yueheng, et al.
Published: (2025)
What Matters in Autonomous Driving Anomaly Detection: A Weakly Supervised Horizon
by: Tiwari, Utkarsh, et al.
Published: (2024)
by: Tiwari, Utkarsh, et al.
Published: (2024)
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets
by: Agrawal, Tanay, et al.
Published: (2025)
by: Agrawal, Tanay, et al.
Published: (2025)
U-Sticker: A Large-Scale Multi-Domain User Sticker Dataset for Retrieval and Personalization
by: Chee, Heng Er Metilda, et al.
Published: (2025)
by: Chee, Heng Er Metilda, et al.
Published: (2025)
Deep Mamba Multi-modal Learning
by: Zhu, Jian, et al.
Published: (2024)
by: Zhu, Jian, et al.
Published: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
by: Zhang, Zhenxing, et al.
Published: (2024)
by: Zhang, Zhenxing, et al.
Published: (2024)
QuMATL: Query-based Multi-annotator Tendency Learning
by: Zhang, Liyun, et al.
Published: (2025)
by: Zhang, Liyun, et al.
Published: (2025)
Multi-Reference Generative Face Video Compression with Contrastive Learning
by: Konuko, Goluck, et al.
Published: (2024)
by: Konuko, Goluck, et al.
Published: (2024)
Robust Multi-generation Learned Compression of Point Cloud Attribute
by: Liu, Xiangzuo, et al.
Published: (2025)
by: Liu, Xiangzuo, et al.
Published: (2025)
Vidformer: Drop-in Declarative Optimization for Rendering Video-Native Query Results
by: Winecki, Dominik, et al.
Published: (2026)
by: Winecki, Dominik, et al.
Published: (2026)
Towards Structure-aware Model for Multi-modal Knowledge Graph Completion
by: Li, Linyu, et al.
Published: (2025)
by: Li, Linyu, et al.
Published: (2025)
Multi-resolution Encoding for HTTP Adaptive Streaming using VVenC
by: Qureshi, Kamran, et al.
Published: (2025)
by: Qureshi, Kamran, et al.
Published: (2025)
A Multi-Embedding Convergence Network on Siamese Architecture for Fake Reviews
by: Dasgupta, Sankarshan, et al.
Published: (2024)
by: Dasgupta, Sankarshan, et al.
Published: (2024)
Latency Effects on Multi-Dimensional QoE in Networked VR Whiteboards
by: Song, Jiarun, et al.
Published: (2026)
by: Song, Jiarun, et al.
Published: (2026)
Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
by: Lu, Jiacheng, et al.
Published: (2025)
by: Lu, Jiacheng, et al.
Published: (2025)
Socially Aware Music Recommendation: A Multi-Modal Graph Neural Networks for Collaborative Music Consumption and Community-Based Engagement
by: Ziaoddini, Kajwan
Published: (2025)
by: Ziaoddini, Kajwan
Published: (2025)
Multi-source Multimodal Progressive Domain Adaption for Audio-Visual Deception Detection
by: Lin, Ronghao, et al.
Published: (2025)
by: Lin, Ronghao, et al.
Published: (2025)
Multi-source Knowledge Enhanced Graph Attention Networks for Multimodal Fact Verification
by: Cao, Han, et al.
Published: (2024)
by: Cao, Han, et al.
Published: (2024)
MViR: Multi-View Visual-Semantic Representation for Fake News Detection
by: Liang, Haochen, et al.
Published: (2026)
by: Liang, Haochen, et al.
Published: (2026)
Self-Training Boosted Multi-Factor Matching Network for Composed Image Retrieval
by: Wen, Haokun, et al.
Published: (2023)
by: Wen, Haokun, et al.
Published: (2023)
Emotional Cues Extraction and Fusion for Multi-modal Emotion Prediction and Recognition in Conversation
by: Shi, Haoxiang, et al.
Published: (2024)
by: Shi, Haoxiang, et al.
Published: (2024)
Block-Partitioning Strategies for Accelerated Multi-rate Encoding in Adaptive VVC Streaming
by: Menon, Vignesh V, et al.
Published: (2025)
by: Menon, Vignesh V, et al.
Published: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
by: Wang, Sen, et al.
Published: (2024)
by: Wang, Sen, et al.
Published: (2024)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
by: Jaspal, Amit, et al.
Published: (2025)
by: Jaspal, Amit, et al.
Published: (2025)
A 3D Framework for Improving Low-Latency Multi-Channel Live Streaming
by: Aiersilan, Aizierjiang, et al.
Published: (2024)
by: Aiersilan, Aizierjiang, et al.
Published: (2024)
Enhancing Few-Shot Classification without Forgetting through Multi-Level Contrastive Constraints
by: Chen, Bingzhi, et al.
Published: (2024)
by: Chen, Bingzhi, et al.
Published: (2024)
Voices, Faces, and Feelings: Multi-modal Emotion-Cognition Captioning for Mental Health Understanding
by: Zhou, Zhiyuan, et al.
Published: (2026)
by: Zhou, Zhiyuan, et al.
Published: (2026)
Routing Experts: Learning to Route Dynamic Experts in Multi-modal Large Language Models
by: Wu, Qiong, et al.
Published: (2024)
by: Wu, Qiong, et al.
Published: (2024)
Accelerating Multi-Condition T2I Generation via Adaptive Condition Offloading and Pruning
by: Kong, Yuxin, et al.
Published: (2026)
by: Kong, Yuxin, et al.
Published: (2026)
Characterizing Multimedia Information Environment through Multi-modal Clustering of YouTube Videos
by: Yousefi, Niloofar, et al.
Published: (2024)
by: Yousefi, Niloofar, et al.
Published: (2024)
Multi Agents Semantic Emotion Aligned Music to Image Generation with Music Derived Captions
by: Shi, Junchang, et al.
Published: (2025)
by: Shi, Junchang, et al.
Published: (2025)
Short-Form Video Viewing Behavior Analysis and Multi-Step Viewing Time Prediction
by: Yen, Vu Thi Hai, et al.
Published: (2026)
by: Yen, Vu Thi Hai, et al.
Published: (2026)
RoSMM: A Robust and Secure Multi-Modal Watermarking Framework for Diffusion Models
by: Fang, ZhongLi, et al.
Published: (2025)
by: Fang, ZhongLi, et al.
Published: (2025)
FineBadminton: A Multi-Level Dataset for Fine-Grained Badminton Video Understanding
by: He, Xusheng, et al.
Published: (2025)
by: He, Xusheng, et al.
Published: (2025)
High-level Codes and Fine-grained Weights for Online Multi-modal Hashing Retrieval
by: Zhan, Yu-Wei, et al.
Published: (2024)
by: Zhan, Yu-Wei, et al.
Published: (2024)
MAR3: Multi-Agent Recognition, Reasoning, and Reflection for Reference Audio-Visual Segmentation
by: Zhao, Yuan, et al.
Published: (2026)
by: Zhao, Yuan, et al.
Published: (2026)
ProMSC-MIS: Prompt-based Multimodal Semantic Communication for Multi-Spectral Image Segmentation
by: Zhang, Haoshuo, et al.
Published: (2025)
by: Zhang, Haoshuo, et al.
Published: (2025)
Multi-view Hypergraph-based Contrastive Learning Model for Cold-Start Micro-video Recommendation
by: Lyu, Sisuo, et al.
Published: (2024)
by: Lyu, Sisuo, et al.
Published: (2024)
Similar Items
-
Identifying Surgical Instruments in Pedagogical Cataract Surgery Videos through an Optimized Aggregation Network
by: Sinha, Sanya, et al.
Published: (2025) -
Transforming Video Subjective Testing with Training, Engagement, and Real-Time Feedback
by: Rahul, Kumar, et al.
Published: (2026) -
Clinical Multi-modal Fusion with Heterogeneous Graph and Disease Correlation Learning for Multi-Disease Prediction
by: Jiang, Yueheng, et al.
Published: (2025) -
What Matters in Autonomous Driving Anomaly Detection: A Weakly Supervised Horizon
by: Tiwari, Utkarsh, et al.
Published: (2024) -
CM3T: Framework for Efficient Multimodal Learning for Inhomogeneous Interaction Datasets
by: Agrawal, Tanay, et al.
Published: (2025)