Sec2Sec Co-attention for Video-Based Apparent Affective Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Mingwei, Zhang, Kunpeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
Decoding the Hook: A Multimodal LLM Framework for Analyzing the Hooking Period of Video Ads
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2026)
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2026)
Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Lu, Jiacheng, et al.
Veröffentlicht: (2025)
Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web
von: Heo, Ryang, et al.
Veröffentlicht: (2026)
von: Heo, Ryang, et al.
Veröffentlicht: (2026)
Predicting Encoding Energy from Low-Pass Anchors for Green Video Streaming
von: Azimi, Zoha, et al.
Veröffentlicht: (2025)
von: Azimi, Zoha, et al.
Veröffentlicht: (2025)
Reply with Sticker: New Dataset and Model for Sticker Retrieval
von: Liang, Bin, et al.
Veröffentlicht: (2024)
von: Liang, Bin, et al.
Veröffentlicht: (2024)
Short-Form Video Viewing Behavior Analysis and Multi-Step Viewing Time Prediction
von: Yen, Vu Thi Hai, et al.
Veröffentlicht: (2026)
von: Yen, Vu Thi Hai, et al.
Veröffentlicht: (2026)
Machine Learning-Based Prediction of Quality Shifts on Video Streaming Over 5G
von: Mustafa, Raza Ul, et al.
Veröffentlicht: (2025)
von: Mustafa, Raza Ul, et al.
Veröffentlicht: (2025)
A Video Steganography for H.265/HEVC Based on Multiple CU Size and Block Structure Distortion
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Traits Run Deep: Enhancing Personality Assessment via Psychology-Guided LLM Representations and Multimodal Apparent Behaviors
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
von: Liang, Feng, et al.
Veröffentlicht: (2023)
von: Liang, Feng, et al.
Veröffentlicht: (2023)
AFL-Net: Integrating Audio, Facial, and Lip Modalities with a Two-step Cross-attention for Robust Speaker Diarization in the Wild
von: Yin, Yongkang, et al.
Veröffentlicht: (2023)
von: Yin, Yongkang, et al.
Veröffentlicht: (2023)
Video Streaming with Kairos: An MPC-Based ABR with Streaming-Aware Throughput Prediction
von: Zhong, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhong, Ziyu, et al.
Veröffentlicht: (2025)
SyMuPe: Affective and Controllable Symbolic Music Performance
von: Borovik, Ilya, et al.
Veröffentlicht: (2025)
von: Borovik, Ilya, et al.
Veröffentlicht: (2025)
Cap2Sum: Learning to Summarize Videos by Generating Captions
von: Zhao, Cairong, et al.
Veröffentlicht: (2024)
von: Zhao, Cairong, et al.
Veröffentlicht: (2024)
EyEar: Learning Audio Synchronized Human Gaze Trajectory Based on Physics-Informed Dynamics
von: Liu, Xiaochuan, et al.
Veröffentlicht: (2025)
von: Liu, Xiaochuan, et al.
Veröffentlicht: (2025)
Advanced Learning-Based Inter Prediction for Future Video Coding
von: Zhao, Yanchen, et al.
Veröffentlicht: (2024)
von: Zhao, Yanchen, et al.
Veröffentlicht: (2024)
Graph-Driven Multimodal Feature Learning Framework for Apparent Personality Assessment
von: Wang, Kangsheng, et al.
Veröffentlicht: (2025)
von: Wang, Kangsheng, et al.
Veröffentlicht: (2025)
Feedback-Driven Rate Control for Learned Video Compression
von: Xu, Zhiheng, et al.
Veröffentlicht: (2026)
von: Xu, Zhiheng, et al.
Veröffentlicht: (2026)
4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer's Diagnosis
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
von: Yang, Dingyi, et al.
Veröffentlicht: (2024)
von: Yang, Dingyi, et al.
Veröffentlicht: (2024)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
von: Wang, Dali, et al.
Veröffentlicht: (2026)
von: Wang, Dali, et al.
Veröffentlicht: (2026)
Joint Optimization of Buffer Delay and HARQ for Video Communications
von: Cheng, Baoping, et al.
Veröffentlicht: (2024)
von: Cheng, Baoping, et al.
Veröffentlicht: (2024)
MIMOSA: Human-AI Co-Creation of Computational Spatial Audio Effects on Videos
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
Contextual Wireless Video Semantic Communication in MIMO-OFDM Systems
von: Xie, Bingyan, et al.
Veröffentlicht: (2026)
von: Xie, Bingyan, et al.
Veröffentlicht: (2026)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
Adaptive Offloading and Enhancement for Low-Light Video Analytics on Mobile Devices
von: He, Yuanyi, et al.
Veröffentlicht: (2024)
von: He, Yuanyi, et al.
Veröffentlicht: (2024)
Towards Open-Vocabulary Video Semantic Segmentation
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
Virbo: Multimodal Multilingual Avatar Video Generation in Digital Marketing
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
von: Zhang, Juan, et al.
Veröffentlicht: (2024)
EV-NVC: Efficient Variable bitrate Neural Video Compression
von: Hu, Yongcun, et al.
Veröffentlicht: (2025)
von: Hu, Yongcun, et al.
Veröffentlicht: (2025)
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
von: Wang, Yisu, et al.
Veröffentlicht: (2025)
von: Wang, Yisu, et al.
Veröffentlicht: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
von: Wang, Sen, et al.
Veröffentlicht: (2024)
von: Wang, Sen, et al.
Veröffentlicht: (2024)
Fine-grained Knowledge Graph-driven Video-Language Learning for Action Recognition
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Music Grounding by Short Video
von: Xin, Zijie, et al.
Veröffentlicht: (2024)
von: Xin, Zijie, et al.
Veröffentlicht: (2024)
Hallucination Localization in Video Captioning
von: Nakada, Shota, et al.
Veröffentlicht: (2025)
von: Nakada, Shota, et al.
Veröffentlicht: (2025)
SVD: Spatial Video Dataset
von: Izadimehr, M. H., et al.
Veröffentlicht: (2025)
von: Izadimehr, M. H., et al.
Veröffentlicht: (2025)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoyuan, et al.
Veröffentlicht: (2024)
Mining the Social Fabric: Unveiling Communities for Fake News Detection in Short Videos
von: Gong, Haisong, et al.
Veröffentlicht: (2025)
von: Gong, Haisong, et al.
Veröffentlicht: (2025)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021) -
Decoding the Hook: A Multimodal LLM Framework for Analyzing the Hooking Period of Video Ads
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2026) -
Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2025) -
Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web
von: Heo, Ryang, et al.
Veröffentlicht: (2026) -
Predicting Encoding Energy from Low-Pass Anchors for Green Video Streaming
von: Azimi, Zoha, et al.
Veröffentlicht: (2025)