SVFAP: Self-supervised Video Facial Affect Perceiver
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Licai, Lian, Zheng, Wang, Kexin, He, Yu, Xu, Mingyu, Sun, Haiyang, Liu, Bin, Tao, Jianhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
von: Sun, Licai, et al.
Veröffentlicht: (2024)
von: Sun, Licai, et al.
Veröffentlicht: (2024)
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
Seeing, Hearing, and Knowing Together: Multimodal Strategies in Deepfake Videos Detection
von: Chen, Chen, et al.
Veröffentlicht: (2026)
von: Chen, Chen, et al.
Veröffentlicht: (2026)
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
von: Zhou, Tian-Yi, et al.
Veröffentlicht: (2026)
von: Zhou, Tian-Yi, et al.
Veröffentlicht: (2026)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
Panonut360: A Head and Eye Tracking Dataset for Panoramic Video
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Body Ownership Affects the Processing of Sensorimotor Contingencies in Virtual Reality
von: Center, Evan G., et al.
Veröffentlicht: (2025)
von: Center, Evan G., et al.
Veröffentlicht: (2025)
MindCross: Fast New Subject Adaptation with Limited Data for Cross-subject Video Reconstruction from Brain Signals
von: Liu, Xuan-Hao, et al.
Veröffentlicht: (2025)
von: Liu, Xuan-Hao, et al.
Veröffentlicht: (2025)
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
G3R: Generating Rich and Fine-grained mmWave Radar Data from 2D Videos for Generalized Gesture Recognition
von: Deng, Kaikai, et al.
Veröffentlicht: (2024)
von: Deng, Kaikai, et al.
Veröffentlicht: (2024)
Self-supervised Spatio-Temporal Graph Mask-Passing Attention Network for Perceptual Importance Prediction of Multi-point Tactility
von: He, Dazhong, et al.
Veröffentlicht: (2024)
von: He, Dazhong, et al.
Veröffentlicht: (2024)
MIMOSA: Human-AI Co-Creation of Computational Spatial Audio Effects on Videos
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
Plug-and-Play Clarifier: A Zero-Shot Multimodal Framework for Egocentric Intent Disambiguation
von: Yang, Sicheng, et al.
Veröffentlicht: (2025)
von: Yang, Sicheng, et al.
Veröffentlicht: (2025)
AffectMachine-Pop: A controllable expert system for real-time pop music generation
von: Agres, Kat R., et al.
Veröffentlicht: (2025)
von: Agres, Kat R., et al.
Veröffentlicht: (2025)
Editing Physiological Signals in Videos Using Latent Representations
von: Zhou, Tianwen, et al.
Veröffentlicht: (2025)
von: Zhou, Tianwen, et al.
Veröffentlicht: (2025)
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
von: Lian, Zheng, et al.
Veröffentlicht: (2025)
Multimodal Infusion Tuning for Large Models
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
Anchorage: Visual Analysis of Satisfaction in Customer Service Videos via Anchor Events
von: Wong, Kam Kwai, et al.
Veröffentlicht: (2023)
von: Wong, Kam Kwai, et al.
Veröffentlicht: (2023)
Color When It Counts: Grayscale-Guided Online Triggering for Always-On Streaming Video Sensing
von: Cai, Weitong, et al.
Veröffentlicht: (2026)
von: Cai, Weitong, et al.
Veröffentlicht: (2026)
Videogenic: Identifying Highlight Moments in Videos with Professional Photographs as a Prior
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
AttentionBender: Manipulating Cross-Attention in Video Diffusion Transformers as a Creative Probe
von: Cole, Adam, et al.
Veröffentlicht: (2026)
von: Cole, Adam, et al.
Veröffentlicht: (2026)
SPICA: Interactive Video Content Exploration through Augmented Audio Descriptions for Blind or Low-Vision Viewers
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
EyeNavGS: A 6-DoF Navigation Dataset and Record-n-Replay Software for Real-World 3DGS Scenes in VR
von: Ding, Zihao, et al.
Veröffentlicht: (2025)
von: Ding, Zihao, et al.
Veröffentlicht: (2025)
MV-Crafter: An Intelligent System for Music-guided Video Generation
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
Self-Guided Virtual Reality Therapy for Anxiety: A Systematic Review
von: Graham, Winona, et al.
Veröffentlicht: (2025)
von: Graham, Winona, et al.
Veröffentlicht: (2025)
FastPerson: Enhancing Video Learning through Effective Video Summarization that Preserves Linguistic and Visual Contexts
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
VisAug: Facilitating Speech-Rich Web Video Navigation and Engagement with Auto-Generated Visual Augmentations
von: Zhao, Baoquan, et al.
Veröffentlicht: (2025)
von: Zhao, Baoquan, et al.
Veröffentlicht: (2025)
Talking-to-Build: How LLM-Assisted Interface Shapes Player Performance and Experience in Minecraft
von: Sun, Xin, et al.
Veröffentlicht: (2025)
von: Sun, Xin, et al.
Veröffentlicht: (2025)
MetaDragonBoat: Exploring Paddling Techniques of Virtual Dragon Boating in a Metaverse Campus
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
MotiBo: The Impact of Interactive Digital Storytelling Robots on Student Motivation through Self-Determination Theory
von: Fung, Ka Yan, et al.
Veröffentlicht: (2026)
von: Fung, Ka Yan, et al.
Veröffentlicht: (2026)
DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos
von: Zheng, Ce, et al.
Veröffentlicht: (2023)
von: Zheng, Ce, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
von: Sun, Licai, et al.
Veröffentlicht: (2024) -
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023) -
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024) -
Seeing, Hearing, and Knowing Together: Multimodal Strategies in Deepfake Videos Detection
von: Chen, Chen, et al.
Veröffentlicht: (2026) -
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
von: Cheng, Luo, et al.
Veröffentlicht: (2025)