Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Choi, Chanhyuk, Kim, Taesoo, Lee, Donggyu, Jung, Siyeol, Kim, Taehwan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Audio-Guided Visual Editing with Complex Multi-Modal Prompts
von: Kim, Hyeonyu, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonyu, et al.
Veröffentlicht: (2025)
Environmental Understanding Vision-Language Model for Embodied Agent
von: Bang, Jinsik, et al.
Veröffentlicht: (2026)
von: Bang, Jinsik, et al.
Veröffentlicht: (2026)
DiffListener: Discrete Diffusion Model for Listener Generation
von: Jung, Siyeol, et al.
Veröffentlicht: (2025)
von: Jung, Siyeol, et al.
Veröffentlicht: (2025)
AdaRank: Adaptive Rank Pruning for Enhanced Model Merging
von: Lee, Chanhyuk, et al.
Veröffentlicht: (2025)
von: Lee, Chanhyuk, et al.
Veröffentlicht: (2025)
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
von: Jeon, Inseok, et al.
Veröffentlicht: (2026)
von: Jeon, Inseok, et al.
Veröffentlicht: (2026)
GCM-Net: Graph-enhanced Cross-Modal Infusion with a Metaheuristic-Driven Network for Video Sentiment and Emotion Analysis
von: Chaudhari, Prasad, et al.
Veröffentlicht: (2024)
von: Chaudhari, Prasad, et al.
Veröffentlicht: (2024)
Machine Pareidolia: Protecting Facial Image with Emotional Editing
von: Le, Binh M., et al.
Veröffentlicht: (2026)
von: Le, Binh M., et al.
Veröffentlicht: (2026)
Spatial-and-Frequency-aware Restoration method for Images based on Diffusion Models
von: Lee, Kyungsung, et al.
Veröffentlicht: (2024)
von: Lee, Kyungsung, et al.
Veröffentlicht: (2024)
Learn2Talk: 3D Talking Face Learns from 2D Talking Face
von: Zhuang, Yixiang, et al.
Veröffentlicht: (2024)
von: Zhuang, Yixiang, et al.
Veröffentlicht: (2024)
Video Face Re-Aging: Toward Temporally Consistent Face Re-Aging
von: Muqeet, Abdul, et al.
Veröffentlicht: (2023)
von: Muqeet, Abdul, et al.
Veröffentlicht: (2023)
SynchroRaMa : Lip-Synchronized and Emotion-Aware Talking Face Generation via Multi-Modal Emotion Embedding
von: Yee, Phyo Thet, et al.
Veröffentlicht: (2025)
von: Yee, Phyo Thet, et al.
Veröffentlicht: (2025)
Crafting Query-Aware Selective Attention for Single Image Super-Resolution
von: Kim, Junyoung, et al.
Veröffentlicht: (2025)
von: Kim, Junyoung, et al.
Veröffentlicht: (2025)
Taming Transformer for Emotion-Controllable Talking Face Generation
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities
von: Mobbs, Rebecca, et al.
Veröffentlicht: (2025)
von: Mobbs, Rebecca, et al.
Veröffentlicht: (2025)
TripleSumm: Adaptive Triple-Modality Fusion for Video Summarization
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
Generation and Editing of Mandrill Faces: Application to Sex Editing and Assessment
von: Dibot, Nicolas M., et al.
Veröffentlicht: (2024)
von: Dibot, Nicolas M., et al.
Veröffentlicht: (2024)
Navigating the Emotion Tree: Hierarchical Hyperbolic RAG for Multimodal Emotion Recognition
von: Wang, Zeheng, et al.
Veröffentlicht: (2026)
von: Wang, Zeheng, et al.
Veröffentlicht: (2026)
Grid Diffusion Models for Text-to-Video Generation
von: Lee, Taegyeong, et al.
Veröffentlicht: (2024)
von: Lee, Taegyeong, et al.
Veröffentlicht: (2024)
Conformal Cross-Modal Active Learning
von: Nguyen, Huy Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Huy Hoang, et al.
Veröffentlicht: (2026)
LES-Talker: Fine-Grained Emotion Editing for Talking Head Generation in Linear Emotion Space
von: Feng, Guanwen, et al.
Veröffentlicht: (2024)
von: Feng, Guanwen, et al.
Veröffentlicht: (2024)
Understanding Flatness in Generative Models: Its Role and Benefits
von: Lee, Taehwan, et al.
Veröffentlicht: (2025)
von: Lee, Taehwan, et al.
Veröffentlicht: (2025)
Take an Emotion Walk: Perceiving Emotions from Gaits Using Hierarchical Attention Pooling and Affective Mapping
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2019)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2019)
Face Generation and Editing with StyleGAN: A Survey
von: Melnik, Andrew, et al.
Veröffentlicht: (2022)
von: Melnik, Andrew, et al.
Veröffentlicht: (2022)
Benchmarking Federated Learning for Semantic Datasets: Federated Scene Graph Generation
von: Ha, SeungBum, et al.
Veröffentlicht: (2024)
von: Ha, SeungBum, et al.
Veröffentlicht: (2024)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
Cross-Modal Few-Shot Learning: a Generative Transfer Learning Framework
von: Yang, Zhengwei, et al.
Veröffentlicht: (2024)
von: Yang, Zhengwei, et al.
Veröffentlicht: (2024)
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
von: Choi, Eunjee, et al.
Veröffentlicht: (2025)
von: Choi, Eunjee, et al.
Veröffentlicht: (2025)
Federated Learning for Face Recognition via Intra-subject Self-supervised Learning
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
Robust Emotion Recognition in Context Debiasing
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
von: Yang, Dingkang, et al.
Veröffentlicht: (2024)
MUST: Modality-Specific Representation-Aware Transformer for Diffusion-Enhanced Survival Prediction with Missing Modality
von: Kim, Kyungwon, et al.
Veröffentlicht: (2026)
von: Kim, Kyungwon, et al.
Veröffentlicht: (2026)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
Domain-Invariant Per-Frame Feature Extraction for Cross-Domain Imitation Learning with Visual Observations
von: Kim, Minung, et al.
Veröffentlicht: (2025)
von: Kim, Minung, et al.
Veröffentlicht: (2025)
Text-Driven Emotionally Continuous Talking Face Generation
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
V-Warper: Appearance-Consistent Video Diffusion Personalization via Value Warping
von: Lee, Hyunkoo, et al.
Veröffentlicht: (2025)
von: Lee, Hyunkoo, et al.
Veröffentlicht: (2025)
Towards Robust Multimodal Emotion Recognition under Missing Modalities and Distribution Shifts
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
Dynamic Modality and View Selection for Multimodal Emotion Recognition with Missing Modalities
von: Menon, Luciana Trinkaus, et al.
Veröffentlicht: (2024)
von: Menon, Luciana Trinkaus, et al.
Veröffentlicht: (2024)
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
von: Lee, Seok Hwan, et al.
Veröffentlicht: (2024)
von: Lee, Seok Hwan, et al.
Veröffentlicht: (2024)
Emotion Recognition Using Convolutional Neural Networks
von: Xu, Shaoyuan, et al.
Veröffentlicht: (2025)
von: Xu, Shaoyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Audio-Guided Visual Editing with Complex Multi-Modal Prompts
von: Kim, Hyeonyu, et al.
Veröffentlicht: (2025) -
Environmental Understanding Vision-Language Model for Embodied Agent
von: Bang, Jinsik, et al.
Veröffentlicht: (2026) -
DiffListener: Discrete Diffusion Model for Listener Generation
von: Jung, Siyeol, et al.
Veröffentlicht: (2025) -
AdaRank: Adaptive Rank Pruning for Enhanced Model Merging
von: Lee, Chanhyuk, et al.
Veröffentlicht: (2025) -
CMTM: Cross-Modal Token Modulation for Unsupervised Video Object Segmentation
von: Jeon, Inseok, et al.
Veröffentlicht: (2026)