DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Jisoo, Cho, Jungbin, Park, Joonho, Hwang, Soonmin, Kim, Da Eun, Kim, Geon, Yu, Youngjae |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
di: Kim, Junhyeok, et al.
Pubblicazione: (2025)
di: Kim, Junhyeok, et al.
Pubblicazione: (2025)
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
di: Cho, Jungbin, et al.
Pubblicazione: (2024)
di: Cho, Jungbin, et al.
Pubblicazione: (2024)
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
di: Cho, Jungbin, et al.
Pubblicazione: (2025)
di: Cho, Jungbin, et al.
Pubblicazione: (2025)
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
di: Lin, Yihong, et al.
Pubblicazione: (2024)
di: Lin, Yihong, et al.
Pubblicazione: (2024)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
di: Kim, Jisoo, et al.
Pubblicazione: (2025)
di: Kim, Jisoo, et al.
Pubblicazione: (2025)
Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
Pri4R: Learning World Dynamics for Vision-Language-Action Models with Privileged 4D Representation
di: Kim, Jisoo, et al.
Pubblicazione: (2026)
di: Kim, Jisoo, et al.
Pubblicazione: (2026)
Lightweight Wasserstein Audio-Visual Model for Unified Speech Enhancement and Separation
di: Park, Jisoo, et al.
Pubblicazione: (2025)
di: Park, Jisoo, et al.
Pubblicazione: (2025)
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
di: EunGi, Han, et al.
Pubblicazione: (2024)
di: EunGi, Han, et al.
Pubblicazione: (2024)
EmoFace: Audio-driven Emotional 3D Face Animation
di: Liu, Chang, et al.
Pubblicazione: (2024)
di: Liu, Chang, et al.
Pubblicazione: (2024)
KMTalk: Speech-Driven 3D Facial Animation with Key Motion Embedding
di: Xu, Zhihao, et al.
Pubblicazione: (2024)
di: Xu, Zhihao, et al.
Pubblicazione: (2024)
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration
di: Jun-Seong, Kim, et al.
Pubblicazione: (2025)
di: Jun-Seong, Kim, et al.
Pubblicazione: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
di: Park, Inkyu, et al.
Pubblicazione: (2023)
di: Park, Inkyu, et al.
Pubblicazione: (2023)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
di: Mao, Yuxiang, et al.
Pubblicazione: (2025)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
ESGaussianFace: Emotional and Stylized Audio-Driven Facial Animation via 3D Gaussian Splatting
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
di: Ma, Chuhang, et al.
Pubblicazione: (2026)
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
Boosting Cross-spectral Unsupervised Domain Adaptation for Thermal Semantic Segmentation
di: Kwon, Seokjun, et al.
Pubblicazione: (2025)
di: Kwon, Seokjun, et al.
Pubblicazione: (2025)
MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding
di: Liu, Chang, et al.
Pubblicazione: (2025)
di: Liu, Chang, et al.
Pubblicazione: (2025)
Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning
di: Park, Byounggun, et al.
Pubblicazione: (2026)
di: Park, Byounggun, et al.
Pubblicazione: (2026)
AVIN-Chat: An Audio-Visual Interactive Chatbot System with Emotional State Tuning
di: Park, Chanhyuk, et al.
Pubblicazione: (2024)
di: Park, Chanhyuk, et al.
Pubblicazione: (2024)
PointCubeNet: 3D Part-level Reasoning with 3x3x3 Point Cloud Blocks
di: Kim, Da-Yeong, et al.
Pubblicazione: (2025)
di: Kim, Da-Yeong, et al.
Pubblicazione: (2025)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
di: Jeon, Subin, et al.
Pubblicazione: (2024)
di: Jeon, Subin, et al.
Pubblicazione: (2024)
Polyglot: Multilingual Style Preserving Speech-Driven Facial Animation
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
di: Nocentini, Federico, et al.
Pubblicazione: (2026)
Joint-Embedding Predictive Architecture for Self-Supervised Learning of Mask Classification Architecture
di: Kim, Dong-Hee, et al.
Pubblicazione: (2024)
di: Kim, Dong-Hee, et al.
Pubblicazione: (2024)
Polarimetric BSSRDF Acquisition of Dynamic Faces
di: Ha, Hyunho, et al.
Pubblicazione: (2024)
di: Ha, Hyunho, et al.
Pubblicazione: (2024)
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
di: Kim, GeonU, et al.
Pubblicazione: (2024)
di: Kim, GeonU, et al.
Pubblicazione: (2024)
A Proof of the Exact Convergence Rate of Gradient Descent
di: Kim, Jungbin
Pubblicazione: (2024)
di: Kim, Jungbin
Pubblicazione: (2024)
A Proof of Exact Convergence Rate of Gradient Descent. Part I. Performance Criterion $\Vert \nabla f(x_N)\Vert^2/(f(x_0)-f_*)$
di: Kim, Jungbin
Pubblicazione: (2024)
di: Kim, Jungbin
Pubblicazione: (2024)
Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video
di: Choi, Chanhyuk, et al.
Pubblicazione: (2026)
di: Choi, Chanhyuk, et al.
Pubblicazione: (2026)
Background-Aware Defect Generation for Robust Industrial Anomaly Detection
di: Cho, Youngjae, et al.
Pubblicazione: (2024)
di: Cho, Youngjae, et al.
Pubblicazione: (2024)
D2E: Scaling Vision-Action Pretraining on Desktop Data for Transfer to Embodied AI
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
SEDTalker: Emotion-Aware 3D Facial Animation Using Frame-Level Speech Emotion Diarization
di: Jafari, Farzaneh, et al.
Pubblicazione: (2026)
di: Jafari, Farzaneh, et al.
Pubblicazione: (2026)
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
di: Zheng, Kai, et al.
Pubblicazione: (2026)
di: Zheng, Kai, et al.
Pubblicazione: (2026)
Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
di: Kim, Youbin, et al.
Pubblicazione: (2026)
di: Kim, Youbin, et al.
Pubblicazione: (2026)
Emotional Speech-driven 3D Body Animation via Disentangled Latent Diffusion
di: Chhatre, Kiran, et al.
Pubblicazione: (2023)
di: Chhatre, Kiran, et al.
Pubblicazione: (2023)
VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding
di: Kim, Kangsan, et al.
Pubblicazione: (2024)
di: Kim, Kangsan, et al.
Pubblicazione: (2024)
ProbTalk3D: Non-Deterministic Emotion Controllable Speech-Driven 3D Facial Animation Synthesis Using VQ-VAE
di: Wu, Sichun, et al.
Pubblicazione: (2024)
di: Wu, Sichun, et al.
Pubblicazione: (2024)
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
di: Liang, Xiangyu, et al.
Pubblicazione: (2024)
di: Liang, Xiangyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
di: Kim, Junhyeok, et al.
Pubblicazione: (2025) -
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
di: Cho, Jungbin, et al.
Pubblicazione: (2024) -
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
di: Cho, Jungbin, et al.
Pubblicazione: (2025) -
EmoFace: Emotion-Content Disentangled Speech-Driven 3D Talking Face Animation
di: Lin, Yihong, et al.
Pubblicazione: (2024) -
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
di: Kim, Jisoo, et al.
Pubblicazione: (2025)