ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Vo, Hoang-Son, Nguyen, Quang-Vinh, Kim, Seungwon, Yang, Hyung-Jeong, Yeom, Soonja, Kim, Soo-Hyung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
by: Vo-Thanh, Hoang-Son, et al.
Published: (2024)
by: Vo-Thanh, Hoang-Son, et al.
Published: (2024)
Anatomical Attention Alignment representation for Radiology Report Generation
by: Nguyen, Quang Vinh, et al.
Published: (2025)
by: Nguyen, Quang Vinh, et al.
Published: (2025)
Polyp-SES: Automatic Polyp Segmentation with Self-Enriched Semantic Model
by: Nguyen, Quang Vinh, et al.
Published: (2024)
by: Nguyen, Quang Vinh, et al.
Published: (2024)
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
by: Nguyen, Quang Vinh, et al.
Published: (2024)
by: Nguyen, Quang Vinh, et al.
Published: (2024)
Adaptation of Distinct Semantics for Uncertain Areas in Polyp Segmentation
by: Nguyen, Quang Vinh, et al.
Published: (2024)
by: Nguyen, Quang Vinh, et al.
Published: (2024)
Mental Workload Estimation with Electroencephalogram Signals by Combining Multi-Space Deep Models
by: Nguyen, Hong-Hai, et al.
Published: (2023)
by: Nguyen, Hong-Hai, et al.
Published: (2023)
Latent Behavior Diffusion for Sequential Reaction Generation in Dyadic Setting
by: Nguyen, Minh-Duc, et al.
Published: (2025)
by: Nguyen, Minh-Duc, et al.
Published: (2025)
Leveraging WaveNet for Dynamic Listening Head Modeling from Speech
by: Nguyen, Minh-Duc, et al.
Published: (2024)
by: Nguyen, Minh-Duc, et al.
Published: (2024)
GlobalizeEd: A Multimodal Translation System that Preserves Speaker Identity in Academic Lectures
by: Vo, Hoang-Son, et al.
Published: (2025)
by: Vo, Hoang-Son, et al.
Published: (2025)
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
by: Kim, Hyung Kyu, et al.
Published: (2025)
by: Kim, Hyung Kyu, et al.
Published: (2025)
TempoSyncDiff: Distilled Temporally-Consistent Diffusion for Low-Latency Audio-Driven Talking Head Generation
by: Mazumdar, Soumya, et al.
Published: (2026)
by: Mazumdar, Soumya, et al.
Published: (2026)
Talking Head Generation via AU-Guided Landmark Prediction
by: Chang, Shao-Yu, et al.
Published: (2025)
by: Chang, Shao-Yu, et al.
Published: (2025)
Transformer with Leveraged Masked Autoencoder for video-based Pain Assessment
by: Nguyen, Minh-Duc, et al.
Published: (2024)
by: Nguyen, Minh-Duc, et al.
Published: (2024)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
by: Flynn, John, et al.
Published: (2026)
by: Flynn, John, et al.
Published: (2026)
Conditional Diffusion Model for Longitudinal Medical Image Generation
by: Dao, Duy-Phuong, et al.
Published: (2024)
by: Dao, Duy-Phuong, et al.
Published: (2024)
Comparative Studies: Cloud-Enabled Adaptive Learning System for Scalable Education in Sub-Saharan
by: Fianyi, Israel, et al.
Published: (2025)
by: Fianyi, Israel, et al.
Published: (2025)
Enhancing tutoring systems by leveraging tailored promptings and domain knowledge with Large Language Models
by: Balavar, Mohsen, et al.
Published: (2025)
by: Balavar, Mohsen, et al.
Published: (2025)
Landmark-guided Diffusion Model for High-fidelity and Temporally Coherent Talking Head Generation
by: Tan, Jintao, et al.
Published: (2024)
by: Tan, Jintao, et al.
Published: (2024)
Audio-Driven Talking Face Generation with Blink Embedding and Hash Grid Landmarks Encoding
by: Zhang, Yuhui, et al.
Published: (2026)
by: Zhang, Yuhui, et al.
Published: (2026)
Posterior Extremity of the Cystic Plate Guided Right Anterior Glissonean Approach and Intraoperative Spatial Validation in Minimally Invasive Liver Surgery (With Video)
by: Jimin Son, et al.
Published: (2026)
by: Jimin Son, et al.
Published: (2026)
Factors determining the survival or mortality of Nyctereutes procyonoides (Canidae) rescued by a wildlife rescue centre: Is urbanisation a threat to wildlife?
by: Kim, Bong Kyun, et al.
Published: (2025)
by: Kim, Bong Kyun, et al.
Published: (2025)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
by: Aneja, Shivangi, et al.
Published: (2023)
by: Aneja, Shivangi, et al.
Published: (2023)
Learning Phonetic Context-Dependent Viseme for Enhancing Speech-Driven 3D Facial Animation
by: Kim, Hyung Kyu, et al.
Published: (2025)
by: Kim, Hyung Kyu, et al.
Published: (2025)
Bridging Technology and Readiness: AI, IoT, and the Effectiveness of Disaster Prevention in Climate-Vulnerable Regions
by: Hoang Minh Quan, et al.
Published: (2025)
by: Hoang Minh Quan, et al.
Published: (2025)
Lower Trapezius Transfer Using the Retrograde Keyhole Technique With an Achilles Tendon–Bone Allograft
by: Hyung‐gyu Cho, et al.
Published: (2025)
by: Hyung‐gyu Cho, et al.
Published: (2025)
Figure 2 from: Kim BK, Shim JH, Kim SS, Eo SH (2025) Morphological and molecular identification of Particolored bat (Vespertilio murinus) in South Korea: A first record. Biodiversity Data Journal 13: e135293. https://doi.org/10.3897/BDJ.12.e135293
by: Kim, Bong Kyun, et al.
Published: (2025)
by: Kim, Bong Kyun, et al.
Published: (2025)
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
by: Chopin, Baptiste, et al.
Published: (2025)
by: Chopin, Baptiste, et al.
Published: (2025)
Intelligent Exercise and Feedback System for Social Healthcare using LLMOps
by: Choi, Yeongrak, et al.
Published: (2025)
by: Choi, Yeongrak, et al.
Published: (2025)
MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
EmoDiffTalk:Emotion-aware Diffusion for Editable 3D Gaussian Talking Head
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
ConsistTalk: Intensity Controllable Temporally Consistent Talking Head Generation with Diffusion Noise Search
by: Liu, Zhenjie, et al.
Published: (2025)
by: Liu, Zhenjie, et al.
Published: (2025)
Evidence-Linked Labeling (ELL) v1.0: 대화형 정성 데이터의 검증 가능한 정량화 파이프라인
by: Kim, HyungChul
Published: (2026)
by: Kim, HyungChul
Published: (2026)
Reformist Muslims in a Yogyakarta Village
by: Kim, Hyung-Jun
Published: (2013)
by: Kim, Hyung-Jun
Published: (2013)
Spin structure of spin-1 charmonium states near $T_c$
by: Kim, HyungJoo
Published: (2025)
by: Kim, HyungJoo
Published: (2025)
Symplectic Homology and 3-dimensional Besse Manifolds with vanishing first Chern class
by: Kim, Do-Hyung
Published: (2024)
by: Kim, Do-Hyung
Published: (2024)
Contrasting views on the role of AMPK in autophagy
by: Do‐Hyung Kim
Published: (2024)
by: Do‐Hyung Kim
Published: (2024)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
by: Agarwal, Madhav, et al.
Published: (2025)
by: Agarwal, Madhav, et al.
Published: (2025)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
by: Wang, Wenqing, et al.
Published: (2024)
by: Wang, Wenqing, et al.
Published: (2024)
Learning Frame-Wise Emotion Intensity for Audio-Driven Talking-Head Generation
by: Xu, Jingyi, et al.
Published: (2024)
by: Xu, Jingyi, et al.
Published: (2024)
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
by: Wang, Haotian, et al.
Published: (2024)
by: Wang, Haotian, et al.
Published: (2024)
Similar Items
-
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
by: Vo-Thanh, Hoang-Son, et al.
Published: (2024) -
Anatomical Attention Alignment representation for Radiology Report Generation
by: Nguyen, Quang Vinh, et al.
Published: (2025) -
Polyp-SES: Automatic Polyp Segmentation with Self-Enriched Semantic Model
by: Nguyen, Quang Vinh, et al.
Published: (2024) -
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
by: Nguyen, Quang Vinh, et al.
Published: (2024) -
Adaptation of Distinct Semantics for Uncertain Areas in Polyp Segmentation
by: Nguyen, Quang Vinh, et al.
Published: (2024)