Toward Fine-Grained Facial Control in 3D Talking Head Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xie, Shaoyang, Cong, Xiaofeng, Yu, Baosheng, Gui, Zhipeng, Gui, Jie, Tang, Yuan Yan, Kwok, James Tin-Yau |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fooling the Image Dehazing Models by First Order Gradient
by: Gui, Jie, et al.
Published: (2023)
by: Gui, Jie, et al.
Published: (2023)
A Survey on Small Sample Imbalance Problem: Metrics, Feature Analysis, and Solutions
by: Zhao, Shuxian, et al.
Published: (2025)
by: Zhao, Shuxian, et al.
Published: (2025)
CFVNet: An End-to-End Cancelable Finger Vein Network for Recognition
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
BitC-3DGS: High-Capacity 3D Gaussian Splatting Watermarking via Bit Compression
by: Bi, Yuquan, et al.
Published: (2026)
by: Bi, Yuquan, et al.
Published: (2026)
ColorVein: Colorful Cancelable Vein Biometrics
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Revisiting Adversarial Training under Hyperspectral Image
by: Zhang, Weihua, et al.
Published: (2025)
by: Zhang, Weihua, et al.
Published: (2025)
Underwater Organism Color Enhancement via Color Code Decomposition, Adaptation and Interpolation
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
Unrevealed Threats: A Comprehensive Study of the Adversarial Robustness of Underwater Image Enhancement Models
by: Zhai, Siyu, et al.
Published: (2024)
by: Zhai, Siyu, et al.
Published: (2024)
Exploring the Coordination of Frequency and Attention in Masked Image Modeling
by: Gui, Jie, et al.
Published: (2022)
by: Gui, Jie, et al.
Published: (2022)
Improving Fast Adversarial Training via Self-Knowledge Guidance
by: Jiang, Chengze, et al.
Published: (2024)
by: Jiang, Chengze, et al.
Published: (2024)
Efficient Diffusion-Based 3D Human Pose Estimation with Hierarchical Temporal Pruning
by: Bi, Yuquan, et al.
Published: (2025)
by: Bi, Yuquan, et al.
Published: (2025)
Deep Learning-Based Point Cloud Registration: A Comprehensive Survey and Taxonomy
by: Zhang, Yu-Xin, et al.
Published: (2024)
by: Zhang, Yu-Xin, et al.
Published: (2024)
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
by: Yao, Ziyu, et al.
Published: (2024)
by: Yao, Ziyu, et al.
Published: (2024)
SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis
by: Zhao, Shuxian, et al.
Published: (2026)
by: Zhao, Shuxian, et al.
Published: (2026)
Towards Localized Fine-Grained Control for Facial Expression Generation
by: Varanka, Tuomas, et al.
Published: (2024)
by: Varanka, Tuomas, et al.
Published: (2024)
Think-Before-Draw: Decomposing Emotion Semantics & Fine-Grained Controllable Expressive Talking Head Generation
by: Shi, Hanlei, et al.
Published: (2025)
by: Shi, Hanlei, et al.
Published: (2025)
FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing
by: Feng, Guanwen, et al.
Published: (2025)
by: Feng, Guanwen, et al.
Published: (2025)
StyleTalk++: A Unified Framework for Controlling the Speaking Styles of Talking Heads
by: Wang, Suzhen, et al.
Published: (2024)
by: Wang, Suzhen, et al.
Published: (2024)
LPIPS-AttnWav2Lip: Generic Audio-Driven lip synchronization for Talking Head Generation in the Wild
by: Chen, Zhipeng, et al.
Published: (2026)
by: Chen, Zhipeng, et al.
Published: (2026)
A Robust and Efficient Boundary Point Detection Method by Measuring Local Direction Dispersion
by: Peng, Dehua, et al.
Published: (2023)
by: Peng, Dehua, et al.
Published: (2023)
TalkCLIP: Talking Head Generation with Text-Guided Expressive Speaking Styles
by: Ma, Yifeng, et al.
Published: (2023)
by: Ma, Yifeng, et al.
Published: (2023)
ConsistTalk: Intensity Controllable Temporally Consistent Talking Head Generation with Diffusion Noise Search
by: Liu, Zhenjie, et al.
Published: (2025)
by: Liu, Zhenjie, et al.
Published: (2025)
A Comprehensive Survey on Underwater Image Enhancement Based on Deep Learning
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
Structure-Aware Fine-Grained Gaussian Splatting for Expressive Avatar Reconstruction
by: Su, Yuze, et al.
Published: (2026)
by: Su, Yuze, et al.
Published: (2026)
IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design
by: Shen, Fei, et al.
Published: (2025)
by: Shen, Fei, et al.
Published: (2025)
LES-Talker: Fine-Grained Emotion Editing for Talking Head Generation in Linear Emotion Space
by: Feng, Guanwen, et al.
Published: (2024)
by: Feng, Guanwen, et al.
Published: (2024)
PixelSmile: Toward Fine-Grained Facial Expression Editing
by: Hua, Jiabin, et al.
Published: (2026)
by: Hua, Jiabin, et al.
Published: (2026)
A Semi-supervised Nighttime Dehazing Baseline with Spatial-Frequency Aware and Realistic Brightness Constraint
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
LokiTalk: Learning Fine-Grained and Generalizable Correspondences to Enhance NeRF-based Talking Head Synthesis
by: Li, Tianqi, et al.
Published: (2024)
by: Li, Tianqi, et al.
Published: (2024)
Remarks on quasilocal mass and fill-ins
by: Tsang, Tin-Yau
Published: (2024)
by: Tsang, Tin-Yau
Published: (2024)
Positive mass theorem for initial data sets with arbitrary ends
by: Tsang, Tin-Yau
Published: (2026)
by: Tsang, Tin-Yau
Published: (2026)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
by: Sun, Zhiyao, et al.
Published: (2023)
by: Sun, Zhiyao, et al.
Published: (2023)
Cafe-Talk: Generating 3D Talking Face Animation with Multimodal Coarse- and Fine-grained Control
by: Chen, Hejia, et al.
Published: (2025)
by: Chen, Hejia, et al.
Published: (2025)
VQTalker: Towards Multilingual Talking Avatars through Facial Motion Tokenization
by: Liu, Tao, et al.
Published: (2024)
by: Liu, Tao, et al.
Published: (2024)
Detecting Deepfake Talking Heads from Facial Biometric Anomalies
by: Norman, Justin D., et al.
Published: (2025)
by: Norman, Justin D., et al.
Published: (2025)
Muse: Towards Reproducible Long-Form Song Generation with Fine-Grained Style Control
by: Jiang, Changhao, et al.
Published: (2026)
by: Jiang, Changhao, et al.
Published: (2026)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
by: Wang, Baiqin, et al.
Published: (2025)
by: Wang, Baiqin, et al.
Published: (2025)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
by: Ling, Jun, et al.
Published: (2024)
by: Ling, Jun, et al.
Published: (2024)
Controllable Talking Face Generation by Implicit Facial Keypoints Editing
by: Zhao, Dong, et al.
Published: (2024)
by: Zhao, Dong, et al.
Published: (2024)
High-Fidelity 3D Facial Avatar Synthesis with Controllable Fine-Grained Expressions
by: He, Yikang, et al.
Published: (2026)
by: He, Yikang, et al.
Published: (2026)
Similar Items
-
Fooling the Image Dehazing Models by First Order Gradient
by: Gui, Jie, et al.
Published: (2023) -
A Survey on Small Sample Imbalance Problem: Metrics, Feature Analysis, and Solutions
by: Zhao, Shuxian, et al.
Published: (2025) -
CFVNet: An End-to-End Cancelable Finger Vein Network for Recognition
by: Wang, Yifan, et al.
Published: (2024) -
BitC-3DGS: High-Capacity 3D Gaussian Splatting Watermarking via Bit Compression
by: Bi, Yuquan, et al.
Published: (2026) -
ColorVein: Colorful Cancelable Vein Biometrics
by: Wang, Yifan, et al.
Published: (2025)