EMOPortraits: Emotion-enhanced Multimodal One-shot Head Avatars
Fuente:
arXiv
Salvato in:
| Autori principali: | Drobyshev, Nikita, Casademunt, Antoni Bigata, Vougioukas, Konstantinos, Landgraf, Zoe, Petridis, Stavros, Pantic, Maja |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame Interpolation
di: Bigata, Antoni, et al.
Pubblicazione: (2025)
di: Bigata, Antoni, et al.
Pubblicazione: (2025)
FaceCrafter: Identity-Conditional Diffusion with Disentangled Control over Facial Pose, Expression, and Emotion
di: Mishima, Kazuaki, et al.
Pubblicazione: (2025)
di: Mishima, Kazuaki, et al.
Pubblicazione: (2025)
FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs
di: Zinonos, Andreas, et al.
Pubblicazione: (2025)
di: Zinonos, Andreas, et al.
Pubblicazione: (2025)
KeySync: A Robust Approach for Leakage-free Lip Synchronization in High Resolution
di: Bigata, Antoni, et al.
Pubblicazione: (2025)
di: Bigata, Antoni, et al.
Pubblicazione: (2025)
Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
di: Haliassos, Alexandros, et al.
Pubblicazione: (2024)
di: Haliassos, Alexandros, et al.
Pubblicazione: (2024)
Dr. SHAP-AV: Decoding Relative Modality Contributions via Shapley Attribution in Audio-Visual Speech Recognition
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2026)
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2026)
Mitigating Attention Sinks and Massive Activations in Audio-Visual Speech Recognition with LLMs
di: Anand, et al.
Pubblicazione: (2025)
di: Anand, et al.
Pubblicazione: (2025)
BRAVEn: Improving Self-Supervised Pre-training for Visual and Auditory Speech Recognition
di: Haliassos, Alexandros, et al.
Pubblicazione: (2024)
di: Haliassos, Alexandros, et al.
Pubblicazione: (2024)
Omni-AVSR: Towards Unified Multimodal Speech Recognition with Large Language Models
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
RT-LA-VocE: Real-Time Low-SNR Audio-Visual Speech Enhancement
di: Chen, Honglie, et al.
Pubblicazione: (2024)
di: Chen, Honglie, et al.
Pubblicazione: (2024)
OMG-Avatar: One-shot Multi-LOD Gaussian Head Avatar
di: Ren, Jianqiang, et al.
Pubblicazione: (2026)
di: Ren, Jianqiang, et al.
Pubblicazione: (2026)
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
di: Yu, Zhixuan, et al.
Pubblicazione: (2024)
di: Yu, Zhixuan, et al.
Pubblicazione: (2024)
MSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
One-shot Compositional 3D Head Avatars with Deformable Hair
di: Sun, Yuan, et al.
Pubblicazione: (2026)
di: Sun, Yuan, et al.
Pubblicazione: (2026)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
di: He, Yisheng, et al.
Pubblicazione: (2025)
di: He, Yisheng, et al.
Pubblicazione: (2025)
Full-Rank No More: Low-Rank Weight Training for Modern Speech Recognition Models
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
MoME: Mixture of Matryoshka Experts for Audio-Visual Speech Recognition
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
di: Li, Peng, et al.
Pubblicazione: (2025)
di: Li, Peng, et al.
Pubblicazione: (2025)
OMEGA-Avatar: One-shot Modeling of 360° Gaussian Avatars
di: Xia, Zehao, et al.
Pubblicazione: (2026)
di: Xia, Zehao, et al.
Pubblicazione: (2026)
Large Language Models are Strong Audio-Visual Speech Recognition Learners
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2024)
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2024)
OHTA: One-shot Hand Avatar via Data-driven Implicit Priors
di: Zheng, Xiaozheng, et al.
Pubblicazione: (2024)
di: Zheng, Xiaozheng, et al.
Pubblicazione: (2024)
Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand Avatars
di: Huang, Xuan, et al.
Pubblicazione: (2024)
di: Huang, Xuan, et al.
Pubblicazione: (2024)
InvertAvatar: Incremental GAN Inversion for Generalized Head Avatars
di: Zhao, Xiaochen, et al.
Pubblicazione: (2023)
di: Zhao, Xiaochen, et al.
Pubblicazione: (2023)
PhysHead: Simulation-Ready Gaussian Head Avatars
di: Kabadayi, Berna, et al.
Pubblicazione: (2026)
di: Kabadayi, Berna, et al.
Pubblicazione: (2026)
LightAvatar: Efficient Head Avatar as Dynamic Neural Light Field
di: Wang, Huan, et al.
Pubblicazione: (2024)
di: Wang, Huan, et al.
Pubblicazione: (2024)
GaussianAvatars: Photorealistic Head Avatars with Rigged 3D Gaussians
di: Qian, Shenhan, et al.
Pubblicazione: (2023)
di: Qian, Shenhan, et al.
Pubblicazione: (2023)
GaussianAvatar-Editor: Photorealistic Animatable Gaussian Head Avatar Editor
di: Liu, Xiangyue, et al.
Pubblicazione: (2025)
di: Liu, Xiangyue, et al.
Pubblicazione: (2025)
Adaptive Audio-Visual Speech Recognition via Matryoshka-Based Multimodal LLMs
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
di: Cappellazzo, Umberto, et al.
Pubblicazione: (2025)
Pay Attention to CTC: Fast and Robust Pseudo-Labelling for Unified Speech Recognition
di: Haliassos, Alexandros, et al.
Pubblicazione: (2026)
di: Haliassos, Alexandros, et al.
Pubblicazione: (2026)
3DRealHead: Few-Shot Detailed Head Avatar
di: Nehvi, Jalees, et al.
Pubblicazione: (2026)
di: Nehvi, Jalees, et al.
Pubblicazione: (2026)
Hearing Loss Detection from Facial Expressions in One-on-one Conversations
di: Yin, Yufeng, et al.
Pubblicazione: (2024)
di: Yin, Yufeng, et al.
Pubblicazione: (2024)
Stable Video-Driven Portraits
di: R., Mallikarjun B., et al.
Pubblicazione: (2025)
di: R., Mallikarjun B., et al.
Pubblicazione: (2025)
FlexAvatar: Learning Complete 3D Head Avatars with Partial Supervision
di: Kirschstein, Tobias, et al.
Pubblicazione: (2025)
di: Kirschstein, Tobias, et al.
Pubblicazione: (2025)
AvatarMakeup: Realistic Makeup Transfer for 3D Animatable Head Avatars
di: Zhong, Yiming, et al.
Pubblicazione: (2025)
di: Zhong, Yiming, et al.
Pubblicazione: (2025)
DiffusionAvatars: Deferred Diffusion for High-fidelity 3D Head Avatars
di: Kirschstein, Tobias, et al.
Pubblicazione: (2023)
di: Kirschstein, Tobias, et al.
Pubblicazione: (2023)
GraphAvatar: Compact Head Avatars with GNN-Generated 3D Gaussians
di: Wei, Xiaobao, et al.
Pubblicazione: (2024)
di: Wei, Xiaobao, et al.
Pubblicazione: (2024)
Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars
di: Gong, Yicheng, et al.
Pubblicazione: (2026)
di: Gong, Yicheng, et al.
Pubblicazione: (2026)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
di: Zhou, Zhenglin, et al.
Pubblicazione: (2025)
di: Zhou, Zhenglin, et al.
Pubblicazione: (2025)
GaussianHead: High-fidelity Head Avatars with Learnable Gaussian Derivation
di: Wang, Jie, et al.
Pubblicazione: (2023)
di: Wang, Jie, et al.
Pubblicazione: (2023)
GPHM: Gaussian Parametric Head Model for Monocular Head Avatar Reconstruction
di: Xu, Yuelang, et al.
Pubblicazione: (2024)
di: Xu, Yuelang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame Interpolation
di: Bigata, Antoni, et al.
Pubblicazione: (2025) -
FaceCrafter: Identity-Conditional Diffusion with Disentangled Control over Facial Pose, Expression, and Emotion
di: Mishima, Kazuaki, et al.
Pubblicazione: (2025) -
FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs
di: Zinonos, Andreas, et al.
Pubblicazione: (2025) -
KeySync: A Robust Approach for Leakage-free Lip Synchronization in High Resolution
di: Bigata, Antoni, et al.
Pubblicazione: (2025) -
Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
di: Haliassos, Alexandros, et al.
Pubblicazione: (2024)