Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bhattacharya, Uttaran, Bera, Aniket, Manocha, Dinesh |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Take an Emotion Walk: Perceiving Emotions from Gaits Using Hierarchical Attention Pooling and Affective Mapping
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
HighlightMe: Detecting Highlights from Human-Centric Videos
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attention
par: Bhattacharya, Uttaran, et autres
Publié: (2022)
par: Bhattacharya, Uttaran, et autres
Publié: (2022)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
par: Mullen Jr, James F., et autres
Publié: (2022)
par: Mullen Jr, James F., et autres
Publié: (2022)
SimMotionEdit: Text-Based Human Motion Editing with Motion Similarity Prediction
par: Li, Zhengyuan, et autres
Publié: (2025)
par: Li, Zhengyuan, et autres
Publié: (2025)
Learning Disentangled Speech- and Expression-Driven Blendshapes for 3D Talking Face Animation
par: Mao, Yuxiang, et autres
Publié: (2025)
par: Mao, Yuxiang, et autres
Publié: (2025)
Driving Animatronic Robot Facial Expression From Speech
par: Li, Boren, et autres
Publié: (2024)
par: Li, Boren, et autres
Publié: (2024)
Efficient and Robust Registration on the 3D Special Euclidean Group
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
OFER: Occluded Face Expression Reconstruction
par: Selvaraju, Pratheba, et autres
Publié: (2024)
par: Selvaraju, Pratheba, et autres
Publié: (2024)
Learning Spatial-Temporal Coherent Correlations for Speech-Preserving Facial Expression Manipulation
par: Chen, Tianshui, et autres
Publié: (2026)
par: Chen, Tianshui, et autres
Publié: (2026)
Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
par: Haliassos, Alexandros, et autres
Publié: (2024)
par: Haliassos, Alexandros, et autres
Publié: (2024)
Co-Speech Gesture and Facial Expression Generation for Non-Photorealistic 3D Characters
par: Omine, Taisei, et autres
Publié: (2025)
par: Omine, Taisei, et autres
Publié: (2025)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
par: Lee, Yonghan, et autres
Publié: (2026)
par: Lee, Yonghan, et autres
Publié: (2026)
Personalized Cross-Modal Emotional Correlation Learning for Speech-Preserving Facial Expression Manipulation
par: Chen, Tianshui, et autres
Publié: (2026)
par: Chen, Tianshui, et autres
Publié: (2026)
LoLep: Single-View View Synthesis with Locally-Learned Planes and Self-Attention Occlusion Inference
par: Wang, Cong, et autres
Publié: (2023)
par: Wang, Cong, et autres
Publié: (2023)
Input-Aware Sparse Attention for Real-Time Co-Speech Video Generation
par: Lu, Beijia, et autres
Publié: (2025)
par: Lu, Beijia, et autres
Publié: (2025)
Unified Multi-Modal Interactive & Reactive 3D Motion Generation via Rectified Flow
par: Gupta, Prerit, et autres
Publié: (2025)
par: Gupta, Prerit, et autres
Publié: (2025)
Text2Gestures: A Transformer-Based Network for Generating Emotive Body Gestures for Virtual Agents
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
Contrastive Decoupled Representation Learning and Regularization for Speech-Preserving Facial Expression Manipulation
par: Chen, Tianshui, et autres
Publié: (2025)
par: Chen, Tianshui, et autres
Publié: (2025)
Exploring Talking Head Models With Adjacent Frame Prior for Speech-Preserving Facial Expression Manipulation
par: Lu, Zhenxuan, et autres
Publié: (2026)
par: Lu, Zhenxuan, et autres
Publié: (2026)
Pairwise Discernment of AffectNet Expressions with ArcFace
par: Waldner, Dylan, et autres
Publié: (2024)
par: Waldner, Dylan, et autres
Publié: (2024)
FaceMixup: Enhancing Facial Expression Recognition through Mixed Face Regularization
par: Faria, Fabio A., et autres
Publié: (2024)
par: Faria, Fabio A., et autres
Publié: (2024)
PACE: Data-Driven Virtual Agent Interaction in Dense and Cluttered Environments
par: Mullen, James, et autres
Publié: (2023)
par: Mullen, James, et autres
Publié: (2023)
Comprehension of Multilingual Expressions Referring to Target Objects in Visual Inputs
par: Nogueira, Francisco, et autres
Publié: (2025)
par: Nogueira, Francisco, et autres
Publié: (2025)
SpeechForensics: Audio-Visual Speech Representation Learning for Face Forgery Detection
par: Liang, Yachao, et autres
Publié: (2025)
par: Liang, Yachao, et autres
Publié: (2025)
Face-StyleSpeech: Enhancing Zero-shot Speech Synthesis from Face Images with Improved Face-to-Speech Mapping
par: Kang, Minki, et autres
Publié: (2023)
par: Kang, Minki, et autres
Publié: (2023)
3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance
par: Kang, Taewon, et autres
Publié: (2024)
par: Kang, Taewon, et autres
Publié: (2024)
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
par: Hogue, Steven, et autres
Publié: (2024)
par: Hogue, Steven, et autres
Publié: (2024)
Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
par: Chen, Bohong, et autres
Publié: (2024)
par: Chen, Bohong, et autres
Publié: (2024)
An Implicit Physical Face Model Driven by Expression and Style
par: Yang, Lingchen, et autres
Publié: (2024)
par: Yang, Lingchen, et autres
Publié: (2024)
Financial Models in Generative Art: Black-Scholes-Inspired Concept Blending in Text-to-Image Diffusion
par: Kothandaraman, Divya, et autres
Publié: (2024)
par: Kothandaraman, Divya, et autres
Publié: (2024)
Differentiable Frequency-based Disentanglement for Aerial Video Action Recognition
par: Kothandaraman, Divya, et autres
Publié: (2022)
par: Kothandaraman, Divya, et autres
Publié: (2022)
SS-SFDA : Self-Supervised Source-Free Domain Adaptation for Road Segmentation in Hazardous Environments
par: Kothandaraman, Divya, et autres
Publié: (2020)
par: Kothandaraman, Divya, et autres
Publié: (2020)
A Unified and Interpretable Emotion Representation and Expression Generation
par: Paskaleva, Reni, et autres
Publié: (2024)
par: Paskaleva, Reni, et autres
Publié: (2024)
Faces of Fairness: Examining Bias in Facial Expression Recognition Datasets and Models
par: Hosseini, Mohammad Mehdi, et autres
Publié: (2025)
par: Hosseini, Mohammad Mehdi, et autres
Publié: (2025)
SLAT-Phys: Fast Material Property Field Prediction from Structured 3D Latents
par: Das, Rocktim Jyoti, et autres
Publié: (2026)
par: Das, Rocktim Jyoti, et autres
Publié: (2026)
DEGAS: Detailed Expressions on Full-Body Gaussian Avatars
par: Shao, Zhijing, et autres
Publié: (2024)
par: Shao, Zhijing, et autres
Publié: (2024)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
par: Jiang, Diqiong, et autres
Publié: (2026)
par: Jiang, Diqiong, et autres
Publié: (2026)
Documents similaires
-
Take an Emotion Walk: Perceiving Emotions from Gaits Using Hierarchical Attention Pooling and Affective Mapping
par: Bhattacharya, Uttaran, et autres
Publié: (2019) -
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
par: Bhattacharya, Uttaran, et autres
Publié: (2019) -
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
par: Bhattacharya, Uttaran, et autres
Publié: (2021) -
HighlightMe: Detecting Highlights from Human-Centric Videos
par: Bhattacharya, Uttaran, et autres
Publié: (2021) -
Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attention
par: Bhattacharya, Uttaran, et autres
Publié: (2022)