A Survey of Body and Face Motion: Datasets, Performance Evaluation Metrics and Generative Techniques
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sookha, Lownish Rai, Pakhale, Nikhil, Ganaie, Mudasir, Dhall, Abhinav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025)
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
End-to-End Motion Capture from Rigid Body Markers with Geodesic Loss
von: Lan, Hai, et al.
Veröffentlicht: (2025)
von: Lan, Hai, et al.
Veröffentlicht: (2025)
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
von: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Veröffentlicht: (2024)
von: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Veröffentlicht: (2024)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair Users
von: Huang, William, et al.
Veröffentlicht: (2024)
von: Huang, William, et al.
Veröffentlicht: (2024)
DeepFace-Attention: Multimodal Face Biometrics for Attention Estimation with Application to e-Learning
von: Daza, Roberto, et al.
Veröffentlicht: (2024)
von: Daza, Roberto, et al.
Veröffentlicht: (2024)
A Multimodal Dataset of Student Oral Presentations with Sensors and Evaluation Data
von: Becerra, Alvaro, et al.
Veröffentlicht: (2026)
von: Becerra, Alvaro, et al.
Veröffentlicht: (2026)
When Less Is More: A Sparse Facial Motion Structure For Listening Motion Learning
von: Nguyen, Tri Tung Nguyen, et al.
Veröffentlicht: (2025)
von: Nguyen, Tri Tung Nguyen, et al.
Veröffentlicht: (2025)
PoseAugment: Generative Human Pose Data Augmentation with Physical Plausibility for IMU-based Motion Capture
von: Li, Zhuojun, et al.
Veröffentlicht: (2024)
von: Li, Zhuojun, et al.
Veröffentlicht: (2024)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
Efficient Expression Neutrality Estimation with Application to Face Recognition Utility Prediction
von: Grimmer, Marcel, et al.
Veröffentlicht: (2024)
von: Grimmer, Marcel, et al.
Veröffentlicht: (2024)
Establishing a Baseline for Gaze-driven Authentication Performance in VR: A Breadth-First Investigation on a Very Large Dataset
von: Lohr, Dillon, et al.
Veröffentlicht: (2024)
von: Lohr, Dillon, et al.
Veröffentlicht: (2024)
Evaluating the Evaluators: Towards Human-aligned Metrics for Missing Markers Reconstruction
von: Kucherenko, Taras, et al.
Veröffentlicht: (2024)
von: Kucherenko, Taras, et al.
Veröffentlicht: (2024)
DEGSTalk: Decomposed Per-Embedding Gaussian Fields for Hair-Preserving Talking Face Synthesis
von: Deng, Kaijun, et al.
Veröffentlicht: (2024)
von: Deng, Kaijun, et al.
Veröffentlicht: (2024)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
von: Hall, Melissa, et al.
Veröffentlicht: (2023)
von: Hall, Melissa, et al.
Veröffentlicht: (2023)
A Convolution-Based Gait Asymmetry Metric for Inter-Limb Synergistic Coordination
von: Fukino, Go, et al.
Veröffentlicht: (2025)
von: Fukino, Go, et al.
Veröffentlicht: (2025)
Motion Generation Review: Exploring Deep Learning for Lifelike Animation with Manifold
von: Zhao, Jiayi, et al.
Veröffentlicht: (2024)
von: Zhao, Jiayi, et al.
Veröffentlicht: (2024)
HaDR: Applying Domain Randomization for Generating Synthetic Multimodal Dataset for Hand Instance Segmentation in Cluttered Industrial Environments
von: Grushko, Stefan, et al.
Veröffentlicht: (2023)
von: Grushko, Stefan, et al.
Veröffentlicht: (2023)
The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility
von: Zhang, Xiantao
Veröffentlicht: (2025)
von: Zhang, Xiantao
Veröffentlicht: (2025)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Efficient Listener: Dyadic Facial Motion Synthesis via Action Diffusion
von: Wang, Zesheng, et al.
Veröffentlicht: (2025)
von: Wang, Zesheng, et al.
Veröffentlicht: (2025)
SeamPose: Repurposing Seams as Capacitive Sensors in a Shirt for Upper-Body Pose Tracking
von: Yu, Tianhong Catherine, et al.
Veröffentlicht: (2024)
von: Yu, Tianhong Catherine, et al.
Veröffentlicht: (2024)
Has the Virtualization of the Face Changed Facial Perception? A Study of the Impact of Photo Editing and Augmented Reality on Facial Perception
von: Conwill, Louisa, et al.
Veröffentlicht: (2023)
von: Conwill, Louisa, et al.
Veröffentlicht: (2023)
Robot Interaction Behavior Generation based on Social Motion Forecasting for Human-Robot Interaction
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2024)
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2024)
Classification Metrics for Image Explanations: Towards Building Reliable XAI-Evaluations
von: Fresz, Benjamin, et al.
Veröffentlicht: (2024)
von: Fresz, Benjamin, et al.
Veröffentlicht: (2024)
A Survey on Drowsiness Detection -- Modern Applications and Methods
von: Fu, Biying, et al.
Veröffentlicht: (2024)
von: Fu, Biying, et al.
Veröffentlicht: (2024)
Motion Sickness Modeling with Visual Vertical Estimation and Its Application to Autonomous Personal Mobility Vehicles
von: Liu, Hailong, et al.
Veröffentlicht: (2022)
von: Liu, Hailong, et al.
Veröffentlicht: (2022)
Collection Space Navigator: An Interactive Visualization Interface for Multidimensional Datasets
von: Ohm, Tillmann, et al.
Veröffentlicht: (2023)
von: Ohm, Tillmann, et al.
Veröffentlicht: (2023)
Weak-Annotation of HAR Datasets using Vision Foundation Models
von: Bock, Marius, et al.
Veröffentlicht: (2024)
von: Bock, Marius, et al.
Veröffentlicht: (2024)
WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition
von: Bock, Marius, et al.
Veröffentlicht: (2023)
von: Bock, Marius, et al.
Veröffentlicht: (2023)
SuDA: Support-based Domain Adaptation for Sim2Real Motion Capture with Flexible Sensors
von: Fang, Jiawei, et al.
Veröffentlicht: (2024)
von: Fang, Jiawei, et al.
Veröffentlicht: (2024)
AI-Enhanced Virtual Reality in Medicine: A Comprehensive Survey
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
von: Liu, Can, et al.
Veröffentlicht: (2025)
von: Liu, Can, et al.
Veröffentlicht: (2025)
VideoA11y: Method and Dataset for Accessible Video Description
von: Li, Chaoyu, et al.
Veröffentlicht: (2025)
von: Li, Chaoyu, et al.
Veröffentlicht: (2025)
MobilePoser: Real-Time Full-Body Pose Estimation and 3D Human Translation from IMUs in Mobile Consumer Devices
von: Xu, Vasco, et al.
Veröffentlicht: (2025)
von: Xu, Vasco, et al.
Veröffentlicht: (2025)
VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents
von: Mazumdar, Amrita, et al.
Veröffentlicht: (2026)
von: Mazumdar, Amrita, et al.
Veröffentlicht: (2026)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2024)
von: Islam, Md Touhidul, et al.
Veröffentlicht: (2024)
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
von: Zhao, Yiming, et al.
Veröffentlicht: (2024)
von: Zhao, Yiming, et al.
Veröffentlicht: (2024)
Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
von: Park, Se Jin, et al.
Veröffentlicht: (2024)
von: Park, Se Jin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MAVEN: Multi-modal Attention for Valence-Arousal Emotion Network
von: Ahire, Vrushank, et al.
Veröffentlicht: (2025) -
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
von: Madan, Surbhi, et al.
Veröffentlicht: (2024) -
End-to-End Motion Capture from Rigid Body Markers with Geodesic Loss
von: Lan, Hai, et al.
Veröffentlicht: (2025) -
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
von: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Veröffentlicht: (2024) -
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)