Multimodal Integration Challenges in Emotionally Expressive Child Avatars for Training Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Salehi, Pegah, Sheshkal, Sajad Amouei, Thambawita, Vajira, Riegler, Michael A., Halvorsen, Pål |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
by: Salehi, Pegah, et al.
Published: (2024)
by: Salehi, Pegah, et al.
Published: (2024)
VideoHEDGE: Entropy-Based Hallucination Detection for Video-VLMs via Semantic Clustering and Spatiotemporal Perturbations
by: Gautam, Sushant, et al.
Published: (2026)
by: Gautam, Sushant, et al.
Published: (2026)
Z-Order Transformer for Feed-Forward Gaussian Splatting
by: Wang, Can, et al.
Published: (2026)
by: Wang, Can, et al.
Published: (2026)
EatGAN: An Edge-Attention Guided Generative Adversarial Network for Single Image Super-Resolution
by: Rao, Penghao, et al.
Published: (2025)
by: Rao, Penghao, et al.
Published: (2025)
TWIG: Two-Step Image Generation using Segmentation Masks in Diffusion Models
by: Rakib, Mazharul Islam, et al.
Published: (2025)
by: Rakib, Mazharul Islam, et al.
Published: (2025)
Isolated Sign Language Recognition with Segmentation and Pose Estimation
by: Perkins, Daniel, et al.
Published: (2025)
by: Perkins, Daniel, et al.
Published: (2025)
SoccerChat: Integrating Multimodal Data for Enhanced Soccer Game Understanding
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
Hybrid SIFT-SNN for Efficient Anomaly Detection of Traffic Flow-Control Infrastructure
by: Rathee, Munish, et al.
Published: (2025)
by: Rathee, Munish, et al.
Published: (2025)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
by: Jin, Hang, et al.
Published: (2025)
by: Jin, Hang, et al.
Published: (2025)
Interactive Image Selection and Training for Brain Tumor Segmentation Network
by: Cerqueira, Matheus A., et al.
Published: (2024)
by: Cerqueira, Matheus A., et al.
Published: (2024)
Performance Decay in Deepfake Detection: The Limitations of Training on Outdated Data
by: Richings, Jack, et al.
Published: (2025)
by: Richings, Jack, et al.
Published: (2025)
Depth Priors in Removal Neural Radiance Fields
by: Guo, Zhihao, et al.
Published: (2024)
by: Guo, Zhihao, et al.
Published: (2024)
How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study
by: Brusnicki, Roberto, et al.
Published: (2026)
by: Brusnicki, Roberto, et al.
Published: (2026)
OptiRoulette Optimizer: A New Stochastic Meta-Optimizer for up to 5.3x Faster Convergence
by: Mastromichalakis, Stamatis
Published: (2026)
by: Mastromichalakis, Stamatis
Published: (2026)
FastGS: Training 3D Gaussian Splatting in 100 Seconds
by: Ren, Shiwei, et al.
Published: (2025)
by: Ren, Shiwei, et al.
Published: (2025)
Real Time Human Detection by Unmanned Aerial Vehicles
by: Guettala, Walid, et al.
Published: (2024)
by: Guettala, Walid, et al.
Published: (2024)
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
by: Françani, André O., et al.
Published: (2024)
by: Françani, André O., et al.
Published: (2024)
Attire-Based Anomaly Detection in Restricted Areas Using YOLOv8 for Enhanced CCTV Security
by: B, Abdul Aziz A., et al.
Published: (2024)
by: B, Abdul Aziz A., et al.
Published: (2024)
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
Multimodal AI-based visualization of strategic leaders' emotional dynamics: a deep behavioral analysis of Trump's trade war discourse
by: Meng, Wei
Published: (2025)
by: Meng, Wei
Published: (2025)
SQUARE: Semantic Query-Augmented Fusion and Efficient Batch Reranking for Training-free Zero-Shot Composed Image Retrieval
by: Wu, Ren-Di, et al.
Published: (2025)
by: Wu, Ren-Di, et al.
Published: (2025)
The Trap of Presumed Equivalence: Artificial General Intelligence Should Not Be Assessed on the Scale of Human Intelligence
by: Dolgikh, Serge
Published: (2024)
by: Dolgikh, Serge
Published: (2024)
Building Brain Tumor Segmentation Networks with User-Assisted Filter Estimation and Selection
by: Cerqueira, Matheus A., et al.
Published: (2024)
by: Cerqueira, Matheus A., et al.
Published: (2024)
SEGS-SLAM: Structure-enhanced 3D Gaussian Splatting SLAM with Appearance Embedding
by: Wen, Tianci, et al.
Published: (2025)
by: Wen, Tianci, et al.
Published: (2025)
Integrating Attendance Tracking and Emotion Detection for Enhanced Student Engagement in Smart Classrooms
by: Ainebyona, Keith, et al.
Published: (2026)
by: Ainebyona, Keith, et al.
Published: (2026)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
by: Rashid, Muhammad, et al.
Published: (2026)
by: Rashid, Muhammad, et al.
Published: (2026)
Poisson Flow Consistency Training
by: Zhang, Anthony, et al.
Published: (2025)
by: Zhang, Anthony, et al.
Published: (2025)
Optimal Transport on the Lie Group of Roto-translations
by: Bon, Daan, et al.
Published: (2024)
by: Bon, Daan, et al.
Published: (2024)
Does CLIP perceive art the same way we do?
by: Asperti, Andrea, et al.
Published: (2025)
by: Asperti, Andrea, et al.
Published: (2025)
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
by: Livieris, Ioannis E., et al.
Published: (2023)
by: Livieris, Ioannis E., et al.
Published: (2023)
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
by: McIntosh, Declan, et al.
Published: (2026)
by: McIntosh, Declan, et al.
Published: (2026)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
by: Gautam, Sushant, et al.
Published: (2025)
by: Gautam, Sushant, et al.
Published: (2025)
Nonverbal Immediacy Analysis in Education: A Multimodal Computational Model
by: Petković, Uroš, et al.
Published: (2024)
by: Petković, Uroš, et al.
Published: (2024)
Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
by: Cooley, Madison, et al.
Published: (2024)
by: Cooley, Madison, et al.
Published: (2024)
Transformer-Based Model for Monocular Visual Odometry: A Video Understanding Approach
by: Françani, André O., et al.
Published: (2023)
by: Françani, André O., et al.
Published: (2023)
Computational Imaging Priors for Wireless Capsule Endoscopy: Monte Carlo-Guided Hemoglobin Mapping for Rare-Anomaly Detection
by: Yang, Chengshuai, et al.
Published: (2026)
by: Yang, Chengshuai, et al.
Published: (2026)
Probabilistic Variational Causal Approach in Observational Studies
by: Faghihi, Usef, et al.
Published: (2022)
by: Faghihi, Usef, et al.
Published: (2022)
Benchmarking Vision Language Models on German Factual Data
by: Peinl, René, et al.
Published: (2025)
by: Peinl, René, et al.
Published: (2025)
Optimizing MoE Routers: Design, Implementation, and Evaluation in Transformer Models
by: Harvey, Daniel Fidel, et al.
Published: (2025)
by: Harvey, Daniel Fidel, et al.
Published: (2025)
Similar Items
-
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
by: Salehi, Pegah, et al.
Published: (2024) -
VideoHEDGE: Entropy-Based Hallucination Detection for Video-VLMs via Semantic Clustering and Spatiotemporal Perturbations
by: Gautam, Sushant, et al.
Published: (2026) -
Z-Order Transformer for Feed-Forward Gaussian Splatting
by: Wang, Can, et al.
Published: (2026) -
EatGAN: An Edge-Attention Guided Generative Adversarial Network for Single Image Super-Resolution
by: Rao, Penghao, et al.
Published: (2025) -
TWIG: Two-Step Image Generation using Segmentation Masks in Diffusion Models
by: Rakib, Mazharul Islam, et al.
Published: (2025)