Development and evaluation of a deep learning algorithm for German word recognition from lip movements
Fuente:
arXiv
Salvato in:
| Autori principali: | Pham, Dinh Nam, Rahne, Torsten |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Transfer Learning from Visual Speech Recognition to Mouthing Recognition in German Sign Language
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025)
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025)
The Importance of Facial Features in Vision-based Sign Language Recognition: Eyes, Mouth or Full Face?
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025)
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025)
MobileMold: A Smartphone-Based Microscopy Dataset for Food Mold Detection
di: Pham, Dinh Nam, et al.
Pubblicazione: (2026)
di: Pham, Dinh Nam, et al.
Pubblicazione: (2026)
A self-supervised learning approach to deep filter banks for texture recognition
di: Florindo, Joao B., et al.
Pubblicazione: (2026)
di: Florindo, Joao B., et al.
Pubblicazione: (2026)
An inclusive review on deep learning techniques and their scope in handwriting recognition
di: Singh, Sukhdeep, et al.
Pubblicazione: (2024)
di: Singh, Sukhdeep, et al.
Pubblicazione: (2024)
Automated facial recognition system using deep learning for pain assessment in adults with cerebral palsy
di: Sabater-Gárriz, Álvaro, et al.
Pubblicazione: (2024)
di: Sabater-Gárriz, Álvaro, et al.
Pubblicazione: (2024)
Cracking the neural code for word recognition in convolutional neural networks
di: Agrawal, Aakash, et al.
Pubblicazione: (2024)
di: Agrawal, Aakash, et al.
Pubblicazione: (2024)
EmMixformer: Mix transformer for eye movement recognition
di: Qin, Huafeng, et al.
Pubblicazione: (2024)
di: Qin, Huafeng, et al.
Pubblicazione: (2024)
Multispectral airborne laser scanning for tree species classification: a benchmark of machine learning and deep learning algorithms
di: Taher, Josef, et al.
Pubblicazione: (2025)
di: Taher, Josef, et al.
Pubblicazione: (2025)
AlertTrap: A study on object detection in remote insects trap monitoring system using on-the-edge deep learning platform
di: Le, An D., et al.
Pubblicazione: (2021)
di: Le, An D., et al.
Pubblicazione: (2021)
Cross-modal learning for plankton recognition
di: Kareinen, Joona, et al.
Pubblicazione: (2026)
di: Kareinen, Joona, et al.
Pubblicazione: (2026)
HadaSmileNet: Hadamard fusion of handcrafted and deep-learning features for enhancing facial emotion recognition of genuine smiles
di: Hasan, Mohammad Junayed, et al.
Pubblicazione: (2025)
di: Hasan, Mohammad Junayed, et al.
Pubblicazione: (2025)
Data standardization for robust lip sync
di: Wang, Chun
Pubblicazione: (2022)
di: Wang, Chun
Pubblicazione: (2022)
A comparative study of generative adversarial networks for image recognition algorithms based on deep learning and traditional methods
di: Zhong, Yihao, et al.
Pubblicazione: (2024)
di: Zhong, Yihao, et al.
Pubblicazione: (2024)
Deeply Dual Supervised learning for melanoma recognition
di: Polma, Rujosh, et al.
Pubblicazione: (2025)
di: Polma, Rujosh, et al.
Pubblicazione: (2025)
Enhancing deep learning performance on burned area delineation from SPOT-6/7 imagery for emergency management
di: Rodriguez, Maria, et al.
Pubblicazione: (2025)
di: Rodriguez, Maria, et al.
Pubblicazione: (2025)
Performance evaluation of deep learning models for image analysis: considerations for visual control and statistical metrics
di: Bertram, Christof A., et al.
Pubblicazione: (2026)
di: Bertram, Christof A., et al.
Pubblicazione: (2026)
Community-aware evaluation and threshold calibration for open-set plankton image recognition
di: Chen, Xi, et al.
Pubblicazione: (2026)
di: Chen, Xi, et al.
Pubblicazione: (2026)
Illicit object detection in X-ray imaging using deep learning techniques: A comparative evaluation
di: Cani, Jorgen, et al.
Pubblicazione: (2025)
di: Cani, Jorgen, et al.
Pubblicazione: (2025)
Multi-model learning by sequential reading of untrimmed videos for action recognition
di: Kamiya, Kodai, et al.
Pubblicazione: (2024)
di: Kamiya, Kodai, et al.
Pubblicazione: (2024)
Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network
di: Dinh, Hiep, et al.
Pubblicazione: (2025)
di: Dinh, Hiep, et al.
Pubblicazione: (2025)
Towards classification-based representation learning for place recognition on LiDAR scans
di: Konoplia, Maksim, et al.
Pubblicazione: (2025)
di: Konoplia, Maksim, et al.
Pubblicazione: (2025)
Link prediction Graph Neural Networks for structure recognition of Handwritten Mathematical Expressions
di: Nguyen, Cuong Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Cuong Tuan, et al.
Pubblicazione: (2025)
OXSeg: Multidimensional attention UNet-based lip segmentation using semi-supervised lip contours
di: Moghaddasi, Hanie, et al.
Pubblicazione: (2025)
di: Moghaddasi, Hanie, et al.
Pubblicazione: (2025)
Restoration algorithms and system performance evaluation for active imagers
di: Gilles, Jerome
Pubblicazione: (2024)
di: Gilles, Jerome
Pubblicazione: (2024)
LoopDraw: a Loop-Based Autoregressive Model for Shape Synthesis and Editing
di: Dinh, Nam Anh, et al.
Pubblicazione: (2022)
di: Dinh, Nam Anh, et al.
Pubblicazione: (2022)
HCR-Net: A deep learning based script independent handwritten character recognition network
di: Chauhan, Vinod Kumar, et al.
Pubblicazione: (2021)
di: Chauhan, Vinod Kumar, et al.
Pubblicazione: (2021)
Multi-objective hybrid knowledge distillation for efficient deep learning in smart agriculture
di: Hoang, Phi-Hung, et al.
Pubblicazione: (2025)
di: Hoang, Phi-Hung, et al.
Pubblicazione: (2025)
Development and validation of an AI foundation model for endoscopic diagnosis of esophagogastric junction adenocarcinoma: a cohort and deep learning study
di: Ma, Yikun, et al.
Pubblicazione: (2025)
di: Ma, Yikun, et al.
Pubblicazione: (2025)
Towards a universal mechanism for successful deep learning
di: Meir, Yuval, et al.
Pubblicazione: (2023)
di: Meir, Yuval, et al.
Pubblicazione: (2023)
Linking in Style: Understanding learned features in deep learning models
di: Wehrheim, Maren H., et al.
Pubblicazione: (2024)
di: Wehrheim, Maren H., et al.
Pubblicazione: (2024)
Leveraging knowledge distillation for partial multi-task learning from multiple remote sensing datasets
di: Lê, Hoàng-Ân, et al.
Pubblicazione: (2024)
di: Lê, Hoàng-Ân, et al.
Pubblicazione: (2024)
A deep learning perspective on Rubens' attribution
di: Afifi, A., et al.
Pubblicazione: (2025)
di: Afifi, A., et al.
Pubblicazione: (2025)
One-shot recognition of any material anywhere using contrastive learning with physics-based rendering
di: Drehwald, Manuel S., et al.
Pubblicazione: (2022)
di: Drehwald, Manuel S., et al.
Pubblicazione: (2022)
Contrastive learning-based video quality assessment-jointed video vision transformer for video recognition
di: Sun, Jian, et al.
Pubblicazione: (2026)
di: Sun, Jian, et al.
Pubblicazione: (2026)
Two-stream joint matching method based on contrastive learning for few-shot action recognition
di: Deng, Long, et al.
Pubblicazione: (2024)
di: Deng, Long, et al.
Pubblicazione: (2024)
Continual-learning-based framework for structural damage recognition
di: Shu, Jiangpeng, et al.
Pubblicazione: (2024)
di: Shu, Jiangpeng, et al.
Pubblicazione: (2024)
One-shot lip-based biometric authentication: extending behavioral features with authentication phrase information
di: Koch, Brando, et al.
Pubblicazione: (2023)
di: Koch, Brando, et al.
Pubblicazione: (2023)
SepFormer: Coarse-to-fine Separator Regression Network for Table Structure Recognition
di: Nguyen, Nam Quan, et al.
Pubblicazione: (2025)
di: Nguyen, Nam Quan, et al.
Pubblicazione: (2025)
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
di: Pham, Hieu Dinh Trung, et al.
Pubblicazione: (2025)
di: Pham, Hieu Dinh Trung, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Transfer Learning from Visual Speech Recognition to Mouthing Recognition in German Sign Language
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025) -
The Importance of Facial Features in Vision-based Sign Language Recognition: Eyes, Mouth or Full Face?
di: Pham, Dinh Nam, et al.
Pubblicazione: (2025) -
MobileMold: A Smartphone-Based Microscopy Dataset for Food Mold Detection
di: Pham, Dinh Nam, et al.
Pubblicazione: (2026) -
A self-supervised learning approach to deep filter banks for texture recognition
di: Florindo, Joao B., et al.
Pubblicazione: (2026) -
An inclusive review on deep learning techniques and their scope in handwriting recognition
di: Singh, Sukhdeep, et al.
Pubblicazione: (2024)