Deep Learning for Visual Speech Analysis: A Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Sheng, Changchong, Kuang, Gangyao, Bai, Liang, Hou, Chenping, Guo, Yulan, Xu, Xin, Pietikäinen, Matti, Liu, Li |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lightweight Pixel Difference Networks for Efficient Visual Representation Learning
by: Su, Zhuo, et al.
Published: (2024)
by: Su, Zhuo, et al.
Published: (2024)
Bottom-Up Scattering Information Perception Network for SAR target recognition
by: Zhao, Chenxi, et al.
Published: (2025)
by: Zhao, Chenxi, et al.
Published: (2025)
Multi-Level Heterogeneous Knowledge Transfer Network on Forward Scattering Center Model for Limited Samples SAR ATR
by: Zhao, Chenxi, et al.
Published: (2025)
by: Zhao, Chenxi, et al.
Published: (2025)
Adaptive Disentangled Representation Learning for Incomplete Multi-View Multi-Label Classification
by: Li, Quanjiang, et al.
Published: (2026)
by: Li, Quanjiang, et al.
Published: (2026)
Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory
by: Li, Quanjiang, et al.
Published: (2026)
by: Li, Quanjiang, et al.
Published: (2026)
MTSGL: Multi-Task Structure Guided Learning for Robust and Interpretable SAR Aircraft Recognition
by: He, Qishan, et al.
Published: (2025)
by: He, Qishan, et al.
Published: (2025)
Rapid Salient Object Detection with Difference Convolutional Neural Networks
by: Su, Zhuo, et al.
Published: (2025)
by: Su, Zhuo, et al.
Published: (2025)
Continuous Subspace Optimization for Continual Learning
by: Cheng, Quan, et al.
Published: (2025)
by: Cheng, Quan, et al.
Published: (2025)
Deep Learning for Cross-Domain Few-Shot Visual Recognition: A Survey
by: Xu, Huali, et al.
Published: (2023)
by: Xu, Huali, et al.
Published: (2023)
Monocular Lane Detection Based on Deep Learning: A Survey
by: He, Xin, et al.
Published: (2024)
by: He, Xin, et al.
Published: (2024)
MonoMobility: Zero-Shot 3D Mobility Analysis from Monocular Videos
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
Progressive Correspondence Regenerator for Robust 3D Registration
by: Zhao, Guiyu, et al.
Published: (2025)
by: Zhao, Guiyu, et al.
Published: (2025)
DuInNet: Dual-Modality Feature Interaction for Point Cloud Completion
by: Liu, Xinpu, et al.
Published: (2024)
by: Liu, Xinpu, et al.
Published: (2024)
SpeechForensics: Audio-Visual Speech Representation Learning for Face Forgery Detection
by: Liang, Yachao, et al.
Published: (2025)
by: Liang, Yachao, et al.
Published: (2025)
MCFCN: Multi-View Clustering via a Fusion-Consensus Graph Convolutional Network
by: Pei, Chenping, et al.
Published: (2025)
by: Pei, Chenping, et al.
Published: (2025)
DropoutGS: Dropping Out Gaussians for Better Sparse-view Rendering
by: Xu, Yexing, et al.
Published: (2025)
by: Xu, Yexing, et al.
Published: (2025)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
by: Chen, Minglin, et al.
Published: (2025)
by: Chen, Minglin, et al.
Published: (2025)
SDSTrack: Self-Distillation Symmetric Adapter Learning for Multi-Modal Visual Object Tracking
by: Hou, Xiaojun, et al.
Published: (2024)
by: Hou, Xiaojun, et al.
Published: (2024)
Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis
by: Zhou, Xin, et al.
Published: (2024)
by: Zhou, Xin, et al.
Published: (2024)
Pluggable Style Representation Learning for Multi-Style Transfer
by: Liu, Hongda, et al.
Published: (2025)
by: Liu, Hongda, et al.
Published: (2025)
A Comprehensive Survey on Underwater Image Enhancement Based on Deep Learning
by: Cong, Xiaofeng, et al.
Published: (2024)
by: Cong, Xiaofeng, et al.
Published: (2024)
Survey on Fundamental Deep Learning 3D Reconstruction Techniques
by: Bai, Yonge, et al.
Published: (2024)
by: Bai, Yonge, et al.
Published: (2024)
Improving Medical Visual Representation Learning with Pathological-level Cross-Modal Alignment and Correlation Exploration
by: Wang, Jun, et al.
Published: (2025)
by: Wang, Jun, et al.
Published: (2025)
Deep Learning Empowered Super-Resolution: A Comprehensive Survey and Future Prospects
by: Zhang, Le, et al.
Published: (2025)
by: Zhang, Le, et al.
Published: (2025)
Omni Survey for Multimodality Analysis in Visual Object Tracking
by: Tang, Zhangyong, et al.
Published: (2025)
by: Tang, Zhangyong, et al.
Published: (2025)
Referring-Aware Visuomotor Policy Learning for Closed-Loop Manipulation
by: Ma, Jiahua, et al.
Published: (2026)
by: Ma, Jiahua, et al.
Published: (2026)
Local Feature Matching Using Deep Learning: A Survey
by: Xu, Shibiao, et al.
Published: (2024)
by: Xu, Shibiao, et al.
Published: (2024)
OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging
by: Tang, Yijie, et al.
Published: (2025)
by: Tang, Yijie, et al.
Published: (2025)
Revisit the Imbalance Optimization in Multi-task Learning: An Experimental Analysis
by: Guo, Yihang, et al.
Published: (2025)
by: Guo, Yihang, et al.
Published: (2025)
Deep Learning-Based Point Cloud Registration: A Comprehensive Survey and Taxonomy
by: Zhang, Yu-Xin, et al.
Published: (2024)
by: Zhang, Yu-Xin, et al.
Published: (2024)
Highly Efficient and Unsupervised Framework for Moving Object Detection in Satellite Videos
by: Xiao, C., et al.
Published: (2024)
by: Xiao, C., et al.
Published: (2024)
Deep Lookup Network
by: Guo, Yulan, et al.
Published: (2025)
by: Guo, Yulan, et al.
Published: (2025)
Visual Bias and Interpretability in Deep Learning for Dermatological Image Analysis
by: Taufik, Enam Ahmed, et al.
Published: (2025)
by: Taufik, Enam Ahmed, et al.
Published: (2025)
Learning Hierarchical Color Guidance for Depth Map Super-Resolution
by: Cong, Runmin, et al.
Published: (2024)
by: Cong, Runmin, et al.
Published: (2024)
CFTrack: Enhancing Lightweight Visual Tracking through Contrastive Learning and Feature Matching
by: Liang, Juntao, et al.
Published: (2025)
by: Liang, Juntao, et al.
Published: (2025)
SAVE: Speech-Aware Video Representation Learning for Video-Text Retrieval
by: Zhao, Ruixiang, et al.
Published: (2026)
by: Zhao, Ruixiang, et al.
Published: (2026)
PdfTable: A Unified Toolkit for Deep Learning-Based Table Extraction
by: Sheng, Lei, et al.
Published: (2024)
by: Sheng, Lei, et al.
Published: (2024)
Deep Learning for Event-based Vision: A Comprehensive Survey and Benchmarks
by: Zheng, Xu, et al.
Published: (2023)
by: Zheng, Xu, et al.
Published: (2023)
Deep Learning For Point Cloud Denoising: A Survey
by: Zhang, Chengwei, et al.
Published: (2025)
by: Zhang, Chengwei, et al.
Published: (2025)
Visual Attention Prompted Prediction and Learning
by: Zhang, Yifei, et al.
Published: (2023)
by: Zhang, Yifei, et al.
Published: (2023)
Similar Items
-
Lightweight Pixel Difference Networks for Efficient Visual Representation Learning
by: Su, Zhuo, et al.
Published: (2024) -
Bottom-Up Scattering Information Perception Network for SAR target recognition
by: Zhao, Chenxi, et al.
Published: (2025) -
Multi-Level Heterogeneous Knowledge Transfer Network on Forward Scattering Center Model for Limited Samples SAR ATR
by: Zhao, Chenxi, et al.
Published: (2025) -
Adaptive Disentangled Representation Learning for Incomplete Multi-View Multi-Label Classification
by: Li, Quanjiang, et al.
Published: (2026) -
Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory
by: Li, Quanjiang, et al.
Published: (2026)