MA-LipNet: Multi-Dimensional Attention Networks for Robust Lipreading
Fuente:
arXiv
Saved in:
| Main Author: | Rossi, Matteo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Attention Fusion of Visual and Geometric Features for Large Vocabulary Arabic Lipreading
by: Daou, Samar, et al.
Published: (2024)
by: Daou, Samar, et al.
Published: (2024)
Learning Speaker-Invariant Visual Features for Lipreading
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Evaluation of End-to-End Continuous Spanish Lipreading in Different Data Conditions
by: Gimeno-Gómez, David, et al.
Published: (2025)
by: Gimeno-Gómez, David, et al.
Published: (2025)
RAL:Redundancy-Aware Lipreading Model Based on Differential Learning with Symmetric Views
by: gu, Zejun, et al.
Published: (2024)
by: gu, Zejun, et al.
Published: (2024)
FSCA-Net: Feature-Separated Cross-Attention Network for Robust Multi-Dataset Training
by: Chen, Yuehai
Published: (2026)
by: Chen, Yuehai
Published: (2026)
A Deformable Attention-Based Detection Transformer with Cross-Scale Feature Fusion for Industrial Coil Spring Inspection
by: Rossi, Matteo, et al.
Published: (2026)
by: Rossi, Matteo, et al.
Published: (2026)
Soft Knowledge Distillation with Multi-Dimensional Cross-Net Attention for Image Restoration Models Compression
by: Zhang, Yongheng, et al.
Published: (2025)
by: Zhang, Yongheng, et al.
Published: (2025)
RADE-Net: Robust Attention Network for Radar-Only Object Detection in Adverse Weather
by: Leitgeb, Christof, et al.
Published: (2026)
by: Leitgeb, Christof, et al.
Published: (2026)
WhisperNetV2: SlowFast Siamese Network For Lip-Based Biometrics
by: Zakeri, Abdollah, et al.
Published: (2024)
by: Zakeri, Abdollah, et al.
Published: (2024)
Enhancing Lip Reading with Multi-Scale Video and Multi-Encoder
by: Wang, He, et al.
Published: (2024)
by: Wang, He, et al.
Published: (2024)
MAVR-Net: Robust Multi-View Learning for MAV Action Recognition with Cross-View Attention
by: Zhang, Nengbo, et al.
Published: (2025)
by: Zhang, Nengbo, et al.
Published: (2025)
Efficient Audio-Visual Speech Separation with Discrete Lip Semantics and Multi-Scale Global-Local Attention
by: Li, Kai, et al.
Published: (2025)
by: Li, Kai, et al.
Published: (2025)
MGCA-Net: Multi-Graph Contextual Attention Network for Two-View Correspondence Learning
by: Lin, Shuyuan, et al.
Published: (2025)
by: Lin, Shuyuan, et al.
Published: (2025)
WMKA-Net: A Weighted Multi-Kernel Attention Network for Retinal Vessel Segmentation
by: Xu, Xinran, et al.
Published: (2025)
by: Xu, Xinran, et al.
Published: (2025)
AHMSA-Net: Adaptive Hierarchical Multi-Scale Attention Network for Micro-Expression Recognition
by: Zhang, Lijun, et al.
Published: (2025)
by: Zhang, Lijun, et al.
Published: (2025)
LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
by: Li, Chunyu, et al.
Published: (2024)
by: Li, Chunyu, et al.
Published: (2024)
MagNet: Multi-Level Attention Graph Network for Predicting High-Resolution Spatial Transcriptomics
by: Zhu, Junchao, et al.
Published: (2025)
by: Zhu, Junchao, et al.
Published: (2025)
BA-Net: Bridge Attention in Deep Neural Networks
by: Zhang, Ronghui, et al.
Published: (2024)
by: Zhang, Ronghui, et al.
Published: (2024)
UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer
by: Sharma, Dhruv, et al.
Published: (2024)
by: Sharma, Dhruv, et al.
Published: (2024)
LASER: Lip Landmark Assisted Speaker Detection for Robustness
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
KeySync: A Robust Approach for Leakage-free Lip Synchronization in High Resolution
by: Bigata, Antoni, et al.
Published: (2025)
by: Bigata, Antoni, et al.
Published: (2025)
Lips Are Lying: Spotting the Temporal Inconsistency between Audio and Visual in Lip-Syncing DeepFakes
by: Liu, Weifeng, et al.
Published: (2024)
by: Liu, Weifeng, et al.
Published: (2024)
LEADER: Lightweight End-to-End Attention-Gated Dual Autoencoder for Robust Minutiae Extraction
by: Cappelli, Raffaele, et al.
Published: (2026)
by: Cappelli, Raffaele, et al.
Published: (2026)
RAUM-Net: Regional Attention and Uncertainty-aware Mamba Network
by: Liu, Mingquan
Published: (2025)
by: Liu, Mingquan
Published: (2025)
BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling
by: Chen, Hao, et al.
Published: (2026)
by: Chen, Hao, et al.
Published: (2026)
MCI-Net: A Robust Multi-Domain Context Integration Network for Point Cloud Registration
by: Lin, Shuyuan, et al.
Published: (2025)
by: Lin, Shuyuan, et al.
Published: (2025)
LipSim: A Provably Robust Perceptual Similarity Metric
by: Ghazanfari, Sara, et al.
Published: (2023)
by: Ghazanfari, Sara, et al.
Published: (2023)
FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs
by: Zinonos, Andreas, et al.
Published: (2025)
by: Zinonos, Andreas, et al.
Published: (2025)
SMR-Net:Robot Snap Detection Based on Multi-Scale Features and Self-Attention Network
by: Hou, Kuanxu
Published: (2026)
by: Hou, Kuanxu
Published: (2026)
MTGA: Multi-View Temporal Granularity Aligned Aggregation for Event-Based Lip-Reading
by: Zhang, Wenhao, et al.
Published: (2024)
by: Zhang, Wenhao, et al.
Published: (2024)
LipGen: Viseme-Guided Lip Video Generation for Enhancing Visual Speech Recognition
by: Hao, Bowen, et al.
Published: (2025)
by: Hao, Bowen, et al.
Published: (2025)
A Streamlined Attention-Based Network for Descriptor Extraction
by: D'Urso, Mattia, et al.
Published: (2026)
by: D'Urso, Mattia, et al.
Published: (2026)
BookNet: Book Image Rectification via Cross-Page Attention Network
by: Liu, Shaokai, et al.
Published: (2026)
by: Liu, Shaokai, et al.
Published: (2026)
LLHA-Net: A Hierarchical Attention Network for Two-View Correspondence Learning
by: Lin, Shuyuan, et al.
Published: (2025)
by: Lin, Shuyuan, et al.
Published: (2025)
DATransNet: Dynamic Attention Transformer Network for Infrared Small Target Detection
by: Hu, Chen, et al.
Published: (2024)
by: Hu, Chen, et al.
Published: (2024)
RAFA-Net: Region Attention Network For Food Items And Agricultural Stress Recognition
by: Bera, Asish, et al.
Published: (2024)
by: Bera, Asish, et al.
Published: (2024)
WaveNets: Wavelet Channel Attention Networks
by: Salman, Hadi, et al.
Published: (2022)
by: Salman, Hadi, et al.
Published: (2022)
DynamicLip: Shape-Independent Continuous Authentication via Lip Articulator Dynamics
by: Chen, Huashan, et al.
Published: (2025)
by: Chen, Huashan, et al.
Published: (2025)
GS-Net: Global Self-Attention Guided CNN for Multi-Stage Glaucoma Classification
by: Das, Dipankar, et al.
Published: (2024)
by: Das, Dipankar, et al.
Published: (2024)
VALLR: Visual ASR Language Model for Lip Reading
by: Thomas, Marshall, et al.
Published: (2025)
by: Thomas, Marshall, et al.
Published: (2025)
Similar Items
-
Cross-Attention Fusion of Visual and Geometric Features for Large Vocabulary Arabic Lipreading
by: Daou, Samar, et al.
Published: (2024) -
Learning Speaker-Invariant Visual Features for Lipreading
by: Li, Yu, et al.
Published: (2025) -
Evaluation of End-to-End Continuous Spanish Lipreading in Different Data Conditions
by: Gimeno-Gómez, David, et al.
Published: (2025) -
RAL:Redundancy-Aware Lipreading Model Based on Differential Learning with Symmetric Views
by: gu, Zejun, et al.
Published: (2024) -
FSCA-Net: Feature-Separated Cross-Attention Network for Robust Multi-Dataset Training
by: Chen, Yuehai
Published: (2026)