Saved in:
| Main Authors: | Taghizadeh, Zahra, Shahverdikondori, Mohammad, Noori, Arian, Dadgarnia, Alireza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.22716 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Integrating Persian Lip Reading in Surena-V Humanoid Robot for Human-Robot Interaction
by: Abbasi, Ali Farshian, et al.
Published: (2025)
by: Abbasi, Ali Farshian, et al.
Published: (2025)
Khayyam Offline Persian Handwriting Dataset
by: Jafarzadeh, Pourya, et al.
Published: (2024)
by: Jafarzadeh, Pourya, et al.
Published: (2024)
A comprehensive Persian offline handwritten database for investigating the effects of heritability and family relationships on handwriting
by: Zohrevand, Abbas, et al.
Published: (2025)
by: Zohrevand, Abbas, et al.
Published: (2025)
STAF: Sinusoidal Trainable Activation Functions for Implicit Neural Representation
by: Morsali, Alireza, et al.
Published: (2025)
by: Morsali, Alireza, et al.
Published: (2025)
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
by: Hegde, Sindhu, et al.
Published: (2024)
by: Hegde, Sindhu, et al.
Published: (2024)
BioLip: Language-Generalizable Lip-Sync Deepfake Detection via Biomechanical Constraint Violation Modeling
by: Chen, Hao, et al.
Published: (2026)
by: Chen, Hao, et al.
Published: (2026)
VALLR: Visual ASR Language Model for Lip Reading
by: Thomas, Marshall, et al.
Published: (2025)
by: Thomas, Marshall, et al.
Published: (2025)
CNN-Based Classification of Persian Miniature Paintings from Five Renowned Schools
by: Shahi, Mojtaba, et al.
Published: (2024)
by: Shahi, Mojtaba, et al.
Published: (2024)
Identification via Retinal Vessels Combining LBP and HOG
by: Noori, Ali
Published: (2022)
by: Noori, Ali
Published: (2022)
Lips Are Lying: Spotting the Temporal Inconsistency between Audio and Visual in Lip-Syncing DeepFakes
by: Liu, Weifeng, et al.
Published: (2024)
by: Liu, Weifeng, et al.
Published: (2024)
SmartWilds: Multimodal Wildlife Monitoring Dataset
by: Kline, Jenna, et al.
Published: (2025)
by: Kline, Jenna, et al.
Published: (2025)
LipGen: Viseme-Guided Lip Video Generation for Enhancing Visual Speech Recognition
by: Hao, Bowen, et al.
Published: (2025)
by: Hao, Bowen, et al.
Published: (2025)
FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs
by: Zinonos, Andreas, et al.
Published: (2025)
by: Zinonos, Andreas, et al.
Published: (2025)
EchoXFlow: A Beamspace Echocardiography Dataset for Cardiac Motion, Flow, and Function
by: Stenhede, Elias, et al.
Published: (2026)
by: Stenhede, Elias, et al.
Published: (2026)
DynamicLip: Shape-Independent Continuous Authentication via Lip Articulator Dynamics
by: Chen, Huashan, et al.
Published: (2025)
by: Chen, Huashan, et al.
Published: (2025)
Caltech Aerial RGB-Thermal Dataset in the Wild
by: Lee, Connor, et al.
Published: (2024)
by: Lee, Connor, et al.
Published: (2024)
ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild
by: Farkhondeh, Arya, et al.
Published: (2024)
by: Farkhondeh, Arya, et al.
Published: (2024)
Multi-Modal Dataset Distillation in the Wild
by: Dang, Zhuohang, et al.
Published: (2025)
by: Dang, Zhuohang, et al.
Published: (2025)
TriLite: Efficient Weakly Supervised Object Localization with Universal Visual Features and Tri-Region Disentanglement
by: Sabaghi, Arian, et al.
Published: (2026)
by: Sabaghi, Arian, et al.
Published: (2026)
Biomarker-Based Pretraining for Chagas Disease Screening in Electrocardiograms
by: Stenhede, Elias, et al.
Published: (2026)
by: Stenhede, Elias, et al.
Published: (2026)
Exposing Lip-syncing Deepfakes from Mouth Inconsistencies
by: Datta, Soumyya Kanti, et al.
Published: (2024)
by: Datta, Soumyya Kanti, et al.
Published: (2024)
Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language
by: Yeo, Jeong Hun, et al.
Published: (2024)
by: Yeo, Jeong Hun, et al.
Published: (2024)
WhisperNetV2: SlowFast Siamese Network For Lip-Based Biometrics
by: Zakeri, Abdollah, et al.
Published: (2024)
by: Zakeri, Abdollah, et al.
Published: (2024)
Excavating in the Wild: The GOOSE-Ex Dataset for Semantic Segmentation
by: Hagmanns, Raphael, et al.
Published: (2024)
by: Hagmanns, Raphael, et al.
Published: (2024)
EV-Flying: an Event-based Dataset for In-The-Wild Recognition of Flying Objects
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
Enhancing Lip Reading with Multi-Scale Video and Multi-Encoder
by: Wang, He, et al.
Published: (2024)
by: Wang, He, et al.
Published: (2024)
Dataset Distillers Are Good Label Denoisers In the Wild
by: Cheng, Lechao, et al.
Published: (2024)
by: Cheng, Lechao, et al.
Published: (2024)
Evaluating Robustness of Vision-Language Models Under Noisy Conditions
by: Purushoth, et al.
Published: (2025)
by: Purushoth, et al.
Published: (2025)
Harmony4D: A Video Dataset for In-The-Wild Close Human Interactions
by: Khirodkar, Rawal, et al.
Published: (2024)
by: Khirodkar, Rawal, et al.
Published: (2024)
WildPPG: A Real-World PPG Dataset of Long Continuous Recordings
by: Meier, Manuel, et al.
Published: (2024)
by: Meier, Manuel, et al.
Published: (2024)
Detect, Classify, Act: Categorizing Industrial Anomalies with Multi-Modal Large Language Models
by: Mokhtar, Sassan, et al.
Published: (2025)
by: Mokhtar, Sassan, et al.
Published: (2025)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
by: Ranjbar, Hossein, et al.
Published: (2025)
by: Ranjbar, Hossein, et al.
Published: (2025)
OmniSync: Towards Universal Lip Synchronization via Diffusion Transformers
by: Peng, Ziqiao, et al.
Published: (2025)
by: Peng, Ziqiao, et al.
Published: (2025)
SayAnything: Audio-Driven Lip Synchronization with Conditional Video Diffusion
by: Ma, Junxian, et al.
Published: (2025)
by: Ma, Junxian, et al.
Published: (2025)
MA-LipNet: Multi-Dimensional Attention Networks for Robust Lipreading
by: Rossi, Matteo
Published: (2026)
by: Rossi, Matteo
Published: (2026)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023)
by: Ki, Taekyung, et al.
Published: (2023)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
by: Wang, Xu, et al.
Published: (2025)
by: Wang, Xu, et al.
Published: (2025)
360 in the Wild: Dataset for Depth Prediction and View Synthesis
by: Park, Kibaek, et al.
Published: (2024)
by: Park, Kibaek, et al.
Published: (2024)
FindingEmo: An Image Dataset for Emotion Recognition in the Wild
by: Mertens, Laurent, et al.
Published: (2024)
by: Mertens, Laurent, et al.
Published: (2024)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
by: Hong, Yan, et al.
Published: (2024)
by: Hong, Yan, et al.
Published: (2024)
Similar Items
-
Integrating Persian Lip Reading in Surena-V Humanoid Robot for Human-Robot Interaction
by: Abbasi, Ali Farshian, et al.
Published: (2025) -
Khayyam Offline Persian Handwriting Dataset
by: Jafarzadeh, Pourya, et al.
Published: (2024) -
A comprehensive Persian offline handwritten database for investigating the effects of heritability and family relationships on handwriting
by: Zohrevand, Abbas, et al.
Published: (2025) -
STAF: Sinusoidal Trainable Activation Functions for Implicit Neural Representation
by: Morsali, Alireza, et al.
Published: (2025) -
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
by: Hegde, Sindhu, et al.
Published: (2024)