WiRD-Gest: Gesture Recognition In The Real World Using Range-Doppler Wi-Fi Sensing on COTS Hardware
Fuente:
arXiv
Saved in:
| Main Authors: | Sanson, Jessica, Shah, Rahul C., Zhu, Yazhou, Rosales, Rafael, Frascolla, Valerio |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LiveSense: A Real-Time Wi-Fi Sensing Platform for Range-Doppler on COTS Laptop
by: Sanson, Jessica, et al.
Published: (2026)
by: Sanson, Jessica, et al.
Published: (2026)
Human Presence Detection via Wi-Fi Range-Filtered Doppler Spectrum on Commodity Laptops
by: Sanson, Jessica, et al.
Published: (2026)
by: Sanson, Jessica, et al.
Published: (2026)
Extracting Range-Doppler Information of Moving Targets from Wi-Fi Channel State Information
by: Sanson, Jessica, et al.
Published: (2025)
by: Sanson, Jessica, et al.
Published: (2025)
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
by: Abreu, Wallace, et al.
Published: (2026)
by: Abreu, Wallace, et al.
Published: (2026)
Gesture-Aware Zero-Shot Speech Recognition for Patients with Language Disorders
by: Kim, Seungbae, et al.
Published: (2025)
by: Kim, Seungbae, et al.
Published: (2025)
Exploiting Foundation Models and Speech Enhancement for Parkinson's Disease Detection from Speech in Real-World Operative Conditions
by: La Quatra, Moreno, et al.
Published: (2024)
by: La Quatra, Moreno, et al.
Published: (2024)
DroFiT: A Lightweight Band-fused Frequency Attention Toward Real-time UAV Speech Enhancement
by: Lee, Jeongmin, et al.
Published: (2025)
by: Lee, Jeongmin, et al.
Published: (2025)
STSM-FiLM: A FiLM-Conditioned Neural Architecture for Time-Scale Modification of Speech
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
Beyond Lips: Integrating Gesture and Lip Cues for Robust Audio-visual Speaker Extraction
by: Pan, Zexu, et al.
Published: (2026)
by: Pan, Zexu, et al.
Published: (2026)
Gesture2Speech: How Far Can Hand Movements Shape Expressive Speech?
by: Kumar, Lokesh, et al.
Published: (2026)
by: Kumar, Lokesh, et al.
Published: (2026)
HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset
by: Langman, Ryan, et al.
Published: (2025)
by: Langman, Ryan, et al.
Published: (2025)
Scalable Frameworks for Real-World Audio-Visual Speech Recognition
by: Kim, Sungnyun
Published: (2025)
by: Kim, Sungnyun
Published: (2025)
ExpGest: Expressive Speaker Generation Using Diffusion Model and Hybrid Audio-Text Guidance
by: Cheng, Yongkang, et al.
Published: (2024)
by: Cheng, Yongkang, et al.
Published: (2024)
Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India
by: Bhogale, Kaushal, et al.
Published: (2026)
by: Bhogale, Kaushal, et al.
Published: (2026)
MMedFD: A Real-world Healthcare Benchmark for Multi-turn Full-Duplex Automatic Speech Recognition
by: Chen, Hongzhao, et al.
Published: (2025)
by: Chen, Hongzhao, et al.
Published: (2025)
Exploring Generative Error Correction for Dysarthric Speech Recognition
by: La Quatra, Moreno, et al.
Published: (2025)
by: La Quatra, Moreno, et al.
Published: (2025)
Robust Nasality Representation Learning for Cleft Palate-Related Velopharyngeal Dysfunction Screening in Real-World Settings
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations
by: Shi, Haohan, et al.
Published: (2025)
by: Shi, Haohan, et al.
Published: (2025)
HRTF-guided Binaural Target Speaker Extraction with Real-World Validation
by: Ellinson, Yoav, et al.
Published: (2026)
by: Ellinson, Yoav, et al.
Published: (2026)
Descriptor:: Extended-Length Audio Dataset for Synthetic Voice Detection and Speaker Recognition (ELAD-SVDSR)
by: Vijaykumar, Rahul, et al.
Published: (2025)
by: Vijaykumar, Rahul, et al.
Published: (2025)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
by: Jiang, Shaohan, et al.
Published: (2025)
by: Jiang, Shaohan, et al.
Published: (2025)
FruitsMusic: A Real-World Corpus of Japanese Idol-Group Songs
by: Suda, Hitoshi, et al.
Published: (2024)
by: Suda, Hitoshi, et al.
Published: (2024)
Meta-learning-based percussion transcription and $t\bar{a}la$ identification from low-resource audio
by: Kodag, Rahul Bapusaheb, et al.
Published: (2025)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2025)
Weakly Supervised Tabla Stroke Transcription via TI-SDRM: A Rhythm-Aware Lattice Rescoring Framework
by: Kodag, Rahul Bapusaheb, et al.
Published: (2026)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2026)
HiFi-Glot: High-Fidelity Neural Formant Synthesis with Differentiable Resonant Filters
by: Gu, Yicheng, et al.
Published: (2024)
by: Gu, Yicheng, et al.
Published: (2024)
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
by: Hou, Yuanbo, et al.
Published: (2024)
by: Hou, Yuanbo, et al.
Published: (2024)
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
by: Watcharasupat, Karn N., et al.
Published: (2025)
by: Watcharasupat, Karn N., et al.
Published: (2025)
A Multiclass Acoustic Dataset and Interactive Tool for Analyzing Drone Signatures in Real-World Environments
by: Wang, Mia Y., et al.
Published: (2025)
by: Wang, Mia Y., et al.
Published: (2025)
Chunkwise Aligners for Streaming Speech Recognition
by: Teo, Wen Shen, et al.
Published: (2026)
by: Teo, Wen Shen, et al.
Published: (2026)
Reshape Dimensions Network for Speaker Recognition
by: Yakovlev, Ivan, et al.
Published: (2024)
by: Yakovlev, Ivan, et al.
Published: (2024)
FastTalker: Jointly Generating Speech and Conversational Gestures from Text
by: Guo, Zixin, et al.
Published: (2024)
by: Guo, Zixin, et al.
Published: (2024)
Reducing the Gap Between Pretrained Speech Enhancement and Recognition Models Using a Real Speech-Trained Bridging Module
by: Cui, Zhongjian, et al.
Published: (2025)
by: Cui, Zhongjian, et al.
Published: (2025)
$T\bar{a}laGen:$ A System for Automatic $T\bar{a}la$ Identification and Generation
by: Kodag, Rahul Bapusaheb, et al.
Published: (2024)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2024)
Identifying and Calibrating Overconfidence in Noisy Speech Recognition
by: Huo, Mingyue, et al.
Published: (2025)
by: Huo, Mingyue, et al.
Published: (2025)
MOVER: Combining Multiple Meeting Recognition Systems
by: Kamo, Naoyuki, et al.
Published: (2025)
by: Kamo, Naoyuki, et al.
Published: (2025)
Group Relative Policy Optimization for Speech Recognition
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2025)
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2025)
Cross-Corpus Validation of Speech Emotion Recognition in Urdu using Domain-Knowledge Acoustic Features
by: Talpur, Unzela, et al.
Published: (2025)
by: Talpur, Unzela, et al.
Published: (2025)
Accelerated Interactive Auralization of Highly Reverberant Spaces using Graphics Hardware
by: Rosseel, Hannes, et al.
Published: (2025)
by: Rosseel, Hannes, et al.
Published: (2025)
Interpreting the Role of Visemes in Audio-Visual Speech Recognition
by: Papadopoulos, Aristeidis, et al.
Published: (2025)
by: Papadopoulos, Aristeidis, et al.
Published: (2025)
Voice Conversion Augmentation for Speaker Recognition on Defective Datasets
by: Tao, Ruijie, et al.
Published: (2024)
by: Tao, Ruijie, et al.
Published: (2024)
Similar Items
-
LiveSense: A Real-Time Wi-Fi Sensing Platform for Range-Doppler on COTS Laptop
by: Sanson, Jessica, et al.
Published: (2026) -
Human Presence Detection via Wi-Fi Range-Filtered Doppler Spectrum on Commodity Laptops
by: Sanson, Jessica, et al.
Published: (2026) -
Extracting Range-Doppler Information of Moving Targets from Wi-Fi Channel State Information
by: Sanson, Jessica, et al.
Published: (2025) -
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
by: Abreu, Wallace, et al.
Published: (2026) -
Gesture-Aware Zero-Shot Speech Recognition for Patients with Language Disorders
by: Kim, Seungbae, et al.
Published: (2025)