NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | , , |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
| _version_ | 1866914384849141760 |
|---|---|
| author | Rekimoto, Jun Nishimura, Yu Yang, Bojian |
| author_facet | Rekimoto, Jun Nishimura, Yu Yang, Bojian |
| contents | Silent and whispered speech offer promise for always-available voice interaction with AI, yet existing methods struggle to balance vocabulary size, wearability, silence, and noise robustness. We present NasoVoce, a nose-bridge-mounted interface that integrates a microphone and a vibration sensor. Positioned at the nasal pads of smart glasses, it unobtrusively captures both acoustic and vibration signals. The nasal bridge, close to the mouth, allows access to bone- and skin-conducted speech and enables reliable capture of low-volume utterances such as whispered speech. While the microphone captures high-quality audio, it is highly sensitive to environmental noise. Conversely, the vibration sensor is robust to noise but yields lower signal quality. By fusing these complementary inputs, NasoVoce generates high-quality speech robust against interference. Evaluation with Whisper Large-v2, PESQ, STOI, and MUSHRA ratings confirms improved recognition and quality. NasoVoce demonstrates the feasibility of a practical interface for always-available, continuous, and discreet AI voice conversations. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2603_10324 |
| institution | arXiv |
| publishDate | 2026 |
| record_format | arxiv |
| spellingShingle | NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction Rekimoto, Jun Nishimura, Yu Yang, Bojian Human-Computer Interaction Artificial Intelligence Machine Learning Sound H.5.2; H.5.5; I.2.1 Silent and whispered speech offer promise for always-available voice interaction with AI, yet existing methods struggle to balance vocabulary size, wearability, silence, and noise robustness. We present NasoVoce, a nose-bridge-mounted interface that integrates a microphone and a vibration sensor. Positioned at the nasal pads of smart glasses, it unobtrusively captures both acoustic and vibration signals. The nasal bridge, close to the mouth, allows access to bone- and skin-conducted speech and enables reliable capture of low-volume utterances such as whispered speech. While the microphone captures high-quality audio, it is highly sensitive to environmental noise. Conversely, the vibration sensor is robust to noise but yields lower signal quality. By fusing these complementary inputs, NasoVoce generates high-quality speech robust against interference. Evaluation with Whisper Large-v2, PESQ, STOI, and MUSHRA ratings confirms improved recognition and quality. NasoVoce demonstrates the feasibility of a practical interface for always-available, continuous, and discreet AI voice conversations. |
| title | NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction |
| topic | Human-Computer Interaction Artificial Intelligence Machine Learning Sound H.5.2; H.5.5; I.2.1 |
| url | https://arxiv.org/abs/2603.10324 |