NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Rekimoto, Jun, Nishimura, Yu, Yang, Bojian
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914384849141760
author Rekimoto, Jun
Nishimura, Yu
Yang, Bojian
author_facet Rekimoto, Jun
Nishimura, Yu
Yang, Bojian
contents Silent and whispered speech offer promise for always-available voice interaction with AI, yet existing methods struggle to balance vocabulary size, wearability, silence, and noise robustness. We present NasoVoce, a nose-bridge-mounted interface that integrates a microphone and a vibration sensor. Positioned at the nasal pads of smart glasses, it unobtrusively captures both acoustic and vibration signals. The nasal bridge, close to the mouth, allows access to bone- and skin-conducted speech and enables reliable capture of low-volume utterances such as whispered speech. While the microphone captures high-quality audio, it is highly sensitive to environmental noise. Conversely, the vibration sensor is robust to noise but yields lower signal quality. By fusing these complementary inputs, NasoVoce generates high-quality speech robust against interference. Evaluation with Whisper Large-v2, PESQ, STOI, and MUSHRA ratings confirms improved recognition and quality. NasoVoce demonstrates the feasibility of a practical interface for always-available, continuous, and discreet AI voice conversations.
format Preprint
id arxiv_https___arxiv_org_abs_2603_10324
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction
Rekimoto, Jun
Nishimura, Yu
Yang, Bojian
Human-Computer Interaction
Artificial Intelligence
Machine Learning
Sound
H.5.2; H.5.5; I.2.1
Silent and whispered speech offer promise for always-available voice interaction with AI, yet existing methods struggle to balance vocabulary size, wearability, silence, and noise robustness. We present NasoVoce, a nose-bridge-mounted interface that integrates a microphone and a vibration sensor. Positioned at the nasal pads of smart glasses, it unobtrusively captures both acoustic and vibration signals. The nasal bridge, close to the mouth, allows access to bone- and skin-conducted speech and enables reliable capture of low-volume utterances such as whispered speech. While the microphone captures high-quality audio, it is highly sensitive to environmental noise. Conversely, the vibration sensor is robust to noise but yields lower signal quality. By fusing these complementary inputs, NasoVoce generates high-quality speech robust against interference. Evaluation with Whisper Large-v2, PESQ, STOI, and MUSHRA ratings confirms improved recognition and quality. NasoVoce demonstrates the feasibility of a practical interface for always-available, continuous, and discreet AI voice conversations.
title NasoVoce: A Nose-Mounted Low-Audibility Speech Interface for Always-Available Speech Interaction
topic Human-Computer Interaction
Artificial Intelligence
Machine Learning
Sound
H.5.2; H.5.5; I.2.1
url https://arxiv.org/abs/2603.10324