Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Li, Yinan, Seifi, Hasti
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866908788224688128
author Li, Yinan
Seifi, Hasti
author_facet Li, Yinan
Seifi, Hasti
contents Environmental sounds like footsteps, keyboard typing, or dog barking carry rich information and emotional context, making them valuable for designing haptics in user applications. Existing audio-to-vibration methods, however, rely on signal-processing rules tuned for music or games and often fail to generalize across diverse sounds. To address this, we first investigated user perception of four existing audio-to-haptic algorithms, then created a data-driven model for environmental sounds. In Study 1, 34 participants rated vibrations generated by the four algorithms for 1,000 sounds, revealing no consistent algorithm preferences. Using this dataset, we trained Sound2Hap, a CNN-based autoencoder, to generate perceptually meaningful vibrations from diverse sounds with low latency. In Study 2, 15 participants rated its output higher than signal-processing baselines on both audio-vibration match and Haptic Experience Index (HXI), finding it more harmonious with diverse sounds. This work demonstrates a perceptually validated approach to audio-haptic translation, broadening the reach of sound-driven haptics.
format Preprint
id arxiv_https___arxiv_org_abs_2601_12245
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
Li, Yinan
Seifi, Hasti
Human-Computer Interaction
Sound
Audio and Speech Processing
Environmental sounds like footsteps, keyboard typing, or dog barking carry rich information and emotional context, making them valuable for designing haptics in user applications. Existing audio-to-vibration methods, however, rely on signal-processing rules tuned for music or games and often fail to generalize across diverse sounds. To address this, we first investigated user perception of four existing audio-to-haptic algorithms, then created a data-driven model for environmental sounds. In Study 1, 34 participants rated vibrations generated by the four algorithms for 1,000 sounds, revealing no consistent algorithm preferences. Using this dataset, we trained Sound2Hap, a CNN-based autoencoder, to generate perceptually meaningful vibrations from diverse sounds with low latency. In Study 2, 15 participants rated its output higher than signal-processing baselines on both audio-vibration match and Haptic Experience Index (HXI), finding it more harmonious with diverse sounds. This work demonstrates a perceptually validated approach to audio-haptic translation, broadening the reach of sound-driven haptics.
title Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
topic Human-Computer Interaction
Sound
Audio and Speech Processing
url https://arxiv.org/abs/2601.12245