EasyEyes: Online hearing research using speakers calibrated by phones
Fuente:
arXiv
Salvato in:
| Autori principali: | Vican, Ivan, De Moraes, Hugo, Liao, Chongjun, Tsegaye, Nathnael H., O'Gara, William, Inamoto, Jasper, Pelli, Denis G. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the calibration of powerset speaker diarization models
di: Plaquet, Alexis, et al.
Pubblicazione: (2024)
di: Plaquet, Alexis, et al.
Pubblicazione: (2024)
Hierarchical speaker representation for target speaker extraction
di: He, Shulin, et al.
Pubblicazione: (2022)
di: He, Shulin, et al.
Pubblicazione: (2022)
Text adaptation for speaker verification with speaker-text factorized embeddings
di: Yang, Yexin, et al.
Pubblicazione: (2025)
di: Yang, Yexin, et al.
Pubblicazione: (2025)
Improving curriculum learning for target speaker extraction with synthetic speakers
di: Liu, Yun, et al.
Pubblicazione: (2024)
di: Liu, Yun, et al.
Pubblicazione: (2024)
Triage knowledge distillation for speaker verification
di: Kim, Ju-ho, et al.
Pubblicazione: (2026)
di: Kim, Ju-ho, et al.
Pubblicazione: (2026)
Privacy-oriented manipulation of speaker representations
di: Teixeira, Francisco, et al.
Pubblicazione: (2023)
di: Teixeira, Francisco, et al.
Pubblicazione: (2023)
Curriculum learning for self-supervised speaker verification
di: Heo, Hee-Soo, et al.
Pubblicazione: (2022)
di: Heo, Hee-Soo, et al.
Pubblicazione: (2022)
Investigation of perception inconsistency in speaker embedding for asynchronous voice anonymization
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Geodesic interpolation of frame-wise speaker embeddings for the diarization of meeting scenarios
di: Cord-Landwehr, Tobias, et al.
Pubblicazione: (2024)
di: Cord-Landwehr, Tobias, et al.
Pubblicazione: (2024)
On the relationship between speech and hearing
di: Umesh, Srinivasan, et al.
Pubblicazione: (2024)
di: Umesh, Srinivasan, et al.
Pubblicazione: (2024)
How phonemes contribute to deep speaker models?
di: Li, Pengqi, et al.
Pubblicazione: (2024)
di: Li, Pengqi, et al.
Pubblicazione: (2024)
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
di: Stafylakis, Themos, et al.
Pubblicazione: (2024)
di: Stafylakis, Themos, et al.
Pubblicazione: (2024)
Tandem spoofing-robust automatic speaker verification based on time-domain embeddings
di: Weizman, Avishai, et al.
Pubblicazione: (2024)
di: Weizman, Avishai, et al.
Pubblicazione: (2024)
DM-ASR: Diarization-aware Multi-speaker ASR with Large Language Models
di: Li, Li, et al.
Pubblicazione: (2026)
di: Li, Li, et al.
Pubblicazione: (2026)
The importance of spatial and spectral information in multiple speaker tracking
di: Beit-On, Hanan, et al.
Pubblicazione: (2024)
di: Beit-On, Hanan, et al.
Pubblicazione: (2024)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
di: Huang, Ziling, et al.
Pubblicazione: (2025)
di: Huang, Ziling, et al.
Pubblicazione: (2025)
Towards predicting binaural audio quality in listeners with normal and impaired hearing
di: Biberger, Thomas, et al.
Pubblicazione: (2025)
di: Biberger, Thomas, et al.
Pubblicazione: (2025)
Wanna hear your voice? A sample is all we need!
di: Pham, The Hieu, et al.
Pubblicazione: (2024)
di: Pham, The Hieu, et al.
Pubblicazione: (2024)
Target speaker anonymization in multi-speaker recordings
di: Tomashenko, Natalia, et al.
Pubblicazione: (2025)
di: Tomashenko, Natalia, et al.
Pubblicazione: (2025)
Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment
di: Boeddeker, Christoph, et al.
Pubblicazione: (2024)
di: Boeddeker, Christoph, et al.
Pubblicazione: (2024)
Efficient training strategies for natural sounding speech synthesis and speaker adaptation based on FastPitch
di: Răgman, Teodora, et al.
Pubblicazione: (2024)
di: Răgman, Teodora, et al.
Pubblicazione: (2024)
On the influence of language similarity in non-target speaker verification trials
di: Reuter, Paul M., et al.
Pubblicazione: (2025)
di: Reuter, Paul M., et al.
Pubblicazione: (2025)
Audio-visual child-adult speaker classification in dyadic interactions
di: Xu, Anfeng, et al.
Pubblicazione: (2023)
di: Xu, Anfeng, et al.
Pubblicazione: (2023)
Spoken language change detection inspired by speaker change detection
di: Mishra, Jagabandhu, et al.
Pubblicazione: (2023)
di: Mishra, Jagabandhu, et al.
Pubblicazione: (2023)
Unsupervised Online Continual Learning for Automatic Speech Recognition
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2024)
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2024)
Online speaker diarization of meetings guided by speech separation
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
End-to-end transfer learning for speaker-independent cross-language and cross-corpus speech emotion recognition
di: Tang, Duowei, et al.
Pubblicazione: (2023)
di: Tang, Duowei, et al.
Pubblicazione: (2023)
PixIT: Joint Training of Speaker Diarization and Speech Separation from Real-world Multi-speaker Recordings
di: Kalda, Joonas, et al.
Pubblicazione: (2024)
di: Kalda, Joonas, et al.
Pubblicazione: (2024)
Head-steered channel selection method for hearing aid applications using remote microphones
di: Sathyapriyan, Vasudha, et al.
Pubblicazione: (2025)
di: Sathyapriyan, Vasudha, et al.
Pubblicazione: (2025)
Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
di: Ma, Yi, et al.
Pubblicazione: (2024)
di: Ma, Yi, et al.
Pubblicazione: (2024)
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
di: Eisenberg, Aviad, et al.
Pubblicazione: (2025)
di: Eisenberg, Aviad, et al.
Pubblicazione: (2025)
Why disentanglement-based speaker anonymization systems fail at preserving emotions?
di: Gaznepoglu, Ünal Ege, et al.
Pubblicazione: (2025)
di: Gaznepoglu, Ünal Ege, et al.
Pubblicazione: (2025)
Speaker-agnostic Emotion Vector for Cross-speaker Emotion Intensity Control
di: Murata, Masato, et al.
Pubblicazione: (2025)
di: Murata, Masato, et al.
Pubblicazione: (2025)
HeightCeleb - an enrichment of VoxCeleb dataset with speaker height information
di: Kacprzak, Stanisław, et al.
Pubblicazione: (2024)
di: Kacprzak, Stanisław, et al.
Pubblicazione: (2024)
Improving fairness in speaker verification via Group-adapted Fusion Network
di: Shen, Hua, et al.
Pubblicazione: (2022)
di: Shen, Hua, et al.
Pubblicazione: (2022)
SPGISpeech 2.0: Transcribed multi-speaker financial audio for speaker-tagged transcription
di: Grossman, Raymond, et al.
Pubblicazione: (2025)
di: Grossman, Raymond, et al.
Pubblicazione: (2025)
X-CrossNet: A complex spectral mapping approach to target speaker extraction with cross attention speaker embedding fusion
di: Sun, Chang, et al.
Pubblicazione: (2024)
di: Sun, Chang, et al.
Pubblicazione: (2024)
A framework of text-dependent speaker verification for chinese numerical string corpus
di: Zheng, Litong, et al.
Pubblicazione: (2024)
di: Zheng, Litong, et al.
Pubblicazione: (2024)
EEND-M2F: Masked-attention mask transformers for speaker diarization
di: Härkönen, Marc, et al.
Pubblicazione: (2024)
di: Härkönen, Marc, et al.
Pubblicazione: (2024)
Some clues to build a sound analysis relevant to hearing
di: Millot, Laurent
Pubblicazione: (2024)
di: Millot, Laurent
Pubblicazione: (2024)
Documenti analoghi
-
On the calibration of powerset speaker diarization models
di: Plaquet, Alexis, et al.
Pubblicazione: (2024) -
Hierarchical speaker representation for target speaker extraction
di: He, Shulin, et al.
Pubblicazione: (2022) -
Text adaptation for speaker verification with speaker-text factorized embeddings
di: Yang, Yexin, et al.
Pubblicazione: (2025) -
Improving curriculum learning for target speaker extraction with synthetic speakers
di: Liu, Yun, et al.
Pubblicazione: (2024) -
Triage knowledge distillation for speaker verification
di: Kim, Ju-ho, et al.
Pubblicazione: (2026)