Enregistré dans:
| Auteurs principaux: | Sun, Anchen, Londono, Juan J, Elbaum, Batya, Estrada, Luis, Lazo, Roberto Jose, Vitale, Laura, Villasanti, Hugo Gonzalez, Fusaroli, Riccardo, Perry, Lynn K, Messinger, Daniel S |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2401.07342 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Who Said What WSW 2.0? Enhanced Automated Analysis of Preschool Classroom Speech
par: Sun, Anchen, et autres
Publié: (2025)
par: Sun, Anchen, et autres
Publié: (2025)
From Who Said What to Who They Are: Modular Training-free Identity-Aware LLM Refinement of Speaker Diarization
par: Chen, Yu-Wen, et autres
Publié: (2025)
par: Chen, Yu-Wen, et autres
Publié: (2025)
VoxSafeBench: Not Just What Is Said, but Who, How, and Where
par: Wang, Yuxiang, et autres
Publié: (2026)
par: Wang, Yuxiang, et autres
Publié: (2026)
Who is Speaking or Who is Depressed? A Controlled Study of Speaker Leakage in Speech-Based Depression Detection
par: Yeh, Hsiang-Chen, et autres
Publié: (2026)
par: Yeh, Hsiang-Chen, et autres
Publié: (2026)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
par: Wang, Pu, et autres
Publié: (2024)
par: Wang, Pu, et autres
Publié: (2024)
Multitask Learning with Capsule Networks for Speech-to-Intent Applications
par: Poncelet, Jakob, et autres
Publié: (2020)
par: Poncelet, Jakob, et autres
Publié: (2020)
Unsupervised Online Continual Learning for Automatic Speech Recognition
par: Eeckt, Steven Vander, et autres
Publié: (2024)
par: Eeckt, Steven Vander, et autres
Publié: (2024)
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
par: Jacobs, Christiaan, et autres
Publié: (2025)
par: Jacobs, Christiaan, et autres
Publié: (2025)
SNIPER Training: Single-Shot Sparse Training for Text-to-Speech
par: Lam, Perry, et autres
Publié: (2022)
par: Lam, Perry, et autres
Publié: (2022)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
par: Eeckt, Steven Vander, et autres
Publié: (2022)
par: Eeckt, Steven Vander, et autres
Publié: (2022)
Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis
par: Carbonneau, Marc-André, et autres
Publié: (2025)
par: Carbonneau, Marc-André, et autres
Publié: (2025)
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
par: Premananth, Gowtham, et autres
Publié: (2025)
par: Premananth, Gowtham, et autres
Publié: (2025)
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch
par: Poncelet, Jakob, et autres
Publié: (2021)
par: Poncelet, Jakob, et autres
Publié: (2021)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
par: Poncelet, Jakob, et autres
Publié: (2025)
par: Poncelet, Jakob, et autres
Publié: (2025)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
par: Eeckt, Steven Vander, et autres
Publié: (2023)
par: Eeckt, Steven Vander, et autres
Publié: (2023)
Beyond Manual Transcripts: The Potential of Automated Speech Recognition Errors in Improving Alzheimer's Disease Detection
par: Liu, Yin-Long, et autres
Publié: (2025)
par: Liu, Yin-Long, et autres
Publié: (2025)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
par: Poncelet, Jakob, et autres
Publié: (2023)
par: Poncelet, Jakob, et autres
Publié: (2023)
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
par: Duret, Jarod, et autres
Publié: (2024)
par: Duret, Jarod, et autres
Publié: (2024)
Efficient Extraction of Noise-Robust Discrete Units from Self-Supervised Speech Models
par: Poncelet, Jakob, et autres
Publié: (2024)
par: Poncelet, Jakob, et autres
Publié: (2024)
What Do Speech Foundation Models Not Learn About Speech?
par: Waheed, Abdul, et autres
Publié: (2024)
par: Waheed, Abdul, et autres
Publié: (2024)
Voice Conversion for Likability Control via Automated Rating of Speech Synthesis Corpora
par: Suda, Hitoshi, et autres
Publié: (2025)
par: Suda, Hitoshi, et autres
Publié: (2025)
TellWhisper: Tell Whisper Who Speaks When
par: Hu, Yifan, et autres
Publié: (2026)
par: Hu, Yifan, et autres
Publié: (2026)
EmoSLLM: Parameter-Efficient Adaptation of LLMs for Speech Emotion Recognition
par: Thimonier, Hugo, et autres
Publié: (2025)
par: Thimonier, Hugo, et autres
Publié: (2025)
RealClass: A Framework for Classroom Speech Simulation with Public Datasets and Game Engines
par: Attia, Ahmed Adel, et autres
Publié: (2025)
par: Attia, Ahmed Adel, et autres
Publié: (2025)
Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens
par: Zhao, Jinzheng, et autres
Publié: (2024)
par: Zhao, Jinzheng, et autres
Publié: (2024)
Do You Hear What I Mean? Quantifying the Instruction-Perception Gap in Instruction-Guided Expressive Text-To-Speech Systems
par: Lin, Yi-Cheng, et autres
Publié: (2025)
par: Lin, Yi-Cheng, et autres
Publié: (2025)
AS-Speech: Adaptive Style For Speech Synthesis
par: Li, Zhipeng, et autres
Publié: (2024)
par: Li, Zhipeng, et autres
Publié: (2024)
What do Speech Foundation Models Learn? Analysis and Applications
par: Pasad, Ankita
Publié: (2025)
par: Pasad, Ankita
Publié: (2025)
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition
par: de Groot, Dimme, et autres
Publié: (2026)
par: de Groot, Dimme, et autres
Publié: (2026)
Scalable Speech Enhancement with Dynamic Channel Pruning
par: Miccini, Riccardo, et autres
Publié: (2024)
par: Miccini, Riccardo, et autres
Publié: (2024)
Speech Quality-Based Localization of Low-Quality Speech and Text-to-Speech Synthesis Artefacts
par: Kuhlmann, Michael, et autres
Publié: (2026)
par: Kuhlmann, Michael, et autres
Publié: (2026)
Psychophysiology-aided Perceptually Fluent Speech Analysis of Children Who Stutter
par: Xiao, Yi, et autres
Publié: (2022)
par: Xiao, Yi, et autres
Publié: (2022)
Who Spoke What When? Evaluating Spoken Language Models for Conversational ASR with Semantic and Overlap-Aware Metrics
par: Tawara, Naohiro, et autres
Publié: (2026)
par: Tawara, Naohiro, et autres
Publié: (2026)
Influence of Clean Speech Characteristics on Speech Enhancement Performance
par: Hou, Mingchi, et autres
Publié: (2025)
par: Hou, Mingchi, et autres
Publié: (2025)
What do neural networks listen to? Exploring the crucial bands in Speech Enhancement using Sinc-convolution
par: Ho, Kuan-Hsun, et autres
Publié: (2024)
par: Ho, Kuan-Hsun, et autres
Publié: (2024)
Analyzing the Impact of Splicing Artifacts in Partially Fake Speech Signals
par: Negroni, Viola, et autres
Publié: (2024)
par: Negroni, Viola, et autres
Publié: (2024)
Expanding and Analyzing ODAQ -- the Open Dataset of Audio Quality
par: Dick, Sascha, et autres
Publié: (2025)
par: Dick, Sascha, et autres
Publié: (2025)
Adaptive Slimming for Scalable and Efficient Speech Enhancement
par: Miccini, Riccardo, et autres
Publié: (2025)
par: Miccini, Riccardo, et autres
Publié: (2025)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
par: Bhattacharjee, Susmita, et autres
Publié: (2025)
par: Bhattacharjee, Susmita, et autres
Publié: (2025)
Assessing the Impact of Noise and Speech Enhancement on the Intelligibility of Speech Codecs
par: Behringer, Lyonel, et autres
Publié: (2026)
par: Behringer, Lyonel, et autres
Publié: (2026)
Documents similaires
-
Who Said What WSW 2.0? Enhanced Automated Analysis of Preschool Classroom Speech
par: Sun, Anchen, et autres
Publié: (2025) -
From Who Said What to Who They Are: Modular Training-free Identity-Aware LLM Refinement of Speaker Diarization
par: Chen, Yu-Wen, et autres
Publié: (2025) -
VoxSafeBench: Not Just What Is Said, but Who, How, and Where
par: Wang, Yuxiang, et autres
Publié: (2026) -
Who is Speaking or Who is Depressed? A Controlled Study of Speaker Leakage in Speech-Based Depression Detection
par: Yeh, Hsiang-Chen, et autres
Publié: (2026) -
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
par: Wang, Pu, et autres
Publié: (2024)