Guardado en:
| Autores principales: | Asaad, Ihab, Jacquelin, Maxime, Perrotin, Olivier, Girin, Laurent, Hueber, Thomas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2405.20101 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025)
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
por: Peng, Junyi, et al.
Publicado: (2025)
por: Peng, Junyi, et al.
Publicado: (2025)
STaR: Distilling Speech Temporal Relation for Lightweight Speech Self-Supervised Learning Models
por: Jang, Kangwook, et al.
Publicado: (2023)
por: Jang, Kangwook, et al.
Publicado: (2023)
Interface Design for Self-Supervised Speech Models
por: Shih, Yi-Jen, et al.
Publicado: (2024)
por: Shih, Yi-Jen, et al.
Publicado: (2024)
Analytic Study of Text-Free Speech Synthesis for Raw Audio using a Self-Supervised Learning Model
por: Park, Joonyong, et al.
Publicado: (2024)
por: Park, Joonyong, et al.
Publicado: (2024)
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
por: Venkateswaran, Nitin, et al.
Publicado: (2025)
por: Venkateswaran, Nitin, et al.
Publicado: (2025)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
por: Wang, Yujin, et al.
Publicado: (2022)
por: Wang, Yujin, et al.
Publicado: (2022)
Fast Word Error Rate Estimation Using Self-Supervised Representations for Speech and Text
por: Park, Chanho, et al.
Publicado: (2023)
por: Park, Chanho, et al.
Publicado: (2023)
Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models
por: Phuong, Tuan Dat, et al.
Publicado: (2025)
por: Phuong, Tuan Dat, et al.
Publicado: (2025)
SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
por: Ashihara, Takanori, et al.
Publicado: (2023)
por: Ashihara, Takanori, et al.
Publicado: (2023)
Textless Acoustic Model with Self-Supervised Distillation for Noise-Robust Expressive Speech-to-Speech Translation
por: Hwang, Min-Jae, et al.
Publicado: (2024)
por: Hwang, Min-Jae, et al.
Publicado: (2024)
Layer-Wise Analysis of Self-Supervised Acoustic Word Embeddings: A Study on Speech Emotion Recognition
por: Saliba, Alexandra, et al.
Publicado: (2024)
por: Saliba, Alexandra, et al.
Publicado: (2024)
Benchmarking Children's ASR with Supervised and Self-supervised Speech Foundation Models
por: Fan, Ruchao, et al.
Publicado: (2024)
por: Fan, Ruchao, et al.
Publicado: (2024)
Do Discrete Self-Supervised Representations of Speech Capture Tone Distinctions?
por: Osakuade, Opeyemi, et al.
Publicado: (2024)
por: Osakuade, Opeyemi, et al.
Publicado: (2024)
BiRQ: Bi-Level Self-Labeling Random Quantization for Self-Supervised Speech Recognition
por: Jiang, Liuyuan, et al.
Publicado: (2025)
por: Jiang, Liuyuan, et al.
Publicado: (2025)
Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers
por: Lin, Tzu-Quan, et al.
Publicado: (2025)
por: Lin, Tzu-Quan, et al.
Publicado: (2025)
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
por: Hernandez, Abner, et al.
Publicado: (2026)
por: Hernandez, Abner, et al.
Publicado: (2026)
LASER: Learning by Aligning Self-supervised Representations of Speech for Improving Content-related Tasks
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
Position-invariant Fine-tuning of Speech Enhancement Models with Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2026)
por: Meghanani, Amit, et al.
Publicado: (2026)
Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models
por: Wang, Haoyu, et al.
Publicado: (2022)
por: Wang, Haoyu, et al.
Publicado: (2022)
What Do Self-Supervised Speech and Speaker Models Learn? New Findings From a Cross Model Layer-Wise Analysis
por: Ashihara, Takanori, et al.
Publicado: (2024)
por: Ashihara, Takanori, et al.
Publicado: (2024)
Leveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative Analysis
por: Xu, Tianyi, et al.
Publicado: (2025)
por: Xu, Tianyi, et al.
Publicado: (2025)
Self-Supervised Learning for Multi-Channel Neural Transducer
por: Kojima, Atsushi
Publicado: (2024)
por: Kojima, Atsushi
Publicado: (2024)
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
por: Lu, Yen-Ju, et al.
Publicado: (2024)
por: Lu, Yen-Ju, et al.
Publicado: (2024)
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
por: Shon, Suwon, et al.
Publicado: (2024)
por: Shon, Suwon, et al.
Publicado: (2024)
A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning
por: minjie, Xiang
Publicado: (2024)
por: minjie, Xiang
Publicado: (2024)
Closing the Modality Reasoning Gap for Speech Large Language Models
por: Wang, Chaoren, et al.
Publicado: (2026)
por: Wang, Chaoren, et al.
Publicado: (2026)
Hierarchical Self-Supervised Representation Learning for Depression Detection from Speech
por: Li, Yuxin, et al.
Publicado: (2025)
por: Li, Yuxin, et al.
Publicado: (2025)
High-Fidelity Simultaneous Speech-To-Speech Translation
por: Labiausse, Tom, et al.
Publicado: (2025)
por: Labiausse, Tom, et al.
Publicado: (2025)
Property Neurons in Self-Supervised Speech Transformers
por: Lin, Tzu-Quan, et al.
Publicado: (2024)
por: Lin, Tzu-Quan, et al.
Publicado: (2024)
Speech-MASSIVE: A Multilingual Speech Dataset for SLU and Beyond
por: Lee, Beomseok, et al.
Publicado: (2024)
por: Lee, Beomseok, et al.
Publicado: (2024)
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
Reduce, Reuse, Recycle: Is Perturbed Data better than Other Language augmentation for Low Resource Self-Supervised Speech Models
por: Ullah, Asad, et al.
Publicado: (2023)
por: Ullah, Asad, et al.
Publicado: (2023)
SKILL: Similarity-aware Knowledge distILLation for Speech Self-Supervised Learning
por: Zampierin, Luca, et al.
Publicado: (2024)
por: Zampierin, Luca, et al.
Publicado: (2024)
SpidR: Learning Fast and Stable Linguistic Units for Spoken Language Models Without Supervision
por: Poli, Maxime, et al.
Publicado: (2025)
por: Poli, Maxime, et al.
Publicado: (2025)
HebDB: a Weakly Supervised Dataset for Hebrew Speech Processing
por: Turetzky, Arnon, et al.
Publicado: (2024)
por: Turetzky, Arnon, et al.
Publicado: (2024)
LESS: Large Language Model Enhanced Semi-Supervised Learning for Speech Foundational Models Using in-the-wild Data
por: Ding, Wen, et al.
Publicado: (2025)
por: Ding, Wen, et al.
Publicado: (2025)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
por: Poli, Maxime, et al.
Publicado: (2026)
por: Poli, Maxime, et al.
Publicado: (2026)
Revisiting Self-supervised Learning of Speech Representation from a Mutual Information Perspective
por: Liu, Alexander H., et al.
Publicado: (2024)
por: Liu, Alexander H., et al.
Publicado: (2024)
Towards Early Prediction of Self-Supervised Speech Model Performance
por: Whetten, Ryan, et al.
Publicado: (2025)
por: Whetten, Ryan, et al.
Publicado: (2025)
Ejemplares similares
-
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
por: Satish, Shree Harsha Bokkahalli, et al.
Publicado: (2025) -
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
por: Peng, Junyi, et al.
Publicado: (2025) -
STaR: Distilling Speech Temporal Relation for Lightweight Speech Self-Supervised Learning Models
por: Jang, Kangwook, et al.
Publicado: (2023) -
Interface Design for Self-Supervised Speech Models
por: Shih, Yi-Jen, et al.
Publicado: (2024) -
Analytic Study of Text-Free Speech Synthesis for Raw Audio using a Self-Supervised Learning Model
por: Park, Joonyong, et al.
Publicado: (2024)