Asymmetric and trial-dependent modeling: the contribution of LIA to SdSV Challenge Task 2
Fuente:
arXiv
Salvato in:
| Autori principali: | Bousquet, Pierre-Michel, Rouvier, Mickael |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Zero-Shot End-To-End Spoken Question Answering In Medical Domain
di: Labrak, Yanis, et al.
Pubblicazione: (2024)
di: Labrak, Yanis, et al.
Pubblicazione: (2024)
MSP-Podcast SER Challenge 2024: L'antenne du Ventoux Multimodal Self-Supervised Learning for Speech Emotion Recognition
di: Duret, Jarod, et al.
Pubblicazione: (2024)
di: Duret, Jarod, et al.
Pubblicazione: (2024)
Probing the Information Encoded in Neural-based Acoustic Models of Automatic Speech Recognition Systems
di: Raymondaud, Quentin, et al.
Pubblicazione: (2024)
di: Raymondaud, Quentin, et al.
Pubblicazione: (2024)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
di: Tomashenko, Natalia, et al.
Pubblicazione: (2026)
di: Tomashenko, Natalia, et al.
Pubblicazione: (2026)
Text-dependent Speaker Verification (TdSV) Challenge 2024: Challenge Evaluation Plan
di: Hossein, Zeinali, et al.
Pubblicazione: (2024)
di: Hossein, Zeinali, et al.
Pubblicazione: (2024)
LastResort at SemEval-2024 Task 3: Exploring Multimodal Emotion Cause Pair Extraction as Sequence Labelling Task
di: Mathur, Suyash Vardhan, et al.
Pubblicazione: (2024)
di: Mathur, Suyash Vardhan, et al.
Pubblicazione: (2024)
Task-Agnostic Structured Pruning of Speech Representation Models
di: Wang, Haoyu, et al.
Pubblicazione: (2023)
di: Wang, Haoyu, et al.
Pubblicazione: (2023)
Communication-Efficient Personalized Federated Learning for Speech-to-Text Tasks
di: Du, Yichao, et al.
Pubblicazione: (2024)
di: Du, Yichao, et al.
Pubblicazione: (2024)
System Description for the Displace Speaker Diarization Challenge 2023
di: Aliyev, Ali
Pubblicazione: (2024)
di: Aliyev, Ali
Pubblicazione: (2024)
Retrieval-Augmented Speech Recognition Approach for Domain Challenges
di: Shen, Peng, et al.
Pubblicazione: (2025)
di: Shen, Peng, et al.
Pubblicazione: (2025)
Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms
di: Konan, Joseph, et al.
Pubblicazione: (2023)
di: Konan, Joseph, et al.
Pubblicazione: (2023)
Stepback: Enhanced Disentanglement for Voice Conversion via Multi-Task Learning
di: Yang, Qian, et al.
Pubblicazione: (2025)
di: Yang, Qian, et al.
Pubblicazione: (2025)
Computational Narrative Understanding for Expressive Text-to-Speech
di: Michel, Gaspard, et al.
Pubblicazione: (2025)
di: Michel, Gaspard, et al.
Pubblicazione: (2025)
The NTNU System at the S&I Challenge 2025 SLA Open Track
di: Lin, Hong-Yun, et al.
Pubblicazione: (2025)
di: Lin, Hong-Yun, et al.
Pubblicazione: (2025)
The FruitShell French synthesis system at the Blizzard 2023 Challenge
di: Qi, Xin, et al.
Pubblicazione: (2023)
di: Qi, Xin, et al.
Pubblicazione: (2023)
The THUEE System Description for the IARPA OpenASR21 Challenge
di: Zhao, Jing, et al.
Pubblicazione: (2022)
di: Zhao, Jing, et al.
Pubblicazione: (2022)
The SVASR System for Text-dependent Speaker Verification (TdSV) AAIC Challenge 2024
di: Molavi, Mohammadreza, et al.
Pubblicazione: (2024)
di: Molavi, Mohammadreza, et al.
Pubblicazione: (2024)
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
di: Žavoronkov, Aleksei, et al.
Pubblicazione: (2025)
di: Žavoronkov, Aleksei, et al.
Pubblicazione: (2025)
LASER: Learning by Aligning Self-supervised Representations of Speech for Improving Content-related Tasks
di: Meghanani, Amit, et al.
Pubblicazione: (2024)
di: Meghanani, Amit, et al.
Pubblicazione: (2024)
TokenVerse: Towards Unifying Speech and NLP Tasks via Transducer-based ASR
di: Kumar, Shashi, et al.
Pubblicazione: (2024)
di: Kumar, Shashi, et al.
Pubblicazione: (2024)
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
di: Arora, Siddhant, et al.
Pubblicazione: (2023)
di: Arora, Siddhant, et al.
Pubblicazione: (2023)
Echotune: A Modular Extractor Leveraging the Variable-Length Nature of Speech in ASR Tasks
di: Chen, Sizhou, et al.
Pubblicazione: (2023)
di: Chen, Sizhou, et al.
Pubblicazione: (2023)
Early Dementia Detection Using Multiple Spontaneous Speech Prompts: The PROCESS Challenge
di: Tao, Fuxiang, et al.
Pubblicazione: (2024)
di: Tao, Fuxiang, et al.
Pubblicazione: (2024)
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models
di: Wang, Dingdong, et al.
Pubblicazione: (2024)
di: Wang, Dingdong, et al.
Pubblicazione: (2024)
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2025)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2025)
LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect
di: Naouara, Hedi, et al.
Pubblicazione: (2025)
di: Naouara, Hedi, et al.
Pubblicazione: (2025)
Acoustic to Articulatory Inversion of Speech; Data Driven Approaches, Challenges, Applications, and Future Scope
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
GTSinger: A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks
di: Zhang, Yu, et al.
Pubblicazione: (2024)
di: Zhang, Yu, et al.
Pubblicazione: (2024)
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving
di: Xie, Jingran, et al.
Pubblicazione: (2025)
di: Xie, Jingran, et al.
Pubblicazione: (2025)
HarmoniFuse: A Component-Selective and Prompt-Adaptive Framework for Multi-Task Speech Language Modeling
di: Si, Yuke, et al.
Pubblicazione: (2025)
di: Si, Yuke, et al.
Pubblicazione: (2025)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
di: Kulkarni, Ajinkya, et al.
Pubblicazione: (2025)
di: Kulkarni, Ajinkya, et al.
Pubblicazione: (2025)
Findings of the 2023 ML-SUPERB Challenge: Pre-Training and Evaluation over More Languages and Beyond
di: Shi, Jiatong, et al.
Pubblicazione: (2023)
di: Shi, Jiatong, et al.
Pubblicazione: (2023)
Speech-Copilot: Leveraging Large Language Models for Speech Processing via Task Decomposition, Modularization, and Program Generation
di: Kuan, Chun-Yi, et al.
Pubblicazione: (2024)
di: Kuan, Chun-Yi, et al.
Pubblicazione: (2024)
Gammatonegram Representation for End-to-End Dysarthric Speech Processing Tasks: Speech Recognition, Speaker Identification, and Intelligibility Assessment
di: Farhadipour, Aref, et al.
Pubblicazione: (2023)
di: Farhadipour, Aref, et al.
Pubblicazione: (2023)
emg2speech: Synthesizing speech from electromyography using self-supervised speech models
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2025)
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2025)
Can Large Audio-Language Models Truly Hear? Tackling Hallucinations with Multi-Task Assessment and Stepwise Audio Reasoning
di: Kuan, Chun-Yi, et al.
Pubblicazione: (2024)
di: Kuan, Chun-Yi, et al.
Pubblicazione: (2024)
Recent Trends in Distant Conversational Speech Recognition: A Review of CHiME-7 and 8 DASR Challenges
di: Cornell, Samuele, et al.
Pubblicazione: (2025)
di: Cornell, Samuele, et al.
Pubblicazione: (2025)
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
di: Parcollet, Titouan, et al.
Pubblicazione: (2023)
di: Parcollet, Titouan, et al.
Pubblicazione: (2023)
Samsung Research China-Beijing at SemEval-2024 Task 3: A multi-stage framework for Emotion-Cause Pair Extraction in Conversations
di: Zhang, Shen, et al.
Pubblicazione: (2024)
di: Zhang, Shen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Zero-Shot End-To-End Spoken Question Answering In Medical Domain
di: Labrak, Yanis, et al.
Pubblicazione: (2024) -
MSP-Podcast SER Challenge 2024: L'antenne du Ventoux Multimodal Self-Supervised Learning for Speech Emotion Recognition
di: Duret, Jarod, et al.
Pubblicazione: (2024) -
Probing the Information Encoded in Neural-based Acoustic Models of Automatic Speech Recognition Systems
di: Raymondaud, Quentin, et al.
Pubblicazione: (2024) -
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
di: Tomashenko, Natalia, et al.
Pubblicazione: (2026) -
Text-dependent Speaker Verification (TdSV) Challenge 2024: Challenge Evaluation Plan
di: Hossein, Zeinali, et al.
Pubblicazione: (2024)