A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | de la Fuente, Antón, Jurafsky, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UMA-Split: unimodal aggregation for both English and Mandarin non-autoregressive speech recognition
von: Fang, Ying, et al.
Veröffentlicht: (2025)
von: Fang, Ying, et al.
Veröffentlicht: (2025)
Tone recognition in low-resource languages of North-East India: peeling the layers of SSL-based speech models
von: Gogoi, Parismita, et al.
Veröffentlicht: (2025)
von: Gogoi, Parismita, et al.
Veröffentlicht: (2025)
Humans overrely on overconfident language models, across languages
von: Rathi, Neil, et al.
Veröffentlicht: (2025)
von: Rathi, Neil, et al.
Veröffentlicht: (2025)
Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models
von: Luo, Yiwei, et al.
Veröffentlicht: (2023)
von: Luo, Yiwei, et al.
Veröffentlicht: (2023)
Introducing MELI: the Mandarin-English Language Interview Corpus
von: Liu, Suyuan, et al.
Veröffentlicht: (2026)
von: Liu, Suyuan, et al.
Veröffentlicht: (2026)
An efficient text augmentation approach for contextualized Mandarin speech recognition
von: Zheng, Naijun, et al.
Veröffentlicht: (2024)
von: Zheng, Naijun, et al.
Veröffentlicht: (2024)
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin
von: Rao, Abhinav, et al.
Veröffentlicht: (2022)
von: Rao, Abhinav, et al.
Veröffentlicht: (2022)
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
von: Zhang, Christine, et al.
Veröffentlicht: (2026)
SumTablets: A Transliteration Dataset of Sumerian Tablets
von: Simmons, Cole, et al.
Veröffentlicht: (2026)
von: Simmons, Cole, et al.
Veröffentlicht: (2026)
Predicting positive transfer for improved low-resource speech recognition using acoustic pseudo-tokens
von: San, Nay, et al.
Veröffentlicht: (2024)
von: San, Nay, et al.
Veröffentlicht: (2024)
HumT DumT: Measuring and controlling human-like language in LLMs
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
Automated evaluation of LLMs for effective machine translation of Mandarin Chinese to English
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
Multilingual Stutter Event Detection for English, German, and Mandarin Speech
von: Haas, Felix, et al.
Veröffentlicht: (2026)
von: Haas, Felix, et al.
Veröffentlicht: (2026)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
CS3-Bench: Evaluating and Enhancing Speech-to-Speech LLMs for Mandarin-English Code-Switching
von: Liu, Heyang, et al.
Veröffentlicht: (2025)
von: Liu, Heyang, et al.
Veröffentlicht: (2025)
Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs
von: Quang, Trung Nguyen, et al.
Veröffentlicht: (2026)
von: Quang, Trung Nguyen, et al.
Veröffentlicht: (2026)
Form and meaning co-determine the realization of tone in Taiwan Mandarin spontaneous speech: the case of T2-T3 and T3-T3 tone sandhi
von: Lu, Yuxin, et al.
Veröffentlicht: (2024)
von: Lu, Yuxin, et al.
Veröffentlicht: (2024)
Advancing Speech Translation: A Corpus of Mandarin-English Conversational Telephone Speech
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
Data Checklist: On Unit-Testing Datasets with Usable Information
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
Learning the meanings of function words from grounded language using a visual question answering model
von: Portelance, Eva, et al.
Veröffentlicht: (2023)
von: Portelance, Eva, et al.
Veröffentlicht: (2023)
A Benchmark for Learning to Translate a New Language from One Grammar Book
von: Tanzer, Garrett, et al.
Veröffentlicht: (2023)
von: Tanzer, Garrett, et al.
Veröffentlicht: (2023)
Beyond Tokens: Concept-Level Training Objectives for LLMs
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
von: Shani, Chen, et al.
Veröffentlicht: (2026)
von: Shani, Chen, et al.
Veröffentlicht: (2026)
What can large language models do for sustainable food?
von: Thomas, Anna T., et al.
Veröffentlicht: (2025)
von: Thomas, Anna T., et al.
Veröffentlicht: (2025)
Dialect prejudice predicts AI decisions about people's character, employability, and criminality
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
Word-specific tonal realizations in Mandarin
von: Chuang, Yu-Ying, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Ying, et al.
Veröffentlicht: (2024)
Measuring Taiwanese Mandarin Language Understanding
von: Chen, Po-Heng, et al.
Veröffentlicht: (2024)
von: Chen, Po-Heng, et al.
Veröffentlicht: (2024)
Acquisition of Recursive Possessives and Recursive Locatives in Mandarin
von: Fu, Chenxi, et al.
Veröffentlicht: (2024)
von: Fu, Chenxi, et al.
Veröffentlicht: (2024)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
The mutual exclusivity bias of bilingual visually grounded speech models
von: Oneata, Dan, et al.
Veröffentlicht: (2025)
von: Oneata, Dan, et al.
Veröffentlicht: (2025)
LLMs for automatic annotation of Mandarin narrative transcripts
von: Zhao, Qingwen, et al.
Veröffentlicht: (2026)
von: Zhao, Qingwen, et al.
Veröffentlicht: (2026)
Generation Space Size: Understanding and Calibrating Open-Endedness of LLM Generations
von: Yu, Sunny, et al.
Veröffentlicht: (2025)
von: Yu, Sunny, et al.
Veröffentlicht: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
Translating speech with just images
von: Oneata, Dan, et al.
Veröffentlicht: (2024)
von: Oneata, Dan, et al.
Veröffentlicht: (2024)
Ara-Best-RQ: Multi Dialectal Arabic SSL
von: Elleuch, Haroun, et al.
Veröffentlicht: (2026)
von: Elleuch, Haroun, et al.
Veröffentlicht: (2026)
CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UMA-Split: unimodal aggregation for both English and Mandarin non-autoregressive speech recognition
von: Fang, Ying, et al.
Veröffentlicht: (2025) -
Tone recognition in low-resource languages of North-East India: peeling the layers of SSL-based speech models
von: Gogoi, Parismita, et al.
Veröffentlicht: (2025) -
Humans overrely on overconfident language models, across languages
von: Rathi, Neil, et al.
Veröffentlicht: (2025) -
Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models
von: Luo, Yiwei, et al.
Veröffentlicht: (2023) -
Introducing MELI: the Mandarin-English Language Interview Corpus
von: Liu, Suyuan, et al.
Veröffentlicht: (2026)