ART: The Alternating Reading Task Corpus for Speech Entrainment and Imitation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Zheng, de Jong, Dorina, Beňuš, Štefan, Nguyen, Noël, Feng, Ruitao, Sabo, Róbert, Fadiga, Luciano, D`Ausilio, Alessandro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Proficiency and F0 Entrainment: A Study of L2 English Imitation in Italian, French, and Slovak Speakers
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
The ART of Conversation: Measuring Phonetic Convergence and Deliberate Imitation in L2-Speech with a Siamese RNN
von: Yuan, Zheng, et al.
Veröffentlicht: (2023)
von: Yuan, Zheng, et al.
Veröffentlicht: (2023)
Measuring Entrainment in Spontaneous Code-switched Speech
von: Bhattacharya, Debasmita, et al.
Veröffentlicht: (2023)
von: Bhattacharya, Debasmita, et al.
Veröffentlicht: (2023)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
von: Zhao, Yan, et al.
Veröffentlicht: (2024)
von: Zhao, Yan, et al.
Veröffentlicht: (2024)
BENYO-S2ST-Corpus-1: A Bilingual English-to-Yoruba Direct Speech-to-Speech Translation Corpus
von: Adetiba, Emmanuel, et al.
Veröffentlicht: (2025)
von: Adetiba, Emmanuel, et al.
Veröffentlicht: (2025)
UrduSpeech: A 156-Hour Urdu Speech Corpus with 12-Dimension Paralinguistic Annotations
von: Haq, Attia Nafees ul, et al.
Veröffentlicht: (2026)
von: Haq, Attia Nafees ul, et al.
Veröffentlicht: (2026)
Emotional Speech Synthesis for Companion Robot to Imitate Professional Caregiver Speech
von: Homma, Takeshi, et al.
Veröffentlicht: (2021)
von: Homma, Takeshi, et al.
Veröffentlicht: (2021)
FeruzaSpeech: A 60 Hour Uzbek Read Speech Corpus with Punctuation, Casing, and Context
von: Povey, Anna, et al.
Veröffentlicht: (2024)
von: Povey, Anna, et al.
Veröffentlicht: (2024)
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving
von: Xie, Jingran, et al.
Veröffentlicht: (2025)
von: Xie, Jingran, et al.
Veröffentlicht: (2025)
WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark
von: Ma, Linhan, et al.
Veröffentlicht: (2024)
von: Ma, Linhan, et al.
Veröffentlicht: (2024)
SaSLaW: Dialogue Speech Corpus with Audio-visual Egocentric Information Toward Environment-adaptive Dialogue Speech Synthesis
von: Take, Osamu, et al.
Veröffentlicht: (2024)
von: Take, Osamu, et al.
Veröffentlicht: (2024)
spINAch: A Diachronic Corpus of French Broadcast Speech Controlled for Speakers' Age and Gender
von: Devauchelle, Simon, et al.
Veröffentlicht: (2026)
von: Devauchelle, Simon, et al.
Veröffentlicht: (2026)
Rethinking Cross-Corpus Speech Emotion Recognition Benchmarking: Are Paralinguistic Pre-Trained Representations Sufficient?
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
Mamba in Speech: Towards an Alternative to Self-Attention
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyu, et al.
Veröffentlicht: (2024)
JVNV: A Corpus of Japanese Emotional Speech with Verbal Content and Nonverbal Expressions
von: Xin, Detai, et al.
Veröffentlicht: (2023)
von: Xin, Detai, et al.
Veröffentlicht: (2023)
JIS: A Speech Corpus of Japanese Idol Speakers with Various Speaking Styles
von: Kondo, Yuto, et al.
Veröffentlicht: (2025)
von: Kondo, Yuto, et al.
Veröffentlicht: (2025)
ÌròyìnSpeech: A multi-purpose Yorùbá Speech Corpus
von: Ogunremi, Tolulope, et al.
Veröffentlicht: (2023)
von: Ogunremi, Tolulope, et al.
Veröffentlicht: (2023)
Enhancing Speaker-Independent Dysarthric Speech Severity Classification with DSSCNet and Cross-Corpus Adaptation
von: Roy, Arnab Kumar, et al.
Veröffentlicht: (2025)
von: Roy, Arnab Kumar, et al.
Veröffentlicht: (2025)
FLEURS-R: A Restored Multilingual Speech Corpus for Generation Tasks
von: Ma, Min, et al.
Veröffentlicht: (2024)
von: Ma, Min, et al.
Veröffentlicht: (2024)
Benchmarking Humans and Machines on Complex Multilingual Speech Understanding Tasks
von: Kankanala, Sai Samrat, et al.
Veröffentlicht: (2025)
von: Kankanala, Sai Samrat, et al.
Veröffentlicht: (2025)
Towards Cross-Task Suicide Risk Detection via Speech LLM
von: Li, Jialun, et al.
Veröffentlicht: (2025)
von: Li, Jialun, et al.
Veröffentlicht: (2025)
Clever Hans Effect Found in Automatic Detection of Alzheimer's Disease through Speech
von: Liu, Yin-Long, et al.
Veröffentlicht: (2024)
von: Liu, Yin-Long, et al.
Veröffentlicht: (2024)
Swedish Whispers; Leveraging a Massive Speech Corpus for Swedish Speech Recognition
von: Vesterbacka, Leonora, et al.
Veröffentlicht: (2025)
von: Vesterbacka, Leonora, et al.
Veröffentlicht: (2025)
Advancing Speech Translation: A Corpus of Mandarin-English Conversational Telephone Speech
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
EmoSpeech: A Corpus of Emotionally Rich and Contextually Detailed Speech Annotations
von: Bian, Weizhen, et al.
Veröffentlicht: (2024)
von: Bian, Weizhen, et al.
Veröffentlicht: (2024)
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2023)
GLOBE: A High-quality English Corpus with Global Accents for Zero-shot Speaker Adaptive Text-to-Speech
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
Privacy Disclosure of Similarity Rank in Speech and Language Processing
von: Bäckström, Tom, et al.
Veröffentlicht: (2025)
von: Bäckström, Tom, et al.
Veröffentlicht: (2025)
UniFlow: Unifying Speech Front-End Tasks via Continuous Generative Modeling
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
LipVoicer: Generating Speech from Silent Videos Guided by Lip Reading
von: Yemini, Yochai, et al.
Veröffentlicht: (2023)
von: Yemini, Yochai, et al.
Veröffentlicht: (2023)
Summary on The Multilingual Conversational Speech Language Model Challenge: Datasets, Tasks, Baselines, and Methods
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
von: Mu, Bingshen, et al.
Veröffentlicht: (2025)
Arabic ASR on the SADA Large-Scale Arabic Speech Corpus with Transformer-Based Models
von: Gerazov, Branislav, et al.
Veröffentlicht: (2025)
von: Gerazov, Branislav, et al.
Veröffentlicht: (2025)
Speech Corpus for Korean Children with Autism Spectrum Disorder: Towards Automatic Assessment Systems
von: Lee, Seonwoo, et al.
Veröffentlicht: (2024)
von: Lee, Seonwoo, et al.
Veröffentlicht: (2024)
SCOREQ: Speech Quality Assessment with Contrastive Regression
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
The MSP-Podcast Corpus
von: Busso, Carlos, et al.
Veröffentlicht: (2025)
von: Busso, Carlos, et al.
Veröffentlicht: (2025)
Audiobook-CC: Controllable Long-context Speech Generation for Multicast Audiobook
von: Liu, Min, et al.
Veröffentlicht: (2025)
von: Liu, Min, et al.
Veröffentlicht: (2025)
Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enhancement
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
von: Kühne, Nikolai Lund, et al.
Veröffentlicht: (2025)
Active Learning of Non-semantic Speech Tasks with Pretrained Models
von: Lee, Harlin, et al.
Veröffentlicht: (2022)
von: Lee, Harlin, et al.
Veröffentlicht: (2022)
Improving Query-by-Vocal Imitation with Contrastive Learning and Audio Pretraining
von: Greif, Jonathan, et al.
Veröffentlicht: (2024)
von: Greif, Jonathan, et al.
Veröffentlicht: (2024)
Exploring Speech Foundation Models for Speaker Diarization Across Lifespan
von: Xu, Anfeng, et al.
Veröffentlicht: (2026)
von: Xu, Anfeng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Language Proficiency and F0 Entrainment: A Study of L2 English Imitation in Italian, French, and Slovak Speakers
von: Yuan, Zheng, et al.
Veröffentlicht: (2024) -
The ART of Conversation: Measuring Phonetic Convergence and Deliberate Imitation in L2-Speech with a Siamese RNN
von: Yuan, Zheng, et al.
Veröffentlicht: (2023) -
Measuring Entrainment in Spontaneous Code-switched Speech
von: Bhattacharya, Debasmita, et al.
Veröffentlicht: (2023) -
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
von: Zhao, Yan, et al.
Veröffentlicht: (2024) -
BENYO-S2ST-Corpus-1: A Bilingual English-to-Yoruba Direct Speech-to-Speech Translation Corpus
von: Adetiba, Emmanuel, et al.
Veröffentlicht: (2025)