EmoTale: An Enacted Speech-emotion Dataset in Danish
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hjuler, Maja J., Skat-Rørdam, Harald V., Clemmensen, Line H., Das, Sneha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Local Interpretable Model-Agnostic Explanations for Speech Emotion Recognition with Distribution-Shift
von: Hjuler, Maja J., et al.
Veröffentlicht: (2025)
von: Hjuler, Maja J., et al.
Veröffentlicht: (2025)
BLSP-Emo: Towards Empathetic Large Speech-Language Models
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
Examining the Interplay Between Privacy and Fairness for Speech Processing: A Review and Perspective
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2024)
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2024)
Towards Emotionally Consistent Text-Based Speech Editing: Introducing EmoCorrector and The ECD-TSE Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2025)
von: Liu, Rui, et al.
Veröffentlicht: (2025)
ArabEmoNet: A Lightweight Hybrid 2D CNN-BiLSTM Model with Attention for Robust Arabic Speech Emotion Recognition
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
EmoFake: An Initial Dataset for Emotion Fake Audio Detection
von: Zhao, Yan, et al.
Veröffentlicht: (2022)
von: Zhao, Yan, et al.
Veröffentlicht: (2022)
Speech-MASSIVE: A Multilingual Speech Dataset for SLU and Beyond
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
EmoQ: Speech Emotion Recognition via Speech-Aware Q-Former and Large Language Model
von: Yang, Yiqing, et al.
Veröffentlicht: (2025)
von: Yang, Yiqing, et al.
Veröffentlicht: (2025)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
von: Zhang, Hezhao, et al.
Veröffentlicht: (2026)
von: Zhang, Hezhao, et al.
Veröffentlicht: (2026)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
von: Das, Sneha, et al.
Veröffentlicht: (2020)
von: Das, Sneha, et al.
Veröffentlicht: (2020)
YODAS: Youtube-Oriented Dataset for Audio and Speech
von: Li, Xinjian, et al.
Veröffentlicht: (2024)
von: Li, Xinjian, et al.
Veröffentlicht: (2024)
nEMO: Dataset of Emotional Speech in Polish
von: Christop, Iwona
Veröffentlicht: (2024)
von: Christop, Iwona
Veröffentlicht: (2024)
EmoShift: Lightweight Activation Steering for Enhanced Emotion-Aware Speech Synthesis
von: Zhou, Li, et al.
Veröffentlicht: (2026)
von: Zhou, Li, et al.
Veröffentlicht: (2026)
CASPER: A Large Scale Spontaneous Speech Dataset
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
von: Xiao, Cihan, et al.
Veröffentlicht: (2025)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
von: Abu, Turi, et al.
Veröffentlicht: (2025)
von: Abu, Turi, et al.
Veröffentlicht: (2025)
HebDB: a Weakly Supervised Dataset for Hebrew Speech Processing
von: Turetzky, Arnon, et al.
Veröffentlicht: (2024)
von: Turetzky, Arnon, et al.
Veröffentlicht: (2024)
CS-FLEURS: A Massively Multilingual and Code-Switched Speech Dataset
von: Yan, Brian, et al.
Veröffentlicht: (2025)
von: Yan, Brian, et al.
Veröffentlicht: (2025)
EmoFormer: A Text-Independent Speech Emotion Recognition using a Hybrid Transformer-CNN model
von: Hasan, Rashedul, et al.
Veröffentlicht: (2025)
von: Hasan, Rashedul, et al.
Veröffentlicht: (2025)
Multilingual Source Tracing of Speech Deepfakes: A First Benchmark
von: Xuan, Xi, et al.
Veröffentlicht: (2025)
von: Xuan, Xi, et al.
Veröffentlicht: (2025)
RoDia: A New Dataset for Romanian Dialect Identification from Speech
von: Rotaru, Codrut, et al.
Veröffentlicht: (2023)
von: Rotaru, Codrut, et al.
Veröffentlicht: (2023)
ChildGuard: A Specialized Dataset for Combatting Child-Targeted Hate Speech
von: Kashyap, Gautam Siddharth, et al.
Veröffentlicht: (2025)
von: Kashyap, Gautam Siddharth, et al.
Veröffentlicht: (2025)
Emilia: A Large-Scale, Extensive, Multilingual, and Diverse Dataset for Speech Generation
von: He, Haorui, et al.
Veröffentlicht: (2025)
von: He, Haorui, et al.
Veröffentlicht: (2025)
LearnerVoice: A Dataset of Non-Native English Learners' Spontaneous Speech
von: Kim, Haechan, et al.
Veröffentlicht: (2024)
von: Kim, Haechan, et al.
Veröffentlicht: (2024)
Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
AnimeScore: A Preference-Based Dataset and Framework for Evaluating Anime-Like Speech Style
von: Park, Joonyong, et al.
Veröffentlicht: (2026)
von: Park, Joonyong, et al.
Veröffentlicht: (2026)
StoryTTS: A Highly Expressive Text-to-Speech Dataset with Rich Textual Expressiveness Annotations
von: Liu, Sen, et al.
Veröffentlicht: (2024)
von: Liu, Sen, et al.
Veröffentlicht: (2024)
LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect
von: Naouara, Hedi, et al.
Veröffentlicht: (2025)
von: Naouara, Hedi, et al.
Veröffentlicht: (2025)
EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
von: Ma, Ziyang, et al.
Veröffentlicht: (2024)
ML-SUPERB 2.0: Benchmarking Multilingual Speech Models Across Modeling Constraints, Languages, and Datasets
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
Idiosyncratic Versus Normative Modeling of Atypical Speech Recognition: Dysarthric Case Studies
von: Raja, Vishnu, et al.
Veröffentlicht: (2025)
von: Raja, Vishnu, et al.
Veröffentlicht: (2025)
EmoTech: A Multi-modal Speech Emotion Recognition Using Multi-source Low-level Information with Hybrid Recurrent Network
von: Avro, Shamin Bin Habib, et al.
Veröffentlicht: (2025)
von: Avro, Shamin Bin Habib, et al.
Veröffentlicht: (2025)
Computational Narrative Understanding for Expressive Text-to-Speech
von: Michel, Gaspard, et al.
Veröffentlicht: (2025)
von: Michel, Gaspard, et al.
Veröffentlicht: (2025)
CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
EmoSpeech: A Corpus of Emotionally Rich and Contextually Detailed Speech Annotations
von: Bian, Weizhen, et al.
Veröffentlicht: (2024)
von: Bian, Weizhen, et al.
Veröffentlicht: (2024)
Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios
von: Subramanian, Aswin Shanmugam, et al.
Veröffentlicht: (2025)
von: Subramanian, Aswin Shanmugam, et al.
Veröffentlicht: (2025)
SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models
von: Zhang, Xin, et al.
Veröffentlicht: (2023)
von: Zhang, Xin, et al.
Veröffentlicht: (2023)
Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs
von: Futami, Hayato, et al.
Veröffentlicht: (2025)
von: Futami, Hayato, et al.
Veröffentlicht: (2025)
Continuous Speech Tokenizer in Text To Speech
von: Li, Yixing, et al.
Veröffentlicht: (2024)
von: Li, Yixing, et al.
Veröffentlicht: (2024)
SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models
von: Peri, Raghuveer, et al.
Veröffentlicht: (2024)
von: Peri, Raghuveer, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring Local Interpretable Model-Agnostic Explanations for Speech Emotion Recognition with Distribution-Shift
von: Hjuler, Maja J., et al.
Veröffentlicht: (2025) -
BLSP-Emo: Towards Empathetic Large Speech-Language Models
von: Wang, Chen, et al.
Veröffentlicht: (2024) -
Examining the Interplay Between Privacy and Fairness for Speech Processing: A Review and Perspective
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2024) -
Towards Emotionally Consistent Text-Based Speech Editing: Introducing EmoCorrector and The ECD-TSE Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2025) -
ArabEmoNet: A Lightweight Hybrid 2D CNN-BiLSTM Model with Attention for Robust Arabic Speech Emotion Recognition
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)