Switchboard-Affect: Emotion Perception Labels from Conversational Speech
Fuente:
arXiv
Guardado en:
| Autores principales: | Romana, Amrit, Narain, Jaya, Tran, Tien Dung, Davis, Andrea, Fong, Jason, Rasipuram, Ramya, Mitra, Vikramjit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Modeling speech emotion with label variance and analyzing performance across speakers and unseen acoustic conditions
por: Mitra, Vikramjit, et al.
Publicado: (2025)
por: Mitra, Vikramjit, et al.
Publicado: (2025)
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
por: Narain, Jaya, et al.
Publicado: (2025)
por: Narain, Jaya, et al.
Publicado: (2025)
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
por: Nie, Jingping, et al.
Publicado: (2025)
por: Nie, Jingping, et al.
Publicado: (2025)
LiveSpeech: Low-Latency Zero-shot Text-to-Speech via Autoregressive Modeling of Audio Discrete Codes
por: Dang, Trung, et al.
Publicado: (2024)
por: Dang, Trung, et al.
Publicado: (2024)
Zero-Shot Text-to-Speech from Continuous Text Streams
por: Dang, Trung, et al.
Publicado: (2024)
por: Dang, Trung, et al.
Publicado: (2024)
ChunkFormer: Masked Chunking Conformer For Long-Form Speech Transcription
por: Le, Khanh, et al.
Publicado: (2025)
por: Le, Khanh, et al.
Publicado: (2025)
SegAug: CTC-Aligned Segmented Augmentation For Robust RNN-Transducer Based Speech Recognition
por: Le, Khanh, et al.
Publicado: (2025)
por: Le, Khanh, et al.
Publicado: (2025)
Model-driven Heart Rate Estimation and Heart Murmur Detection based on Phonocardiogram
por: Nie, Jingping, et al.
Publicado: (2024)
por: Nie, Jingping, et al.
Publicado: (2024)
VC-ENHANCE: Speech Restoration with Integrated Noise Suppression and Voice Conversion
por: Byun, Kyungguen, et al.
Publicado: (2024)
por: Byun, Kyungguen, et al.
Publicado: (2024)
MAGE: A Coarse-to-Fine Speech Enhancer with Masked Generative Model
por: Pham, The Hieu, et al.
Publicado: (2025)
por: Pham, The Hieu, et al.
Publicado: (2025)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
por: Qi, Tianhua, et al.
Publicado: (2026)
por: Qi, Tianhua, et al.
Publicado: (2026)
Voice-ENHANCE: Speech Restoration using a Diffusion-based Voice Conversion Framework
por: Byun, Kyungguen, et al.
Publicado: (2025)
por: Byun, Kyungguen, et al.
Publicado: (2025)
PARROT: Synergizing Mamba and Attention-based SSL Pre-Trained Models via Parallel Branch Hadamard Optimal Transport for Speech Emotion Recognition
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
por: Phukan, Orchid Chetia, et al.
Publicado: (2025)
Emotion Neural Transducer for Fine-Grained Speech Emotion Recognition
por: Shen, Siyuan, et al.
Publicado: (2024)
por: Shen, Siyuan, et al.
Publicado: (2024)
Textless and Non-Parallel Speech-to-Speech Emotion Style Transfer
por: Dutta, Soumya, et al.
Publicado: (2025)
por: Dutta, Soumya, et al.
Publicado: (2025)
Speech Emotion Recognition with ASR Integration
por: Li, Yuanchao
Publicado: (2026)
por: Li, Yuanchao
Publicado: (2026)
Hierarchical Control of Emotion Rendering in Speech Synthesis
por: Inoue, Sho, et al.
Publicado: (2024)
por: Inoue, Sho, et al.
Publicado: (2024)
A Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion
por: Geng, Haopeng, et al.
Publicado: (2025)
por: Geng, Haopeng, et al.
Publicado: (2025)
ED-TTS: Multi-Scale Emotion Modeling using Cross-Domain Emotion Diarization for Emotional Speech Synthesis
por: Tang, Haobin, et al.
Publicado: (2024)
por: Tang, Haobin, et al.
Publicado: (2024)
MIKU-PAL: An Automated and Standardized Multi-Modal Method for Speech Paralinguistic and Affect Labeling
por: Cheng, Yifan, et al.
Publicado: (2025)
por: Cheng, Yifan, et al.
Publicado: (2025)
Hierarchical Emotion Prediction and Control in Text-to-Speech Synthesis
por: Inoue, Sho, et al.
Publicado: (2024)
por: Inoue, Sho, et al.
Publicado: (2024)
Fine-Grained Quantitative Emotion Editing for Speech Generation
por: Inoue, Sho, et al.
Publicado: (2024)
por: Inoue, Sho, et al.
Publicado: (2024)
EMO-SUPERB: An In-depth Look at Speech Emotion Recognition
por: Wu, Haibin, et al.
Publicado: (2024)
por: Wu, Haibin, et al.
Publicado: (2024)
Dataset-Distillation Generative Model for Speech Emotion Recognition
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2024)
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2024)
THAI Speech Emotion Recognition (THAI-SER) corpus
por: Wongpithayadisai, Jilamika, et al.
Publicado: (2025)
por: Wongpithayadisai, Jilamika, et al.
Publicado: (2025)
Iterative Prototype Refinement for Ambiguous Speech Emotion Recognition
por: Sun, Haoqin, et al.
Publicado: (2024)
por: Sun, Haoqin, et al.
Publicado: (2024)
Prosody Labeling with Phoneme-BERT and Speech Foundation Models
por: Koriyama, Tomoki
Publicado: (2025)
por: Koriyama, Tomoki
Publicado: (2025)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
por: Zhao, Yan, et al.
Publicado: (2024)
por: Zhao, Yan, et al.
Publicado: (2024)
LLM supervised Pre-training for Multimodal Emotion Recognition in Conversations
por: Dutta, Soumya, et al.
Publicado: (2025)
por: Dutta, Soumya, et al.
Publicado: (2025)
Semantic-Emotional Resonance Embedding: A Semi-Supervised Paradigm for Cross-Lingual Speech Emotion Recognition
por: Zhao, Ya, et al.
Publicado: (2026)
por: Zhao, Ya, et al.
Publicado: (2026)
PCQ: Emotion Recognition in Speech via Progressive Channel Querying
por: Wang, Xincheng, et al.
Publicado: (2024)
por: Wang, Xincheng, et al.
Publicado: (2024)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023)
por: Derington, Anna, et al.
Publicado: (2023)
Adaptive Speech Emotion Representation Learning Based On Dynamic Graph
por: Gao, Yingxue, et al.
Publicado: (2024)
por: Gao, Yingxue, et al.
Publicado: (2024)
EME-TTS: Unlocking the Emphasis and Emotion Link in Speech Synthesis
por: Li, Haoxun, et al.
Publicado: (2025)
por: Li, Haoxun, et al.
Publicado: (2025)
SELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios
por: Bukhari, Hazim, et al.
Publicado: (2024)
por: Bukhari, Hazim, et al.
Publicado: (2024)
Mitigating Subgroup Disparities in Multi-Label Speech Emotion Recognition: A Pseudo-Labeling and Unsupervised Learning Approach
por: Lin, Yi-Cheng, et al.
Publicado: (2025)
por: Lin, Yi-Cheng, et al.
Publicado: (2025)
EmoAttack: Utilizing Emotional Voice Conversion for Speech Backdoor Attacks on Deep Speech Classification Models
por: Yao, Wenhan, et al.
Publicado: (2024)
por: Yao, Wenhan, et al.
Publicado: (2024)
EmoQ: Speech Emotion Recognition via Speech-Aware Q-Former and Large Language Model
por: Yang, Yiqing, et al.
Publicado: (2025)
por: Yang, Yiqing, et al.
Publicado: (2025)
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
por: Cha, Jun-Hyeok, et al.
Publicado: (2025)
por: Cha, Jun-Hyeok, et al.
Publicado: (2025)
MSP-Conversation: A Corpus for Naturalistic, Time-Continuous Emotion Recognition
por: Martinez-Lucas, Luz, et al.
Publicado: (2026)
por: Martinez-Lucas, Luz, et al.
Publicado: (2026)
Ejemplares similares
-
Modeling speech emotion with label variance and analyzing performance across speakers and unseen acoustic conditions
por: Mitra, Vikramjit, et al.
Publicado: (2025) -
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
por: Narain, Jaya, et al.
Publicado: (2025) -
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
por: Nie, Jingping, et al.
Publicado: (2025) -
LiveSpeech: Low-Latency Zero-shot Text-to-Speech via Autoregressive Modeling of Audio Discrete Codes
por: Dang, Trung, et al.
Publicado: (2024) -
Zero-Shot Text-to-Speech from Continuous Text Streams
por: Dang, Trung, et al.
Publicado: (2024)