BERSting at the Screams: A Benchmark for Distanced, Emotional and Shouted Speech Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Tuttösí, Paige, Dhillon, Mantaj, Sang, Luna, Eastwood, Shane, Bhatia, Poorvi, Dinh, Quang Minh, Kapoor, Avni, Jin, Yewon, Lim, Angelica |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Good Things Come in Trees: Emotion and Context Aware Behaviour Trees for Ethical Robotic Decision-Making
by: Tuttösí, Paige, et al.
Published: (2024)
by: Tuttösí, Paige, et al.
Published: (2024)
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025)
by: Tuttösí, Paige
Published: (2025)
Read the Room: Adapting a Robot's Voice to Ambient and Social Contexts
by: Tuttosi, Paige, et al.
Published: (2022)
by: Tuttosi, Paige, et al.
Published: (2022)
Covertly improving intelligibility with data-driven adaptations of speech timing
by: Tuttösí, Paige, et al.
Published: (2026)
by: Tuttösí, Paige, et al.
Published: (2026)
You Sound a Little Tense: L2 Tailored Clear TTS Using Durational Vowel Properties
by: Tuttösí, Paige, et al.
Published: (2025)
by: Tuttösí, Paige, et al.
Published: (2025)
Vectorlike lepton imprints at lepton $g-2$ measurements and $e^+e^-$ colliders
by: Dinh, Sang Quang, et al.
Published: (2025)
by: Dinh, Sang Quang, et al.
Published: (2025)
Salsa as a Nonverbal Embodied Language -- The CoMPAS3D Dataset and Benchmarks
by: Burkanova, Bermet, et al.
Published: (2025)
by: Burkanova, Bermet, et al.
Published: (2025)
EmojiVoice: Towards long-term controllable expressivity in robot speech
by: Tuttösí, Paige, et al.
Published: (2025)
by: Tuttösí, Paige, et al.
Published: (2025)
Textless and Non-Parallel Speech-to-Speech Emotion Style Transfer
by: Dutta, Soumya, et al.
Published: (2025)
by: Dutta, Soumya, et al.
Published: (2025)
RISC: A Corpus for Shout Type Classification and Shout Intensity Prediction
by: Fukumori, Takahiro, et al.
Published: (2023)
by: Fukumori, Takahiro, et al.
Published: (2023)
"I Just Need GPT to Refine My Prompts": Rethinking Onboarding and Help-Seeking with Generative 3D Modeling Tools
by: Gautam, Kanak, et al.
Published: (2026)
by: Gautam, Kanak, et al.
Published: (2026)
South End Shout
by: House, Roger
Published: (2023)
by: House, Roger
Published: (2023)
Shout Dakota's badlands
Published: (1981)
Published: (1981)
Mmm whatcha say? Uncovering distal and proximal context effects in first and second-language word perception using psychophysical reverse correlation
by: Tuttösí, Paige, et al.
Published: (2024)
by: Tuttösí, Paige, et al.
Published: (2024)
Emotional Vietnamese Speech-Based Depression Diagnosis Using Dynamic Attention Mechanism
by: D., Quang-Anh N., et al.
Published: (2024)
by: D., Quang-Anh N., et al.
Published: (2024)
Tracing Everyday AI Literacy Discussions at Scale: How Online Creative Communities Make Sense of Generative AI
by: Liu, Haidan, et al.
Published: (2026)
by: Liu, Haidan, et al.
Published: (2026)
Past, Present, and Future: A Survey of The Evolution of Affective Robotics For Well-being
by: Spitale, Micol, et al.
Published: (2024)
by: Spitale, Micol, et al.
Published: (2024)
Expected statistical uncertainties at future $e^+e^-$ colliders
by: Tran, Hieu Minh, et al.
Published: (2025)
by: Tran, Hieu Minh, et al.
Published: (2025)
Speech-based Multimodel Pipeline for Vietnamese Services Quality Assessment
by: D., Quang-Anh N., et al.
Published: (2024)
by: D., Quang-Anh N., et al.
Published: (2024)
Britain. Screaming at the umpire
Published: (1995)
Published: (1995)
Gender Bias in Emotion Recognition by Large Language Models
by: Herbert, Maureen, et al.
Published: (2025)
by: Herbert, Maureen, et al.
Published: (2025)
Latent Equivariant Operators for Robust Object Recognition: Promises and Challenges
by: Dinh, Minh, et al.
Published: (2026)
by: Dinh, Minh, et al.
Published: (2026)
IgnitionInnovators at "Discharge Me!": Chain-of-Thought Instruction Finetuning Large Language Models for Discharge Summaries
by: Tang, An Quang, et al.
Published: (2024)
by: Tang, An Quang, et al.
Published: (2024)
PIXHELL: When Pixels Learn to Scream
by: Guri, Mordechai
Published: (2025)
by: Guri, Mordechai
Published: (2025)
Contextual Emotion Recognition using Large Vision Language Models
by: Etesam, Yasaman, et al.
Published: (2024)
by: Etesam, Yasaman, et al.
Published: (2024)
Time Series Based Network Intrusion Detection using MTF-Aided Transformer
by: Joshi, Poorvi, et al.
Published: (2025)
by: Joshi, Poorvi, et al.
Published: (2025)
Context-Aware Information Transfer via Digital Semantic Communication in UAV-Based Networks
by: Joshi, Poorvi, et al.
Published: (2026)
by: Joshi, Poorvi, et al.
Published: (2026)
Iowa City Reads! The Reading Event Worth Shouting About.
by: Donham van Deusen, Jean, et al.
Published: (1997)
by: Donham van Deusen, Jean, et al.
Published: (1997)
MuSE-SVS: Multi-Singer Emotional Singing Voice Synthesizer that Controls Emotional Intensity
by: Kim, Sungjae, et al.
Published: (2022)
by: Kim, Sungjae, et al.
Published: (2022)
Behind the Scenes: Mechanistic Interpretability of LoRA-adapted Whisper for Speech Emotion Recognition
by: Ma, Yujian, et al.
Published: (2025)
by: Ma, Yujian, et al.
Published: (2025)
Transfer Learning from Visual Speech Recognition to Mouthing Recognition in German Sign Language
by: Pham, Dinh Nam, et al.
Published: (2025)
by: Pham, Dinh Nam, et al.
Published: (2025)
Color-based Emotion Representation for Speech Emotion Recognition
by: Nagase, Ryotaro, et al.
Published: (2026)
by: Nagase, Ryotaro, et al.
Published: (2026)
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
by: Ma, Ziyang, et al.
Published: (2023)
by: Ma, Ziyang, et al.
Published: (2023)
The Undergraduate Library: A Screaming Success As Study Halls
by: Wilkinson, Billy R.
Published: (1971)
by: Wilkinson, Billy R.
Published: (1971)
Driving Bots with a Neuroevolved Brain: Screaming Racers
by: Faraón Llorens
Published: (2005)
by: Faraón Llorens
Published: (2005)
Speech Emotion Recognition with ASR Integration
by: Li, Yuanchao
Published: (2026)
by: Li, Yuanchao
Published: (2026)
TSPC: A Two-Stage Phoneme-Centric Architecture for code-switching Vietnamese-English Speech Recognition
by: Anh, Tran Nguyen, et al.
Published: (2025)
by: Anh, Tran Nguyen, et al.
Published: (2025)
The number of cut-edges and conflict-free connection number in planar graphs
by: Ha, Pham Hoang, et al.
Published: (2026)
by: Ha, Pham Hoang, et al.
Published: (2026)
QQSUM: A Novel Task and Model of Quantitative Query-Focused Summarization for Review-based Product Question Answering
by: Tang, An Quang, et al.
Published: (2025)
by: Tang, An Quang, et al.
Published: (2025)
ARQUSUMM: Argument-aware Quantitative Summarization of Online Conversations
by: Tang, An Quang, et al.
Published: (2025)
by: Tang, An Quang, et al.
Published: (2025)
Similar Items
-
Good Things Come in Trees: Emotion and Context Aware Behaviour Trees for Ethical Robotic Decision-Making
by: Tuttösí, Paige, et al.
Published: (2024) -
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025) -
Read the Room: Adapting a Robot's Voice to Ambient and Social Contexts
by: Tuttosi, Paige, et al.
Published: (2022) -
Covertly improving intelligibility with data-driven adaptations of speech timing
by: Tuttösí, Paige, et al.
Published: (2026) -
You Sound a Little Tense: L2 Tailored Clear TTS Using Durational Vowel Properties
by: Tuttösí, Paige, et al.
Published: (2025)