Beyond IVR Touch-Tones: Customer Intent Routing using LLMs
Fuente:
arXiv
Guardado en:
| Autor principal: | Rojas-Galeano, Sergio |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
por: Dietrich, Juergen
Publicado: (2026)
por: Dietrich, Juergen
Publicado: (2026)
Beamforming-LLM: What, Where and When Did I Miss?
por: Choudhari, Vishal
Publicado: (2025)
por: Choudhari, Vishal
Publicado: (2025)
Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization
por: Andrusenko, Andrei, et al.
Publicado: (2026)
por: Andrusenko, Andrei, et al.
Publicado: (2026)
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
por: Chen, Qian, et al.
Publicado: (2025)
por: Chen, Qian, et al.
Publicado: (2025)
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
por: Huang, Ailin, et al.
Publicado: (2025)
por: Huang, Ailin, et al.
Publicado: (2025)
Step-Audio-EditX Technical Report
por: Yan, Chao, et al.
Publicado: (2025)
por: Yan, Chao, et al.
Publicado: (2025)
AAD-LLM: Neural Attention-Driven Auditory Scene Understanding
por: Jiang, Xilin, et al.
Publicado: (2025)
por: Jiang, Xilin, et al.
Publicado: (2025)
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
por: Han, Zhichen, et al.
Publicado: (2024)
por: Han, Zhichen, et al.
Publicado: (2024)
Language Model Can Listen While Speaking
por: Ma, Ziyang, et al.
Publicado: (2024)
por: Ma, Ziyang, et al.
Publicado: (2024)
EmoKnob: Enhance Voice Cloning with Fine-Grained Emotion Control
por: Chen, Haozhe, et al.
Publicado: (2024)
por: Chen, Haozhe, et al.
Publicado: (2024)
Toward a Realistic Encoding Model of Auditory Affective Understanding in the Brain
por: Pan, Guandong, et al.
Publicado: (2025)
por: Pan, Guandong, et al.
Publicado: (2025)
I Hear, Therefore I Trust: A Socio-Technical Investigation of Humans as Synthetic Speech Detectors
por: Erscoi, Lelia, et al.
Publicado: (2026)
por: Erscoi, Lelia, et al.
Publicado: (2026)
InsightPulse: An IoT-based System for User Experience Interview Analysis
por: Lyu, Dian, et al.
Publicado: (2024)
por: Lyu, Dian, et al.
Publicado: (2024)
Open-Source Conversational AI with SpeechBrain 1.0
por: Ravanelli, Mirco, et al.
Publicado: (2024)
por: Ravanelli, Mirco, et al.
Publicado: (2024)
MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes
por: Chen, Maximillian, et al.
Publicado: (2026)
por: Chen, Maximillian, et al.
Publicado: (2026)
Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding
por: Wang, Yuchen, et al.
Publicado: (2026)
por: Wang, Yuchen, et al.
Publicado: (2026)
Super Kawaii Vocalics: Amplifying the "Cute" Factor in Computer Voice
por: Mandai, Yuto, et al.
Publicado: (2025)
por: Mandai, Yuto, et al.
Publicado: (2025)
Inter(sectional) Alia(s): Ambiguity in Voice Agent Identity via Intersectional Japanese Self-Referents
por: Fujii, Takao, et al.
Publicado: (2025)
por: Fujii, Takao, et al.
Publicado: (2025)
Qualitative Approaches to Voice UX
por: Seaborn, Katie, et al.
Publicado: (2024)
por: Seaborn, Katie, et al.
Publicado: (2024)
Multimodal Large Language Models with Fusion Low Rank Adaptation for Device Directed Speech Detection
por: Palaskar, Shruti, et al.
Publicado: (2024)
por: Palaskar, Shruti, et al.
Publicado: (2024)
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
por: Xu, Ziqing, et al.
Publicado: (2025)
por: Xu, Ziqing, et al.
Publicado: (2025)
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
por: Han, Runduo, et al.
Publicado: (2025)
por: Han, Runduo, et al.
Publicado: (2025)
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
por: Kotowski, Błażej, et al.
Publicado: (2025)
por: Kotowski, Błażej, et al.
Publicado: (2025)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
por: Wu, Liang-Yuan, et al.
Publicado: (2025)
por: Wu, Liang-Yuan, et al.
Publicado: (2025)
Reimagining Dance: Real-time Music Co-creation between Dancers and AI
por: Vechtomova, Olga, et al.
Publicado: (2025)
por: Vechtomova, Olga, et al.
Publicado: (2025)
MCP2OSC: Parametric Control by Natural Language
por: Fan, Yuan-Yi
Publicado: (2025)
por: Fan, Yuan-Yi
Publicado: (2025)
Human Perception of Audio Deepfakes
por: Müller, Nicolas M., et al.
Publicado: (2021)
por: Müller, Nicolas M., et al.
Publicado: (2021)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
por: Brade, Stephen, et al.
Publicado: (2023)
por: Brade, Stephen, et al.
Publicado: (2023)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
por: Kwon, Joonwoo, et al.
Publicado: (2024)
por: Kwon, Joonwoo, et al.
Publicado: (2024)
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
por: Guo, Yiwei, et al.
Publicado: (2023)
por: Guo, Yiwei, et al.
Publicado: (2023)
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
por: Damacharla, Praveen, et al.
Publicado: (2023)
por: Damacharla, Praveen, et al.
Publicado: (2023)
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
por: Benster, Tyler, et al.
Publicado: (2024)
por: Benster, Tyler, et al.
Publicado: (2024)
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
por: Chang, Yi, et al.
Publicado: (2024)
por: Chang, Yi, et al.
Publicado: (2024)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
por: Yan, Hui, et al.
Publicado: (2024)
por: Yan, Hui, et al.
Publicado: (2024)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
por: Fong, Hortense, et al.
Publicado: (2024)
por: Fong, Hortense, et al.
Publicado: (2024)
Interactive Melody Generation System for Enhancing the Creativity of Musicians
por: Hirawata, So, et al.
Publicado: (2024)
por: Hirawata, So, et al.
Publicado: (2024)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
por: Wilson, Elizabeth, et al.
Publicado: (2024)
por: Wilson, Elizabeth, et al.
Publicado: (2024)
Tipping Points, Pulse Elasticity and Tonal Tension: An Empirical Study on What Generates Tipping Points
por: Naik, Canishk, et al.
Publicado: (2024)
por: Naik, Canishk, et al.
Publicado: (2024)
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
por: Sarmento, Pedro, et al.
Publicado: (2024)
por: Sarmento, Pedro, et al.
Publicado: (2024)
A Penny for Your Thoughts: Decoding Speech from Inexpensive Brain Signals
por: Auster, Quentin, et al.
Publicado: (2025)
por: Auster, Quentin, et al.
Publicado: (2025)
Ejemplares similares
-
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
por: Dietrich, Juergen
Publicado: (2026) -
Beamforming-LLM: What, Where and When Did I Miss?
por: Choudhari, Vishal
Publicado: (2025) -
Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization
por: Andrusenko, Andrei, et al.
Publicado: (2026) -
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
por: Chen, Qian, et al.
Publicado: (2025) -
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
por: Huang, Ailin, et al.
Publicado: (2025)