LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Damacharla, Praveen, Rajabalipanah, Hamid, Fakheri, Mohammad Hosein |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Human Perception of Audio Deepfakes
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
von: Manukalpa, J. M. Chan Sri, et al.
Veröffentlicht: (2025)
von: Manukalpa, J. M. Chan Sri, et al.
Veröffentlicht: (2025)
Step-Audio-EditX Technical Report
von: Yan, Chao, et al.
Veröffentlicht: (2025)
von: Yan, Chao, et al.
Veröffentlicht: (2025)
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
von: Huang, Ailin, et al.
Veröffentlicht: (2025)
von: Huang, Ailin, et al.
Veröffentlicht: (2025)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
von: Brade, Stephen, et al.
Veröffentlicht: (2023)
von: Brade, Stephen, et al.
Veröffentlicht: (2023)
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
von: Guo, Yiwei, et al.
Veröffentlicht: (2023)
von: Guo, Yiwei, et al.
Veröffentlicht: (2023)
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
von: Han, Runduo, et al.
Veröffentlicht: (2025)
von: Han, Runduo, et al.
Veröffentlicht: (2025)
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
von: Kotowski, Błażej, et al.
Veröffentlicht: (2025)
von: Kotowski, Błażej, et al.
Veröffentlicht: (2025)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
von: Wu, Liang-Yuan, et al.
Veröffentlicht: (2025)
von: Wu, Liang-Yuan, et al.
Veröffentlicht: (2025)
Reimagining Dance: Real-time Music Co-creation between Dancers and AI
von: Vechtomova, Olga, et al.
Veröffentlicht: (2025)
von: Vechtomova, Olga, et al.
Veröffentlicht: (2025)
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
von: Benster, Tyler, et al.
Veröffentlicht: (2024)
von: Benster, Tyler, et al.
Veröffentlicht: (2024)
MCP2OSC: Parametric Control by Natural Language
von: Fan, Yuan-Yi
Veröffentlicht: (2025)
von: Fan, Yuan-Yi
Veröffentlicht: (2025)
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
von: Chang, Yi, et al.
Veröffentlicht: (2024)
von: Chang, Yi, et al.
Veröffentlicht: (2024)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
von: Yan, Hui, et al.
Veröffentlicht: (2024)
von: Yan, Hui, et al.
Veröffentlicht: (2024)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
von: Fong, Hortense, et al.
Veröffentlicht: (2024)
von: Fong, Hortense, et al.
Veröffentlicht: (2024)
Interactive Melody Generation System for Enhancing the Creativity of Musicians
von: Hirawata, So, et al.
Veröffentlicht: (2024)
von: Hirawata, So, et al.
Veröffentlicht: (2024)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
von: Wilson, Elizabeth, et al.
Veröffentlicht: (2024)
von: Wilson, Elizabeth, et al.
Veröffentlicht: (2024)
Tipping Points, Pulse Elasticity and Tonal Tension: An Empirical Study on What Generates Tipping Points
von: Naik, Canishk, et al.
Veröffentlicht: (2024)
von: Naik, Canishk, et al.
Veröffentlicht: (2024)
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
von: Sarmento, Pedro, et al.
Veröffentlicht: (2024)
von: Sarmento, Pedro, et al.
Veröffentlicht: (2024)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
von: Liu, Ailin, et al.
Veröffentlicht: (2024)
von: Liu, Ailin, et al.
Veröffentlicht: (2024)
Learning Relationships Between Separate Audio Tracks for Creative Applications
von: Bujard, Balthazar, et al.
Veröffentlicht: (2025)
von: Bujard, Balthazar, et al.
Veröffentlicht: (2025)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
von: Blum'e, Ashlae
Veröffentlicht: (2025)
von: Blum'e, Ashlae
Veröffentlicht: (2025)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
von: Wang, Hongbin, et al.
Veröffentlicht: (2025)
von: Wang, Hongbin, et al.
Veröffentlicht: (2025)
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2025)
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
von: Zheng, Shuoyang, et al.
Veröffentlicht: (2024)
von: Zheng, Shuoyang, et al.
Veröffentlicht: (2024)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
von: Li, Yinan, et al.
Veröffentlicht: (2026)
von: Li, Yinan, et al.
Veröffentlicht: (2026)
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
von: Dietrich, Juergen
Veröffentlicht: (2026)
von: Dietrich, Juergen
Veröffentlicht: (2026)
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
von: Liu, Haoxuan, et al.
Veröffentlicht: (2024)
von: Liu, Haoxuan, et al.
Veröffentlicht: (2024)
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
von: Bryan-Kinns, Nick, et al.
Veröffentlicht: (2024)
von: Bryan-Kinns, Nick, et al.
Veröffentlicht: (2024)
SounDiT: Geo-Contextual Soundscape-to-Landscape Generation
von: Wang, Junbo, et al.
Veröffentlicht: (2025)
von: Wang, Junbo, et al.
Veröffentlicht: (2025)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
von: Yang, Sicheng, et al.
Veröffentlicht: (2024)
von: Yang, Sicheng, et al.
Veröffentlicht: (2024)
G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition
von: Peng, Jing, et al.
Veröffentlicht: (2026)
von: Peng, Jing, et al.
Veröffentlicht: (2026)
Zero-Shot KWS for Children's Speech using Layer-Wise Features from SSL Models
von: Kutum, Subham, et al.
Veröffentlicht: (2025)
von: Kutum, Subham, et al.
Veröffentlicht: (2025)
Layer-Wise Analysis of Self-Supervised Representations for Age and Gender Classification in Children's Speech
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
Accelerating Audio Research with Robotic Dummy Heads
von: Lu, Austin, et al.
Veröffentlicht: (2025)
von: Lu, Austin, et al.
Veröffentlicht: (2025)
Lla-VAP: LSTM Ensemble of Llama and VAP for Turn-Taking Prediction
von: Jeon, Hyunbae, et al.
Veröffentlicht: (2024)
von: Jeon, Hyunbae, et al.
Veröffentlicht: (2024)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
von: Li, Lu, et al.
Veröffentlicht: (2025)
von: Li, Lu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Human Perception of Audio Deepfakes
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021) -
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
von: Manukalpa, J. M. Chan Sri, et al.
Veröffentlicht: (2025) -
Step-Audio-EditX Technical Report
von: Yan, Chao, et al.
Veröffentlicht: (2025) -
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
von: Huang, Ailin, et al.
Veröffentlicht: (2025) -
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
von: Brade, Stephen, et al.
Veröffentlicht: (2023)