Extracting Urban Sound Information for Residential Areas in Smart Cities Using an End-to-End IoT System
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Ee-Leng, Karnapi, Furi Andi, Ng, Linus Junjia, Ooi, Kenneth, Gan, Woon-Seng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
von: Yeow, Jun Wei, et al.
Veröffentlicht: (2024)
von: Yeow, Jun Wei, et al.
Veröffentlicht: (2024)
Enhancing Situational Awareness in Wearable Audio Devices Using a Lightweight Sound Event Localization and Detection System
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
Transformer-based End-to-End Control Filter Generation for Active Noise Control
von: Yang, Ziyi, et al.
Veröffentlicht: (2026)
von: Yang, Ziyi, et al.
Veröffentlicht: (2026)
Automating Urban Soundscape Enhancements with AI: In-situ Assessment of Quality and Restorativeness in Traffic-Exposed Residential Areas
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
Acoustic Scene Classification Using CNN-GRU Model Without Knowledge Distillation
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2025)
von: Tan, Ee-Leng, et al.
Veröffentlicht: (2025)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction
von: Yeo, Jin Jie Sean, et al.
Veröffentlicht: (2024)
von: Yeo, Jin Jie Sean, et al.
Veröffentlicht: (2024)
Joint Feature and Output Distillation for Low-complexity Acoustic Scene Classification
von: Li, Haowen, et al.
Veröffentlicht: (2025)
von: Li, Haowen, et al.
Veröffentlicht: (2025)
Do neonates hear what we measure? Assessing neonatal ward soundscapes at the neonates ears
von: Lam, Bhan, et al.
Veröffentlicht: (2025)
von: Lam, Bhan, et al.
Veröffentlicht: (2025)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
IoT-based Noise Monitoring using Mobile Nodes for Smart Cities
von: Manthina, Bhima Sankar, et al.
Veröffentlicht: (2025)
von: Manthina, Bhima Sankar, et al.
Veröffentlicht: (2025)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
Mixed-gradients Distributed Filtered Reference Least Mean Square Algorithm -- A Robust Distributed Multichannel Active Noise Control Algorithm
von: Ji, Junwei, et al.
Veröffentlicht: (2025)
von: Ji, Junwei, et al.
Veröffentlicht: (2025)
VBx for End-to-End Neural and Clustering-based Diarization
von: Pálka, Petr, et al.
Veröffentlicht: (2025)
von: Pálka, Petr, et al.
Veröffentlicht: (2025)
DNCASR: End-to-End Training for Speaker-Attributed ASR
von: Zheng, Xianrui, et al.
Veröffentlicht: (2025)
von: Zheng, Xianrui, et al.
Veröffentlicht: (2025)
An End-To-End Stuttering Detection Method Based On Conformer And BILSTM
von: Liu, Xiaokang, et al.
Veröffentlicht: (2024)
von: Liu, Xiaokang, et al.
Veröffentlicht: (2024)
AudioLog: LLMs-Powered Long Audio Logging with Hybrid Token-Semantic Contrastive Learning
von: Bai, Jisheng, et al.
Veröffentlicht: (2023)
von: Bai, Jisheng, et al.
Veröffentlicht: (2023)
End-to-End Speech Recognition with Pre-trained Masked Language Model
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2024)
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2024)
Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT
von: Dai, Dongyang, et al.
Veröffentlicht: (2025)
von: Dai, Dongyang, et al.
Veröffentlicht: (2025)
Self-Boosted Weight-Constrained FxLMS: A Robustness Distributed Active Noise Control Algorithm Without Internode Communication
von: Ji, Junwei, et al.
Veröffentlicht: (2025)
von: Ji, Junwei, et al.
Veröffentlicht: (2025)
Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
A Stabilized Hybrid Active Noise Control Algorithm of GFANC and FxNLMS with Online Clustering
von: Luo, Zhengding, et al.
Veröffentlicht: (2026)
von: Luo, Zhengding, et al.
Veröffentlicht: (2026)
End-to-End Direction-Aware Keyword Spotting with Spatial Priors in Noisy Environments
von: Wang, Rui, et al.
Veröffentlicht: (2026)
von: Wang, Rui, et al.
Veröffentlicht: (2026)
End-to-End DOA-Guided Speech Extraction in Noisy Multi-Talker Scenarios
von: Jing, Kangqi, et al.
Veröffentlicht: (2025)
von: Jing, Kangqi, et al.
Veröffentlicht: (2025)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
Lightweight and Robust Multi-Channel End-to-End Speech Recognition with Spherical Harmonic Transform
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2025)
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2025)
End-to-End Diarization utilizing Attractor Deep Clustering
von: Palzer, David, et al.
Veröffentlicht: (2025)
von: Palzer, David, et al.
Veröffentlicht: (2025)
Speaker Adaptation for Quantised End-to-End ASR Models
von: Zhao, Qiuming, et al.
Veröffentlicht: (2024)
von: Zhao, Qiuming, et al.
Veröffentlicht: (2024)
An Investigation on Speaker Augmentation for End-to-End Speaker Extraction
von: You, Zhenghai, et al.
Veröffentlicht: (2025)
von: You, Zhenghai, et al.
Veröffentlicht: (2025)
Reference Channel Selection by Multi-Channel Masking for End-to-End Multi-Channel Speech Enhancement
von: Dai, Wang, et al.
Veröffentlicht: (2024)
von: Dai, Wang, et al.
Veröffentlicht: (2024)
Breaking Walls: Pioneering Automatic Speech Recognition for Central Kurdish: End-to-End Transformer Paradigm
von: Abdullah, Abdulhady Abas, et al.
Veröffentlicht: (2024)
von: Abdullah, Abdulhady Abas, et al.
Veröffentlicht: (2024)
SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription
von: Dai, Yuhang, et al.
Veröffentlicht: (2026)
von: Dai, Yuhang, et al.
Veröffentlicht: (2026)
Exploring an Inter-Pausal Unit (IPU) based Approach for Indic End-to-End TTS Systems
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
Dissecting the Segmentation Model of End-to-End Diarization with Vector Clustering
von: Plaquet, Alexis, et al.
Veröffentlicht: (2025)
von: Plaquet, Alexis, et al.
Veröffentlicht: (2025)
CSSinger: End-to-End Chunkwise Streaming Singing Voice Synthesis System Based on Conditional Variational Autoencoder
von: Cui, Jianwei, et al.
Veröffentlicht: (2024)
von: Cui, Jianwei, et al.
Veröffentlicht: (2024)
CUSIDE-array: A Streaming Multi-Channel End-to-End Speech Recognition System with Realistic Evaluations
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2024)
von: Kong, Xiangzhu, et al.
Veröffentlicht: (2024)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
von: Wang, Pu, et al.
Veröffentlicht: (2024)
von: Wang, Pu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Improving Stereo 3D Sound Event Localization and Detection: Perceptual Features, Stereo-specific Data Augmentation, and Distance Normalization
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025) -
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025) -
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
von: Yeow, Jun Wei, et al.
Veröffentlicht: (2024) -
Enhancing Situational Awareness in Wearable Audio Devices Using a Lightweight Sound Event Localization and Detection System
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025) -
Transformer-based End-to-End Control Filter Generation for Active Noise Control
von: Yang, Ziyi, et al.
Veröffentlicht: (2026)