Real-Time Emergency Vehicle Siren Detection with Efficient CNNs on Embedded Hardware
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Giordano, Marco, Giacomelli, Stefano, Rinaldi, Claudia, Graziosi, Fabio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Large-scale Audio Tagging to Real-Time Explainable Emergency Vehicle Sirens Detection
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2025)
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2025)
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024)
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024)
The OCON model: an old but gold solution for distributable supervised classification
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024)
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024)
Detecting and Preventing Latent Risk Accumulation in High-Performance Software Systems
von: Arafat, Jahidul, et al.
Veröffentlicht: (2025)
von: Arafat, Jahidul, et al.
Veröffentlicht: (2025)
The evolution of inharmonicity and noisiness in contemporary popular music
von: Deruty, Emmanuel, et al.
Veröffentlicht: (2024)
von: Deruty, Emmanuel, et al.
Veröffentlicht: (2024)
Neural Proxies for Sound Synthesizers: Learning Perceptually Informed Preset Representations
von: Combes, Paolo, et al.
Veröffentlicht: (2025)
von: Combes, Paolo, et al.
Veröffentlicht: (2025)
HELIX: Scaling Raw Audio Understanding with Hybrid Mamba-Attention Beyond the Quadratic Limit
von: Khushiyant, et al.
Veröffentlicht: (2026)
von: Khushiyant, et al.
Veröffentlicht: (2026)
TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection
von: Siddiqui, Yousuf Ahmed, et al.
Veröffentlicht: (2025)
von: Siddiqui, Yousuf Ahmed, et al.
Veröffentlicht: (2025)
Graph Connectionist Temporal Classification for Phoneme Recognition
von: Grafé, Henry, et al.
Veröffentlicht: (2025)
von: Grafé, Henry, et al.
Veröffentlicht: (2025)
STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution
von: Firc, Anton, et al.
Veröffentlicht: (2025)
von: Firc, Anton, et al.
Veröffentlicht: (2025)
Leveraging large multimodal models for audio-video deepfake detection: a pilot study
von: Cao, Songjun, et al.
Veröffentlicht: (2026)
von: Cao, Songjun, et al.
Veröffentlicht: (2026)
M2D-CLAP: Exploring General-purpose Audio-Language Representations Beyond CLAP
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
von: Hori, Takaaki, et al.
Veröffentlicht: (2025)
von: Hori, Takaaki, et al.
Veröffentlicht: (2025)
Audio-based Kinship Verification Using Age Domain Conversion
von: Sun, Qiyang, et al.
Veröffentlicht: (2024)
von: Sun, Qiyang, et al.
Veröffentlicht: (2024)
Implementation and Evaluation of Fast Raft for Hierarchical Consensus
von: Melnychuk, Anton, et al.
Veröffentlicht: (2025)
von: Melnychuk, Anton, et al.
Veröffentlicht: (2025)
Passive Underwater Acoustic Signal Separation based on Feature Decoupling Dual-path Network
von: Liu, Yucheng, et al.
Veröffentlicht: (2025)
von: Liu, Yucheng, et al.
Veröffentlicht: (2025)
Polarization-Based Eye Tracking with Personalized Siamese Architectures
von: Kalkanli, Beyza, et al.
Veröffentlicht: (2026)
von: Kalkanli, Beyza, et al.
Veröffentlicht: (2026)
Splitformer: An improved early-exit architecture for automatic speech recognition on edge devices
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2025)
Prevailing Research Areas for Music AI in the Era of Foundation Models
von: Wei, Megan, et al.
Veröffentlicht: (2024)
von: Wei, Megan, et al.
Veröffentlicht: (2024)
OBHS: An Optimized Block Huffman Scheme for Real-Time Audio Compression
von: Mahfi, Muntahi Safwan, et al.
Veröffentlicht: (2025)
von: Mahfi, Muntahi Safwan, et al.
Veröffentlicht: (2025)
Quantization for OpenAI's Whisper Models: A Comparative Analysis
von: Andreyev, Allison
Veröffentlicht: (2025)
von: Andreyev, Allison
Veröffentlicht: (2025)
Fine-Tuning Large Audio-Language Models with LoRA for Precise Temporal Localization of Prolonged Exposure Therapy Elements
von: BN, Suhas, et al.
Veröffentlicht: (2025)
von: BN, Suhas, et al.
Veröffentlicht: (2025)
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake Dataset
von: Hou, Yang, et al.
Veröffentlicht: (2024)
von: Hou, Yang, et al.
Veröffentlicht: (2024)
Transforming faces into video stories -- VideoFace2.0
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
Depth Priors in Removal Neural Radiance Fields
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
Contract-Driven QoE Auditing for Speech and Singing Services: From MOS Regression to Service Graphs
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
Deep Learning Approaches for Medical Imaging Under Varying Degrees of Label Availability: A Comprehensive Survey
von: Ma, Siteng, et al.
Veröffentlicht: (2025)
von: Ma, Siteng, et al.
Veröffentlicht: (2025)
BEC: Bit-Level Static Analysis for Reliability against Soft Errors
von: Ko, Yousun, et al.
Veröffentlicht: (2024)
von: Ko, Yousun, et al.
Veröffentlicht: (2024)
Task-Aligned Self-Supervised Learning for Medical Image Analysis: A Systematic Review and Practical Design Guidelines
von: Wimalasiri, Chathura
Veröffentlicht: (2026)
von: Wimalasiri, Chathura
Veröffentlicht: (2026)
A Cost-Effective Eye-Tracker for Early Detection of Mild Cognitive Impairment
von: Greco, Danilo, et al.
Veröffentlicht: (2024)
von: Greco, Danilo, et al.
Veröffentlicht: (2024)
Person detection and re-identification in open-world settings of retail stores and public spaces
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
Developing an aeroponic smart experimental greenhouse for controlling irrigation and plant disease detection using deep learning and IoT
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
Simultaneous source separation of unknown numbers of single-channel underwater acoustic signals based on deep neural networks with separator-decoder structure
von: Sun, Qinggang, et al.
Veröffentlicht: (2022)
von: Sun, Qinggang, et al.
Veröffentlicht: (2022)
Make Some Noise: Towards LLM audio reasoning and generation using sound tokens
von: Mehta, Shivam, et al.
Veröffentlicht: (2025)
von: Mehta, Shivam, et al.
Veröffentlicht: (2025)
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment
von: Mehta, Shivam, et al.
Veröffentlicht: (2025)
von: Mehta, Shivam, et al.
Veröffentlicht: (2025)
Position: Age Estimation Models Do Not Process Biometric Data
von: Marshalkin, Nikita
Veröffentlicht: (2026)
von: Marshalkin, Nikita
Veröffentlicht: (2026)
HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level and Fidelity-Rich Conditions in Diffusion Models
von: Zhang, Shengkai, et al.
Veröffentlicht: (2024)
von: Zhang, Shengkai, et al.
Veröffentlicht: (2024)
Heart Failure Prediction using Modal Decomposition and Masked Autoencoders for Scarce Echocardiography Databases
von: Bell-Navas, Andrés, et al.
Veröffentlicht: (2025)
von: Bell-Navas, Andrés, et al.
Veröffentlicht: (2025)
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
From Large-scale Audio Tagging to Real-Time Explainable Emergency Vehicle Sirens Detection
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2025) -
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024) -
The OCON model: an old but gold solution for distributable supervised classification
von: Giacomelli, Stefano, et al.
Veröffentlicht: (2024) -
Detecting and Preventing Latent Risk Accumulation in High-Performance Software Systems
von: Arafat, Jahidul, et al.
Veröffentlicht: (2025) -
The evolution of inharmonicity and noisiness in contemporary popular music
von: Deruty, Emmanuel, et al.
Veröffentlicht: (2024)