Fundamental Survey on Neuromorphic Based Audio Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Basu, Amlan, Chaudhari, Pranav, Di Caterina, Gaetano |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
von: Singh, Shubhr, et al.
Veröffentlicht: (2025)
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
Raw Audio Classification with Cosine Convolutional Neural Network (CosCovNN)
von: Haque, Kazi Nazmul, et al.
Veröffentlicht: (2024)
von: Haque, Kazi Nazmul, et al.
Veröffentlicht: (2024)
Studying the Effect of Audio Filters in Pre-Trained Models for Environmental Sound Classification
von: Dawn, Aditya, et al.
Veröffentlicht: (2024)
von: Dawn, Aditya, et al.
Veröffentlicht: (2024)
4,500 Seconds: Small Data Training Approaches for Deep UAV Audio Classification
von: Berg, Andrew P., et al.
Veröffentlicht: (2025)
von: Berg, Andrew P., et al.
Veröffentlicht: (2025)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
Domain Adaptation Method and Modality Gap Impact in Audio-Text Models for Prototypical Sound Classification
von: Acevedo, Emiliano, et al.
Veröffentlicht: (2025)
von: Acevedo, Emiliano, et al.
Veröffentlicht: (2025)
Audio Mamba: Pretrained Audio State Space Model For Audio Tagging
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
Audio Atlas: Visualizing and Exploring Audio Datasets
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2024)
Region-Based Optimization in Continual Learning for Audio Deepfake Detection
von: Chen, Yujie, et al.
Veröffentlicht: (2024)
von: Chen, Yujie, et al.
Veröffentlicht: (2024)
Example-Based Framework for Perceptually Guided Audio Texture Generation
von: Kamath, Purnima, et al.
Veröffentlicht: (2023)
von: Kamath, Purnima, et al.
Veröffentlicht: (2023)
Representation-Regularized Convolutional Audio Transformer for Audio Understanding
von: Han, Bing, et al.
Veröffentlicht: (2026)
von: Han, Bing, et al.
Veröffentlicht: (2026)
Audio Spatially-Guided Fusion for Audio-Visual Navigation
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
EnCLAP: Combining Neural Audio Codec and Audio-Text Joint Embedding for Automated Audio Captioning
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
Compressing Quaternion Convolutional Neural Networks for Audio Classification
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
von: Zhao, Junqi, et al.
Veröffentlicht: (2025)
DreamAudio: Customized Text-to-Audio Generation with Diffusion Models
von: Yuan, Yi, et al.
Veröffentlicht: (2025)
von: Yuan, Yi, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation of CNN-Based Audio Tagging Models on Resource-Constrained Devices
von: Grau-Haro, Jordi, et al.
Veröffentlicht: (2025)
von: Grau-Haro, Jordi, et al.
Veröffentlicht: (2025)
AudioScene: Integrating Object-Event Audio into 3D Scenes
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2025)
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2025)
Audio Mamba: Bidirectional State Space Model for Audio Representation Learning
von: Erol, Mehmet Hamza, et al.
Veröffentlicht: (2024)
von: Erol, Mehmet Hamza, et al.
Veröffentlicht: (2024)
Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations
von: Yadav, Sarthak, et al.
Veröffentlicht: (2024)
von: Yadav, Sarthak, et al.
Veröffentlicht: (2024)
Stable Audio Open
von: Evans, Zach, et al.
Veröffentlicht: (2024)
von: Evans, Zach, et al.
Veröffentlicht: (2024)
Audio Jailbreak: An Open Comprehensive Benchmark for Jailbreaking Large Audio-Language Models
von: Song, Zirui, et al.
Veröffentlicht: (2025)
von: Song, Zirui, et al.
Veröffentlicht: (2025)
SLAP: Scalable Language-Audio Pretraining with Variable-Duration Audio and Multi-Objective Training
von: Mei, Xinhao, et al.
Veröffentlicht: (2026)
von: Mei, Xinhao, et al.
Veröffentlicht: (2026)
AudioRouter: Data Efficient Audio Understanding via RL based Dual Reasoning
von: Chen, Liyang, et al.
Veröffentlicht: (2026)
von: Chen, Liyang, et al.
Veröffentlicht: (2026)
UltraEval-Audio: A Unified Framework for Comprehensive Evaluation of Audio Foundation Models
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
EditGen: Harnessing Cross-Attention Control for Instruction-Based Auto-Regressive Audio Editing
von: Sioros, Vassilis, et al.
Veröffentlicht: (2025)
von: Sioros, Vassilis, et al.
Veröffentlicht: (2025)
The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition
von: Gao, Ming, et al.
Veröffentlicht: (2025)
von: Gao, Ming, et al.
Veröffentlicht: (2025)
Estimating Musical Surprisal in Audio
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
von: Bjare, Mathias Rose, et al.
Veröffentlicht: (2025)
Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
von: Xie, Yuankun, et al.
Veröffentlicht: (2024)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
FusionAudio-1.2M: Towards Fine-grained Audio Captioning with Multimodal Contextual Fusion
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
von: Chen, Shunian, et al.
Veröffentlicht: (2025)
DASB - Discrete Audio and Speech Benchmark
von: Mousavi, Pooneh, et al.
Veröffentlicht: (2024)
von: Mousavi, Pooneh, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Audio Deepfake Detection
von: Kang, Zuheng, et al.
Veröffentlicht: (2024)
von: Kang, Zuheng, et al.
Veröffentlicht: (2024)
Towards Controllable Audio Texture Morphing
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
MAT-SED: A Masked Audio Transformer with Masked-Reconstruction Based Pre-training for Sound Event Detection
von: Cai, Pengfei, et al.
Veröffentlicht: (2024)
von: Cai, Pengfei, et al.
Veröffentlicht: (2024)
Replay Attacks Against Audio Deepfake Detection
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
von: Müller, Nicolas, et al.
Veröffentlicht: (2025)
ViSAGe: Video-to-Spatial Audio Generation
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2025)
Perceptual Musical Features for Interpretable Audio Tagging
von: Lyberatos, Vassilis, et al.
Veröffentlicht: (2023)
von: Lyberatos, Vassilis, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
von: Singh, Shubhr, et al.
Veröffentlicht: (2025) -
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024) -
Raw Audio Classification with Cosine Convolutional Neural Network (CosCovNN)
von: Haque, Kazi Nazmul, et al.
Veröffentlicht: (2024) -
Studying the Effect of Audio Filters in Pre-Trained Models for Environmental Sound Classification
von: Dawn, Aditya, et al.
Veröffentlicht: (2024) -
4,500 Seconds: Small Data Training Approaches for Deep UAV Audio Classification
von: Berg, Andrew P., et al.
Veröffentlicht: (2025)