LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Shubhr, Benetos, Emmanouil, Phan, Huy, Stowell, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GraFPrint: A GNN-Based Approach for Audio Identification
von: Bhattacharjee, Aditya, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Aditya, et al.
Veröffentlicht: (2024)
ST-ITO: Controlling Audio Effects for Style Transfer with Inference-Time Optimization
von: Steinmetz, Christian J., et al.
Veröffentlicht: (2024)
von: Steinmetz, Christian J., et al.
Veröffentlicht: (2024)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
Acoustic identification of individual animals with hierarchical contrastive learning
von: Nolasco, Ines, et al.
Veröffentlicht: (2024)
von: Nolasco, Ines, et al.
Veröffentlicht: (2024)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
Compressing Quaternion Convolutional Neural Networks for Audio Classification
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
Audio Mamba: Pretrained Audio State Space Model For Audio Tagging
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
von: Lin, Jiaju, et al.
Veröffentlicht: (2024)
Perceptual Musical Features for Interpretable Audio Tagging
von: Lyberatos, Vassilis, et al.
Veröffentlicht: (2023)
von: Lyberatos, Vassilis, et al.
Veröffentlicht: (2023)
Integrating IP Broadcasting with Audio Tags: Workflow and Challenges
von: Burchett-Vass, Rhys, et al.
Veröffentlicht: (2024)
von: Burchett-Vass, Rhys, et al.
Veröffentlicht: (2024)
Raw Audio Classification with Cosine Convolutional Neural Network (CosCovNN)
von: Haque, Kazi Nazmul, et al.
Veröffentlicht: (2024)
von: Haque, Kazi Nazmul, et al.
Veröffentlicht: (2024)
Classification of Spontaneous and Scripted Speech for Multilingual Audio
von: Elisha, Shahar, et al.
Veröffentlicht: (2024)
von: Elisha, Shahar, et al.
Veröffentlicht: (2024)
Comprehensive Evaluation of CNN-Based Audio Tagging Models on Resource-Constrained Devices
von: Grau-Haro, Jordi, et al.
Veröffentlicht: (2025)
von: Grau-Haro, Jordi, et al.
Veröffentlicht: (2025)
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
Heterogeneous bimodal attention fusion for speech emotion recognition
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following
von: Ma, Yinghao, et al.
Veröffentlicht: (2025)
von: Ma, Yinghao, et al.
Veröffentlicht: (2025)
Fundamental Survey on Neuromorphic Based Audio Classification
von: Basu, Amlan, et al.
Veröffentlicht: (2025)
von: Basu, Amlan, et al.
Veröffentlicht: (2025)
EnCLAP: Combining Neural Audio Codec and Audio-Text Joint Embedding for Automated Audio Captioning
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection
von: Liang, Jinhua, et al.
Veröffentlicht: (2024)
von: Liang, Jinhua, et al.
Veröffentlicht: (2024)
Audio-to-Image Encoding for Improved Voice Characteristic Detection Using Deep Convolutional Neural Networks
von: Atif, Youness
Veröffentlicht: (2025)
von: Atif, Youness
Veröffentlicht: (2025)
In-the-wild Audio Spatialization with Flexible Text-guided Localization
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
HyperPotter: Spell the Charm of High-Order Interactions in Audio Deepfake Detection
von: Wen, Qing, et al.
Veröffentlicht: (2026)
von: Wen, Qing, et al.
Veröffentlicht: (2026)
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
von: Huang, Jiawen, et al.
Veröffentlicht: (2024)
von: Huang, Jiawen, et al.
Veröffentlicht: (2024)
Domain-Invariant Representation Learning of Bird Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
Studying the Effect of Audio Filters in Pre-Trained Models for Environmental Sound Classification
von: Dawn, Aditya, et al.
Veröffentlicht: (2024)
von: Dawn, Aditya, et al.
Veröffentlicht: (2024)
Enhancing Partially Spoofed Audio Localization with Boundary-aware Attention Mechanism
von: Zhong, Jiafeng, et al.
Veröffentlicht: (2024)
von: Zhong, Jiafeng, et al.
Veröffentlicht: (2024)
SpectroStream: A Versatile Neural Codec for General Audio
von: Li, Yunpeng, et al.
Veröffentlicht: (2025)
von: Li, Yunpeng, et al.
Veröffentlicht: (2025)
GraphMuse: A Library for Symbolic Music Graph Processing
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2024)
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2024)
Temporal Information Reconstruction and Non-Aligned Residual in Spiking Neural Networks for Speech Classification
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
4,500 Seconds: Small Data Training Approaches for Deep UAV Audio Classification
von: Berg, Andrew P., et al.
Veröffentlicht: (2025)
von: Berg, Andrew P., et al.
Veröffentlicht: (2025)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
AND: Audio Network Dissection for Interpreting Deep Acoustic Models
von: Wu, Tung-Yu, et al.
Veröffentlicht: (2024)
von: Wu, Tung-Yu, et al.
Veröffentlicht: (2024)
How Do Neural Spoofing Countermeasures Detect Partially Spoofed Audio?
von: Liu, Tianchi, et al.
Veröffentlicht: (2024)
von: Liu, Tianchi, et al.
Veröffentlicht: (2024)
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
von: Grötschla, Florian, et al.
Veröffentlicht: (2024)
Automatic acoustic detection of birds through deep learning: the first Bird Audio Detection challenge
von: Stowell, Dan, et al.
Veröffentlicht: (2018)
von: Stowell, Dan, et al.
Veröffentlicht: (2018)
Quantum-Inspired Audio Unlearning: Towards Privacy-Preserving Voice Biometrics
von: Pathak, Shreyansh, et al.
Veröffentlicht: (2025)
von: Pathak, Shreyansh, et al.
Veröffentlicht: (2025)
Domain Adaptation Method and Modality Gap Impact in Audio-Text Models for Prototypical Sound Classification
von: Acevedo, Emiliano, et al.
Veröffentlicht: (2025)
von: Acevedo, Emiliano, et al.
Veröffentlicht: (2025)
Audio Deepfake Detection in the Age of Advanced Text-to-Speech models
von: Singh, Robin, et al.
Veröffentlicht: (2026)
von: Singh, Robin, et al.
Veröffentlicht: (2026)
The Sounds of Home: A Speech-Removed Residential Audio Dataset for Sound Event Detection
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
von: Bibbó, Gabriel, et al.
Veröffentlicht: (2024)
MFAAN: Unveiling Audio Deepfakes with a Multi-Feature Authenticity Network
von: Krishnan, Karthik Sivarama, et al.
Veröffentlicht: (2023)
von: Krishnan, Karthik Sivarama, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GraFPrint: A GNN-Based Approach for Audio Identification
von: Bhattacharjee, Aditya, et al.
Veröffentlicht: (2024) -
ST-ITO: Controlling Audio Effects for Style Transfer with Inference-Time Optimization
von: Steinmetz, Christian J., et al.
Veröffentlicht: (2024) -
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024) -
Acoustic identification of individual animals with hierarchical contrastive learning
von: Nolasco, Ines, et al.
Veröffentlicht: (2024) -
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)