Energy-based features and bi-LSTM neural network for EEG-based music and voice classification
Fuente:
arXiv
Saved in:
| Main Authors: | Ariza, Isaac, Barbancho, Ana M., Tardon, Lorenzo J., Barbancho, Isabel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bi-LSTM neural network for EEG-based error detection in musicians' performance
by: Ariza, Isaac, et al.
Published: (2024)
by: Ariza, Isaac, et al.
Published: (2024)
Building music with Lego bricks and Raspberry Pi
by: Barbancho, Ana M., et al.
Published: (2024)
by: Barbancho, Ana M., et al.
Published: (2024)
Implementation of tools for lessening the influence of artifacts in EEG signal analysis
by: Molina-Molina, Mario, et al.
Published: (2024)
by: Molina-Molina, Mario, et al.
Published: (2024)
Enhanced average for event-related potential analysis using dynamic time warping
by: Molina, Mario, et al.
Published: (2024)
by: Molina, Mario, et al.
Published: (2024)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
by: Wang, Bo, et al.
Published: (2024)
by: Wang, Bo, et al.
Published: (2024)
Preprocessing for lessening the influence of eye artifacts in eeg analysis
by: Villena, Alejandro, et al.
Published: (2024)
by: Villena, Alejandro, et al.
Published: (2024)
Deep learning classification system for coconut maturity levels based on acoustic signals
by: Caladcad, June Anne, et al.
Published: (2024)
by: Caladcad, June Anne, et al.
Published: (2024)
Proposal of protocols for speech materials acquisition and presentation assisted by tools based on structured test signals
by: Kawahara, Hideki, et al.
Published: (2024)
by: Kawahara, Hideki, et al.
Published: (2024)
Deep learning-based filtering of cross-spectral matrices using generative adversarial networks
by: Puhle, Christof
Published: (2025)
by: Puhle, Christof
Published: (2025)
Dynamic Prediction of Full-Ocean Depth SSP by Hierarchical LSTM: An Experimental Result
by: Lu, Jiajun, et al.
Published: (2023)
by: Lu, Jiajun, et al.
Published: (2023)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
by: Zhu, Haolin, et al.
Published: (2024)
by: Zhu, Haolin, et al.
Published: (2024)
SSM2Mel: State Space Model to Reconstruct Mel Spectrogram from the EEG
by: Fan, Cunhang, et al.
Published: (2025)
by: Fan, Cunhang, et al.
Published: (2025)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
by: Wilroth, Johanna, et al.
Published: (2026)
by: Wilroth, Johanna, et al.
Published: (2026)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
by: Liao, Yuan, et al.
Published: (2025)
by: Liao, Yuan, et al.
Published: (2025)
Synthetic training set generation using text-to-audio models for environmental sound classification
by: Ronchini, Francesca, et al.
Published: (2024)
by: Ronchini, Francesca, et al.
Published: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
by: Berghi, Davide, et al.
Published: (2025)
by: Berghi, Davide, et al.
Published: (2025)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
by: Yi, Jayeon, et al.
Published: (2024)
by: Yi, Jayeon, et al.
Published: (2024)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
by: Torabi, Yasaman, et al.
Published: (2025)
by: Torabi, Yasaman, et al.
Published: (2025)
QHARMA-GAN: Quasi-Harmonic Neural Vocoder based on Autoregressive Moving Average Model
by: Chen, Shaowen, et al.
Published: (2025)
by: Chen, Shaowen, et al.
Published: (2025)
EMOCONV-DIFF: Diffusion-based Speech Emotion Conversion for Non-parallel and In-the-wild Data
by: Prabhu, Navin Raj, et al.
Published: (2023)
by: Prabhu, Navin Raj, et al.
Published: (2023)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
by: Fejgin, Daniel, et al.
Published: (2025)
by: Fejgin, Daniel, et al.
Published: (2025)
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
by: Zhuang, Xuanyu, et al.
Published: (2024)
by: Zhuang, Xuanyu, et al.
Published: (2024)
An incremental algorithm based on multichannel non-negative matrix partial co-factorization for ambient denoising in auscultation
by: Cruz, Juan De La Torre, et al.
Published: (2024)
by: Cruz, Juan De La Torre, et al.
Published: (2024)
Acousto-optic reconstruction of exterior sound field based on concentric circle sampling with circular harmonic expansion
by: Nguyen, Phuc Duc, et al.
Published: (2023)
by: Nguyen, Phuc Duc, et al.
Published: (2023)
Implicit neural representation with physics-informed neural networks for the reconstruction of the early part of room impulse responses
by: Pezzoli, Mirco, et al.
Published: (2023)
by: Pezzoli, Mirco, et al.
Published: (2023)
Predicting Heart Activity from Speech using Data-driven and Knowledge-based features
by: Elbanna, Gasser, et al.
Published: (2024)
by: Elbanna, Gasser, et al.
Published: (2024)
STAR: Speech-to-Audio Generation via Representation Learning
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
PicoAudio2: Temporal Controllable Text-to-Audio Generation with Natural Language Description
by: Zheng, Zihao, et al.
Published: (2025)
by: Zheng, Zihao, et al.
Published: (2025)
FakeSound: Deepfake General Audio Detection
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
by: Picard, Emilio, et al.
Published: (2026)
by: Picard, Emilio, et al.
Published: (2026)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
by: Hiroe, Atsuo, et al.
Published: (2024)
by: Hiroe, Atsuo, et al.
Published: (2024)
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
by: Lu, Jiajun, et al.
Published: (2023)
by: Lu, Jiajun, et al.
Published: (2023)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
by: Meng, Hanyu, et al.
Published: (2025)
by: Meng, Hanyu, et al.
Published: (2025)
What is Learnt by the LEArnable Front-end (LEAF)? Adapting Per-Channel Energy Normalisation (PCEN) to Noisy Conditions
by: Meng, Hanyu, et al.
Published: (2024)
by: Meng, Hanyu, et al.
Published: (2024)
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
by: Ikuma, Takeshi, et al.
Published: (2025)
by: Ikuma, Takeshi, et al.
Published: (2025)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
by: Niizumi, Daisuke, et al.
Published: (2026)
by: Niizumi, Daisuke, et al.
Published: (2026)
Similar Items
-
Bi-LSTM neural network for EEG-based error detection in musicians' performance
by: Ariza, Isaac, et al.
Published: (2024) -
Building music with Lego bricks and Raspberry Pi
by: Barbancho, Ana M., et al.
Published: (2024) -
Implementation of tools for lessening the influence of artifacts in EEG signal analysis
by: Molina-Molina, Mario, et al.
Published: (2024) -
Enhanced average for event-related potential analysis using dynamic time warping
by: Molina, Mario, et al.
Published: (2024) -
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
by: Wang, Bo, et al.
Published: (2024)