EMVD dataset: a dataset of extreme vocal distortion techniques used in heavy metal
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Tailleur, Modan, Pinquier, Julien, Millot, Laurent, Vogel, Corsin, Lagrange, Mathieu |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
par: Tailleur, Modan, et autres
Publié: (2025)
par: Tailleur, Modan, et autres
Publié: (2025)
Towards better visualizations of urban sound environments: insights from interviews
par: Tailleur, Modan, et autres
Publié: (2024)
par: Tailleur, Modan, et autres
Publié: (2024)
Sound Scene Synthesis at the DCASE 2024 Challenge
par: Lagrange, Mathieu, et autres
Publié: (2025)
par: Lagrange, Mathieu, et autres
Publié: (2025)
Using perceptive subbands analysis to perform audio scenes cartography
par: Millot, Laurent, et autres
Publié: (2024)
par: Millot, Laurent, et autres
Publié: (2024)
Challenge on Sound Scene Synthesis: Evaluating Text-to-Audio Generation
par: Lee, Junwon, et autres
Publié: (2024)
par: Lee, Junwon, et autres
Publié: (2024)
Machine listening in a neonatal intensive care unit
par: Tailleur, Modan, et autres
Publié: (2024)
par: Tailleur, Modan, et autres
Publié: (2024)
Detection of Deepfake Environmental Audio
par: Ouajdi, Hafsa, et autres
Publié: (2024)
par: Ouajdi, Hafsa, et autres
Publié: (2024)
Is Phase Really Needed for Weakly-Supervised Dereverberation ?
par: Rodrigues, Marius, et autres
Publié: (2025)
par: Rodrigues, Marius, et autres
Publié: (2025)
A proposal for a minimal model of free reed
par: Millot, Laurent
Publié: (2024)
par: Millot, Laurent
Publié: (2024)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
par: Tailleur, Modan, et autres
Publié: (2024)
par: Tailleur, Modan, et autres
Publié: (2024)
Artificial intelligence in creating, representing or expressing an immersive soundscape
par: Ayoubi, Rima, et autres
Publié: (2025)
par: Ayoubi, Rima, et autres
Publié: (2025)
Modèle physique variationnel pour l'estimation de réponses impulsionnelles de salles
par: Lalay, Louis, et autres
Publié: (2025)
par: Lalay, Louis, et autres
Publié: (2025)
Contribution of soundscape appropriateness to soundscape quality assessment in space: a mediating variable affecting acoustic comfort
par: Yang, Xinhao, et autres
Publié: (2024)
par: Yang, Xinhao, et autres
Publié: (2024)
Retrieving Effective Acoustic Impedance and Refractive Index for Size Mismatch Samples
par: Khodaei, Mohammad Javad, et autres
Publié: (2021)
par: Khodaei, Mohammad Javad, et autres
Publié: (2021)
METAMAT 01: A semi-analytic Solution for Benchmarking Wave Propagation Simulations of homogeneous Absorbers in 1D/3D and 2D
par: Schoder, Stefan, et autres
Publié: (2024)
par: Schoder, Stefan, et autres
Publié: (2024)
Intuitive Control of Scraping and Rubbing Through Audio-tactile Synthesis
par: Aramaki, Mitsuko, et autres
Publié: (2024)
par: Aramaki, Mitsuko, et autres
Publié: (2024)
Correlation and Spectral Density Functions in Mode-Stirred Reverberation -- II. Spectral Moments, Sampling, Noise, EMI and Understirring
par: Arnaut, Luk R., et autres
Publié: (2024)
par: Arnaut, Luk R., et autres
Publié: (2024)
In situ sound absorption estimation with the discrete complex image source method
par: Brandao, Eric, et autres
Publié: (2024)
par: Brandao, Eric, et autres
Publié: (2024)
ECOSoundSet: a finely annotated dataset for the automated acoustic identification of Orthoptera and Cicadidae in North, Central and temperate Western Europe
par: Funosas, David, et autres
Publié: (2025)
par: Funosas, David, et autres
Publié: (2025)
Echoes: A semantically-aligned music deepfake detection dataset
par: Pascu, Octavian, et autres
Publié: (2026)
par: Pascu, Octavian, et autres
Publié: (2026)
On the de-duplication of the Lakh MIDI dataset
par: Choi, Eunjin, et autres
Publié: (2025)
par: Choi, Eunjin, et autres
Publié: (2025)
Some clues to build a sound analysis relevant to hearing
par: Millot, Laurent
Publié: (2024)
par: Millot, Laurent
Publié: (2024)
AS-70: A Mandarin stuttered speech dataset for automatic speech recognition and stuttering event detection
par: Gong, Rong, et autres
Publié: (2024)
par: Gong, Rong, et autres
Publié: (2024)
Fully Reversing the Shoebox Image Source Method: From Impulse Responses to Room Parameters
par: Sprunck, Tom, et autres
Publié: (2024)
par: Sprunck, Tom, et autres
Publié: (2024)
Fitting Auditory Filterbanks with Multiresolution Neural Networks
par: Lostanlen, Vincent, et autres
Publié: (2023)
par: Lostanlen, Vincent, et autres
Publié: (2023)
Developing vocal system impaired patient-aimed voice quality assessment approach using ASR representation-included multiple features
par: Dang, Shaoxiang, et autres
Publié: (2024)
par: Dang, Shaoxiang, et autres
Publié: (2024)
Universality of physical neural networks with multivariate nonlinearity
par: Savinson, Benjamin, et autres
Publié: (2025)
par: Savinson, Benjamin, et autres
Publié: (2025)
From Kepler to Newton: Inductive Biases Guide Learned World Models in Transformers
par: Liu, Ziming, et autres
Publié: (2026)
par: Liu, Ziming, et autres
Publié: (2026)
A hybrid numerical methodology coupling Reduced Order Modeling and Graph Neural Networks for non-parametric geometries: applications to structural dynamics problems
par: Matray, Victor, et autres
Publié: (2024)
par: Matray, Victor, et autres
Publié: (2024)
A new economic and financial theory of money
par: Glinsky, Michael E., et autres
Publié: (2023)
par: Glinsky, Michael E., et autres
Publié: (2023)
The Equalizer: Introducing Shape-Gain Decomposition in Neural Audio Codecs
par: Sadok, Samir, et autres
Publié: (2026)
par: Sadok, Samir, et autres
Publié: (2026)
HHL with a Coherent Fourier Oracle: A Proof-of-Concept Quantum Architecture for Joint Melody-Harmony Generation
par: Kirke, Alexis
Publié: (2026)
par: Kirke, Alexis
Publié: (2026)
Towards measuring fairness in speech recognition: Fair-Speech dataset
par: Veliche, Irina-Elena, et autres
Publié: (2024)
par: Veliche, Irina-Elena, et autres
Publié: (2024)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
par: Schäfer-Zimmermann, Julian C., et autres
Publié: (2024)
par: Schäfer-Zimmermann, Julian C., et autres
Publié: (2024)
Efficient and Fast Generative-Based Singing Voice Separation using a Latent Diffusion Model
par: Plaja-Roglans, Genís, et autres
Publié: (2025)
par: Plaja-Roglans, Genís, et autres
Publié: (2025)
Listen Like a Teacher: Mitigating Whisper Hallucinations using Adaptive Layer Attention and Knowledge Distillation
par: Tripathi, Kumud, et autres
Publié: (2025)
par: Tripathi, Kumud, et autres
Publié: (2025)
Adaptive Knowledge Distillation using a Device-Aware Teacher for Low-Complexity Acoustic Scene Classification
par: Jeong, Seung Gyu, et autres
Publié: (2025)
par: Jeong, Seung Gyu, et autres
Publié: (2025)
Continuous Telemonitoring of Heart Failure using Personalised Speech Dynamics
par: Pan, Yue, et autres
Publié: (2026)
par: Pan, Yue, et autres
Publié: (2026)
Robust Neural Audio Fingerprinting using Music Foundation Models
par: Singh, Shubhr, et autres
Publié: (2025)
par: Singh, Shubhr, et autres
Publié: (2025)
Keyword spotting using convolutional neural network for speech recognition in Hindi
par: Bharti, Saru, et autres
Publié: (2026)
par: Bharti, Saru, et autres
Publié: (2026)
Documents similaires
-
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
par: Tailleur, Modan, et autres
Publié: (2025) -
Towards better visualizations of urban sound environments: insights from interviews
par: Tailleur, Modan, et autres
Publié: (2024) -
Sound Scene Synthesis at the DCASE 2024 Challenge
par: Lagrange, Mathieu, et autres
Publié: (2025) -
Using perceptive subbands analysis to perform audio scenes cartography
par: Millot, Laurent, et autres
Publié: (2024) -
Challenge on Sound Scene Synthesis: Evaluating Text-to-Audio Generation
par: Lee, Junwon, et autres
Publié: (2024)