Saved in:
| Main Authors: | Fiala, David, Pugin, Laurent, van Berchum, Marnix, Thomae, Martha, Roger, Kévin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.15991 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Rest is Silence: Leveraging Unseen Species Models for Computational Musicology
by: Moss, Fabian C., et al.
Published: (2025)
by: Moss, Fabian C., et al.
Published: (2025)
Distilling a speech and music encoder with task arithmetic
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
A correlation-permutation approach for speech-music encoders model merging
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
DDFAD: Dataset Distillation Framework for Audio Data
by: Jiang, Wenbo, et al.
Published: (2024)
by: Jiang, Wenbo, et al.
Published: (2024)
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
by: Peladeau, Côme, et al.
Published: (2023)
by: Peladeau, Côme, et al.
Published: (2023)
Pièces de viole des Cinq Livres and their statistical signatures: the musical work of Marin Marais and Jordi Savall
by: Lugo, Igor, et al.
Published: (2024)
by: Lugo, Igor, et al.
Published: (2024)
Neural Ambisonics encoding for compact irregular microphone arrays
by: Heikkinen, Mikko, et al.
Published: (2024)
by: Heikkinen, Mikko, et al.
Published: (2024)
Effect of laboratory conditions on the perception of virtual stages for music
by: Accolti, Ernesto
Published: (2025)
by: Accolti, Ernesto
Published: (2025)
Musical composition and 2D cellular automata based on music intervals
by: Lugo, Igor, et al.
Published: (2024)
by: Lugo, Igor, et al.
Published: (2024)
Scaling up masked audio encoder learning for general audio classification
by: Dinkel, Heinrich, et al.
Published: (2024)
by: Dinkel, Heinrich, et al.
Published: (2024)
LiveScaler: Live control of the harmony of an electronic music track
by: Rixte, Alice
Published: (2024)
by: Rixte, Alice
Published: (2024)
Investigation of perceptual music similarity focusing on each instrumental part
by: Hashizume, Yuka, et al.
Published: (2025)
by: Hashizume, Yuka, et al.
Published: (2025)
STASE: A spatialized text-to-audio synthesis engine for music generation
by: Chi, Tutti, et al.
Published: (2025)
by: Chi, Tutti, et al.
Published: (2025)
PAGURI: a user experience study of creative interaction with text-to-music models
by: Ronchini, Francesca, et al.
Published: (2024)
by: Ronchini, Francesca, et al.
Published: (2024)
MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling
by: Rouard, Simon, et al.
Published: (2025)
by: Rouard, Simon, et al.
Published: (2025)
Leave-One-EquiVariant: Alleviating invariance-related information loss in contrastive music representations
by: Guinot, Julien, et al.
Published: (2024)
by: Guinot, Julien, et al.
Published: (2024)
The CHiME-7 UDASE task: Unsupervised domain adaptation for conversational speech enhancement
by: Leglaive, Simon, et al.
Published: (2023)
by: Leglaive, Simon, et al.
Published: (2023)
DualSep: A Light-weight dual-encoder convolutional recurrent network for real-time in-car speech separation
by: Wang, Ziqian, et al.
Published: (2024)
by: Wang, Ziqian, et al.
Published: (2024)
Automatic Melody Reduction via Shortest Path Finding
by: Wang, Ziyu, et al.
Published: (2025)
by: Wang, Ziyu, et al.
Published: (2025)
TEAdapter: Supply abundant guidance for controllable text-to-music generation
by: Zou, Jialing, et al.
Published: (2024)
by: Zou, Jialing, et al.
Published: (2024)
musif: a Python package for symbolic music feature extraction
by: Llorens, Ana, et al.
Published: (2023)
by: Llorens, Ana, et al.
Published: (2023)
Exploring compressibility of transformer based text-to-music (TTM) models
by: Moschopoulos, Vasileios, et al.
Published: (2024)
by: Moschopoulos, Vasileios, et al.
Published: (2024)
Some clues to build a sound analysis relevant to hearing
by: Millot, Laurent
Published: (2024)
by: Millot, Laurent
Published: (2024)
Short-term cognitive fatigue of spatial selective attention after face-to-face conversations in virtual noisy environments
by: Hládek, Ľuboš, et al.
Published: (2025)
by: Hládek, Ľuboš, et al.
Published: (2025)
Dance2MIDI: Dance-driven multi-instruments music generation
by: Han, Bo, et al.
Published: (2023)
by: Han, Bo, et al.
Published: (2023)
Machine Unlearning in Speech Emotion Recognition via Forget Set Alone
by: Ren, Zhao, et al.
Published: (2025)
by: Ren, Zhao, et al.
Published: (2025)
An Experiment with Electric Guitar Signals for Exploring the Virtuosity based on the Entropy of Music
by: Lugo, Igor, et al.
Published: (2024)
by: Lugo, Igor, et al.
Published: (2024)
Diff-ETS: Learning a Diffusion Probabilistic Model for Electromyography-to-Speech Conversion
by: Ren, Zhao, et al.
Published: (2024)
by: Ren, Zhao, et al.
Published: (2024)
AdaProj: Adaptively Scaled Angular Margin Subspace Projections for Anomalous Sound Detection with Auxiliary Classification Tasks
by: Wilkinghoff, Kevin
Published: (2024)
by: Wilkinghoff, Kevin
Published: (2024)
Detecting music deepfakes is easy but actually hard
by: Afchar, Darius, et al.
Published: (2024)
by: Afchar, Darius, et al.
Published: (2024)
Long-form music generation with latent diffusion
by: Evans, Zach, et al.
Published: (2024)
by: Evans, Zach, et al.
Published: (2024)
Computational music analysis from first principles
by: Tymoczko, Dmitri, et al.
Published: (2024)
by: Tymoczko, Dmitri, et al.
Published: (2024)
Zipformer: A faster and better encoder for automatic speech recognition
by: Yao, Zengwei, et al.
Published: (2023)
by: Yao, Zengwei, et al.
Published: (2023)
Testing chatbots on the creation of encoders for audio conditioned image generation
by: León, Jorge E., et al.
Published: (2025)
by: León, Jorge E., et al.
Published: (2025)
Codec-Based Deepfake Source Tracing via Neural Audio Codec Taxonomy
by: Chen, Xuanjun, et al.
Published: (2025)
by: Chen, Xuanjun, et al.
Published: (2025)
Audio Dialogues: Dialogues dataset for audio and music understanding
by: Goel, Arushi, et al.
Published: (2024)
by: Goel, Arushi, et al.
Published: (2024)
StemGen: A music generation model that listens
by: Parker, Julian D., et al.
Published: (2023)
by: Parker, Julian D., et al.
Published: (2023)
Echoes: A semantically-aligned music deepfake detection dataset
by: Pascu, Octavian, et al.
Published: (2026)
by: Pascu, Octavian, et al.
Published: (2026)
Learning and composing of classical music using restricted Boltzmann machines
by: Kobayashi, Mutsumi, et al.
Published: (2025)
by: Kobayashi, Mutsumi, et al.
Published: (2025)
Deep learning for music generation. Four approaches and their comparative evaluation
by: Paroiu, Razvan, et al.
Published: (2025)
by: Paroiu, Razvan, et al.
Published: (2025)
Similar Items
-
The Rest is Silence: Leveraging Unseen Species Models for Computational Musicology
by: Moss, Fabian C., et al.
Published: (2025) -
Distilling a speech and music encoder with task arithmetic
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025) -
A correlation-permutation approach for speech-music encoders model merging
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025) -
DDFAD: Dataset Distillation Framework for Audio Data
by: Jiang, Wenbo, et al.
Published: (2024) -
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
by: Peladeau, Côme, et al.
Published: (2023)