The Rest is Silence: Leveraging Unseen Species Models for Computational Musicology
Fuente:
arXiv
Saved in:
| Main Authors: | Moss, Fabian C., Hajič jr., Jan, Nachtwey, Adrian, Pugin, Laurent |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Modeling the Difficulty of Saxophone Music
by: Libřický, Šimon, et al.
Published: (2025)
by: Libřický, Šimon, et al.
Published: (2025)
Complexity of frequency fluctuations and the interpretive style in the bass viola da gamba
by: Lugo, Igor, et al.
Published: (2025)
by: Lugo, Igor, et al.
Published: (2025)
Generative AI-based data augmentation for improved bioacoustic classification in noisy environments
by: Gibbons, Anthony, et al.
Published: (2024)
by: Gibbons, Anthony, et al.
Published: (2024)
Pièces de viole des Cinq Livres and their statistical signatures: the musical work of Marin Marais and Jordi Savall
by: Lugo, Igor, et al.
Published: (2024)
by: Lugo, Igor, et al.
Published: (2024)
Deep functional multiple index models with an application to SER
by: Saumard, Matthieu, et al.
Published: (2024)
by: Saumard, Matthieu, et al.
Published: (2024)
A new XML conversion process for mensural music encoding : CMME\_to\_MEI (via Verovio)
by: Fiala, David, et al.
Published: (2025)
by: Fiala, David, et al.
Published: (2025)
Multi-Representation Attention Framework for Underwater Bioacoustic Denoising and Recognition
by: Razig, Amine, et al.
Published: (2025)
by: Razig, Amine, et al.
Published: (2025)
Musical composition and 2D cellular automata based on music intervals
by: Lugo, Igor, et al.
Published: (2024)
by: Lugo, Igor, et al.
Published: (2024)
Dirichlet process mixture model based on topologically augmented signal representation for clustering infant vocalizations
by: Bonafos, Guillem, et al.
Published: (2024)
by: Bonafos, Guillem, et al.
Published: (2024)
Bayesian Restoration of Audio Degraded by Low-Frequency Pulses Modeled via Gaussian Process
by: de Carvalho, Hugo Tremonte, et al.
Published: (2020)
by: de Carvalho, Hugo Tremonte, et al.
Published: (2020)
Unseen but not Unknown: Using Dataset Concealment to Robustly Evaluate Speech Quality Estimation Models
by: Pieper, Jaden, et al.
Published: (2026)
by: Pieper, Jaden, et al.
Published: (2026)
CLaMP 3: Universal Music Information Retrieval Across Unaligned Modalities and Unseen Languages
by: Wu, Shangda, et al.
Published: (2025)
by: Wu, Shangda, et al.
Published: (2025)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
by: Farhadipour, Aref, et al.
Published: (2024)
by: Farhadipour, Aref, et al.
Published: (2024)
Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model
by: Ren, Yong, et al.
Published: (2025)
by: Ren, Yong, et al.
Published: (2025)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
by: Zhang, Xueyao, et al.
Published: (2023)
by: Zhang, Xueyao, et al.
Published: (2023)
BrainWhisperer: Leveraging Large-Scale ASR Models for Neural Speech Decoding
by: Boccato, Tommaso, et al.
Published: (2026)
by: Boccato, Tommaso, et al.
Published: (2026)
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
by: Shi, Yu-Fei, et al.
Published: (2024)
by: Shi, Yu-Fei, et al.
Published: (2024)
Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis
by: Liao, Shijia, et al.
Published: (2024)
by: Liao, Shijia, et al.
Published: (2024)
AeroGPT: Leveraging Large-Scale Audio Model for Aero-Engine Bearing Fault Diagnosis
by: Liu, Jiale, et al.
Published: (2025)
by: Liu, Jiale, et al.
Published: (2025)
Leveraging Real Electric Guitar Tones and Effects to Improve Robustness in Guitar Tablature Transcription Modeling
by: Pedroza, Hegel, et al.
Published: (2024)
by: Pedroza, Hegel, et al.
Published: (2024)
Audio-Based Classification of Insect Species Using Machine Learning Models: Cicada, Beetle, Termite, and Cricket
by: Shetty, Manas V, et al.
Published: (2025)
by: Shetty, Manas V, et al.
Published: (2025)
Explainable anomaly detection for sound spectrograms using pooling statistics with quantile differences
by: Thewes, Nicolas, et al.
Published: (2025)
by: Thewes, Nicolas, et al.
Published: (2025)
Some clues to build a sound analysis relevant to hearing
by: Millot, Laurent
Published: (2024)
by: Millot, Laurent
Published: (2024)
Leveraging Language Information for Target Language Extraction
by: Yıldırım, Mehmet Sinan, et al.
Published: (2025)
by: Yıldırım, Mehmet Sinan, et al.
Published: (2025)
Leveraging Self-Supervised Learning for Speaker Diarization
by: Han, Jiangyu, et al.
Published: (2024)
by: Han, Jiangyu, et al.
Published: (2024)
Leveraging Sound Source Trajectories for Universal Sound Separation
by: Wu, Donghang, et al.
Published: (2024)
by: Wu, Donghang, et al.
Published: (2024)
MACE: Leveraging Audio for Evaluating Audio Captioning Systems
by: Dixit, Satvik, et al.
Published: (2024)
by: Dixit, Satvik, et al.
Published: (2024)
Cross-Dialect Bird Species Recognition with Dialect-Calibrated Augmentation
by: Ding, Jiani, et al.
Published: (2025)
by: Ding, Jiani, et al.
Published: (2025)
Leveraging AM and FM Rhythm Spectrograms for Dementia Classification and Assessment
by: Gogoi, Parismita, et al.
Published: (2025)
by: Gogoi, Parismita, et al.
Published: (2025)
Leveraging Prompt Learning and Pause Encoding for Alzheimer's Disease Detection
by: Liu, Yin-Long, et al.
Published: (2024)
by: Liu, Yin-Long, et al.
Published: (2024)
Leveraging Multimodal Methods and Spontaneous Speech for Alzheimer's Disease Identification
by: Gao, Yifan, et al.
Published: (2024)
by: Gao, Yifan, et al.
Published: (2024)
Self-Supervised Speech Quality Assessment (S3QA): Leveraging Speech Foundation Models for a Scalable Speech Quality Metric
by: Ogg, Mattson, et al.
Published: (2025)
by: Ogg, Mattson, et al.
Published: (2025)
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
by: Eisenberg, Aviad, et al.
Published: (2025)
by: Eisenberg, Aviad, et al.
Published: (2025)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
by: Chao, Rong, et al.
Published: (2025)
by: Chao, Rong, et al.
Published: (2025)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
by: Rimon, Inbal, et al.
Published: (2025)
by: Rimon, Inbal, et al.
Published: (2025)
TF-CorrNet: Leveraging Spatial Correlation for Continuous Speech Separation
by: Shin, Ui-Hyeop, et al.
Published: (2025)
by: Shin, Ui-Hyeop, et al.
Published: (2025)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
by: Poncelet, Jakob, et al.
Published: (2025)
by: Poncelet, Jakob, et al.
Published: (2025)
Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
by: Saijo, Kohei, et al.
Published: (2024)
by: Saijo, Kohei, et al.
Published: (2024)
Leveraging Joint Spectral and Spatial Learning with MAMBA for Multichannel Speech Enhancement
by: Ren, Wenze, et al.
Published: (2024)
by: Ren, Wenze, et al.
Published: (2024)
Similar Items
-
Modeling the Difficulty of Saxophone Music
by: Libřický, Šimon, et al.
Published: (2025) -
Complexity of frequency fluctuations and the interpretive style in the bass viola da gamba
by: Lugo, Igor, et al.
Published: (2025) -
Generative AI-based data augmentation for improved bioacoustic classification in noisy environments
by: Gibbons, Anthony, et al.
Published: (2024) -
Pièces de viole des Cinq Livres and their statistical signatures: the musical work of Marin Marais and Jordi Savall
by: Lugo, Igor, et al.
Published: (2024) -
Deep functional multiple index models with an application to SER
by: Saumard, Matthieu, et al.
Published: (2024)