S-KEY: Self-supervised Learning of Major and Minor Keys from Audio
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kong, Yuexuan, Meseguer-Brocal, Gabriel, Lostanlen, Vincent, Lagrange, Mathieu, Hennequin, Romain |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STONE: Self-supervised Tonality Estimator
von: Kong, Yuexuan, et al.
Veröffentlicht: (2024)
von: Kong, Yuexuan, et al.
Veröffentlicht: (2024)
Emergent musical properties of a transformer under contrastive self-supervised learning
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
An Experimental Comparison Of Multi-view Self-supervised Methods For Music Tagging
von: Meseguer-Brocal, Gabriel, et al.
Veröffentlicht: (2024)
von: Meseguer-Brocal, Gabriel, et al.
Veröffentlicht: (2024)
AI-Generated Music Detection and its Challenges
von: Afchar, Darius, et al.
Veröffentlicht: (2025)
von: Afchar, Darius, et al.
Veröffentlicht: (2025)
Learning to Solve Inverse Problems for Perceptual Sound Matching
von: Han, Han, et al.
Veröffentlicht: (2023)
von: Han, Han, et al.
Veröffentlicht: (2023)
Detecting music deepfakes is easy but actually hard
von: Afchar, Darius, et al.
Veröffentlicht: (2024)
von: Afchar, Darius, et al.
Veröffentlicht: (2024)
Multi-Class-Token Transformer for Multitask Self-supervised Music Information Retrieval
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
STraDa: A Singer Traits Dataset
von: Kong, Yuexuan, et al.
Veröffentlicht: (2024)
von: Kong, Yuexuan, et al.
Veröffentlicht: (2024)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
von: Mitcheltree, Christopher, et al.
Veröffentlicht: (2026)
von: Mitcheltree, Christopher, et al.
Veröffentlicht: (2026)
From Real to Cloned Singer Identification
von: Desblancs, Dorian, et al.
Veröffentlicht: (2024)
von: Desblancs, Dorian, et al.
Veröffentlicht: (2024)
Learning Linearity in Audio Consistency Autoencoders via Implicit Regularization
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
Musical Metamerism with Time--Frequency Scattering
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2026)
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2026)
Detection of Deepfake Environmental Audio
von: Ouajdi, Hafsa, et al.
Veröffentlicht: (2024)
von: Ouajdi, Hafsa, et al.
Veröffentlicht: (2024)
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
von: Tailleur, Modan, et al.
Veröffentlicht: (2025)
von: Tailleur, Modan, et al.
Veröffentlicht: (2025)
Instabilities in Convnets for Raw Audio
von: Haider, Daniel, et al.
Veröffentlicht: (2023)
von: Haider, Daniel, et al.
Veröffentlicht: (2023)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
Double Entendre: Robust Audio-Based AI-Generated Lyrics Detection via Multi-View Fusion
von: Frohmann, Markus, et al.
Veröffentlicht: (2025)
von: Frohmann, Markus, et al.
Veröffentlicht: (2025)
Mixture of Mixups for Multi-label Classification of Rare Anuran Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
Towards better visualizations of urban sound environments: insights from interviews
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
Fitting Auditory Filterbanks with Multiresolution Neural Networks
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
von: Cai, Yiqiang, et al.
Veröffentlicht: (2024)
von: Cai, Yiqiang, et al.
Veröffentlicht: (2024)
Angular Distance Distribution Loss for Audio Classification
von: Almudévar, Antonio, et al.
Veröffentlicht: (2024)
von: Almudévar, Antonio, et al.
Veröffentlicht: (2024)
WAKE: Watermarking Audio with Key Enrichment
von: Xu, Yaoxun, et al.
Veröffentlicht: (2025)
von: Xu, Yaoxun, et al.
Veröffentlicht: (2025)
AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining
von: Liu, Haohe, et al.
Veröffentlicht: (2023)
von: Liu, Haohe, et al.
Veröffentlicht: (2023)
Self-supervised Reflective Learning through Self-distillation and Online Clustering for Speaker Representation Learning
von: Cai, Danwei, et al.
Veröffentlicht: (2024)
von: Cai, Danwei, et al.
Veröffentlicht: (2024)
Probing Self-supervised Learning Models with Target Speech Extraction
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
Target Speech Extraction with Pre-trained Self-supervised Learning Models
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
SCDNet: Self-supervised Learning Feature-based Speaker Change Detection
von: Li, Yue, et al.
Veröffentlicht: (2024)
von: Li, Yue, et al.
Veröffentlicht: (2024)
Latent Watermarking of Audio Generative Models
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
Adapter Incremental Continual Learning of Efficient Audio Spectrogram Transformers
von: Selvaraj, Nithish Muthuchamy, et al.
Veröffentlicht: (2023)
von: Selvaraj, Nithish Muthuchamy, et al.
Veröffentlicht: (2023)
Selection of Layers from Self-supervised Learning Models for Predicting Mean-Opinion-Score of Speech
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
Hold Me Tight: Stable Encoder-Decoder Design for Speech Enhancement
von: Haider, Daniel, et al.
Veröffentlicht: (2024)
von: Haider, Daniel, et al.
Veröffentlicht: (2024)
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
von: Wilkins, Julia, et al.
Veröffentlicht: (2024)
von: Wilkins, Julia, et al.
Veröffentlicht: (2024)
Robust Lossy Audio Compression Identification
von: Koops, Hendrik Vincent, et al.
Veröffentlicht: (2024)
von: Koops, Hendrik Vincent, et al.
Veröffentlicht: (2024)
SemanticAudio: Audio Generation and Editing in Semantic Space
von: Dai, Zheqi, et al.
Veröffentlicht: (2026)
von: Dai, Zheqi, et al.
Veröffentlicht: (2026)
Causal Speech Enhancement with Predicting Semantics based on Quantized Self-supervised Learning Features
von: Tsunoo, Emiru, et al.
Veröffentlicht: (2024)
von: Tsunoo, Emiru, et al.
Veröffentlicht: (2024)
MMM: Multi-Layer Multi-Residual Multi-Stream Discrete Speech Representation from Self-supervised Learning Model
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
PESTO: Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2023)
von: Riou, Alain, et al.
Veröffentlicht: (2023)
Compositional Audio Representation Learning
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2024)
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2024)
Region-Specific Audio Tagging for Spatial Sound
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
STONE: Self-supervised Tonality Estimator
von: Kong, Yuexuan, et al.
Veröffentlicht: (2024) -
Emergent musical properties of a transformer under contrastive self-supervised learning
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025) -
An Experimental Comparison Of Multi-view Self-supervised Methods For Music Tagging
von: Meseguer-Brocal, Gabriel, et al.
Veröffentlicht: (2024) -
AI-Generated Music Detection and its Challenges
von: Afchar, Darius, et al.
Veröffentlicht: (2025) -
Learning to Solve Inverse Problems for Perceptual Sound Matching
von: Han, Han, et al.
Veröffentlicht: (2023)