Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
Fuente:
arXiv
Salvato in:
| Autori principali: | Sechaud, Victor, Jacques, Laurent, Abry, Patrice, Tachella, Julián |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
di: Oncescu, Andreea-Maria, et al.
Pubblicazione: (2024)
di: Oncescu, Andreea-Maria, et al.
Pubblicazione: (2024)
SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
Learning to reconstruct from saturated data: audio declipping and high-dynamic range imaging
di: Sechaud, Victor, et al.
Pubblicazione: (2026)
di: Sechaud, Victor, et al.
Pubblicazione: (2026)
Emergent musical properties of a transformer under contrastive self-supervised learning
di: Kong, Yuexuan, et al.
Pubblicazione: (2025)
di: Kong, Yuexuan, et al.
Pubblicazione: (2025)
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
di: Voran, Stephen D.
Pubblicazione: (2024)
di: Voran, Stephen D.
Pubblicazione: (2024)
Deep learning classification system for coconut maturity levels based on acoustic signals
di: Caladcad, June Anne, et al.
Pubblicazione: (2024)
di: Caladcad, June Anne, et al.
Pubblicazione: (2024)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
di: Kim, Haven, et al.
Pubblicazione: (2026)
di: Kim, Haven, et al.
Pubblicazione: (2026)
Compositional nonlinear audio signal processing with Volterra series
di: Araujo-Simon, Jake
Pubblicazione: (2023)
di: Araujo-Simon, Jake
Pubblicazione: (2023)
Self-supervised learning for phase retrieval
di: Sechaud, Victor, et al.
Pubblicazione: (2025)
di: Sechaud, Victor, et al.
Pubblicazione: (2025)
Language-based Audio Retrieval with Co-Attention Networks
di: Sun, Haoran, et al.
Pubblicazione: (2024)
di: Sun, Haoran, et al.
Pubblicazione: (2024)
AxLSTMs: learning self-supervised audio representations with xLSTMs
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
di: Xin, Yifei, et al.
Pubblicazione: (2024)
di: Xin, Yifei, et al.
Pubblicazione: (2024)
Evaluating Interval-based Tokenization for Pitch Representation in Symbolic Music Analysis
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2025)
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2025)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
di: Wang, Ziqian, et al.
Pubblicazione: (2025)
di: Wang, Ziqian, et al.
Pubblicazione: (2025)
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Using perceptive subbands analysis to perform audio scenes cartography
di: Millot, Laurent, et al.
Pubblicazione: (2024)
di: Millot, Laurent, et al.
Pubblicazione: (2024)
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
di: Rackauckas, Zackary, et al.
Pubblicazione: (2025)
di: Rackauckas, Zackary, et al.
Pubblicazione: (2025)
Exploring Diverse Sounds: Identifying Outliers in a Music Corpus
di: Cai, Le, et al.
Pubblicazione: (2024)
di: Cai, Le, et al.
Pubblicazione: (2024)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
di: Doh, SeungHeon, et al.
Pubblicazione: (2024)
di: Doh, SeungHeon, et al.
Pubblicazione: (2024)
Track Role Prediction of Single-Instrumental Sequences
di: Han, Changheon, et al.
Pubblicazione: (2024)
di: Han, Changheon, et al.
Pubblicazione: (2024)
Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning
di: Zhang, Dengming, et al.
Pubblicazione: (2024)
di: Zhang, Dengming, et al.
Pubblicazione: (2024)
LARP: Language Audio Relational Pre-training for Cold-Start Playlist Continuation
di: Salganik, Rebecca, et al.
Pubblicazione: (2024)
di: Salganik, Rebecca, et al.
Pubblicazione: (2024)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
di: Chowdhury, Shreyan, et al.
Pubblicazione: (2024)
di: Chowdhury, Shreyan, et al.
Pubblicazione: (2024)
Exploring GPT's Ability as a Judge in Music Understanding
di: Fang, Kun, et al.
Pubblicazione: (2025)
di: Fang, Kun, et al.
Pubblicazione: (2025)
Multi-Sample Dynamic Time Warping for Few-Shot Keyword Spotting
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2024)
di: Wilkinghoff, Kevin, et al.
Pubblicazione: (2024)
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
di: Wang, Qian, et al.
Pubblicazione: (2024)
di: Wang, Qian, et al.
Pubblicazione: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
di: Lee, Jinwoo, et al.
Pubblicazione: (2024)
di: Lee, Jinwoo, et al.
Pubblicazione: (2024)
Automatic Estimation of Singing Voice Musical Dynamics
di: Narang, Jyoti, et al.
Pubblicazione: (2024)
di: Narang, Jyoti, et al.
Pubblicazione: (2024)
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
di: Pan, Xiaofeng, et al.
Pubblicazione: (2025)
di: Pan, Xiaofeng, et al.
Pubblicazione: (2025)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
di: Doh, Seungheon, et al.
Pubblicazione: (2025)
di: Doh, Seungheon, et al.
Pubblicazione: (2025)
Evaluating High-Resolution Piano Sustain Pedal Depth Estimation with Musically Informed Metrics
di: Zhang, Hanwen, et al.
Pubblicazione: (2025)
di: Zhang, Hanwen, et al.
Pubblicazione: (2025)
Engraving Oriented Joint Estimation of Pitch Spelling and Local and Global Keys
di: Bouquillard, Augustin, et al.
Pubblicazione: (2024)
di: Bouquillard, Augustin, et al.
Pubblicazione: (2024)
Towards Computational Analysis of Pansori Singing
di: Park, Sangheon, et al.
Pubblicazione: (2024)
di: Park, Sangheon, et al.
Pubblicazione: (2024)
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
di: Premananth, Gowtham, et al.
Pubblicazione: (2024)
Analytical model for the relation between signal bandwidth and spatial resolution in Steered-Response Power Phase Transform (SRP-PHAT) maps
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
di: Garcia-Barrios, Guillermo, et al.
Pubblicazione: (2024)
Audio signal interpolation using optimal transportation of spectrograms
di: Valdivia, David, et al.
Pubblicazione: (2025)
di: Valdivia, David, et al.
Pubblicazione: (2025)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
di: Wang, Bo, et al.
Pubblicazione: (2024)
di: Wang, Bo, et al.
Pubblicazione: (2024)
Zema Dataset: A Comprehensive Study of Yaredawi Zema with a Focus on Horologium Chants
di: Muluneh, Mequanent Argaw, et al.
Pubblicazione: (2024)
di: Muluneh, Mequanent Argaw, et al.
Pubblicazione: (2024)
Deep learning-based filtering of cross-spectral matrices using generative adversarial networks
di: Puhle, Christof
Pubblicazione: (2025)
di: Puhle, Christof
Pubblicazione: (2025)
Benchmarking multi-component signal processing methods in the time-frequency plane
di: Miramont, Juan M., et al.
Pubblicazione: (2024)
di: Miramont, Juan M., et al.
Pubblicazione: (2024)
Documenti analoghi
-
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
di: Oncescu, Andreea-Maria, et al.
Pubblicazione: (2024) -
SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription
di: Zang, Yongyi, et al.
Pubblicazione: (2023) -
Learning to reconstruct from saturated data: audio declipping and high-dynamic range imaging
di: Sechaud, Victor, et al.
Pubblicazione: (2026) -
Emergent musical properties of a transformer under contrastive self-supervised learning
di: Kong, Yuexuan, et al.
Pubblicazione: (2025) -
Why some audio signal short-time Fourier transform coefficients have nonuniform phase distributions
di: Voran, Stephen D.
Pubblicazione: (2024)