Aligned Contrastive Predictive Coding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chorowski, Jan, Ciesielski, Grzegorz, Dzikowski, Jarosław, Łańcucki, Adrian, Marxer, Ricard, Opala, Mateusz, Pusz, Piotr, Rychlikowski, Paweł, Stypułkowski, Michał |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Information Retrieval for ZeroSpeech 2021: The Submission by University of Wroclaw
von: Chorowski, Jan, et al.
Veröffentlicht: (2021)
von: Chorowski, Jan, et al.
Veröffentlicht: (2021)
Contrastive prediction strategies for unsupervised segmentation and categorization of phonemes and words
von: Cuervo, Santiago, et al.
Veröffentlicht: (2021)
von: Cuervo, Santiago, et al.
Veröffentlicht: (2021)
Transfer Learning from Whisper for Microscopic Intelligibility Prediction
von: Best, Paul, et al.
Veröffentlicht: (2024)
von: Best, Paul, et al.
Veröffentlicht: (2024)
Speech foundation models on intelligibility prediction for hearing-impaired listeners
von: Cuervo, Santiago, et al.
Veröffentlicht: (2024)
von: Cuervo, Santiago, et al.
Veröffentlicht: (2024)
SCRAPS: Speech Contrastive Representations of Acoustic and Phonetic Spaces
von: Vallés-Pérez, Ivan, et al.
Veröffentlicht: (2023)
von: Vallés-Pérez, Ivan, et al.
Veröffentlicht: (2023)
Clustering-based hard negative sampling for supervised contrastive speaker verification
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
von: Masztalski, Piotr, et al.
Veröffentlicht: (2025)
Emotion-Aligned Contrastive Learning Between Images and Music
von: Stewart, Shanti, et al.
Veröffentlicht: (2023)
von: Stewart, Shanti, et al.
Veröffentlicht: (2023)
SCOREQ: Speech Quality Assessment with Contrastive Regression
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
Symbotunes: unified hub for symbolic music generative models
von: Skierś, Paweł, et al.
Veröffentlicht: (2024)
von: Skierś, Paweł, et al.
Veröffentlicht: (2024)
MidiTok Visualizer: a tool for visualization and analysis of tokenized MIDI symbolic music
von: Wiszenko, Michał, et al.
Veröffentlicht: (2024)
von: Wiszenko, Michał, et al.
Veröffentlicht: (2024)
SOI: Scaling Down Computational Complexity by Estimating Partial States of the Model
von: Stefański, Grzegorz, et al.
Veröffentlicht: (2024)
von: Stefański, Grzegorz, et al.
Veröffentlicht: (2024)
Are audio DeepFake detection models polyglots?
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech
von: Nakazawa, Kazushi
Veröffentlicht: (2026)
von: Nakazawa, Kazushi
Veröffentlicht: (2026)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
Deep learning based spatial aliasing reduction in beamforming for audio capture
von: Guzik, Mateusz, et al.
Veröffentlicht: (2025)
von: Guzik, Mateusz, et al.
Veröffentlicht: (2025)
SGPA: Spectrogram-Guided Phonetic Alignment for Feasible Shapley Value Explanations in Multimodal Large Language Models
von: Pozorski, Paweł, et al.
Veröffentlicht: (2026)
von: Pozorski, Paweł, et al.
Veröffentlicht: (2026)
Quality of Automatic Speech Recognition -- Polish Language case study -- from Wav2Vec to Scribe ElevenLabs
von: Pietroń, Marcin, et al.
Veröffentlicht: (2026)
von: Pietroń, Marcin, et al.
Veröffentlicht: (2026)
CVSM: Contrastive Vocal Similarity Modeling
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
von: Garoufis, Christos, et al.
Veröffentlicht: (2025)
Short-term cognitive fatigue of spatial selective attention after face-to-face conversations in virtual noisy environments
von: Hládek, Ľuboš, et al.
Veröffentlicht: (2025)
von: Hládek, Ľuboš, et al.
Veröffentlicht: (2025)
Cacophony: An Improved Contrastive Audio-Text Model
von: Zhu, Ge, et al.
Veröffentlicht: (2024)
von: Zhu, Ge, et al.
Veröffentlicht: (2024)
Domain Adaptation for Contrastive Audio-Language Models
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
Noise-Aware Speech Separation with Contrastive Learning
von: Zhang, Zizheng, et al.
Veröffentlicht: (2023)
von: Zhang, Zizheng, et al.
Veröffentlicht: (2023)
Speaker Contrastive Learning for Source Speaker Tracing
von: Wang, Qing, et al.
Veröffentlicht: (2024)
von: Wang, Qing, et al.
Veröffentlicht: (2024)
URGENT-PK: Perceptually-Aligned Ranking Model Designed for Speech Enhancement Competition
von: Wang, Jiahe, et al.
Veröffentlicht: (2025)
von: Wang, Jiahe, et al.
Veröffentlicht: (2025)
SegAug: CTC-Aligned Segmented Augmentation For Robust RNN-Transducer Based Speech Recognition
von: Le, Khanh, et al.
Veröffentlicht: (2025)
von: Le, Khanh, et al.
Veröffentlicht: (2025)
BrainWhisperer: Leveraging Large-Scale ASR Models for Neural Speech Decoding
von: Boccato, Tommaso, et al.
Veröffentlicht: (2026)
von: Boccato, Tommaso, et al.
Veröffentlicht: (2026)
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
FGCL: Fine-grained Contrastive Learning For Mandarin Stuttering Event Detection
von: Jiang, Han, et al.
Veröffentlicht: (2024)
von: Jiang, Han, et al.
Veröffentlicht: (2024)
Temporally Heterogeneous Graph Contrastive Learning for Multimodal Acoustic event Classification
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
von: Chen, Yuanjian, et al.
Veröffentlicht: (2025)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Learning Disentangled Speech Representations with Contrastive Learning and Time-Invariant Retrieval
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2024)
Semantic Proximity Alignment: Towards Human Perception-consistent Audio Tagging by Aligning with Label Text Description
von: Liu, Wuyang, et al.
Veröffentlicht: (2023)
von: Liu, Wuyang, et al.
Veröffentlicht: (2023)
Phoneme-Level Contrastive Learning for User-Defined Keyword Spotting with Flexible Enrollment
von: Kewei, Li, et al.
Veröffentlicht: (2024)
von: Kewei, Li, et al.
Veröffentlicht: (2024)
Noise-Robust Contrastive Learning with an MFCC-Conformer For Coronary Artery Disease Detection
von: Marocchi, Milan, et al.
Veröffentlicht: (2026)
von: Marocchi, Milan, et al.
Veröffentlicht: (2026)
Boosting Multi-Speaker Expressive Speech Synthesis with Semi-supervised Contrastive Learning
von: Zhu, Xinfa, et al.
Veröffentlicht: (2023)
von: Zhu, Xinfa, et al.
Veröffentlicht: (2023)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
von: Guan, Yadong, et al.
Veröffentlicht: (2024)
Combining Masked Language Modeling and Cross-Modal Contrastive Learning for Prosody-Aware TTS
von: Borodin, Kirill, et al.
Veröffentlicht: (2026)
von: Borodin, Kirill, et al.
Veröffentlicht: (2026)
SmoothCLAP: Soft-Target Enhanced Contrastive Language\--Audio Pretraining for Affective Computing
von: Jing, Xin, et al.
Veröffentlicht: (2026)
von: Jing, Xin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Information Retrieval for ZeroSpeech 2021: The Submission by University of Wroclaw
von: Chorowski, Jan, et al.
Veröffentlicht: (2021) -
Contrastive prediction strategies for unsupervised segmentation and categorization of phonemes and words
von: Cuervo, Santiago, et al.
Veröffentlicht: (2021) -
Transfer Learning from Whisper for Microscopic Intelligibility Prediction
von: Best, Paul, et al.
Veröffentlicht: (2024) -
Speech foundation models on intelligibility prediction for hearing-impaired listeners
von: Cuervo, Santiago, et al.
Veröffentlicht: (2024) -
SCRAPS: Speech Contrastive Representations of Acoustic and Phonetic Spaces
von: Vallés-Pérez, Ivan, et al.
Veröffentlicht: (2023)