Phonetic and Lexical Discovery of a Canine Language using HuBERT
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Xingyuan, Wang, Sinong, Xie, Zeyu, Wu, Mengyue, Zhu, Kenny Q. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MelHuBERT: A simplified HuBERT on Mel spectrograms
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2022)
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2022)
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition
von: Yoon, Ji Won, et al.
Veröffentlicht: (2022)
von: Yoon, Ji Won, et al.
Veröffentlicht: (2022)
Self-Supervised Syllable Discovery Based on Speaker-Disentangled HuBERT
von: Komatsu, Ryota, et al.
Veröffentlicht: (2024)
von: Komatsu, Ryota, et al.
Veröffentlicht: (2024)
mHuBERT-147: A Compact Multilingual HuBERT Model
von: Boito, Marcely Zanon, et al.
Veröffentlicht: (2024)
von: Boito, Marcely Zanon, et al.
Veröffentlicht: (2024)
DiceHuBERT: Distilling HuBERT with a Self-Supervised Learning Objective
von: Chi, Hyung Gun, et al.
Veröffentlicht: (2025)
von: Chi, Hyung Gun, et al.
Veröffentlicht: (2025)
Target Speech Extraction with Pre-trained AV-HuBERT and Mask-And-Recover Strategy
von: Wu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wu, Wenxuan, et al.
Veröffentlicht: (2024)
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
von: Huo, Robin, et al.
Veröffentlicht: (2025)
von: Huo, Robin, et al.
Veröffentlicht: (2025)
Exploiting Audio-Visual Features with Pretrained AV-HuBERT for Multi-Modal Dysarthric Speech Reconstruction
von: Chen, Xueyuan, et al.
Veröffentlicht: (2024)
von: Chen, Xueyuan, et al.
Veröffentlicht: (2024)
RobustSVC: HuBERT-based Melody Extractor and Adversarial Learning for Robust Singing Voice Conversion
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
Speaker Emotion Recognition: Leveraging Self-Supervised Models for Feature Extraction Using Wav2Vec2 and HuBERT
von: Jafarzadeh, Pourya, et al.
Veröffentlicht: (2024)
von: Jafarzadeh, Pourya, et al.
Veröffentlicht: (2024)
Multi-resolution HuBERT: Multi-resolution Speech Self-Supervised Learning with Masked Unit Prediction
von: Shi, Jiatong, et al.
Veröffentlicht: (2023)
von: Shi, Jiatong, et al.
Veröffentlicht: (2023)
SD-HuBERT: Sentence-Level Self-Distillation Induces Syllabic Organization in HuBERT
von: Cho, Cheol Jun, et al.
Veröffentlicht: (2023)
von: Cho, Cheol Jun, et al.
Veröffentlicht: (2023)
BLAT: Bootstrapping Language-Audio Pre-training based on AudioSet Tag-guided Synthetic Data
von: Xu, Xuenan, et al.
Veröffentlicht: (2023)
von: Xu, Xuenan, et al.
Veröffentlicht: (2023)
PAST: Phonetic-Acoustic Speech Tokenizer
von: Har-Tuv, Nadav, et al.
Veröffentlicht: (2025)
von: Har-Tuv, Nadav, et al.
Veröffentlicht: (2025)
VisG AV-HuBERT: Viseme-Guided AV-HuBERT
von: Papadopoulos, Aristeidis, et al.
Veröffentlicht: (2026)
von: Papadopoulos, Aristeidis, et al.
Veröffentlicht: (2026)
A Technique for Isolating Lexically-Independent Phonetic Dependencies in Generative CNNs
von: Šegedin, Bruno Ferenc
Veröffentlicht: (2025)
von: Šegedin, Bruno Ferenc
Veröffentlicht: (2025)
BabyHuBERT: Multilingual Self-Supervised Learning for Segmenting Speakers in Child-Centered Long-Form Recordings
von: Charlot, Théo, et al.
Veröffentlicht: (2025)
von: Charlot, Théo, et al.
Veröffentlicht: (2025)
ISPA: Inter-Species Phonetic Alphabet for Transcribing Animal Sounds
von: Hagiwara, Masato, et al.
Veröffentlicht: (2024)
von: Hagiwara, Masato, et al.
Veröffentlicht: (2024)
Self-Supervised Speech Models Encode Phonetic Context via Position-dependent Orthogonal Subspaces
von: Choi, Kwanghee, et al.
Veröffentlicht: (2026)
von: Choi, Kwanghee, et al.
Veröffentlicht: (2026)
The ART of Conversation: Measuring Phonetic Convergence and Deliberate Imitation in L2-Speech with a Siamese RNN
von: Yuan, Zheng, et al.
Veröffentlicht: (2023)
von: Yuan, Zheng, et al.
Veröffentlicht: (2023)
The Mason-Alberta Phonetic Segmenter: A forced alignment system based on deep neural networks and interpolation
von: Kelley, Matthew C., et al.
Veröffentlicht: (2023)
von: Kelley, Matthew C., et al.
Veröffentlicht: (2023)
AfriHuBERT: A self-supervised speech representation model for African languages
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2024)
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2024)
Phonetic Segmentation of the UCLA Phonetics Lab Archive
von: Chodroff, Eleanor, et al.
Veröffentlicht: (2024)
von: Chodroff, Eleanor, et al.
Veröffentlicht: (2024)
How Does a Deep Neural Network Look at Lexical Stress in English Words?
von: Allouche, Itai, et al.
Veröffentlicht: (2025)
von: Allouche, Itai, et al.
Veröffentlicht: (2025)
HuBERT-VIC: Improving Noise-Robust Automatic Speech Recognition of Speech Foundation Model via Variance-Invariance-Covariance Regularization
von: Ahn, Hyebin, et al.
Veröffentlicht: (2025)
von: Ahn, Hyebin, et al.
Veröffentlicht: (2025)
AV-Lip-Sync+: Leveraging AV-HuBERT to Exploit Multimodal Inconsistency for Deepfake Detection of Frontal Face Videos
von: Shahzad, Sahibzada Adil, et al.
Veröffentlicht: (2023)
von: Shahzad, Sahibzada Adil, et al.
Veröffentlicht: (2023)
Phonetic Enhanced Language Modeling for Text-to-Speech Synthesis
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
Enhance Temporal Relations in Audio Captioning with Sound Event Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
von: Xie, Zeyu, et al.
Veröffentlicht: (2023)
A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance
von: Park, Jiyun, et al.
Veröffentlicht: (2024)
von: Park, Jiyun, et al.
Veröffentlicht: (2024)
GTR-Voice: Articulatory Phonetics Informed Controllable Expressive Speech Synthesis
von: Li, Zehua Kcriss, et al.
Veröffentlicht: (2024)
von: Li, Zehua Kcriss, et al.
Veröffentlicht: (2024)
BERT-LID: Leveraging BERT to Improve Spoken Language Identification
von: Nie, Yuting, et al.
Veröffentlicht: (2022)
von: Nie, Yuting, et al.
Veröffentlicht: (2022)
A Detailed Audio-Text Data Simulation Pipeline using Single-Event Sounds
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
HuLA: Prosody-Aware Anti-Spoofing with Multi-Task Learning for Expressive and Emotional Synthetic Speech
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
Predicting User Intents and Musical Attributes from Music Discovery Conversations
von: Kwon, Daeyong, et al.
Veröffentlicht: (2024)
von: Kwon, Daeyong, et al.
Veröffentlicht: (2024)
Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models
von: Li, Weiqin, et al.
Veröffentlicht: (2024)
von: Li, Weiqin, et al.
Veröffentlicht: (2024)
Multimodal Input Aids a Bayesian Model of Phonetic Learning
von: Zhi, Sophia, et al.
Veröffentlicht: (2024)
von: Zhi, Sophia, et al.
Veröffentlicht: (2024)
SepMamba: State-space models for speaker separation using Mamba
von: Avenstrup, Thor Højhus, et al.
Veröffentlicht: (2024)
von: Avenstrup, Thor Højhus, et al.
Veröffentlicht: (2024)
Unified Pathological Speech Analysis with Prompt Tuning
von: Yang, Fei, et al.
Veröffentlicht: (2024)
von: Yang, Fei, et al.
Veröffentlicht: (2024)
Textually Pretrained Speech Language Models
von: Hassid, Michael, et al.
Veröffentlicht: (2023)
von: Hassid, Michael, et al.
Veröffentlicht: (2023)
Exploring Fine-Tuning of Large Audio Language Models for Spoken Language Understanding under Limited Speech Data
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MelHuBERT: A simplified HuBERT on Mel spectrograms
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2022) -
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition
von: Yoon, Ji Won, et al.
Veröffentlicht: (2022) -
Self-Supervised Syllable Discovery Based on Speaker-Disentangled HuBERT
von: Komatsu, Ryota, et al.
Veröffentlicht: (2024) -
mHuBERT-147: A Compact Multilingual HuBERT Model
von: Boito, Marcely Zanon, et al.
Veröffentlicht: (2024) -
DiceHuBERT: Distilling HuBERT with a Self-Supervised Learning Objective
von: Chi, Hyung Gun, et al.
Veröffentlicht: (2025)