Acoustic-to-articulatory Inversion of the Complete Vocal Tract from RT-MRI with Various Audio Embeddings and Dataset Sizes
Fuente:
arXiv
Saved in:
| Main Authors: | Azzouz, Sofiane, Vuissoz, Pierre-André, Laprie, Yves |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reconstruction of the Complete Vocal Tract Contour Through Acoustic to Articulatory Inversion Using Real-Time MRI Data
by: Azzouz, Sofiane, et al.
Published: (2025)
by: Azzouz, Sofiane, et al.
Published: (2025)
Complete reconstruction of the tongue contour through acoustic to articulatory inversion using real-time MRI data
by: Azzouz, Sofiane, et al.
Published: (2024)
by: Azzouz, Sofiane, et al.
Published: (2024)
Reconstruction of the Vocal Tract from Speech via Phonetic Representations Using MRI Data
by: Azzouz, Sofiane, et al.
Published: (2026)
by: Azzouz, Sofiane, et al.
Published: (2026)
Acoustic-to-Articulatory Inversion of Clean Speech Using an MRI-Trained Model
by: Azzouz, Sofiane, et al.
Published: (2026)
by: Azzouz, Sofiane, et al.
Published: (2026)
Decoding Vocal Articulations from Acoustic Latent Representations
by: Cámara, Mateo, et al.
Published: (2024)
by: Cámara, Mateo, et al.
Published: (2024)
Auditory Representation Effective for Estimating Vocal Tract Information
by: Irino, Toshio, et al.
Published: (2023)
by: Irino, Toshio, et al.
Published: (2023)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
by: Maharana, Sarthak Kumar, et al.
Published: (2023)
Improving Query-by-Vocal Imitation with Contrastive Learning and Audio Pretraining
by: Greif, Jonathan, et al.
Published: (2024)
by: Greif, Jonathan, et al.
Published: (2024)
Learning Vocal-Tract Area and Radiation with a Physics-Informed Webster Model
by: Lu, Minhui, et al.
Published: (2026)
by: Lu, Minhui, et al.
Published: (2026)
AAT: Adapting Audio Transformer for Various Acoustics Recognition Tasks
by: Liang, Yun, et al.
Published: (2024)
by: Liang, Yun, et al.
Published: (2024)
Diff-VS: Efficient Audio-Aware Diffusion U-Net for Vocals Separation
by: Yun-Ning, et al.
Published: (2026)
by: Yun-Ning, et al.
Published: (2026)
Acoustic to Articulatory Speech Inversion for Children with Velopharyngeal Insufficiency
by: Tabatabaee, Saba, et al.
Published: (2025)
by: Tabatabaee, Saba, et al.
Published: (2025)
Enhancing Acoustic-to-Articulatory Speech Inversion by Incorporating Nasality
by: Tabatabaee, Saba, et al.
Published: (2025)
by: Tabatabaee, Saba, et al.
Published: (2025)
Improving Speech Inversion Through Self-Supervised Embeddings and Enhanced Tract Variables
by: Attia, Ahmed Adel, et al.
Published: (2023)
by: Attia, Ahmed Adel, et al.
Published: (2023)
From perception to production: how acoustic invariance facilitates articulatory learning in a self-supervised vocal imitation model
by: Lavechin, Marvin, et al.
Published: (2025)
by: Lavechin, Marvin, et al.
Published: (2025)
Towards Evaluating Generative Audio: Insights from Neural Audio Codec Embedding Distances
by: Biswas, Arijit, et al.
Published: (2025)
by: Biswas, Arijit, et al.
Published: (2025)
SAGA-SR: Semantically and Acoustically Guided Audio Super-Resolution
by: Im, Jaekwon, et al.
Published: (2025)
by: Im, Jaekwon, et al.
Published: (2025)
Acoustic Teleportation via Disentangled Neural Audio Codec Representations
by: Grundhuber, Philipp, et al.
Published: (2025)
by: Grundhuber, Philipp, et al.
Published: (2025)
XANE Background Acoustic Embeddings: Ablation and Clustering Analysis
by: Sharma, Dushyant, et al.
Published: (2024)
by: Sharma, Dushyant, et al.
Published: (2024)
ODAQ: Open Dataset of Audio Quality
by: Torcoli, Matteo, et al.
Published: (2023)
by: Torcoli, Matteo, et al.
Published: (2023)
Coding Speech through Vocal Tract Kinematics
by: Cho, Cheol Jun, et al.
Published: (2024)
by: Cho, Cheol Jun, et al.
Published: (2024)
UrBAN: Urban Beehive Acoustics and PheNotyping Dataset
by: Abdollahi, Mahsa, et al.
Published: (2024)
by: Abdollahi, Mahsa, et al.
Published: (2024)
WoW-Bench: Evaluating Fine-Grained Acoustic Perception in Audio-Language Models via Marine Mammal Vocalizations
by: Kim, Jaeyeon, et al.
Published: (2025)
by: Kim, Jaeyeon, et al.
Published: (2025)
Enhancing Child Vocalization Classification with Phonetically-Tuned Embeddings for Assisting Autism Diagnosis
by: Li, Jialu, et al.
Published: (2023)
by: Li, Jialu, et al.
Published: (2023)
OmniCodec: Low Frame Rate Universal Audio Codec with Semantic-Acoustic Disentanglement
by: Hu, Jingbin, et al.
Published: (2026)
by: Hu, Jingbin, et al.
Published: (2026)
AudioSetCaps: An Enriched Audio-Caption Dataset using Automated Generation Pipeline with Large Audio and Language Models
by: Bai, Jisheng, et al.
Published: (2024)
by: Bai, Jisheng, et al.
Published: (2024)
Past, Present, and Future of Spatial Audio and Room Acoustics
by: Koyama, Shoichi, et al.
Published: (2025)
by: Koyama, Shoichi, et al.
Published: (2025)
Modeling and Estimation of Vocal Tract and Glottal Source Parameters Using ARMAX-LF Model
by: Lia, Kai, et al.
Published: (2024)
by: Lia, Kai, et al.
Published: (2024)
Speech Separation using Neural Audio Codecs with Embedding Loss
by: Yip, Jia Qi, et al.
Published: (2024)
by: Yip, Jia Qi, et al.
Published: (2024)
Expressive Timing in Hindustani Vocal Music
by: Bhake, Yash, et al.
Published: (2025)
by: Bhake, Yash, et al.
Published: (2025)
Expanding and Analyzing ODAQ -- the Open Dataset of Audio Quality
by: Dick, Sascha, et al.
Published: (2025)
by: Dick, Sascha, et al.
Published: (2025)
Layer-Aware Early Fusion of Acoustic and Linguistic Embeddings for Cognitive Status Classification
by: Novotny, Krystof, et al.
Published: (2026)
by: Novotny, Krystof, et al.
Published: (2026)
Matching Reverberant Speech Through Learned Acoustic Embeddings and Feedback Delay Networks
by: Götz, Philipp, et al.
Published: (2025)
by: Götz, Philipp, et al.
Published: (2025)
XANE: eXplainable Acoustic Neural Embeddings
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
by: Wang, Ju-Chiang, et al.
Published: (2024)
by: Wang, Ju-Chiang, et al.
Published: (2024)
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
by: Ritu, Jarin, et al.
Published: (2025)
by: Ritu, Jarin, et al.
Published: (2025)
CoPlay: Audio-agnostic Cognitive Scaling for Acoustic Sensing
by: Li, Yin, et al.
Published: (2024)
by: Li, Yin, et al.
Published: (2024)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
by: Götz, Philipp, et al.
Published: (2024)
by: Götz, Philipp, et al.
Published: (2024)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
by: Shi, Runwu, et al.
Published: (2024)
by: Shi, Runwu, et al.
Published: (2024)
Auto-Landmark: Acoustic Landmark Dataset and Open-Source Toolkit for Landmark Extraction
by: Zhang, Xiangyu, et al.
Published: (2024)
by: Zhang, Xiangyu, et al.
Published: (2024)
Similar Items
-
Reconstruction of the Complete Vocal Tract Contour Through Acoustic to Articulatory Inversion Using Real-Time MRI Data
by: Azzouz, Sofiane, et al.
Published: (2025) -
Complete reconstruction of the tongue contour through acoustic to articulatory inversion using real-time MRI data
by: Azzouz, Sofiane, et al.
Published: (2024) -
Reconstruction of the Vocal Tract from Speech via Phonetic Representations Using MRI Data
by: Azzouz, Sofiane, et al.
Published: (2026) -
Acoustic-to-Articulatory Inversion of Clean Speech Using an MRI-Trained Model
by: Azzouz, Sofiane, et al.
Published: (2026) -
Decoding Vocal Articulations from Acoustic Latent Representations
by: Cámara, Mateo, et al.
Published: (2024)