VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Daiqi, Enk, Johannes, Stone, Maureen, Xing, Fangxu, Arias-Vergara, Tomás, Prince, Jerry L., Hutter, Jana, Woo, Jonghye, Maier, Andreas, Pérez-Toro, Paula Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Speech-to-Video Synthesis Approach Using Spatio-Temporal Diffusion for Vocal Tract MRI
by: Pérez-Toro, Paula Andrea, et al.
Published: (2025)
by: Pérez-Toro, Paula Andrea, et al.
Published: (2025)
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
by: Liu, Daiqi, et al.
Published: (2026)
by: Liu, Daiqi, et al.
Published: (2026)
Speech motion anomaly detection via cross-modal translation of 4D motion fields from tagged MRI
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
by: Hasan, Md, et al.
Published: (2026)
by: Hasan, Md, et al.
Published: (2026)
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Semi-Supervised Bone Marrow Lesion Detection from Knee MRI Segmentation Using Mask Inpainting Models
by: Qin, Shihua, et al.
Published: (2024)
by: Qin, Shihua, et al.
Published: (2024)
Speech Audio Generation from dynamic MRI via a Knowledge Enhanced Conditional Variational Autoencoder
by: Li, Yaxuan, et al.
Published: (2025)
by: Li, Yaxuan, et al.
Published: (2025)
Multilingual Phonological Feature Recognition with Self-Supervised Speech Models
by: Hernandez, Abner, et al.
Published: (2026)
by: Hernandez, Abner, et al.
Published: (2026)
Principled Feature Disentanglement for High-Fidelity Unified Brain MRI Synthesis
by: Cho, Jihoon, et al.
Published: (2024)
by: Cho, Jihoon, et al.
Published: (2024)
Brightness-Invariant Tracking Estimation in Tagged MRI
by: Bian, Zhangxing, et al.
Published: (2025)
by: Bian, Zhangxing, et al.
Published: (2025)
LoGSAM: Parameter-Efficient Cross-Modal Grounding for MRI Segmentation
by: Bhuiyan, Mohammad Robaitul Islam, et al.
Published: (2026)
by: Bhuiyan, Mohammad Robaitul Islam, et al.
Published: (2026)
Anatomy-based quality metric of diffusion-weighted MRI data for accurate derivation of muscle fiber orientation
by: Shusharina, Nadya, et al.
Published: (2024)
by: Shusharina, Nadya, et al.
Published: (2024)
Ethics of Generating Synthetic MRI Vocal Tract Views from the Face
by: Shahid, Muhammad Suhaib, et al.
Published: (2024)
by: Shahid, Muhammad Suhaib, et al.
Published: (2024)
A Category-Fragment Segmentation Framework for Pelvic Fracture Segmentation in X-ray Images
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Multimodal Segmentation for Vocal Tract Modeling
by: Jain, Rishi, et al.
Published: (2024)
by: Jain, Rishi, et al.
Published: (2024)
Reconstruction of the Vocal Tract from Speech via Phonetic Representations Using MRI Data
by: Azzouz, Sofiane, et al.
Published: (2026)
by: Azzouz, Sofiane, et al.
Published: (2026)
Speech2rtMRI: Speech-Guided Diffusion Model for Real-time MRI Video of the Vocal Tract during Speech
by: Nguyen, Hong, et al.
Published: (2024)
by: Nguyen, Hong, et al.
Published: (2024)
BreastSegNet: Multi-label Segmentation of Breast MRI
by: Li, Qihang, et al.
Published: (2025)
by: Li, Qihang, et al.
Published: (2025)
Segmentation of spinal rootlets across MRI contrasts with RootletSeg
by: Krejci, Katerina, et al.
Published: (2025)
by: Krejci, Katerina, et al.
Published: (2025)
RATNUS: Rapid, Automatic Thalamic Nuclei Segmentation using Multimodal MRI inputs
by: Feng, Anqi, et al.
Published: (2024)
by: Feng, Anqi, et al.
Published: (2024)
Towards contrast- and pathology-agnostic clinical fetal brain MRI segmentation using SynthSeg
by: Shang, Ziyao, et al.
Published: (2025)
by: Shang, Ziyao, et al.
Published: (2025)
Acoustic-to-articulatory Inversion of the Complete Vocal Tract from RT-MRI with Various Audio Embeddings and Dataset Sizes
by: Azzouz, Sofiane, et al.
Published: (2026)
by: Azzouz, Sofiane, et al.
Published: (2026)
Reconstruction of the Complete Vocal Tract Contour Through Acoustic to Articulatory Inversion Using Real-Time MRI Data
by: Azzouz, Sofiane, et al.
Published: (2025)
by: Azzouz, Sofiane, et al.
Published: (2025)
Treatment-wise Glioblastoma Survival Inference with Multi-parametric Preoperative MRI
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
DeepGI: An Automated Approach for Gastrointestinal Tract Segmentation in MRI Scans
by: Zhang, Ye, et al.
Published: (2024)
by: Zhang, Ye, et al.
Published: (2024)
Diffusion-Driven Generation of Minimally Preprocessed Brain MRI
by: Remedios, Samuel W., et al.
Published: (2025)
by: Remedios, Samuel W., et al.
Published: (2025)
Confidence-Guided Error Correction for Disordered Speech Recognition
by: Hernandez, Abner, et al.
Published: (2025)
by: Hernandez, Abner, et al.
Published: (2025)
Novel-view X-ray Projection Synthesis through Geometry-Integrated Deep Learning
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Vocal tract physiology and its MRI evaluation
by: Bruno Murmura
Published: (2021)
by: Bruno Murmura
Published: (2021)
Real-time MRI-based fetal femur length measurement
by: Barcsay, Johannes, et al.
Published: (2025)
by: Barcsay, Johannes, et al.
Published: (2025)
Disentangled Multimodal Brain MR Image Translation via Transformer-based Modality Infuser
by: Cho, Jihoon, et al.
Published: (2024)
by: Cho, Jihoon, et al.
Published: (2024)
CATNUS: Coordinate-Aware Thalamic Nuclei Segmentation Using T1-Weighted MRI
by: Feng, Anqi, et al.
Published: (2025)
by: Feng, Anqi, et al.
Published: (2025)
Unique MS Lesion Identification from MRI
by: Rivas, Carlos A., et al.
Published: (2024)
by: Rivas, Carlos A., et al.
Published: (2024)
OpticNerveSeg: Public Dataset for Cranial Nerve II Segmentation from Multimodal MRI
by: Diakite, Alou, et al.
Published: (2026)
by: Diakite, Alou, et al.
Published: (2026)
Label-Efficient 3D Brain Segmentation via Complementary 2D Diffusion Models with Orthogonal Views
by: Cho, Jihoon, et al.
Published: (2024)
by: Cho, Jihoon, et al.
Published: (2024)
Diffusing the Blind Spot: Uterine MRI Synthesis with Diffusion Models
by: Müller, Johanna P., et al.
Published: (2025)
by: Müller, Johanna P., et al.
Published: (2025)
Open-Source Manually Annotated Vocal Tract Database for Automatic Segmentation from 3D MRI Using Deep Learning: Benchmarking 2D and 3D Convolutional and Transformer Networks
by: Erattakulangara, Subin, et al.
Published: (2025)
by: Erattakulangara, Subin, et al.
Published: (2025)
AtlasSeg: Atlas Prior Guided Dual-U-Net for Cortical Segmentation in Fetal Brain MRI
by: Xu, Haoan, et al.
Published: (2024)
by: Xu, Haoan, et al.
Published: (2024)
Point-supervised Brain Tumor Segmentation with Box-prompted MedSAM
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
Coding Speech through Vocal Tract Kinematics
by: Cho, Cheol Jun, et al.
Published: (2024)
by: Cho, Cheol Jun, et al.
Published: (2024)
Similar Items
-
A Speech-to-Video Synthesis Approach Using Spatio-Temporal Diffusion for Vocal Tract MRI
by: Pérez-Toro, Paula Andrea, et al.
Published: (2025) -
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
by: Liu, Daiqi, et al.
Published: (2026) -
Speech motion anomaly detection via cross-modal translation of 4D motion fields from tagged MRI
by: Liu, Xiaofeng, et al.
Published: (2024) -
SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
by: Hasan, Md, et al.
Published: (2026) -
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025)