SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Hasan, Md, Castro, Nyvenn, Liu, Daiqi, Mulzer, Lukas, Hutter, Jana, Woo, Jonghye, Zaiss, Moritz, Maier, Andreas, Perez-Toro, Paula A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
by: Liu, Daiqi, et al.
Published: (2026)
by: Liu, Daiqi, et al.
Published: (2026)
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Multi-Parameter Molecular MRI Quantification using Physics-Informed Self-Supervised Learning
by: Finkelstein, Alex, et al.
Published: (2024)
by: Finkelstein, Alex, et al.
Published: (2024)
Speech Audio Generation from dynamic MRI via a Knowledge Enhanced Conditional Variational Autoencoder
by: Li, Yaxuan, et al.
Published: (2025)
by: Li, Yaxuan, et al.
Published: (2025)
VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
A Deep Risk Estimator for Known Operator Learning
by: Maier, Andreas, et al.
Published: (2026)
by: Maier, Andreas, et al.
Published: (2026)
Anatomy-based quality metric of diffusion-weighted MRI data for accurate derivation of muscle fiber orientation
by: Shusharina, Nadya, et al.
Published: (2024)
by: Shusharina, Nadya, et al.
Published: (2024)
Speech motion anomaly detection via cross-modal translation of 4D motion fields from tagged MRI
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
Treatment-wise Glioblastoma Survival Inference with Multi-parametric Preoperative MRI
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
Decoding the human brain tissue response to radiofrequency excitation using a biophysical-model-free deep MRI on a chip framework
by: Nagar, Dinor, et al.
Published: (2024)
by: Nagar, Dinor, et al.
Published: (2024)
Optimization of pulsed saturation transfer MR fingerprinting (ST MRF) acquisition using the Cramér-Rao bound and sequential quadratic programming
by: Vladimirov, Nikita, et al.
Published: (2025)
by: Vladimirov, Nikita, et al.
Published: (2025)
Towards Orthographically-Informed Evaluation of Speech Recognition Systems for Indian Languages
by: Bhogale, Kaushal Santosh, et al.
Published: (2026)
by: Bhogale, Kaushal Santosh, et al.
Published: (2026)
Standard audiogram classification from loudness scaling data using unsupervised, supervised, and explainable machine learning techniques
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Cross-modal characterization of infant cry: validation of a chest-surface accelerometer in extracting acoustic vocal function measures
by: An, Winko W., et al.
Published: (2026)
by: An, Winko W., et al.
Published: (2026)
A Speech-to-Video Synthesis Approach Using Spatio-Temporal Diffusion for Vocal Tract MRI
by: Pérez-Toro, Paula Andrea, et al.
Published: (2025)
by: Pérez-Toro, Paula Andrea, et al.
Published: (2025)
End-to-End Simultaneous Dysarthric Speech Reconstruction with Frame-Level Adaptor and Multiple Wait-k Knowledge Distillation
by: Wu, Minghui, et al.
Published: (2026)
by: Wu, Minghui, et al.
Published: (2026)
UNIT-DSR: Dysarthric Speech Reconstruction System Using Speech Unit Normalization
by: Wang, Yuejiao, et al.
Published: (2024)
by: Wang, Yuejiao, et al.
Published: (2024)
An Implantable Piezofilm Middle Ear Microphone: Performance in Human Cadaveric Temporal Bones
by: Zhang, John Z., et al.
Published: (2023)
by: Zhang, John Z., et al.
Published: (2023)
Multilingual Phonological Feature Recognition with Self-Supervised Speech Models
by: Hernandez, Abner, et al.
Published: (2026)
by: Hernandez, Abner, et al.
Published: (2026)
The UmboMic: A PVDF Cantilever Microphone
by: Yeiser, Aaron J., et al.
Published: (2023)
by: Yeiser, Aaron J., et al.
Published: (2023)
SpeechT: Findings of the First Mentorship in Speech Translation
by: Moslem, Yasmin, et al.
Published: (2025)
by: Moslem, Yasmin, et al.
Published: (2025)
Improving X-Codec-2.0 for Multi-Lingual Speech: 25 Hz Latent Rate and 24 kHz Sampling
by: Zolkepli, Husein
Published: (2026)
by: Zolkepli, Husein
Published: (2026)
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
by: Hernandez, Abner, et al.
Published: (2026)
by: Hernandez, Abner, et al.
Published: (2026)
Artificial Neural Networks to Recognize Speakers Division from Continuous Bengali Speech
by: Ali, Hasmot, et al.
Published: (2024)
by: Ali, Hasmot, et al.
Published: (2024)
Speaker- and Text-Independent Estimation of Articulatory Movements and Phoneme Alignments from Speech
by: Weise, Tobias, et al.
Published: (2024)
by: Weise, Tobias, et al.
Published: (2024)
Finding My Voice: Generative Reconstruction of Disordered Speech for Automated Clinical Evaluation
by: Rosero, Karen, et al.
Published: (2025)
by: Rosero, Karen, et al.
Published: (2025)
Point-supervised Brain Tumor Segmentation with Box-prompted MedSAM
by: Liu, Xiaofeng, et al.
Published: (2024)
by: Liu, Xiaofeng, et al.
Published: (2024)
Listening or Reading? Evaluating Speech Awareness in Chain-of-Thought Speech-to-Text Translation
by: Romero-Díaz, Jacobo, et al.
Published: (2025)
by: Romero-Díaz, Jacobo, et al.
Published: (2025)
POTSA: A Cross-Lingual Speech Alignment Framework for Speech-to-Text Translation
by: Li, Xuanchen, et al.
Published: (2025)
by: Li, Xuanchen, et al.
Published: (2025)
Speech-Worthy Alignment for Japanese SpeechLLMs via Direct Preference Optimization
by: Zhao, Mengjie, et al.
Published: (2026)
by: Zhao, Mengjie, et al.
Published: (2026)
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs
by: Song, Yuhan, et al.
Published: (2025)
by: Song, Yuhan, et al.
Published: (2025)
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition
by: Yoon, Ji Won, et al.
Published: (2022)
by: Yoon, Ji Won, et al.
Published: (2022)
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
by: Paar, Ferdinand, et al.
Published: (2026)
by: Paar, Ferdinand, et al.
Published: (2026)
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
by: Mei, Yuxiang, et al.
Published: (2026)
by: Mei, Yuxiang, et al.
Published: (2026)
Revisiting Direct Speech-to-Text Translation with Speech LLMs: Better Scaling than CoT Prompting?
by: Pareras, Oriol, et al.
Published: (2025)
by: Pareras, Oriol, et al.
Published: (2025)
Multi-Teacher Language-Aware Knowledge Distillation for Multilingual Speech Emotion Recognition
by: Bijoy, Mehedi Hasan, et al.
Published: (2025)
by: Bijoy, Mehedi Hasan, et al.
Published: (2025)
Deep End-to-end Adaptive k-Space Sampling, Reconstruction, and Registration for Dynamic MRI
by: Yiasemis, George, et al.
Published: (2024)
by: Yiasemis, George, et al.
Published: (2024)
WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing
by: Dai, Yuhang, et al.
Published: (2025)
by: Dai, Yuhang, et al.
Published: (2025)
Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs
by: Wang, Dingdong, et al.
Published: (2025)
by: Wang, Dingdong, et al.
Published: (2025)
Generative Modeling of Complex-Valued Brain MRI Data
by: Schlimbach, Marco, et al.
Published: (2026)
by: Schlimbach, Marco, et al.
Published: (2026)
Similar Items
-
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
by: Liu, Daiqi, et al.
Published: (2026) -
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025) -
Multi-Parameter Molecular MRI Quantification using Physics-Informed Self-Supervised Learning
by: Finkelstein, Alex, et al.
Published: (2024) -
Speech Audio Generation from dynamic MRI via a Knowledge Enhanced Conditional Variational Autoencoder
by: Li, Yaxuan, et al.
Published: (2025) -
VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI
by: Liu, Daiqi, et al.
Published: (2025)