Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Maghsoudi, Maryam, Shamma, Shihab |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses
par: Maghsoudi, Maryam, et autres
Publié: (2025)
par: Maghsoudi, Maryam, et autres
Publié: (2025)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
par: Maghsoudi, Maryam, et autres
Publié: (2026)
par: Maghsoudi, Maryam, et autres
Publié: (2026)
Listenable Maps for Zero-Shot Audio Classifiers
par: Paissan, Francesco, et autres
Publié: (2024)
par: Paissan, Francesco, et autres
Publié: (2024)
Imagined Speech State Classification for Robust Brain-Computer Interface
par: Ko, Byung-Kwan, et autres
Publié: (2024)
par: Ko, Byung-Kwan, et autres
Publié: (2024)
Zero-Shot Mono-to-Binaural Speech Synthesis
par: Levkovitch, Alon, et autres
Publié: (2024)
par: Levkovitch, Alon, et autres
Publié: (2024)
Prompt Amplification and Zero-Shot Late Fusion in Audio-Language Models for Speech Emotion Recognition
par: Kataria, Saurabh, et autres
Publié: (2026)
par: Kataria, Saurabh, et autres
Publié: (2026)
RosettaSpeech: Zero-Shot Speech-to-Speech Translation without Parallel Speech
par: Zheng, Zhisheng, et autres
Publié: (2025)
par: Zheng, Zhisheng, et autres
Publié: (2025)
Listenable Maps for Audio Classifiers
par: Paissan, Francesco, et autres
Publié: (2024)
par: Paissan, Francesco, et autres
Publié: (2024)
DiffSoundStream: Efficient Speech Tokenization via Diffusion Decoding
par: Yang, Yang, et autres
Publié: (2025)
par: Yang, Yang, et autres
Publié: (2025)
PRESENT: Zero-Shot Text-to-Prosody Control
par: Lam, Perry, et autres
Publié: (2024)
par: Lam, Perry, et autres
Publié: (2024)
Principled Coarse-Grained Acceptance for Speculative Decoding in Speech
par: Yanuka, Moran, et autres
Publié: (2025)
par: Yanuka, Moran, et autres
Publié: (2025)
MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
par: Jiang, Ziyue, et autres
Publié: (2025)
par: Jiang, Ziyue, et autres
Publié: (2025)
Biomimetic Frontend for Differentiable Audio Processing
par: Famularo, Ruolan Leslie, et autres
Publié: (2024)
par: Famularo, Ruolan Leslie, et autres
Publié: (2024)
Magnetoencephalography (MEG) Based Non-Invasive Chinese Speech Decoding
par: Jia, Zhihong, et autres
Publié: (2025)
par: Jia, Zhihong, et autres
Publié: (2025)
GenVC: Self-Supervised Zero-Shot Voice Conversion
par: Cai, Zexin, et autres
Publié: (2025)
par: Cai, Zexin, et autres
Publié: (2025)
Listening, Imagining & Refining: A Heuristic Optimized ASR Correction Framework with LLMs
par: Liu, Yutong, et autres
Publié: (2025)
par: Liu, Yutong, et autres
Publié: (2025)
Towards Unified Neural Decoding of Perceived, Spoken and Imagined Speech from EEG Signals
par: Lee, Jung-Sun, et autres
Publié: (2024)
par: Lee, Jung-Sun, et autres
Publié: (2024)
Towards Decoding Brain Activity During Passive Listening of Speech
par: Fodor, Milán András, et autres
Publié: (2024)
par: Fodor, Milán András, et autres
Publié: (2024)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
par: Glazer, Neta, et autres
Publié: (2025)
par: Glazer, Neta, et autres
Publié: (2025)
Zero-Shot Multi-Lingual Speaker Verification in Clinical Trials
par: Akram, Ali, et autres
Publié: (2024)
par: Akram, Ali, et autres
Publié: (2024)
Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification
par: Sims, Ysobel, et autres
Publié: (2024)
par: Sims, Ysobel, et autres
Publié: (2024)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
par: Dutta, Soumya, et autres
Publié: (2024)
par: Dutta, Soumya, et autres
Publié: (2024)
Multi-label Zero-Shot Audio Classification with Temporal Attention
par: Dogan, Duygu, et autres
Publié: (2024)
par: Dogan, Duygu, et autres
Publié: (2024)
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
par: Janiczek, John, et autres
Publié: (2024)
par: Janiczek, John, et autres
Publié: (2024)
Kinetic-Optimal Scheduling with Moment Correction for Metric-Induced Discrete Flow Matching in Zero-Shot Text-to-Speech
par: Yang, Dong, et autres
Publié: (2026)
par: Yang, Dong, et autres
Publié: (2026)
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
par: Wang, Yuancheng, et autres
Publié: (2024)
par: Wang, Yuancheng, et autres
Publié: (2024)
FlashSpeech: Efficient Zero-Shot Speech Synthesis
par: Ye, Zhen, et autres
Publié: (2024)
par: Ye, Zhen, et autres
Publié: (2024)
Hold Me Tight: Stable Encoder-Decoder Design for Speech Enhancement
par: Haider, Daniel, et autres
Publié: (2024)
par: Haider, Daniel, et autres
Publié: (2024)
Chunked Attention-based Encoder-Decoder Model for Streaming Speech Recognition
par: Zeineldeen, Mohammad, et autres
Publié: (2023)
par: Zeineldeen, Mohammad, et autres
Publié: (2023)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
par: Tavares, Tiago, et autres
Publié: (2024)
par: Tavares, Tiago, et autres
Publié: (2024)
Information Retrieval for ZeroSpeech 2021: The Submission by University of Wroclaw
par: Chorowski, Jan, et autres
Publié: (2021)
par: Chorowski, Jan, et autres
Publié: (2021)
kNN Retrieval for Simple and Effective Zero-Shot Multi-speaker Text-to-Speech
par: Hajal, Karl El, et autres
Publié: (2024)
par: Hajal, Karl El, et autres
Publié: (2024)
Making Flow-Matching-Based Zero-Shot Text-to-Speech Laugh as You Like
par: Kanda, Naoyuki, et autres
Publié: (2024)
par: Kanda, Naoyuki, et autres
Publié: (2024)
Can Masked Autoencoders Also Listen to Birds?
par: Rauch, Lukas, et autres
Publié: (2025)
par: Rauch, Lukas, et autres
Publié: (2025)
ControlSpeech: Towards Simultaneous and Independent Zero-shot Speaker Cloning and Zero-shot Language Style Control
par: Ji, Shengpeng, et autres
Publié: (2024)
par: Ji, Shengpeng, et autres
Publié: (2024)
Blind Audio Bandwidth Extension: A Diffusion-Based Zero-Shot Approach
par: Moliner, Eloi, et autres
Publié: (2023)
par: Moliner, Eloi, et autres
Publié: (2023)
Zero-Shot Unsupervised and Text-Based Audio Editing Using DDPM Inversion
par: Manor, Hila, et autres
Publié: (2024)
par: Manor, Hila, et autres
Publié: (2024)
AV-CrossNet: an Audiovisual Complex Spectral Mapping Network for Speech Separation By Leveraging Narrow- and Cross-Band Modeling
par: Kalkhorani, Vahid Ahmadi, et autres
Publié: (2024)
par: Kalkhorani, Vahid Ahmadi, et autres
Publié: (2024)
Multi-Channel Replay Speech Detection using Acoustic Maps
par: Neri, Michael, et autres
Publié: (2026)
par: Neri, Michael, et autres
Publié: (2026)
CJST: CTC Compressor based Joint Speech and Text Training for Decoder-Only ASR
par: Zhou, Wei, et autres
Publié: (2024)
par: Zhou, Wei, et autres
Publié: (2024)
Documents similaires
-
A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses
par: Maghsoudi, Maryam, et autres
Publié: (2025) -
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
par: Maghsoudi, Maryam, et autres
Publié: (2026) -
Listenable Maps for Zero-Shot Audio Classifiers
par: Paissan, Francesco, et autres
Publié: (2024) -
Imagined Speech State Classification for Robust Brain-Computer Interface
par: Ko, Byung-Kwan, et autres
Publié: (2024) -
Zero-Shot Mono-to-Binaural Speech Synthesis
par: Levkovitch, Alon, et autres
Publié: (2024)