Measuring Robustness of Speech Recognition from MEG Signals Under Distribution Shift
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chien, Sheng-You, Mao, Bo-Yi, Chang, Yi-Ning, Kuo, Po-Chih |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
von: Ismail, Saifelden M.
Veröffentlicht: (2025)
von: Ismail, Saifelden M.
Veröffentlicht: (2025)
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
von: Kim, Minu, et al.
Veröffentlicht: (2025)
von: Kim, Minu, et al.
Veröffentlicht: (2025)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
von: Bhadra, Dipayan, et al.
Veröffentlicht: (2025)
von: Bhadra, Dipayan, et al.
Veröffentlicht: (2025)
Alternative Local Discriminant Bases Using Empirical Expectation and Variance Estimation
von: Fossgaard, Eirik
Veröffentlicht: (1999)
von: Fossgaard, Eirik
Veröffentlicht: (1999)
Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization
von: Alamr, Meshal, et al.
Veröffentlicht: (2026)
von: Alamr, Meshal, et al.
Veröffentlicht: (2026)
Improving Omics-Based Classification: The Role of Feature Selection and Synthetic Data Generation
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
von: Hori, Takaaki, et al.
Veröffentlicht: (2025)
von: Hori, Takaaki, et al.
Veröffentlicht: (2025)
Uncovering Population PK Covariates from VAE-Generated Latent Spaces
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
VoiceSHIELD-Small: Real-Time Malicious Speech Detection and Transcription
von: Ranjan, Sumit, et al.
Veröffentlicht: (2026)
von: Ranjan, Sumit, et al.
Veröffentlicht: (2026)
Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device
von: Kozak, Nazar
Veröffentlicht: (2026)
von: Kozak, Nazar
Veröffentlicht: (2026)
VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio
von: Basha, Maris, et al.
Veröffentlicht: (2025)
von: Basha, Maris, et al.
Veröffentlicht: (2025)
What Would GPT Click: Practical Effects of Human-AI Behavioral Misalignment and the Cost of Synthetic Participants in User Experience
von: Kuric, Eduard, et al.
Veröffentlicht: (2026)
von: Kuric, Eduard, et al.
Veröffentlicht: (2026)
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages
von: Zaragozá, Lucía Gómez, et al.
Veröffentlicht: (2024)
von: Zaragozá, Lucía Gómez, et al.
Veröffentlicht: (2024)
Improving Speech Recognition Accuracy Using Custom Language Models with the Vosk Toolkit
von: Soni, Aniket Abhishek
Veröffentlicht: (2025)
von: Soni, Aniket Abhishek
Veröffentlicht: (2025)
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
von: Kim, Minu, et al.
Veröffentlicht: (2025)
von: Kim, Minu, et al.
Veröffentlicht: (2025)
Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
von: Perera, Amal S., et al.
Veröffentlicht: (2025)
von: Perera, Amal S., et al.
Veröffentlicht: (2025)
IMUVIE: Pickup Timeline Action Localization via Motion Movies
von: Clapham, John, et al.
Veröffentlicht: (2024)
von: Clapham, John, et al.
Veröffentlicht: (2024)
OrganicHAR: Towards Activity Discovery in Organic Settings for Privacy Preserving Sensors Using Efficient Video Analysis
von: Patidar, Prasoon, et al.
Veröffentlicht: (2026)
von: Patidar, Prasoon, et al.
Veröffentlicht: (2026)
Prevailing Research Areas for Music AI in the Era of Foundation Models
von: Wei, Megan, et al.
Veröffentlicht: (2024)
von: Wei, Megan, et al.
Veröffentlicht: (2024)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
von: Kajare, Prajwal Vijay, et al.
Veröffentlicht: (2026)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
von: Idahl, Maximilian, et al.
Veröffentlicht: (2026)
von: Idahl, Maximilian, et al.
Veröffentlicht: (2026)
LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
Pipeline and Dataset Generation for Automated Fact-checking in Almost Any Language
von: Drchal, Jan, et al.
Veröffentlicht: (2023)
von: Drchal, Jan, et al.
Veröffentlicht: (2023)
Automating Clinical Information Retrieval from Finnish Electronic Health Records Using Large Language Models
von: Saukkoriipi, Mikko, et al.
Veröffentlicht: (2026)
von: Saukkoriipi, Mikko, et al.
Veröffentlicht: (2026)
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
von: Dar, Daniyal Kabir, et al.
Veröffentlicht: (2025)
von: Dar, Daniyal Kabir, et al.
Veröffentlicht: (2025)
Named entity recognition for Serbian legal documents: Design, methodology and dataset development
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2025)
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2025)
Precision at Scale: Domain-Specific Datasets On-Demand
von: Rodríguez-de-Vera, Jesús M, et al.
Veröffentlicht: (2024)
von: Rodríguez-de-Vera, Jesús M, et al.
Veröffentlicht: (2024)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
A Spatio-Temporal Deep Learning Approach For High-Resolution Gridded Monsoon Prediction
von: Borah, Parashjyoti, et al.
Veröffentlicht: (2026)
von: Borah, Parashjyoti, et al.
Veröffentlicht: (2026)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
von: Hua, Wenjie, et al.
Veröffentlicht: (2025)
von: Hua, Wenjie, et al.
Veröffentlicht: (2025)
An End-to-End Approach for Korean Wakeword Systems with Speaker Authentication
von: Seo, Geonwoo
Veröffentlicht: (2025)
von: Seo, Geonwoo
Veröffentlicht: (2025)
STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts
von: Opria, Joshua
Veröffentlicht: (2026)
von: Opria, Joshua
Veröffentlicht: (2026)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
von: Banthia, Saumya, et al.
Veröffentlicht: (2020)
von: Banthia, Saumya, et al.
Veröffentlicht: (2020)
The Binding Effect: Analyzing How Multi-Dimensional Cues Form Gender Bias in Instruction TTS
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2026)
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2026)
Fine-Tuning Large Audio-Language Models with LoRA for Precise Temporal Localization of Prolonged Exposure Therapy Elements
von: BN, Suhas, et al.
Veröffentlicht: (2025)
von: BN, Suhas, et al.
Veröffentlicht: (2025)
Evaluating Prompting Strategies for Chart Question Answering with Large Language Models
von: Naikar, Ruthuparna, et al.
Veröffentlicht: (2026)
von: Naikar, Ruthuparna, et al.
Veröffentlicht: (2026)
Quantum-Enhanced Analysis and Grading of Vocal Performance
von: Agarwal, Rohan
Veröffentlicht: (2025)
von: Agarwal, Rohan
Veröffentlicht: (2025)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
Membership Inference Attacks against Large Audio Language Models
von: Dong, Jia-Kai, et al.
Veröffentlicht: (2026)
von: Dong, Jia-Kai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
von: Ismail, Saifelden M.
Veröffentlicht: (2025) -
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
von: Kim, Minu, et al.
Veröffentlicht: (2025) -
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
von: Bhadra, Dipayan, et al.
Veröffentlicht: (2025) -
Alternative Local Discriminant Bases Using Empirical Expectation and Variance Estimation
von: Fossgaard, Eirik
Veröffentlicht: (1999) -
Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization
von: Alamr, Meshal, et al.
Veröffentlicht: (2026)