BELT-2: Bootstrapping EEG-to-Language representation alignment for multi-task brain decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Jinzhao, Duan, Yiqun, Chang, Fred, Do, Thomas, Wang, Yu-Kai, Lin, Chin-Teng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
by: Wang, Bo, et al.
Published: (2024)
by: Wang, Bo, et al.
Published: (2024)
Auditory Attention Decoding from Ear-EEG Signals: A Dataset with Dynamic Attention Switching and Rigorous Cross-Validation
by: Zhang, Yuanming, et al.
Published: (2025)
by: Zhang, Yuanming, et al.
Published: (2025)
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
by: Fan, Shitong, et al.
Published: (2024)
by: Fan, Shitong, et al.
Published: (2024)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
by: Geirnaert, Simon, et al.
Published: (2024)
by: Geirnaert, Simon, et al.
Published: (2024)
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
by: Yu, Chin-Yun, et al.
Published: (2022)
by: Yu, Chin-Yun, et al.
Published: (2022)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
by: Zhu, Haolin, et al.
Published: (2024)
by: Zhu, Haolin, et al.
Published: (2024)
Enhancing dysarthria speech feature representation with empirical mode decomposition and Walsh-Hadamard transform
by: Zhu, Ting, et al.
Published: (2023)
by: Zhu, Ting, et al.
Published: (2023)
SSM2Mel: State Space Model to Reconstruct Mel Spectrogram from the EEG
by: Fan, Cunhang, et al.
Published: (2025)
by: Fan, Cunhang, et al.
Published: (2025)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
by: Yu, Chin-Yun, et al.
Published: (2024)
by: Yu, Chin-Yun, et al.
Published: (2024)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
by: Wilroth, Johanna, et al.
Published: (2026)
by: Wilroth, Johanna, et al.
Published: (2026)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
by: Liao, Yuan, et al.
Published: (2025)
by: Liao, Yuan, et al.
Published: (2025)
Exploring Audio-Visual Information Fusion for Sound Event Localization and Detection In Low-Resource Realistic Scenarios
by: Jiang, Ya, et al.
Published: (2024)
by: Jiang, Ya, et al.
Published: (2024)
Neuro2Semantic: A Transfer Learning Framework for Semantic Reconstruction of Continuous Language from Human Intracranial EEG
by: Shams, Siavash, et al.
Published: (2025)
by: Shams, Siavash, et al.
Published: (2025)
SELM: Speech Enhancement Using Discrete Tokens and Language Models
by: Wang, Ziqian, et al.
Published: (2023)
by: Wang, Ziqian, et al.
Published: (2023)
A Novel Numerical Method for Relaxing the Minimal Configurations of TOA-Based Joint Sensors and Sources Localization
by: Cao, Faxian, et al.
Published: (2024)
by: Cao, Faxian, et al.
Published: (2024)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
by: Yao, Shengshi, et al.
Published: (2025)
by: Yao, Shengshi, et al.
Published: (2025)
ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction
by: Yang, Shu-wen, et al.
Published: (2025)
by: Yang, Shu-wen, et al.
Published: (2025)
Scalable-Complexity Steered Response Power Mapping based on Low-Rank and Sparse Interpolation
by: Dietzen, Thomas, et al.
Published: (2023)
by: Dietzen, Thomas, et al.
Published: (2023)
Tracking of Spatially Dynamic Room Impulse Responses Along Locally Linearized Trajectories
by: MacWilliam, Kathleen, et al.
Published: (2025)
by: MacWilliam, Kathleen, et al.
Published: (2025)
Cross-Linguistic Rhythmic and Spectral Feature-Based Analysis of Nyishi and Adi: Two Under-Resourced Languages of Arunachal Pradesh
by: Gogoi, Deepshikha, et al.
Published: (2026)
by: Gogoi, Deepshikha, et al.
Published: (2026)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
by: MacWilliam, Kathleen, et al.
Published: (2024)
by: MacWilliam, Kathleen, et al.
Published: (2024)
Benchmarking multi-component signal processing methods in the time-frequency plane
by: Miramont, Juan M., et al.
Published: (2024)
by: Miramont, Juan M., et al.
Published: (2024)
Accelerating Automatic Differentiation of Direct Form Digital Filters
by: Yu, Chin-Yun, et al.
Published: (2025)
by: Yu, Chin-Yun, et al.
Published: (2025)
MASSLOC: A Massive Sound Source Localization System based on Direction-of-Arrival Estimation
by: Fischer, Georg K. J., et al.
Published: (2025)
by: Fischer, Georg K. J., et al.
Published: (2025)
MusicHiFi: Fast High-Fidelity Stereo Vocoding
by: Zhu, Ge, et al.
Published: (2024)
by: Zhu, Ge, et al.
Published: (2024)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
by: Qi, Tianhua, et al.
Published: (2025)
by: Qi, Tianhua, et al.
Published: (2025)
SuperM2M: Supervised and Mixture-to-Mixture Co-Learning for Speech Enhancement and Noise-Robust ASR
by: Wang, Zhong-Qiu
Published: (2024)
by: Wang, Zhong-Qiu
Published: (2024)
Joint Spectrogram Separation and TDOA Estimation using Optimal Transport
by: Fabiani, Linda, et al.
Published: (2025)
by: Fabiani, Linda, et al.
Published: (2025)
Harmonics to the Rescue: Why Voiced Speech is Not a Wss Process
by: Bologni, Giovanni, et al.
Published: (2025)
by: Bologni, Giovanni, et al.
Published: (2025)
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
by: Bhattacharjee, Susmita, et al.
Published: (2025)
by: Bhattacharjee, Susmita, et al.
Published: (2025)
RIR-Mega: a large-scale simulated room impulse response dataset for machine learning and room acoustics modeling
by: Goswami, Mandip
Published: (2025)
by: Goswami, Mandip
Published: (2025)
SUNAC: Source-aware Unified Neural Audio Codec
by: Aihara, Ryo, et al.
Published: (2025)
by: Aihara, Ryo, et al.
Published: (2025)
Semi-Blind Channel Estimation and Hybrid Receiver Beamforming in the Tera-Hertz Multi-User Massive MIMO Uplink
by: Garg, Abhisha, et al.
Published: (2026)
by: Garg, Abhisha, et al.
Published: (2026)
Unsupervised Face-Masked Speech Enhancement Using Generative Adversarial Networks With Human-in-the-Loop Assessment Metrics
by: Wang, Syu-Siang, et al.
Published: (2024)
by: Wang, Syu-Siang, et al.
Published: (2024)
Efficient Personalization of Amplification in Hearing Aids via Multi-band Bayesian Machine Learning
by: Ni, Aoxin, et al.
Published: (2024)
by: Ni, Aoxin, et al.
Published: (2024)
ERes2NetV2: Boosting Short-Duration Speaker Verification Performance with Computational Efficiency
by: Chen, Yafeng, et al.
Published: (2024)
by: Chen, Yafeng, et al.
Published: (2024)
Classification of Adventitious Sounds Combining Cochleogram and Vision Transformers
by: Mang, Loredana Daria, et al.
Published: (2024)
by: Mang, Loredana Daria, et al.
Published: (2024)
Advanced Signal Analysis in Detecting Replay Attacks for Automatic Speaker Verification Systems
by: Kuang, Lee Shih
Published: (2024)
by: Kuang, Lee Shih
Published: (2024)
Cough-E: A multimodal, privacy-preserving cough detection algorithm for the edge
by: Albini, Stefano, et al.
Published: (2024)
by: Albini, Stefano, et al.
Published: (2024)
Cyclic Multichannel Wiener Filter for Acoustic Beamforming
by: Bologni, Giovanni, et al.
Published: (2025)
by: Bologni, Giovanni, et al.
Published: (2025)
Similar Items
-
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
by: Wang, Bo, et al.
Published: (2024) -
Auditory Attention Decoding from Ear-EEG Signals: A Dataset with Dynamic Attention Switching and Rigorous Cross-Validation
by: Zhang, Yuanming, et al.
Published: (2025) -
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
by: Fan, Shitong, et al.
Published: (2024) -
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
by: Geirnaert, Simon, et al.
Published: (2024) -
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
by: Yu, Chin-Yun, et al.
Published: (2022)