Neck-Learn: Attention-Based Multiple Instance Learning and Ensemble Framework for Ecological Momentary Assessment
Fuente:
arXiv
Guardado en:
| Autor principal: | Cheema, Ahsan Jamal |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
por: Cheema, Ahsan J., et al.
Publicado: (2025)
por: Cheema, Ahsan J., et al.
Publicado: (2025)
Uncertainty-Based Ensemble Learning For Speech Classification
por: Atmaja, Bagus Tris, et al.
Publicado: (2024)
por: Atmaja, Bagus Tris, et al.
Publicado: (2024)
Utilizing Information Theoretic Approach to Study Cochlear Neural Degeneration
por: Cheema, Ahsan J., et al.
Publicado: (2025)
por: Cheema, Ahsan J., et al.
Publicado: (2025)
EffortNet: A Deep Learning Framework for Objective Assessment of Speech Enhancement Technologies Using EEG-Based Alpha Oscillations
por: Sung, Ching-Chih, et al.
Publicado: (2025)
por: Sung, Ching-Chih, et al.
Publicado: (2025)
Bridging Attribution and Open-Set Detection using Graph-Augmented Instance Learning in Synthetic Speech
por: Akhtar, Mohd Mujtaba, et al.
Publicado: (2026)
por: Akhtar, Mohd Mujtaba, et al.
Publicado: (2026)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
por: Huang, Wen, et al.
Publicado: (2024)
por: Huang, Wen, et al.
Publicado: (2024)
Automatic Assessment of Dysarthria Using Audio-visual Vowel Graph Attention Network
por: Liu, Xiaokang, et al.
Publicado: (2024)
por: Liu, Xiaokang, et al.
Publicado: (2024)
A Multimodal Framework for the Assessment of the Schizophrenia Spectrum
por: Premananth, Gowtham, et al.
Publicado: (2024)
por: Premananth, Gowtham, et al.
Publicado: (2024)
Learning Representation of Therapist Empathy in Counseling Conversation Using Siamese Hierarchical Attention Network
por: Tao, Dehua, et al.
Publicado: (2023)
por: Tao, Dehua, et al.
Publicado: (2023)
Tracking Listener Attention: Gaze-Guided Audio-Visual Speech Enhancement Framework
por: Yang, Hsiang-Cheng, et al.
Publicado: (2026)
por: Yang, Hsiang-Cheng, et al.
Publicado: (2026)
Unifying Listener Scoring Scales: Comparison Learning Framework for Speech Quality Assessment and Continuous Speech Emotion Recognition
por: Hu, Cheng-Hung, et al.
Publicado: (2025)
por: Hu, Cheng-Hung, et al.
Publicado: (2025)
First Deep Learning Approach to Hammering Acoustics for Stem Stability Assessment in Total Hip Arthroplasty
por: Zhu, Dongqi, et al.
Publicado: (2025)
por: Zhu, Dongqi, et al.
Publicado: (2025)
SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations
por: Yang, Xiaoyu, et al.
Publicado: (2025)
por: Yang, Xiaoyu, et al.
Publicado: (2025)
Estimating the Number and Locations of Boundaries in Reverberant Environments with Deep Learning
por: Arikan, Toros, et al.
Publicado: (2024)
por: Arikan, Toros, et al.
Publicado: (2024)
Bayesian Speech Synthesizers Can Learn from Multiple Teachers
por: Zhang, Ziyang, et al.
Publicado: (2025)
por: Zhang, Ziyang, et al.
Publicado: (2025)
A Pre-training Framework that Encodes Noise Information for Speech Quality Assessment
por: Sultana, Subrina, et al.
Publicado: (2024)
por: Sultana, Subrina, et al.
Publicado: (2024)
Coherence-Based Frequency Subset Selection For Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2022)
por: Fejgin, Daniel, et al.
Publicado: (2022)
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
por: Yang, Yifan, et al.
Publicado: (2024)
por: Yang, Yifan, et al.
Publicado: (2024)
Contextual Biasing for LLM-Based ASR with Hotword Retrieval and Reinforcement Learning
por: Kong, YuXiang, et al.
Publicado: (2025)
por: Kong, YuXiang, et al.
Publicado: (2025)
Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations
por: Shi, Haohan, et al.
Publicado: (2025)
por: Shi, Haohan, et al.
Publicado: (2025)
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning
por: Ouyang, Zhicheng, et al.
Publicado: (2026)
por: Ouyang, Zhicheng, et al.
Publicado: (2026)
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
por: Hashizume, Yuka, et al.
Publicado: (2024)
por: Hashizume, Yuka, et al.
Publicado: (2024)
Automatic Detection of Depression in Speech Using Ensemble Convolutional Neural Networks
por: Vázquez-Romero, Adrián, et al.
Publicado: (2024)
por: Vázquez-Romero, Adrián, et al.
Publicado: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
por: Shen, Yu-Han, et al.
Publicado: (2018)
por: Shen, Yu-Han, et al.
Publicado: (2018)
IR-UWB Radar-Based Contactless Silent Speech Recognition with Attention-Enhanced Temporal Convolutional Networks
por: Lee, Sunghwa, et al.
Publicado: (2025)
por: Lee, Sunghwa, et al.
Publicado: (2025)
Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models
por: Zucatelli, Guilherme, et al.
Publicado: (2025)
por: Zucatelli, Guilherme, et al.
Publicado: (2025)
AudioCIL: A Python Toolbox for Audio Class-Incremental Learning with Multiple Scenes
por: Xu, Qisheng, et al.
Publicado: (2024)
por: Xu, Qisheng, et al.
Publicado: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
por: Liao, Yuan, et al.
Publicado: (2025)
por: Liao, Yuan, et al.
Publicado: (2025)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
por: Zhu, Haolin, et al.
Publicado: (2024)
por: Zhu, Haolin, et al.
Publicado: (2024)
Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment
por: Wang, Wei, et al.
Publicado: (2025)
por: Wang, Wei, et al.
Publicado: (2025)
Attention-Based Audio Embeddings for Query-by-Example
por: Singh, Anup, et al.
Publicado: (2022)
por: Singh, Anup, et al.
Publicado: (2022)
I-DCCRN-VAE: An Improved Deep Representation Learning Framework for Complex VAE-based Single-channel Speech Enhancement
por: Li, Jiatong, et al.
Publicado: (2025)
por: Li, Jiatong, et al.
Publicado: (2025)
Lightweight Resolution-Aware Audio Deepfake Detection via Cross-Scale Attention and Consistency Learning
por: Shahriar, K. A.
Publicado: (2026)
por: Shahriar, K. A.
Publicado: (2026)
How Attention Shapes Emotion: A Comparative Study of Attention Mechanisms for Speech Emotion Recognition
por: Casals-Salvador, Marc, et al.
Publicado: (2026)
por: Casals-Salvador, Marc, et al.
Publicado: (2026)
AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning
por: Tatarjitzky, Michael, et al.
Publicado: (2025)
por: Tatarjitzky, Michael, et al.
Publicado: (2025)
Deep Learning-Based Prediction of Energy Decay Curves from Room Geometry and Material Properties
por: Muhammad, Imran, et al.
Publicado: (2025)
por: Muhammad, Imran, et al.
Publicado: (2025)
Controlling Contrastive Self-Supervised Learning with Knowledge-Driven Multiple Hypothesis: Application to Beat Tracking
por: Gagnere, Antonin, et al.
Publicado: (2025)
por: Gagnere, Antonin, et al.
Publicado: (2025)
Attention-based Interactive Disentangling Network for Instance-level Emotional Voice Conversion
por: Chen, Yun, et al.
Publicado: (2023)
por: Chen, Yun, et al.
Publicado: (2023)
Music Source Separation Based on a Lightweight Deep Learning Framework (DTTNET: DUAL-PATH TFC-TDF UNET)
por: Chen, Junyu, et al.
Publicado: (2023)
por: Chen, Junyu, et al.
Publicado: (2023)
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
por: Pierotti, Francesco, et al.
Publicado: (2025)
por: Pierotti, Francesco, et al.
Publicado: (2025)
Ejemplares similares
-
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
por: Cheema, Ahsan J., et al.
Publicado: (2025) -
Uncertainty-Based Ensemble Learning For Speech Classification
por: Atmaja, Bagus Tris, et al.
Publicado: (2024) -
Utilizing Information Theoretic Approach to Study Cochlear Neural Degeneration
por: Cheema, Ahsan J., et al.
Publicado: (2025) -
EffortNet: A Deep Learning Framework for Objective Assessment of Speech Enhancement Technologies Using EEG-Based Alpha Oscillations
por: Sung, Ching-Chih, et al.
Publicado: (2025) -
Bridging Attribution and Open-Set Detection using Graph-Augmented Instance Learning in Synthetic Speech
por: Akhtar, Mohd Mujtaba, et al.
Publicado: (2026)