Neck-Learn: Attention-Based Multiple Instance Learning and Ensemble Framework for Ecological Momentary Assessment
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Cheema, Ahsan Jamal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
Uncertainty-Based Ensemble Learning For Speech Classification
von: Atmaja, Bagus Tris, et al.
Veröffentlicht: (2024)
von: Atmaja, Bagus Tris, et al.
Veröffentlicht: (2024)
Utilizing Information Theoretic Approach to Study Cochlear Neural Degeneration
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025)
EffortNet: A Deep Learning Framework for Objective Assessment of Speech Enhancement Technologies Using EEG-Based Alpha Oscillations
von: Sung, Ching-Chih, et al.
Veröffentlicht: (2025)
von: Sung, Ching-Chih, et al.
Veröffentlicht: (2025)
Bridging Attribution and Open-Set Detection using Graph-Augmented Instance Learning in Synthetic Speech
von: Akhtar, Mohd Mujtaba, et al.
Veröffentlicht: (2026)
von: Akhtar, Mohd Mujtaba, et al.
Veröffentlicht: (2026)
Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification
von: Huang, Wen, et al.
Veröffentlicht: (2024)
von: Huang, Wen, et al.
Veröffentlicht: (2024)
Automatic Assessment of Dysarthria Using Audio-visual Vowel Graph Attention Network
von: Liu, Xiaokang, et al.
Veröffentlicht: (2024)
von: Liu, Xiaokang, et al.
Veröffentlicht: (2024)
A Multimodal Framework for the Assessment of the Schizophrenia Spectrum
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
Learning Representation of Therapist Empathy in Counseling Conversation Using Siamese Hierarchical Attention Network
von: Tao, Dehua, et al.
Veröffentlicht: (2023)
von: Tao, Dehua, et al.
Veröffentlicht: (2023)
Tracking Listener Attention: Gaze-Guided Audio-Visual Speech Enhancement Framework
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
von: Yang, Hsiang-Cheng, et al.
Veröffentlicht: (2026)
Unifying Listener Scoring Scales: Comparison Learning Framework for Speech Quality Assessment and Continuous Speech Emotion Recognition
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
First Deep Learning Approach to Hammering Acoustics for Stem Stability Assessment in Total Hip Arthroplasty
von: Zhu, Dongqi, et al.
Veröffentlicht: (2025)
von: Zhu, Dongqi, et al.
Veröffentlicht: (2025)
SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2025)
Estimating the Number and Locations of Boundaries in Reverberant Environments with Deep Learning
von: Arikan, Toros, et al.
Veröffentlicht: (2024)
von: Arikan, Toros, et al.
Veröffentlicht: (2024)
Bayesian Speech Synthesizers Can Learn from Multiple Teachers
von: Zhang, Ziyang, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyang, et al.
Veröffentlicht: (2025)
A Pre-training Framework that Encodes Noise Information for Speech Quality Assessment
von: Sultana, Subrina, et al.
Veröffentlicht: (2024)
von: Sultana, Subrina, et al.
Veröffentlicht: (2024)
Coherence-Based Frequency Subset Selection For Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2022)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2022)
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Contextual Biasing for LLM-Based ASR with Hotword Retrieval and Reinforcement Learning
von: Kong, YuXiang, et al.
Veröffentlicht: (2025)
von: Kong, YuXiang, et al.
Veröffentlicht: (2025)
Multi-Granularity Adaptive Time-Frequency Attention Framework for Audio Deepfake Detection under Real-World Communication Degradations
von: Shi, Haohan, et al.
Veröffentlicht: (2025)
von: Shi, Haohan, et al.
Veröffentlicht: (2025)
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning
von: Ouyang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Ouyang, Zhicheng, et al.
Veröffentlicht: (2026)
Learning Multidimensional Disentangled Representations of Instrumental Sounds for Musical Similarity Assessment
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2024)
Automatic Detection of Depression in Speech Using Ensemble Convolutional Neural Networks
von: Vázquez-Romero, Adrián, et al.
Veröffentlicht: (2024)
von: Vázquez-Romero, Adrián, et al.
Veröffentlicht: (2024)
Learning How to Listen: A Temporal-Frequential Attention Model for Sound Event Detection
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
von: Shen, Yu-Han, et al.
Veröffentlicht: (2018)
IR-UWB Radar-Based Contactless Silent Speech Recognition with Attention-Enhanced Temporal Convolutional Networks
von: Lee, Sunghwa, et al.
Veröffentlicht: (2025)
von: Lee, Sunghwa, et al.
Veröffentlicht: (2025)
Acoustic Non-Stationarity Objective Assessment with Hard Label Criteria for Supervised Learning Models
von: Zucatelli, Guilherme, et al.
Veröffentlicht: (2025)
von: Zucatelli, Guilherme, et al.
Veröffentlicht: (2025)
AudioCIL: A Python Toolbox for Audio Class-Incremental Learning with Multiple Scenes
von: Xu, Qisheng, et al.
Veröffentlicht: (2024)
von: Xu, Qisheng, et al.
Veröffentlicht: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment
von: Wang, Wei, et al.
Veröffentlicht: (2025)
von: Wang, Wei, et al.
Veröffentlicht: (2025)
Attention-Based Audio Embeddings for Query-by-Example
von: Singh, Anup, et al.
Veröffentlicht: (2022)
von: Singh, Anup, et al.
Veröffentlicht: (2022)
I-DCCRN-VAE: An Improved Deep Representation Learning Framework for Complex VAE-based Single-channel Speech Enhancement
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
Lightweight Resolution-Aware Audio Deepfake Detection via Cross-Scale Attention and Consistency Learning
von: Shahriar, K. A.
Veröffentlicht: (2026)
von: Shahriar, K. A.
Veröffentlicht: (2026)
How Attention Shapes Emotion: A Comparative Study of Attention Mechanisms for Speech Emotion Recognition
von: Casals-Salvador, Marc, et al.
Veröffentlicht: (2026)
von: Casals-Salvador, Marc, et al.
Veröffentlicht: (2026)
AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
von: Tatarjitzky, Michael, et al.
Veröffentlicht: (2025)
Deep Learning-Based Prediction of Energy Decay Curves from Room Geometry and Material Properties
von: Muhammad, Imran, et al.
Veröffentlicht: (2025)
von: Muhammad, Imran, et al.
Veröffentlicht: (2025)
Controlling Contrastive Self-Supervised Learning with Knowledge-Driven Multiple Hypothesis: Application to Beat Tracking
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
von: Gagnere, Antonin, et al.
Veröffentlicht: (2025)
Attention-based Interactive Disentangling Network for Instance-level Emotional Voice Conversion
von: Chen, Yun, et al.
Veröffentlicht: (2023)
von: Chen, Yun, et al.
Veröffentlicht: (2023)
Music Source Separation Based on a Lightweight Deep Learning Framework (DTTNET: DUAL-PATH TFC-TDF UNET)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
von: Chen, Junyu, et al.
Veröffentlicht: (2023)
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
von: Pierotti, Francesco, et al.
Veröffentlicht: (2025)
von: Pierotti, Francesco, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Using Neurogram Similarity Index Measure (NSIM) to Model Hearing Loss and Cochlear Neural Degeneration
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025) -
Uncertainty-Based Ensemble Learning For Speech Classification
von: Atmaja, Bagus Tris, et al.
Veröffentlicht: (2024) -
Utilizing Information Theoretic Approach to Study Cochlear Neural Degeneration
von: Cheema, Ahsan J., et al.
Veröffentlicht: (2025) -
EffortNet: A Deep Learning Framework for Objective Assessment of Speech Enhancement Technologies Using EEG-Based Alpha Oscillations
von: Sung, Ching-Chih, et al.
Veröffentlicht: (2025) -
Bridging Attribution and Open-Set Detection using Graph-Augmented Instance Learning in Synthetic Speech
von: Akhtar, Mohd Mujtaba, et al.
Veröffentlicht: (2026)