AuditoryHuM: Auditory Scene Label Generation and Clustering using Human-MLLM Collaboration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhong, Henry, Buchholz, Jörg M., Maclaren, Julian, Carlile, Simon, Lyon, Richard F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A dataset and model for auditory scene recognition for hearing devices: AHEAD-DS and OpenYAMNet
von: Zhong, Henry, et al.
Veröffentlicht: (2025)
von: Zhong, Henry, et al.
Veröffentlicht: (2025)
Identifying Hearing Difficulty Moments in Conversational Audio
von: Collins, Jack, et al.
Veröffentlicht: (2025)
von: Collins, Jack, et al.
Veröffentlicht: (2025)
End-to-end Topographic Auditory Models Replicate Signatures of Human Auditory Cortex
von: Al-Tahan, Haider, et al.
Veröffentlicht: (2025)
von: Al-Tahan, Haider, et al.
Veröffentlicht: (2025)
Speech Denoising with Auditory Models
von: Saddler, Mark R., et al.
Veröffentlicht: (2020)
von: Saddler, Mark R., et al.
Veröffentlicht: (2020)
DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)
AuditoryBench++: Can Language Models Understand Auditory Knowledge without Hearing?
von: Ok, Hyunjong, et al.
Veröffentlicht: (2025)
von: Ok, Hyunjong, et al.
Veröffentlicht: (2025)
AAD-LLM: Neural Attention-Driven Auditory Scene Understanding
von: Jiang, Xilin, et al.
Veröffentlicht: (2025)
von: Jiang, Xilin, et al.
Veröffentlicht: (2025)
Efficient Solutions for Mitigating Initialization Bias in Unsupervised Self-Adaptive Auditory Attention Decoding
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
Auditory Representation Effective for Estimating Vocal Tract Information
von: Irino, Toshio, et al.
Veröffentlicht: (2023)
von: Irino, Toshio, et al.
Veröffentlicht: (2023)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
Auditory Intelligence: Understanding the World Through Sound
von: Nam, Hyeonuk
Veröffentlicht: (2025)
von: Nam, Hyeonuk
Veröffentlicht: (2025)
Moravec's Paradox: Towards an Auditory Turing Test
von: Noever, David, et al.
Veröffentlicht: (2025)
von: Noever, David, et al.
Veröffentlicht: (2025)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
Performance Modeling for Correlation-based Neural Decoding of Auditory Attention to Speech
von: Geirnaert, Simon, et al.
Veröffentlicht: (2025)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2025)
CA-TCN: A Causal-Anticausal Temporal Convolutional Network for Direct Auditory Attention Decoding
von: García-Ugarte, Iñigo, et al.
Veröffentlicht: (2026)
von: García-Ugarte, Iñigo, et al.
Veröffentlicht: (2026)
A General Close-loop Predictive Coding Framework for Auditory Working Memory
von: Yuan, Zhongju, et al.
Veröffentlicht: (2025)
von: Yuan, Zhongju, et al.
Veröffentlicht: (2025)
ISAC: An Invertible and Stable Auditory Filter Bank with Customizable Kernels for ML Integration
von: Haider, Daniel, et al.
Veröffentlicht: (2025)
von: Haider, Daniel, et al.
Veröffentlicht: (2025)
Audio-Visual Target Speaker Extraction with Reverse Selective Auditory Attention
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
von: Tao, Ruijie, et al.
Veröffentlicht: (2024)
StreamAAD: Decoding Spatial Auditory Attention with a Streaming Architecture
von: Qiu, Zelin, et al.
Veröffentlicht: (2024)
von: Qiu, Zelin, et al.
Veröffentlicht: (2024)
Auditory Violence
von: Smirnov, Dimitri
Veröffentlicht: (2025)
von: Smirnov, Dimitri
Veröffentlicht: (2025)
Detect All-Type Deepfake Audio: Wavelet Prompt Tuning for Enhanced Auditory Perception
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
Training a Perceptual Model for Evaluating Auditory Similarity in Music Adversarial Attack
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
Harmonic Detection from Noisy Speech with Auditory Frame Gain for Intelligibility Enhancement
von: Queiroz, A., et al.
Veröffentlicht: (2024)
von: Queiroz, A., et al.
Veröffentlicht: (2024)
Fitting Auditory Filterbanks with Multiresolution Neural Networks
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
von: Lostanlen, Vincent, et al.
Veröffentlicht: (2023)
Auditory Filter Behavior and Updated Estimated Constants
von: Alkhairy, Samiya A
Veröffentlicht: (2026)
von: Alkhairy, Samiya A
Veröffentlicht: (2026)
Evaluating Spatialized Auditory Cues for Rapid Attention Capture in XR
von: Kim, Yoonsang, et al.
Veröffentlicht: (2026)
von: Kim, Yoonsang, et al.
Veröffentlicht: (2026)
Réduire le bruit grâce à la réalité augmentée sonore -- Auditory Concealer
von: Boukhemia, Clara
Veröffentlicht: (2025)
von: Boukhemia, Clara
Veröffentlicht: (2025)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
APG-MOS: Auditory Perception Guided-MOS Predictor for Synthetic Speech
von: Lian, Zhicheng, et al.
Veröffentlicht: (2025)
von: Lian, Zhicheng, et al.
Veröffentlicht: (2025)
Enabling Auditory Large Language Models for Automatic Speech Quality Evaluation
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
von: Wang, Siyin, et al.
Veröffentlicht: (2024)
Papez: Resource-Efficient Speech Separation with Auditory Working Memory
von: Oh, Hyunseok, et al.
Veröffentlicht: (2024)
von: Oh, Hyunseok, et al.
Veröffentlicht: (2024)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
von: Li, Lu, et al.
Veröffentlicht: (2025)
von: Li, Lu, et al.
Veröffentlicht: (2025)
The MUSE Benchmark: Probing Music Perception and Auditory Relational Reasoning in Audio LLMS
von: Carone, Brandon James, et al.
Veröffentlicht: (2025)
von: Carone, Brandon James, et al.
Veröffentlicht: (2025)
Scaling Auditory Cognition via Test-Time Compute in Audio Language Models
von: Dang, Ting, et al.
Veröffentlicht: (2025)
von: Dang, Ting, et al.
Veröffentlicht: (2025)
Listen, Chat, and Remix: Text-Guided Soundscape Remixing for Enhanced Auditory Experience
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
Imagine to Hear: Auditory Knowledge Generation can be an Effective Assistant for Language Models
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
von: Chen, Meng, et al.
Veröffentlicht: (2026)
von: Chen, Meng, et al.
Veröffentlicht: (2026)
Spiking-LEAF: A Learnable Auditory front-end for Spiking Neural Networks
von: Song, Zeyang, et al.
Veröffentlicht: (2023)
von: Song, Zeyang, et al.
Veröffentlicht: (2023)
Interfacing PDM MEMS microphones with PFM spiking systems: Application for Neuromorphic Auditory Sensors
von: Jimenez-Fernandez, Angel, et al.
Veröffentlicht: (2019)
von: Jimenez-Fernandez, Angel, et al.
Veröffentlicht: (2019)
NAACA: Training-Free NeuroAuditory Attentive Cognitive Architecture with Oscillatory Working Memory for Salience-Driven Attention Gating
von: Yuan, Zhongju, et al.
Veröffentlicht: (2026)
von: Yuan, Zhongju, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A dataset and model for auditory scene recognition for hearing devices: AHEAD-DS and OpenYAMNet
von: Zhong, Henry, et al.
Veröffentlicht: (2025) -
Identifying Hearing Difficulty Moments in Conversational Audio
von: Collins, Jack, et al.
Veröffentlicht: (2025) -
End-to-end Topographic Auditory Models Replicate Signatures of Human Auditory Cortex
von: Al-Tahan, Haider, et al.
Veröffentlicht: (2025) -
Speech Denoising with Auditory Models
von: Saddler, Mark R., et al.
Veröffentlicht: (2020) -
DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis
von: Lee, Dongheon, et al.
Veröffentlicht: (2025)