NAACA: Training-Free NeuroAuditory Attentive Cognitive Architecture with Oscillatory Working Memory for Salience-Driven Attention Gating
Fuente:
arXiv
Guardado en:
| Autores principales: | Yuan, Zhongju, Wiggins, Geraint, Botteldooren, Dick |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A General Close-loop Predictive Coding Framework for Auditory Working Memory
por: Yuan, Zhongju, et al.
Publicado: (2025)
por: Yuan, Zhongju, et al.
Publicado: (2025)
BioOSS: A Bio-Inspired Oscillatory State System with Spatio-Temporal Dynamics
por: Yuan, Zhongju, et al.
Publicado: (2025)
por: Yuan, Zhongju, et al.
Publicado: (2025)
A novel Reservoir Architecture for Periodic Time Series Prediction
por: Yuan, Zhongju, et al.
Publicado: (2024)
por: Yuan, Zhongju, et al.
Publicado: (2024)
A Reservoir-based Model for Human-like Perception of Complex Rhythm Pattern
por: Yuan, Zhongju, et al.
Publicado: (2025)
por: Yuan, Zhongju, et al.
Publicado: (2025)
A Dynamic Systems Approach to Modelling Human-Machine Rhythm Interaction
por: Yuan, Zhongju, et al.
Publicado: (2024)
por: Yuan, Zhongju, et al.
Publicado: (2024)
Yin-Yang: Developing Motifs With Long-Term Structure And Controllability
por: Bhandari, Keshav, et al.
Publicado: (2025)
por: Bhandari, Keshav, et al.
Publicado: (2025)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
por: Wilson, Elizabeth, et al.
Publicado: (2024)
por: Wilson, Elizabeth, et al.
Publicado: (2024)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
por: Tang, Jingjing, et al.
Publicado: (2025)
por: Tang, Jingjing, et al.
Publicado: (2025)
NeuroSpex: Neuro-Guided Speaker Extraction with Cross-Modal Attention
por: De Silva, Dashanka, et al.
Publicado: (2024)
por: De Silva, Dashanka, et al.
Publicado: (2024)
Scaling Auditory Cognition via Test-Time Compute in Audio Language Models
por: Dang, Ting, et al.
Publicado: (2025)
por: Dang, Ting, et al.
Publicado: (2025)
DRASP: A Dual-Resolution Attentive Statistics Pooling Framework for Automatic MOS Prediction
por: Yang, Cheng-Yeh, et al.
Publicado: (2025)
por: Yang, Cheng-Yeh, et al.
Publicado: (2025)
AAD-LLM: Neural Attention-Driven Auditory Scene Understanding
por: Jiang, Xilin, et al.
Publicado: (2025)
por: Jiang, Xilin, et al.
Publicado: (2025)
Fusing Memory and Attention: A study on LSTM, Transformer and Hybrid Architectures for Symbolic Music Generation
por: Ghoshal, Soudeep, et al.
Publicado: (2026)
por: Ghoshal, Soudeep, et al.
Publicado: (2026)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
por: Sun, Zhe, et al.
Publicado: (2025)
por: Sun, Zhe, et al.
Publicado: (2025)
DARNet: Dual Attention Refinement Network with Spatiotemporal Construction for Auditory Attention Detection
por: Yan, Sheng, et al.
Publicado: (2024)
por: Yan, Sheng, et al.
Publicado: (2024)
Detect All-Type Deepfake Audio: Wavelet Prompt Tuning for Enhanced Auditory Perception
por: Xie, Yuankun, et al.
Publicado: (2025)
por: Xie, Yuankun, et al.
Publicado: (2025)
LoopGen: Training-Free Loopable Music Generation
por: Marincione, Davide, et al.
Publicado: (2025)
por: Marincione, Davide, et al.
Publicado: (2025)
AST: Adaptive, Seamless, and Training-Free Precise Speech Editing
por: Lv, Sihan, et al.
Publicado: (2026)
por: Lv, Sihan, et al.
Publicado: (2026)
Scattering Transformer: A Training-Free Transformer Architecture for Heart Murmur Detection
por: Zewail, Rami
Publicado: (2025)
por: Zewail, Rami
Publicado: (2025)
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation with SpeechLLMs
por: Papi, Sara, et al.
Publicado: (2026)
por: Papi, Sara, et al.
Publicado: (2026)
AuditoryBench++: Can Language Models Understand Auditory Knowledge without Hearing?
por: Ok, Hyunjong, et al.
Publicado: (2025)
por: Ok, Hyunjong, et al.
Publicado: (2025)
Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models
por: Li, Yanda, et al.
Publicado: (2026)
por: Li, Yanda, et al.
Publicado: (2026)
Flamed-TTS: Flow Matching Attention-Free Models for Efficient Generating and Dynamic Pacing Zero-shot Text-to-Speech
por: Huynh-Nguyen, Hieu-Nghia, et al.
Publicado: (2025)
por: Huynh-Nguyen, Hieu-Nghia, et al.
Publicado: (2025)
Parallel Delayed Memory Units for Enhanced Temporal Modeling in Biomedical and Bioacoustic Signal Analysis
por: Sun, Pengfei, et al.
Publicado: (2025)
por: Sun, Pengfei, et al.
Publicado: (2025)
Auditory Intelligence: Understanding the World Through Sound
por: Nam, Hyeonuk
Publicado: (2025)
por: Nam, Hyeonuk
Publicado: (2025)
Moravec's Paradox: Towards an Auditory Turing Test
por: Noever, David, et al.
Publicado: (2025)
por: Noever, David, et al.
Publicado: (2025)
ASoBO: Attentive Beamformer Selection for Distant Speaker Diarization in Meetings
por: Mariotte, Theo, et al.
Publicado: (2024)
por: Mariotte, Theo, et al.
Publicado: (2024)
AUREXA-SE: Audio-Visual Unified Representation Exchange Architecture with Cross-Attention and Squeezeformer for Speech Enhancement
por: Sajid, M., et al.
Publicado: (2025)
por: Sajid, M., et al.
Publicado: (2025)
SWIM: Short-Window CNN Integrated with Mamba for EEG-Based Auditory Spatial Attention Decoding
por: Zhang, Ziyang, et al.
Publicado: (2024)
por: Zhang, Ziyang, et al.
Publicado: (2024)
TFGA-Net: Temporal-Frequency Graph Attention Network for Brain-Controlled Speaker Extraction
por: Si, Youhao, et al.
Publicado: (2025)
por: Si, Youhao, et al.
Publicado: (2025)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
por: Chen, Meng, et al.
Publicado: (2026)
por: Chen, Meng, et al.
Publicado: (2026)
DGFNet: End-to-End Audio-Visual Source Separation Based on Dynamic Gating Fusion
por: Yu, Yinfeng, et al.
Publicado: (2025)
por: Yu, Yinfeng, et al.
Publicado: (2025)
APG-MOS: Auditory Perception Guided-MOS Predictor for Synthetic Speech
por: Lian, Zhicheng, et al.
Publicado: (2025)
por: Lian, Zhicheng, et al.
Publicado: (2025)
Neuro-MSBG: An End-to-End Neural Model for Hearing Loss Simulation
por: Yuan, Hui-Guan, et al.
Publicado: (2025)
por: Yuan, Hui-Guan, et al.
Publicado: (2025)
End-to-end Topographic Auditory Models Replicate Signatures of Human Auditory Cortex
por: Al-Tahan, Haider, et al.
Publicado: (2025)
por: Al-Tahan, Haider, et al.
Publicado: (2025)
The MUSE Benchmark: Probing Music Perception and Auditory Relational Reasoning in Audio LLMS
por: Carone, Brandon James, et al.
Publicado: (2025)
por: Carone, Brandon James, et al.
Publicado: (2025)
Lina-Speech: Gated Linear Attention and Initial-State Tuning for Multi-Sample Prompting Text-To-Speech Synthesis
por: Lemerle, Théodor, et al.
Publicado: (2024)
por: Lemerle, Théodor, et al.
Publicado: (2024)
Evaluating Neural Networks Architectures for Spring Reverb Modelling
por: Papaleo, Francesco, et al.
Publicado: (2024)
por: Papaleo, Francesco, et al.
Publicado: (2024)
Automatic Speech Recognition in the Modern Era: Architectures, Training, and Evaluation
por: Nayeem, Md., et al.
Publicado: (2025)
por: Nayeem, Md., et al.
Publicado: (2025)
Speech Emotion Recognition Leveraging OpenAI's Whisper Representations and Attentive Pooling Methods
por: Shendabadi, Ali, et al.
Publicado: (2026)
por: Shendabadi, Ali, et al.
Publicado: (2026)
Ejemplares similares
-
A General Close-loop Predictive Coding Framework for Auditory Working Memory
por: Yuan, Zhongju, et al.
Publicado: (2025) -
BioOSS: A Bio-Inspired Oscillatory State System with Spatio-Temporal Dynamics
por: Yuan, Zhongju, et al.
Publicado: (2025) -
A novel Reservoir Architecture for Periodic Time Series Prediction
por: Yuan, Zhongju, et al.
Publicado: (2024) -
A Reservoir-based Model for Human-like Perception of Complex Rhythm Pattern
por: Yuan, Zhongju, et al.
Publicado: (2025) -
A Dynamic Systems Approach to Modelling Human-Machine Rhythm Interaction
por: Yuan, Zhongju, et al.
Publicado: (2024)