Saved in:
| Main Authors: | Muhammad, Imran, Schuller, Gerald |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.24769 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Room Impulse Response Prediction with Neural Networks: From Energy Decay Curves to Perceptual Validation
by: Muhammad, Imran, et al.
Published: (2025)
by: Muhammad, Imran, et al.
Published: (2025)
3D Room Geometry Inference from Multichannel Room Impulse Response using Deep Neural Network
by: Yeon, Inmo, et al.
Published: (2024)
by: Yeon, Inmo, et al.
Published: (2024)
Deep Room Impulse Response Completion
by: Lin, Jackie, et al.
Published: (2024)
by: Lin, Jackie, et al.
Published: (2024)
Multimodal Deep Learning Method for Real-Time Spatial Room Impulse Response Computing
by: Li, Zhiyu, et al.
Published: (2026)
by: Li, Zhiyu, et al.
Published: (2026)
AttentiveMOS: A Lightweight Attention-Only Model for Speech Quality Prediction
by: Kibria, Imran E, et al.
Published: (2024)
by: Kibria, Imran E, et al.
Published: (2024)
Predicting Global HRTFs From Scanned Head Geometry Using Deep Learning and Compact Representations
by: Wang, Yuxiang, et al.
Published: (2022)
by: Wang, Yuxiang, et al.
Published: (2022)
Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers
by: Latif, Siddique, et al.
Published: (2023)
by: Latif, Siddique, et al.
Published: (2023)
Curved Worlds, Clear Boundaries: Generalizing Speech Deepfake Detection using Hyperbolic and Spherical Geometry Spaces
by: Sheth, Farhan, et al.
Published: (2025)
by: Sheth, Farhan, et al.
Published: (2025)
Quantifying Dimensional Independence in Speech: An Information-Theoretic Framework for Disentangled Representation Learning
by: Kashyap, Bipasha, et al.
Published: (2026)
by: Kashyap, Bipasha, et al.
Published: (2026)
EchoScan: Scanning Complex Room Geometries via Acoustic Echoes
by: Yeon, Inmo, et al.
Published: (2023)
by: Yeon, Inmo, et al.
Published: (2023)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
by: Rajapakshe, Thejan, et al.
Published: (2022)
by: Rajapakshe, Thejan, et al.
Published: (2022)
Learning Filters in Feedback Delay Networks from Noisy Room Impulse Responses
by: Santo, Gloria Dal, et al.
Published: (2025)
by: Santo, Gloria Dal, et al.
Published: (2025)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
by: Jing, Xin, et al.
Published: (2024)
by: Jing, Xin, et al.
Published: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
by: Jing, Xin, et al.
Published: (2024)
by: Jing, Xin, et al.
Published: (2024)
Room Impulse Responses help attackers to evade Deep Fake Detection
by: Luong, Hieu-Thi, et al.
Published: (2024)
by: Luong, Hieu-Thi, et al.
Published: (2024)
Room Impulse Response Completion Using Signal-Prediction Diffusion Models Conditioned on Simulated Early Reflections
by: Xu, Zeyu, et al.
Published: (2026)
by: Xu, Zeyu, et al.
Published: (2026)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
Computer Audition: From Task-Specific Machine Learning to Foundation Models
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
RGI-Net: 3D Room Geometry Inference from Room Impulse Responses With Hidden First-Order Reflections
by: Yeon, Inmo, et al.
Published: (2023)
by: Yeon, Inmo, et al.
Published: (2023)
A Comprehensive Survey on Heart Sound Analysis in the Deep Learning Era
by: Ren, Zhao, et al.
Published: (2023)
by: Ren, Zhao, et al.
Published: (2023)
Low-Rank Adaptation of Deep Prior Neural Networks For Room Impulse Response Reconstruction
by: Pezzoli, Mirco, et al.
Published: (2025)
by: Pezzoli, Mirco, et al.
Published: (2025)
Audio-based Step-count Estimation for Running -- Windowing and Neural Network Baselines
by: Wagner, Philipp, et al.
Published: (2024)
by: Wagner, Philipp, et al.
Published: (2024)
An Adaptive Method for Target Curve Selection
by: Ravizza, Gabriele, et al.
Published: (2025)
by: Ravizza, Gabriele, et al.
Published: (2025)
Blind Identification of Binaural Room Impulse Responses from Smart Glasses
by: Deppisch, Thomas, et al.
Published: (2024)
by: Deppisch, Thomas, et al.
Published: (2024)
Cross-Dialect Bird Species Recognition with Dialect-Calibrated Augmentation
by: Ding, Jiani, et al.
Published: (2025)
by: Ding, Jiani, et al.
Published: (2025)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
by: Li, Yupei, et al.
Published: (2024)
by: Li, Yupei, et al.
Published: (2024)
Intelligent Cardiac Auscultation for Murmur Detection via Parallel-Attentive Models with Uncertainty Estimation
by: Zhang, Zixing, et al.
Published: (2024)
by: Zhang, Zixing, et al.
Published: (2024)
Discovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models
by: Zhao, Xiutian, et al.
Published: (2026)
by: Zhao, Xiutian, et al.
Published: (2026)
Room compensation for loudspeaker reproduction using a supporting source
by: Brooks-Park, James, et al.
Published: (2026)
by: Brooks-Park, James, et al.
Published: (2026)
DARAS: Dynamic Audio-Room Acoustic Synthesis for Blind Room Impulse Response Estimation
by: Wang, Chunxi, et al.
Published: (2025)
by: Wang, Chunxi, et al.
Published: (2025)
Abusive Speech Detection in Indic Languages Using Acoustic Features
by: Spiesberger, Anika A., et al.
Published: (2024)
by: Spiesberger, Anika A., et al.
Published: (2024)
An automatic analysis of ultrasound vocalisations for the prediction of interaction context in captive Egyptian fruit bats
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
by: Rampp, Simon, et al.
Published: (2024)
by: Rampp, Simon, et al.
Published: (2024)
Adapting a Text-to-Audio Model for Room Impulse Response Generation
by: Kim, Kirak, et al.
Published: (2026)
by: Kim, Kirak, et al.
Published: (2026)
Exploring the Power of Pure Attention Mechanisms in Blind Room Parameter Estimation
by: Wang, Chunxi, et al.
Published: (2024)
by: Wang, Chunxi, et al.
Published: (2024)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
by: MacWilliam, Kathleen, et al.
Published: (2024)
by: MacWilliam, Kathleen, et al.
Published: (2024)
Multiple Speaker Separation from Noisy Sources in Reverberant Rooms using Relative Transfer Matrix
by: Manamperi, Wageesha N., et al.
Published: (2025)
by: Manamperi, Wageesha N., et al.
Published: (2025)
Explainable Detection of Machine Generated Music and Early Systematic Evaluation
by: Li, Yupei, et al.
Published: (2024)
by: Li, Yupei, et al.
Published: (2024)
AnyRIR: Robust Non-intrusive Room Impulse Response Estimation in the Wild
by: Lee, Kyung Yun, et al.
Published: (2025)
by: Lee, Kyung Yun, et al.
Published: (2025)
StreamMark: A Deep Learning-Based Semi-Fragile Audio Watermarking for Proactive Deepfake Detection
by: Liu, Zhentao, et al.
Published: (2026)
by: Liu, Zhentao, et al.
Published: (2026)
Similar Items
-
Room Impulse Response Prediction with Neural Networks: From Energy Decay Curves to Perceptual Validation
by: Muhammad, Imran, et al.
Published: (2025) -
3D Room Geometry Inference from Multichannel Room Impulse Response using Deep Neural Network
by: Yeon, Inmo, et al.
Published: (2024) -
Deep Room Impulse Response Completion
by: Lin, Jackie, et al.
Published: (2024) -
Multimodal Deep Learning Method for Real-Time Spatial Room Impulse Response Computing
by: Li, Zhiyu, et al.
Published: (2026) -
AttentiveMOS: A Lightweight Attention-Only Model for Speech Quality Prediction
by: Kibria, Imran E, et al.
Published: (2024)