Quantifying Spatial Audio Quality Impairment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Watcharasupat, Karn N., Lerch, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Separate This, and All of these Things Around It: Music Source Separation via Hyperellipsoidal Queries
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022)
Facing the Music: Tackling Singing Voice Separation in Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
Remastering Divide and Remaster: A Cinematic Audio Source Separation Dataset with Multilingual Support
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)
A Generalized Bandsplit Neural Network for Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
ASPED: An Audio Dataset for Detecting Pedestrians
von: Seshadri, Pavan, et al.
Veröffentlicht: (2023)
von: Seshadri, Pavan, et al.
Veröffentlicht: (2023)
Automating Urban Soundscape Enhancements with AI: In-situ Assessment of Quality and Restorativeness in Traffic-Exposed Residential Areas
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
Universal Spatial Audio Transcoder
von: Sagasti, Amaia, et al.
Veröffentlicht: (2024)
von: Sagasti, Amaia, et al.
Veröffentlicht: (2024)
Exploring Perceptual Audio Quality Measurement on Stereo Processing Using the Open Dataset of Audio Quality
von: Delgado, Pablo M., et al.
Veröffentlicht: (2025)
von: Delgado, Pablo M., et al.
Veröffentlicht: (2025)
PAM: Prompting Audio-Language Models for Audio Quality Assessment
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
von: Pierotti, Francesco, et al.
Veröffentlicht: (2025)
von: Pierotti, Francesco, et al.
Veröffentlicht: (2025)
Parameter-Efficient Transfer Learning for Music Foundation Models
von: Ding, Yiwei, et al.
Veröffentlicht: (2024)
von: Ding, Yiwei, et al.
Veröffentlicht: (2024)
AudioSpa: Spatializing Sound Events with Text
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
von: Feng, Linfeng, et al.
Veröffentlicht: (2025)
Region-Specific Audio Tagging for Spatial Sound
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2025)
Past, Present, and Future of Spatial Audio and Room Acoustics
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
von: Koyama, Shoichi, et al.
Veröffentlicht: (2025)
Towards Spatial Audio Understanding via Question Answering
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
von: Sudarsanam, Parthasaarathy, et al.
Veröffentlicht: (2025)
ASAudio: A Survey of Advanced Spatial Audio Research
von: Zhu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhu, Zhiyuan, et al.
Veröffentlicht: (2025)
Can Large Language Models Understand Spatial Audio?
von: Tang, Changli, et al.
Veröffentlicht: (2024)
von: Tang, Changli, et al.
Veröffentlicht: (2024)
Exploring the Potential of Data-Driven Spatial Audio Enhancement Using a Single-Channel Model
von: Santos, Arthur N. dos, et al.
Veröffentlicht: (2024)
von: Santos, Arthur N. dos, et al.
Veröffentlicht: (2024)
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
von: Poole, Katarina C., et al.
Veröffentlicht: (2025)
SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
von: Hu, Jinbo, et al.
Veröffentlicht: (2025)
Diff-SAGe: End-to-End Spatial Audio Generation Using Diffusion Models
von: Kushwaha, Saksham Singh, et al.
Veröffentlicht: (2024)
von: Kushwaha, Saksham Singh, et al.
Veröffentlicht: (2024)
VR-PTOLEMAIC: A Virtual Environment for the Perceptual Testing of Spatial Audio Algorithms
von: Ostan, Paolo, et al.
Veröffentlicht: (2025)
von: Ostan, Paolo, et al.
Veröffentlicht: (2025)
Room Impulse Response Synthesis via Differentiable Feedback Delay Networks for Efficient Spatial Audio Rendering
von: Gerami, Armin, et al.
Veröffentlicht: (2025)
von: Gerami, Armin, et al.
Veröffentlicht: (2025)
UniAudio: An Audio Foundation Model Toward Universal Audio Generation
von: Yang, Dongchao, et al.
Veröffentlicht: (2023)
von: Yang, Dongchao, et al.
Veröffentlicht: (2023)
NOMAD: Unsupervised Learning of Perceptual Embeddings for Speech Enhancement and Non-matching Reference Audio Quality Assessment
von: Ragano, Alessandro, et al.
Veröffentlicht: (2023)
von: Ragano, Alessandro, et al.
Veröffentlicht: (2023)
Stereo Audio Rendering for Personal Sound Zones Using a Binaural Spatially Adaptive Neural Network (BSANN)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
MOSS-Audio-Tokenizer: Scaling Audio Tokenizers for Future Audio Foundation Models
von: Gong, Yitian, et al.
Veröffentlicht: (2026)
von: Gong, Yitian, et al.
Veröffentlicht: (2026)
SAVGBench: Benchmarking Spatially Aligned Audio-Video Generation
von: Shimada, Kazuki, et al.
Veröffentlicht: (2024)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2024)
BickGraphing: Web-Based Application for Visual Inspection of Audio Recordings
von: Seow, Kayley, et al.
Veröffentlicht: (2026)
von: Seow, Kayley, et al.
Veröffentlicht: (2026)
Streaming Audio Transformers for Online Audio Tagging
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2023)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2023)
Pengi: An Audio Language Model for Audio Tasks
von: Deshmukh, Soham, et al.
Veröffentlicht: (2023)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2023)
Discrete Audio Representations for Automated Audio Captioning
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
von: Tian, Jingguang, et al.
Veröffentlicht: (2025)
AV-SSAN: Audio-Visual Selective DoA Estimation through Explicit Multi-Band Semantic-Spatial Alignment
von: Chen, Yu, et al.
Veröffentlicht: (2025)
von: Chen, Yu, et al.
Veröffentlicht: (2025)
Audio-based Step-count Estimation for Running -- Windowing and Neural Network Baselines
von: Wagner, Philipp, et al.
Veröffentlicht: (2024)
von: Wagner, Philipp, et al.
Veröffentlicht: (2024)
MACE: Leveraging Audio for Evaluating Audio Captioning Systems
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
von: Dixit, Satvik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Separate This, and All of these Things Around It: Music Source Separation via Hyperellipsoidal Queries
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025) -
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024) -
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025) -
FRCRN: Boosting Feature Representation using Frequency Recurrence for Monaural Speech Enhancement
von: Zhao, Shengkui, et al.
Veröffentlicht: (2022) -
Facing the Music: Tackling Singing Voice Separation in Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2024)