Real-time auralization for performers on virtual stages
Fuente:
arXiv
Guardado en:
| Autores principales: | Accolti, Ernesto, Aspöck, Lukas, Yadav, Manuj, Vorländer, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Effect of laboratory conditions on the perception of virtual stages for music
por: Accolti, Ernesto
Publicado: (2025)
por: Accolti, Ernesto
Publicado: (2025)
Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality
por: Ermert, Cosima A., et al.
Publicado: (2024)
por: Ermert, Cosima A., et al.
Publicado: (2024)
Noise disturbance and lack of privacy: Modeling acoustic dissatisfaction in open-plan offices
por: Yadav, Manuj, et al.
Publicado: (2025)
por: Yadav, Manuj, et al.
Publicado: (2025)
Two-stage Audio-Visual Target Speaker Extraction System for Real-Time Processing On Edge Device
por: Li, Zixuan, et al.
Publicado: (2025)
por: Li, Zixuan, et al.
Publicado: (2025)
Real-time implementation of vibrato transfer as an audio effect
por: Hyrkas, Jeremy
Publicado: (2025)
por: Hyrkas, Jeremy
Publicado: (2025)
BFA: Real-time Multilingual Text-to-speech Forced Alignment
por: Rehman, Abdul, et al.
Publicado: (2025)
por: Rehman, Abdul, et al.
Publicado: (2025)
Neural Speech Coding for Real-time Communications using Constant Bitrate Scalar Quantization
por: Brendel, Andreas, et al.
Publicado: (2024)
por: Brendel, Andreas, et al.
Publicado: (2024)
StreamVoice: Streamable Context-Aware Language Modeling for Real-time Zero-Shot Voice Conversion
por: Wang, Zhichao, et al.
Publicado: (2024)
por: Wang, Zhichao, et al.
Publicado: (2024)
Noisy Disentanglement with Tri-stage Training for Noise-Robust Speech Recognition
por: Chen, Shuangyuan, et al.
Publicado: (2025)
por: Chen, Shuangyuan, et al.
Publicado: (2025)
A Multi-stage Low-latency Enhancement System for Hearing Aids
por: Ouyang, Chengwei, et al.
Publicado: (2025)
por: Ouyang, Chengwei, et al.
Publicado: (2025)
DroFiT: A Lightweight Band-fused Frequency Attention Toward Real-time UAV Speech Enhancement
por: Lee, Jeongmin, et al.
Publicado: (2025)
por: Lee, Jeongmin, et al.
Publicado: (2025)
Communication conditions in virtual acoustic scenes in an underground station
por: Hládek, Ľuboš, et al.
Publicado: (2021)
por: Hládek, Ľuboš, et al.
Publicado: (2021)
Analysing the Masked predictive coding training criterion for pre-training a Speech Representation Model
por: Yadav, Hemant, et al.
Publicado: (2023)
por: Yadav, Hemant, et al.
Publicado: (2023)
AxLSTMs: learning self-supervised audio representations with xLSTMs
por: Yadav, Sarthak, et al.
Publicado: (2024)
por: Yadav, Sarthak, et al.
Publicado: (2024)
Temporal Pooling Strategies for Training-Free Anomalous Sound Detection with Self-Supervised Audio Embeddings
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
PiCoGen: Generate Piano Covers with a Two-stage Approach
por: Tan, Chih-Pin, et al.
Publicado: (2024)
por: Tan, Chih-Pin, et al.
Publicado: (2024)
Real-time multichannel deep speech enhancement in hearing aids: Comparing monaural and binaural processing in complex acoustic scenarios
por: Westhausen, Nils L., et al.
Publicado: (2024)
por: Westhausen, Nils L., et al.
Publicado: (2024)
A toolbox for rendering virtual acoustic environments in the context of audiology
por: Grimm, Giso, et al.
Publicado: (2018)
por: Grimm, Giso, et al.
Publicado: (2018)
M2R-Whisper: Multi-stage and Multi-scale Retrieval Augmentation for Enhancing Whisper
por: Zhou, Jiaming, et al.
Publicado: (2024)
por: Zhou, Jiaming, et al.
Publicado: (2024)
Real-time Speech Extraction Using Spatially Regularized Independent Low-rank Matrix Analysis and Rank-constrained Spatial Covariance Matrix Estimation
por: Ishikawa, Yuto, et al.
Publicado: (2024)
por: Ishikawa, Yuto, et al.
Publicado: (2024)
Effective User-defined Keyword Spotting with Dual-stage Matching, Multi-modal Enrollment, and Continual Adaptation
por: Ai, Zhiqi, et al.
Publicado: (2026)
por: Ai, Zhiqi, et al.
Publicado: (2026)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
por: Serre, Thomas, et al.
Publicado: (2024)
por: Serre, Thomas, et al.
Publicado: (2024)
A circular microphone array with virtual microphones based on acoustics-informed neural networks
por: Zhao, Sipei, et al.
Publicado: (2024)
por: Zhao, Sipei, et al.
Publicado: (2024)
Efficient learning-based sound propagation for virtual and real-world audio processing applications
por: Ratnarajah, Anton Jeran
Publicado: (2024)
por: Ratnarajah, Anton Jeran
Publicado: (2024)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
por: Yang, Bing, et al.
Publicado: (2024)
por: Yang, Bing, et al.
Publicado: (2024)
A two-stage transliteration approach to improve performance of a multilingual ASR
por: Kumar, Rohit
Publicado: (2024)
por: Kumar, Rohit
Publicado: (2024)
SPGM: Prioritizing Local Features for enhanced speech separation performance
por: Yip, Jia Qi, et al.
Publicado: (2023)
por: Yip, Jia Qi, et al.
Publicado: (2023)
Short-term cognitive fatigue of spatial selective attention after face-to-face conversations in virtual noisy environments
por: Hládek, Ľuboš, et al.
Publicado: (2025)
por: Hládek, Ľuboš, et al.
Publicado: (2025)
Complexity boosted adaptive training for better low resource ASR performance
por: Lu, Hongxuan, et al.
Publicado: (2024)
por: Lu, Hongxuan, et al.
Publicado: (2024)
RaD-Net 2: A causal two-stage repairing and denoising speech enhancement network with knowledge distillation and complex axial self-attention
por: Liu, Mingshuai, et al.
Publicado: (2024)
por: Liu, Mingshuai, et al.
Publicado: (2024)
DiaPer: End-to-End Neural Diarization with Perceiver-Based Attractors
por: Landini, Federico, et al.
Publicado: (2023)
por: Landini, Federico, et al.
Publicado: (2023)
Diffuse Sound Field Synthesis
por: Zotter, Franz, et al.
Publicado: (2024)
por: Zotter, Franz, et al.
Publicado: (2024)
Optimal Real-Weighted Beamforming With Application to Linear and Spherical Arrays
por: Tourbabin, V., et al.
Publicado: (2024)
por: Tourbabin, V., et al.
Publicado: (2024)
nlm: Real-Time Non-linear Modal Synthesis in Max
por: Diaz, Rodrigo, et al.
Publicado: (2026)
por: Diaz, Rodrigo, et al.
Publicado: (2026)
Audiosockets: A Python socket package for Real-Time Audio Processing
por: Shu, Nicolas, et al.
Publicado: (2024)
por: Shu, Nicolas, et al.
Publicado: (2024)
HRTF-guided Binaural Target Speaker Extraction with Real-World Validation
por: Ellinson, Yoav, et al.
Publicado: (2026)
por: Ellinson, Yoav, et al.
Publicado: (2026)
Comparative Analysis of ASR Methods for Speech Deepfake Detection
por: Salvi, Davide, et al.
Publicado: (2024)
por: Salvi, Davide, et al.
Publicado: (2024)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
por: Jiang, Shaohan, et al.
Publicado: (2025)
por: Jiang, Shaohan, et al.
Publicado: (2025)
Real-Time Scream Detection and Position Estimation for Worker Safety in Construction Sites
por: Gautam, Bikalpa, et al.
Publicado: (2024)
por: Gautam, Bikalpa, et al.
Publicado: (2024)
FruitsMusic: A Real-World Corpus of Japanese Idol-Group Songs
por: Suda, Hitoshi, et al.
Publicado: (2024)
por: Suda, Hitoshi, et al.
Publicado: (2024)
Ejemplares similares
-
Effect of laboratory conditions on the perception of virtual stages for music
por: Accolti, Ernesto
Publicado: (2025) -
Audiovisual angle and voice incongruence do not affect audiovisual verbal short-term memory in virtual reality
por: Ermert, Cosima A., et al.
Publicado: (2024) -
Noise disturbance and lack of privacy: Modeling acoustic dissatisfaction in open-plan offices
por: Yadav, Manuj, et al.
Publicado: (2025) -
Two-stage Audio-Visual Target Speaker Extraction System for Real-Time Processing On Edge Device
por: Li, Zixuan, et al.
Publicado: (2025) -
Real-time implementation of vibrato transfer as an audio effect
por: Hyrkas, Jeremy
Publicado: (2025)