Salvato in:
| Autori principali: | Zang, Yongyi, Kong, Qiuqiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2503.17866 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Training-Free Multi-Step Audio Source Separation
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
Piano Transcription by Hierarchical Language Modeling with Pretrained Roll-based Encoders
di: Li, Dichucheng, et al.
Pubblicazione: (2025)
di: Li, Dichucheng, et al.
Pubblicazione: (2025)
Ambisonizer: Neural Upmixing as Spherical Harmonics Generation
di: Zang, Yongyi, et al.
Pubblicazione: (2024)
di: Zang, Yongyi, et al.
Pubblicazione: (2024)
Music Source Restoration
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
HiFi-HARP: A High-Fidelity 7th-Order Ambisonic Room Impulse Response Dataset
di: Saini, Shivam, et al.
Pubblicazione: (2025)
di: Saini, Shivam, et al.
Pubblicazione: (2025)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2026)
di: Masuyama, Yoshiki, et al.
Pubblicazione: (2026)
Voices of Civilizations: A Multilingual QA Benchmark for Global Music Understanding
di: Wu, Shangda, et al.
Pubblicazione: (2026)
di: Wu, Shangda, et al.
Pubblicazione: (2026)
MSRBench: A Benchmarking Dataset for Music Source Restoration
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
HARP: A Large-Scale Higher-Order Ambisonic Room Impulse Response Dataset
di: Saini, Shivam, et al.
Pubblicazione: (2024)
di: Saini, Shivam, et al.
Pubblicazione: (2024)
PromptReverb: Multimodal Room Impulse Response Generation Through Latent Rectified Flow Matching
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
Summary of The Inaugural Music Source Restoration Challenge
di: Zang, Yongyi, et al.
Pubblicazione: (2026)
di: Zang, Yongyi, et al.
Pubblicazione: (2026)
Direction-Aware Neural Acoustic Fields for Few-Shot Interpolation of Ambisonic Impulse Responses
di: Ick, Christopher, et al.
Pubblicazione: (2025)
di: Ick, Christopher, et al.
Pubblicazione: (2025)
SHroom: A Python Framework for Ambisonics Room Acoustics Simulation and Binaural Rendering
di: Gayer, Yhonatan
Pubblicazione: (2026)
di: Gayer, Yhonatan
Pubblicazione: (2026)
FlowSynth: Instrument Generation Through Distributional Flow Matching and Test-Time Search
di: Yang, Qihui, et al.
Pubblicazione: (2025)
di: Yang, Qihui, et al.
Pubblicazione: (2025)
Efficient Vocal Source Separation Through Windowed Sink Attention
di: Benetatos, Christodoulos, et al.
Pubblicazione: (2025)
di: Benetatos, Christodoulos, et al.
Pubblicazione: (2025)
The Interpretation Gap in Text-to-Music Generation Models
di: Zang, Yongyi, et al.
Pubblicazione: (2024)
di: Zang, Yongyi, et al.
Pubblicazione: (2024)
Direct and Residual Subspace Decomposition of Spatial Room Impulse Responses
di: Deppisch, Thomas, et al.
Pubblicazione: (2022)
di: Deppisch, Thomas, et al.
Pubblicazione: (2022)
MusicScore: A Dataset for Music Score Modeling and Generation
di: Lin, Yuheng, et al.
Pubblicazione: (2024)
di: Lin, Yuheng, et al.
Pubblicazione: (2024)
FOA Tokenizer: Low-bitrate Neural Codec for First Order Ambisonics with Spatial Consistency Loss
di: Sudarsanam, Parthasaarathy, et al.
Pubblicazione: (2025)
di: Sudarsanam, Parthasaarathy, et al.
Pubblicazione: (2025)
Accelerated Interactive Auralization of Highly Reverberant Spaces using Graphics Hardware
di: Rosseel, Hannes, et al.
Pubblicazione: (2025)
di: Rosseel, Hannes, et al.
Pubblicazione: (2025)
Room Impulse Responses help attackers to evade Deep Fake Detection
di: Luong, Hieu-Thi, et al.
Pubblicazione: (2024)
di: Luong, Hieu-Thi, et al.
Pubblicazione: (2024)
AuralNet: Hierarchical Attention-based 3D Binaural Localization of Overlapping Speakers
di: Fu, Linya, et al.
Pubblicazione: (2025)
di: Fu, Linya, et al.
Pubblicazione: (2025)
Learning Interpretable Features in Audio Latent Spaces via Sparse Autoencoders
di: Paek, Nathan, et al.
Pubblicazione: (2025)
di: Paek, Nathan, et al.
Pubblicazione: (2025)
Residual Learning for Neural Ambisonics Encoders
di: Deppisch, Thomas, et al.
Pubblicazione: (2026)
di: Deppisch, Thomas, et al.
Pubblicazione: (2026)
Perceptually Transparent Binaural Auralization of Simulated Sound Fields
di: Ahrens, Jens
Pubblicazione: (2024)
di: Ahrens, Jens
Pubblicazione: (2024)
Region-Specific Audio Tagging for Spatial Sound
di: Zhao, Jinzheng, et al.
Pubblicazione: (2025)
di: Zhao, Jinzheng, et al.
Pubblicazione: (2025)
Ambisonics Networks -- The Effect Of Radial Functions Regularization
di: Shaybet, Bar, et al.
Pubblicazione: (2024)
di: Shaybet, Bar, et al.
Pubblicazione: (2024)
Blind Spatial Impulse Response Generation from Separate Room- and Scene-Specific Information
di: Lluís, Francesc, et al.
Pubblicazione: (2024)
di: Lluís, Francesc, et al.
Pubblicazione: (2024)
Room Impulse Response Synthesis via Differentiable Feedback Delay Networks for Efficient Spatial Audio Rendering
di: Gerami, Armin, et al.
Pubblicazione: (2025)
di: Gerami, Armin, et al.
Pubblicazione: (2025)
Neural Ambisonics encoding for compact irregular microphone arrays
di: Heikkinen, Mikko, et al.
Pubblicazione: (2024)
di: Heikkinen, Mikko, et al.
Pubblicazione: (2024)
FineLAP: Taming Heterogeneous Supervision for Fine-grained Language-Audio Pretraining
di: Li, Xiquan, et al.
Pubblicazione: (2026)
di: Li, Xiquan, et al.
Pubblicazione: (2026)
On the Usefulness of Diffusion-Based Room Impulse Response Interpolation to Microphone Array Processing
di: Della Torre, Sagi, et al.
Pubblicazione: (2026)
di: Della Torre, Sagi, et al.
Pubblicazione: (2026)
SingFake: Singing Voice Deepfake Detection
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
Spatial Analysis and Synthesis Methods: Subjective and Objective Evaluations Using Various Microphone Arrays in the Auralization of a Critical Listening Room
di: Pawlak, Alan, et al.
Pubblicazione: (2024)
di: Pawlak, Alan, et al.
Pubblicazione: (2024)
Sensitivity of Room Impulse Responses in Changing Acoustic Environment
di: Prawda, Karolina
Pubblicazione: (2025)
di: Prawda, Karolina
Pubblicazione: (2025)
Room Impulse Response Generation Conditioned on Acoustic Parameters
di: Arellano, Silvia, et al.
Pubblicazione: (2025)
di: Arellano, Silvia, et al.
Pubblicazione: (2025)
Acoustic Volume Rendering for Neural Impulse Response Fields
di: Lan, Zitong, et al.
Pubblicazione: (2024)
di: Lan, Zitong, et al.
Pubblicazione: (2024)
Are you really listening? Boosting Perceptual Awareness in Music-QA Benchmarks
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
di: Zang, Yongyi, et al.
Pubblicazione: (2025)
DiffAU: Diffusion-Based Ambisonics Upscaling
di: Milstein, Amit, et al.
Pubblicazione: (2025)
di: Milstein, Amit, et al.
Pubblicazione: (2025)
Ambisonics Super-Resolution Using A Waveform-Domain Neural Network
di: Nawfal, Ismael, et al.
Pubblicazione: (2025)
di: Nawfal, Ismael, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Training-Free Multi-Step Audio Source Separation
di: Zang, Yongyi, et al.
Pubblicazione: (2025) -
Piano Transcription by Hierarchical Language Modeling with Pretrained Roll-based Encoders
di: Li, Dichucheng, et al.
Pubblicazione: (2025) -
Ambisonizer: Neural Upmixing as Spherical Harmonics Generation
di: Zang, Yongyi, et al.
Pubblicazione: (2024) -
Music Source Restoration
di: Zang, Yongyi, et al.
Pubblicazione: (2025) -
HiFi-HARP: A High-Fidelity 7th-Order Ambisonic Room Impulse Response Dataset
di: Saini, Shivam, et al.
Pubblicazione: (2025)