Auptimize: Optimal Placement of Spatial Audio Cues for Extended Reality
Fuente:
arXiv
Salvato in:
| Autori principali: | Cho, Hyunsung, Wang, Alexander, Kartik, Divya, Xie, Emily Liying, Yan, Yukang, Lindlbauer, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AudioMiXR: Spatial Audio Object Manipulation with 6DoF for Sound Design in Augmented Reality
di: Woodard, Brandon, et al.
Pubblicazione: (2025)
di: Woodard, Brandon, et al.
Pubblicazione: (2025)
SonoHaptics: An Audio-Haptic Cursor for Gaze-Based Object Selection in XR
di: Cho, Hyunsung, et al.
Pubblicazione: (2024)
di: Cho, Hyunsung, et al.
Pubblicazione: (2024)
A Framework for Multimodal Medical Image Interaction
di: Schütz, Laura, et al.
Pubblicazione: (2024)
di: Schütz, Laura, et al.
Pubblicazione: (2024)
Sonify Anything: Towards Context-Aware Sonic Interactions in AR
di: Schütz, Laura, et al.
Pubblicazione: (2025)
di: Schütz, Laura, et al.
Pubblicazione: (2025)
Dichotic harmony for the musical practice
di: Madgazin, Vadim R.
Pubblicazione: (2010)
di: Madgazin, Vadim R.
Pubblicazione: (2010)
Enhanced DareFightingICE Competitions: Sound Design and AI Competitions
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
Score Distillation Sampling for Audio: Source Separation, Synthesis, and Beyond
di: Richter-Powell, Jessie, et al.
Pubblicazione: (2025)
di: Richter-Powell, Jessie, et al.
Pubblicazione: (2025)
Self-Improvement for Audio Large Language Model using Unlabeled Speech
di: Wang, Shaowen, et al.
Pubblicazione: (2025)
di: Wang, Shaowen, et al.
Pubblicazione: (2025)
OBHS: An Optimized Block Huffman Scheme for Real-Time Audio Compression
di: Mahfi, Muntahi Safwan, et al.
Pubblicazione: (2025)
di: Mahfi, Muntahi Safwan, et al.
Pubblicazione: (2025)
Compositional Phoneme Approximation for L1-Grounded L2 Pronunciation Training
di: Park, Jisang, et al.
Pubblicazione: (2024)
di: Park, Jisang, et al.
Pubblicazione: (2024)
Audio Foundation Models Outperform Symbolic Representations for Piano Performance Evaluation
di: Dhiman, Jai
Pubblicazione: (2026)
di: Dhiman, Jai
Pubblicazione: (2026)
Adaptive Background Music for a Fighting Game: A Multi-Instrument Volume Modulation Approach
di: Khan, Ibrahim, et al.
Pubblicazione: (2023)
di: Khan, Ibrahim, et al.
Pubblicazione: (2023)
Fighting Game Adaptive Background Music for Improved Gameplay
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
Generation of Musical Timbres using a Text-Guided Diffusion Model
di: Yuan, Weixuan, et al.
Pubblicazione: (2025)
di: Yuan, Weixuan, et al.
Pubblicazione: (2025)
MAIN-VC: Lightweight Speech Representation Disentanglement for One-shot Voice Conversion
di: Li, Pengcheng, et al.
Pubblicazione: (2024)
di: Li, Pengcheng, et al.
Pubblicazione: (2024)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
di: Bhadra, Dipayan, et al.
Pubblicazione: (2025)
di: Bhadra, Dipayan, et al.
Pubblicazione: (2025)
If You Hold Me Without Hurting Me: Pathways to Designing Game Audio for Healthy Escapism and Player Well-being
di: Nunes, Caio, et al.
Pubblicazione: (2025)
di: Nunes, Caio, et al.
Pubblicazione: (2025)
Masked Contrastive Pre-Training Improves Music Audio Key Detection
di: Yonay, Ori, et al.
Pubblicazione: (2026)
di: Yonay, Ori, et al.
Pubblicazione: (2026)
Taming Audio VAEs via Target-KL Regularization
di: Seetharaman, Prem, et al.
Pubblicazione: (2026)
di: Seetharaman, Prem, et al.
Pubblicazione: (2026)
GraFPrint: A GNN-Based Approach for Audio Identification
di: Bhattacharjee, Aditya, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Aditya, et al.
Pubblicazione: (2024)
Scalable Evaluation for Audio Identification via Synthetic Latent Fingerprint Generation
di: Bhattacharjee, Aditya, et al.
Pubblicazione: (2025)
di: Bhattacharjee, Aditya, et al.
Pubblicazione: (2025)
GestoBrush: Facilitating Graffiti Artists' Digital Creation Experiences through Embodied AR Interactions
di: Chen, Ruiqi, et al.
Pubblicazione: (2025)
di: Chen, Ruiqi, et al.
Pubblicazione: (2025)
CompanionCast: Toward Social Collaboration with Multi-Agent Systems in Shared Experiences
di: Wang, Yiyang, et al.
Pubblicazione: (2025)
di: Wang, Yiyang, et al.
Pubblicazione: (2025)
MaskClip: Detachable Clip-on Piezoelectric Sensing of Mask Surface Vibrations for Real-time Noise-Robust Speech Input
di: Hiraki, Hirotaka, et al.
Pubblicazione: (2025)
di: Hiraki, Hirotaka, et al.
Pubblicazione: (2025)
MoXaRt: Audio-Visual Object-Guided Sound Interaction for XR
di: Xu, Tianyu, et al.
Pubblicazione: (2026)
di: Xu, Tianyu, et al.
Pubblicazione: (2026)
Can pre-trained Deep Learning models predict groove ratings?
di: Marmoret, Axel, et al.
Pubblicazione: (2026)
di: Marmoret, Axel, et al.
Pubblicazione: (2026)
Acoustic Wave Modeling Using 2D FDTD: Applications in Unreal Engine For Dynamic Sound Rendering
di: Samsurya, Bilkent
Pubblicazione: (2025)
di: Samsurya, Bilkent
Pubblicazione: (2025)
Quantum-Enhanced Analysis and Grading of Vocal Performance
di: Agarwal, Rohan
Pubblicazione: (2025)
di: Agarwal, Rohan
Pubblicazione: (2025)
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
di: Kim, Minu, et al.
Pubblicazione: (2025)
di: Kim, Minu, et al.
Pubblicazione: (2025)
Spatial Audio Rendering for Real-Time Speech Translation in Virtual Meetings
di: Geleta, Margarita, et al.
Pubblicazione: (2025)
di: Geleta, Margarita, et al.
Pubblicazione: (2025)
BemaGANv2: Discriminator Combination Strategies for GAN-based Vocoders in Long-Term Audio Generation
di: Park, Taesoo, et al.
Pubblicazione: (2025)
di: Park, Taesoo, et al.
Pubblicazione: (2025)
SABER: Spatial Attention, Brain, Extended Reality
di: Bullock, Tom, et al.
Pubblicazione: (2026)
di: Bullock, Tom, et al.
Pubblicazione: (2026)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
di: Aristorenas, Aris J.
Pubblicazione: (2024)
di: Aristorenas, Aris J.
Pubblicazione: (2024)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
di: Chen, Kuan-Yu, et al.
Pubblicazione: (2025)
di: Chen, Kuan-Yu, et al.
Pubblicazione: (2025)
SFMS-ALR: Script-First Multilingual Speech Synthesis with Adaptive Locale Resolution
di: Donepudi, Dharma Teja
Pubblicazione: (2025)
di: Donepudi, Dharma Teja
Pubblicazione: (2025)
REMAST: Real-time Emotion-based Music Arrangement with Soft Transition
di: Wang, Zihao, et al.
Pubblicazione: (2023)
di: Wang, Zihao, et al.
Pubblicazione: (2023)
Embodied Exploration of Latent Spaces and Explainable AI
di: Wilson, Elizabeth, et al.
Pubblicazione: (2024)
di: Wilson, Elizabeth, et al.
Pubblicazione: (2024)
Two Sonification Methods for the MindCube
di: Liu, Fangzheng, et al.
Pubblicazione: (2025)
di: Liu, Fangzheng, et al.
Pubblicazione: (2025)
AI Harmonizer: Expanding Vocal Expression with a Generative Neurosymbolic Music AI System
di: Blanchard, Lancelot, et al.
Pubblicazione: (2025)
di: Blanchard, Lancelot, et al.
Pubblicazione: (2025)
MineXR: Mining Personalized Extended Reality Interfaces
di: Cho, Hyunsung, et al.
Pubblicazione: (2024)
di: Cho, Hyunsung, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AudioMiXR: Spatial Audio Object Manipulation with 6DoF for Sound Design in Augmented Reality
di: Woodard, Brandon, et al.
Pubblicazione: (2025) -
SonoHaptics: An Audio-Haptic Cursor for Gaze-Based Object Selection in XR
di: Cho, Hyunsung, et al.
Pubblicazione: (2024) -
A Framework for Multimodal Medical Image Interaction
di: Schütz, Laura, et al.
Pubblicazione: (2024) -
Sonify Anything: Towards Context-Aware Sonic Interactions in AR
di: Schütz, Laura, et al.
Pubblicazione: (2025) -
Dichotic harmony for the musical practice
di: Madgazin, Vadim R.
Pubblicazione: (2010)