Real-Time Auralization for First-Person Vocal Interaction in Immersive Virtual Environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Flores-Vargas, Mauricio, Bates, Enda, McDonnell, Rachel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Inferring trust in recommendation systems from brain, behavioural, and physiological data
por: Cheung, Vincent K. M., et al.
Publicado: (2025)
por: Cheung, Vincent K. M., et al.
Publicado: (2025)
Auditory Attention Decoding without Spatial Information: A Diotic EEG Study
por: Yoshino, Masahiro, et al.
Publicado: (2026)
por: Yoshino, Masahiro, et al.
Publicado: (2026)
Subject Disentanglement Neural Network for Speech Envelope Reconstruction from EEG
por: Zhang, Li, et al.
Publicado: (2025)
por: Zhang, Li, et al.
Publicado: (2025)
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
por: Sato, Motoshige, et al.
Publicado: (2026)
por: Sato, Motoshige, et al.
Publicado: (2026)
Active noise cancellation on open-ear smart glasses
por: Yuan, Kuang, et al.
Publicado: (2026)
por: Yuan, Kuang, et al.
Publicado: (2026)
Zero-Shot KWS for Children's Speech using Layer-Wise Features from SSL Models
por: Kutum, Subham, et al.
Publicado: (2025)
por: Kutum, Subham, et al.
Publicado: (2025)
Sound Source Localization for Human-Robot Interaction in Outdoor Environments
por: Liu, Victor, et al.
Publicado: (2025)
por: Liu, Victor, et al.
Publicado: (2025)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
por: Li, Yue, et al.
Publicado: (2024)
por: Li, Yue, et al.
Publicado: (2024)
Toward using Speech to Sense Student Emotion in Remote Learning Environments
por: Vyas, Sargam, et al.
Publicado: (2026)
por: Vyas, Sargam, et al.
Publicado: (2026)
Collecting Prosody in the Wild: A Content-Controlled, Privacy-First Smartphone Protocol and Empirical Evaluation
por: Koch, Timo K., et al.
Publicado: (2026)
por: Koch, Timo K., et al.
Publicado: (2026)
ASE: Practical Acoustic Speed Estimation Beyond Doppler via Sound Diffusion Field
por: Lyu, Sheng, et al.
Publicado: (2024)
por: Lyu, Sheng, et al.
Publicado: (2024)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
por: Mishra, Ruchik, et al.
Publicado: (2024)
por: Mishra, Ruchik, et al.
Publicado: (2024)
Comparative Analysis of Personalized Voice Activity Detection Systems: Assessing Real-World Effectiveness
por: Kumar, Satyam, et al.
Publicado: (2024)
por: Kumar, Satyam, et al.
Publicado: (2024)
Leveraging Unlabeled Audio-Visual Data in Speech Emotion Recognition using Knowledge Distillation
por: Pendyala, Varsha, et al.
Publicado: (2025)
por: Pendyala, Varsha, et al.
Publicado: (2025)
Collection: UAV-Based RSS Measurements from the AFAR Challenge in Digital Twin and Real-World Environments
por: Masrur, Saad, et al.
Publicado: (2025)
por: Masrur, Saad, et al.
Publicado: (2025)
Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation
por: Kim, Miseul, et al.
Publicado: (2024)
por: Kim, Miseul, et al.
Publicado: (2024)
Springboard, Roadblock or "Crutch"?: How Transgender Users Leverage Voice Changers for Gender Presentation in Social Virtual Reality
por: Povinelli, Kassie, et al.
Publicado: (2024)
por: Povinelli, Kassie, et al.
Publicado: (2024)
Single-word Auditory Attention Decoding Using Deep Learning Model
por: Nguyen, Nhan Duc Thanh, et al.
Publicado: (2024)
por: Nguyen, Nhan Duc Thanh, et al.
Publicado: (2024)
Real-Time Word-Level Temporal Segmentation in Streaming Speech Recognition
por: Nishida, Naoto, et al.
Publicado: (2025)
por: Nishida, Naoto, et al.
Publicado: (2025)
Interactive Sonification for Health and Energy using ChucK and Unity
por: Zhao, Yichun, et al.
Publicado: (2024)
por: Zhao, Yichun, et al.
Publicado: (2024)
Evolving Performance Practices in Beethoven's Cello Sonatas: Tempo, Portamento, and Historical Interpretation of the First Movements
por: Sole, Ignasi
Publicado: (2025)
por: Sole, Ignasi
Publicado: (2025)
Live Vocal Extraction from K-pop Performances
por: Kim, Yujin, et al.
Publicado: (2025)
por: Kim, Yujin, et al.
Publicado: (2025)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System
por: Kato, Kazushi, et al.
Publicado: (2025)
por: Kato, Kazushi, et al.
Publicado: (2025)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
por: Piao, Ziyue, et al.
Publicado: (2024)
por: Piao, Ziyue, et al.
Publicado: (2024)
WSCoach: Wearable Real-time Auditory Feedback for Reducing Unwanted Words in Daily Communication
por: Youpeng, Zhang, et al.
Publicado: (2025)
por: Youpeng, Zhang, et al.
Publicado: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
por: Zheng, Shuoyang, et al.
Publicado: (2024)
por: Zheng, Shuoyang, et al.
Publicado: (2024)
BioSonix: Can Physics-Based Sonification Perceptualize Tissue Deformations From Tool Interactions?
por: Ruozzi, Veronica, et al.
Publicado: (2025)
por: Ruozzi, Veronica, et al.
Publicado: (2025)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
por: Liu, Ailin, et al.
Publicado: (2024)
por: Liu, Ailin, et al.
Publicado: (2024)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
por: Ma, Yong, et al.
Publicado: (2025)
por: Ma, Yong, et al.
Publicado: (2025)
Poster: Recognizing Hidden-in-the-Ear Private Key for Reliable Silent Speech Interface Using Multi-Task Learning
por: Dong, Xuefu, et al.
Publicado: (2025)
por: Dong, Xuefu, et al.
Publicado: (2025)
Cluster-to-Predict Affect Contours from Speech
por: Kuşçu, Gökhan, et al.
Publicado: (2024)
por: Kuşçu, Gökhan, et al.
Publicado: (2024)
Silent Speech Sentence Recognition with Six-Axis Accelerometers using Conformer and CTC Algorithm
por: Xie, Yudong, et al.
Publicado: (2025)
por: Xie, Yudong, et al.
Publicado: (2025)
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
por: Arya, Lalaram, et al.
Publicado: (2026)
por: Arya, Lalaram, et al.
Publicado: (2026)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
por: Shi, Runwu, et al.
Publicado: (2024)
por: Shi, Runwu, et al.
Publicado: (2024)
Transhuman Ansambl - Voice Beyond Language
por: Ivsic, Lucija, et al.
Publicado: (2024)
por: Ivsic, Lucija, et al.
Publicado: (2024)
High-Density MIMO Localization Using a 32x64 Ultrasonic Transducer-Microphone Array with Real-Time Data Streaming
por: Baeyens, Rens, et al.
Publicado: (2025)
por: Baeyens, Rens, et al.
Publicado: (2025)
Chord Colourizer: A Near Real-Time System for Visualizing Musical Key
por: Haimes, Paul
Publicado: (2025)
por: Haimes, Paul
Publicado: (2025)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
por: Wu, Liang-Yuan, et al.
Publicado: (2025)
por: Wu, Liang-Yuan, et al.
Publicado: (2025)
Transferable Selective Virtual Sensing Active Noise Control Technique Based on Metric Learning
por: Wang, Boxiang, et al.
Publicado: (2024)
por: Wang, Boxiang, et al.
Publicado: (2024)
Ejemplares similares
-
Inferring trust in recommendation systems from brain, behavioural, and physiological data
por: Cheung, Vincent K. M., et al.
Publicado: (2025) -
Auditory Attention Decoding without Spatial Information: A Diotic EEG Study
por: Yoshino, Masahiro, et al.
Publicado: (2026) -
Subject Disentanglement Neural Network for Speech Envelope Reconstruction from EEG
por: Zhang, Li, et al.
Publicado: (2025) -
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
por: Sato, Motoshige, et al.
Publicado: (2026) -
Active noise cancellation on open-ear smart glasses
por: Yuan, Kuang, et al.
Publicado: (2026)