Embodied Exploration of Latent Spaces and Explainable AI
Fuente:
arXiv
Guardado en:
| Autores principales: | Wilson, Elizabeth, Satomi, Mika, McLean, Alex, Schubert, Deva, Gonzalez, Juan Felipe Amaya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AI Harmonizer: Expanding Vocal Expression with a Generative Neurosymbolic Music AI System
por: Blanchard, Lancelot, et al.
Publicado: (2025)
por: Blanchard, Lancelot, et al.
Publicado: (2025)
Acoustic Wave Modeling Using 2D FDTD: Applications in Unreal Engine For Dynamic Sound Rendering
por: Samsurya, Bilkent
Publicado: (2025)
por: Samsurya, Bilkent
Publicado: (2025)
Two Sonification Methods for the MindCube
por: Liu, Fangzheng, et al.
Publicado: (2025)
por: Liu, Fangzheng, et al.
Publicado: (2025)
AudioMiXR: Spatial Audio Object Manipulation with 6DoF for Sound Design in Augmented Reality
por: Woodard, Brandon, et al.
Publicado: (2025)
por: Woodard, Brandon, et al.
Publicado: (2025)
Auptimize: Optimal Placement of Spatial Audio Cues for Extended Reality
por: Cho, Hyunsung, et al.
Publicado: (2024)
por: Cho, Hyunsung, et al.
Publicado: (2024)
Enhanced DareFightingICE Competitions: Sound Design and AI Competitions
por: Khan, Ibrahim, et al.
Publicado: (2024)
por: Khan, Ibrahim, et al.
Publicado: (2024)
Tailors: New Music Timbre Visualizer to Entertain Music Through Imagery
por: Lee, ChungHa
Publicado: (2024)
por: Lee, ChungHa
Publicado: (2024)
A Framework for Multimodal Medical Image Interaction
por: Schütz, Laura, et al.
Publicado: (2024)
por: Schütz, Laura, et al.
Publicado: (2024)
WhisperMask: A Noise Suppressive Mask-Type Microphone for Whisper Speech
por: Hiraki, Hirotaka, et al.
Publicado: (2024)
por: Hiraki, Hirotaka, et al.
Publicado: (2024)
Whisphone: Whispering Input Earbuds
por: Fukumoto, Masaaki
Publicado: (2025)
por: Fukumoto, Masaaki
Publicado: (2025)
Compositional Phoneme Approximation for L1-Grounded L2 Pronunciation Training
por: Park, Jisang, et al.
Publicado: (2024)
por: Park, Jisang, et al.
Publicado: (2024)
Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
por: Mehta, Shivam, et al.
Publicado: (2024)
por: Mehta, Shivam, et al.
Publicado: (2024)
M6(GPT)3: Generating Multitrack Modifiable Multi-Minute MIDI Music from Text using Genetic algorithms, Probabilistic methods and GPT Models in any Progression and Time Signature
por: Poćwiardowski, Jakub, et al.
Publicado: (2024)
por: Poćwiardowski, Jakub, et al.
Publicado: (2024)
Sonify Anything: Towards Context-Aware Sonic Interactions in AR
por: Schütz, Laura, et al.
Publicado: (2025)
por: Schütz, Laura, et al.
Publicado: (2025)
Insights on Harmonic Tones from a Generative Music Experiment
por: Deruty, Emmanuel, et al.
Publicado: (2025)
por: Deruty, Emmanuel, et al.
Publicado: (2025)
MaskClip: Detachable Clip-on Piezoelectric Sensing of Mask Surface Vibrations for Real-time Noise-Robust Speech Input
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
Dichotic harmony for the musical practice
por: Madgazin, Vadim R.
Publicado: (2010)
por: Madgazin, Vadim R.
Publicado: (2010)
Matcha-TTS: A fast TTS architecture with conditional flow matching
por: Mehta, Shivam, et al.
Publicado: (2023)
por: Mehta, Shivam, et al.
Publicado: (2023)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
por: Park, Seohyun, et al.
Publicado: (2025)
por: Park, Seohyun, et al.
Publicado: (2025)
Spatial Audio Rendering for Real-Time Speech Translation in Virtual Meetings
por: Geleta, Margarita, et al.
Publicado: (2025)
por: Geleta, Margarita, et al.
Publicado: (2025)
Towards LLM-Empowered Fine-Grained Speech Descriptors for Explainable Emotion Recognition
por: Chen, Youjun, et al.
Publicado: (2025)
por: Chen, Youjun, et al.
Publicado: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
por: Zheng, Shuoyang, et al.
Publicado: (2024)
por: Zheng, Shuoyang, et al.
Publicado: (2024)
A Framework for AI assisted Musical Devices
por: Civit, Miguel, et al.
Publicado: (2024)
por: Civit, Miguel, et al.
Publicado: (2024)
Less Stress, More Privacy: Stress Detection on Anonymized Speech of Air Traffic Controllers
por: Viswanathan, Janaki, et al.
Publicado: (2025)
por: Viswanathan, Janaki, et al.
Publicado: (2025)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
por: Wilson, Elizabeth, et al.
Publicado: (2024)
por: Wilson, Elizabeth, et al.
Publicado: (2024)
Real-Time Emergency Vehicle Detection using Mel Spectrograms and Regular Expressions
por: Pacheco-Gonzalez, Alberto, et al.
Publicado: (2023)
por: Pacheco-Gonzalez, Alberto, et al.
Publicado: (2023)
Scalable Evaluation for Audio Identification via Synthetic Latent Fingerprint Generation
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
por: Khanday, Owais Mujtaba, et al.
Publicado: (2025)
por: Khanday, Owais Mujtaba, et al.
Publicado: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
por: Khanday, Owais Mujtaba, et al.
Publicado: (2025)
por: Khanday, Owais Mujtaba, et al.
Publicado: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
por: Li, Yinan, et al.
Publicado: (2026)
por: Li, Yinan, et al.
Publicado: (2026)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
por: Piao, Ziyue, et al.
Publicado: (2024)
por: Piao, Ziyue, et al.
Publicado: (2024)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
por: Han, Hyewon, et al.
Publicado: (2024)
por: Han, Hyewon, et al.
Publicado: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
por: Zhao, Yichun, et al.
Publicado: (2024)
por: Zhao, Yichun, et al.
Publicado: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
por: Li, Yue, et al.
Publicado: (2024)
por: Li, Yue, et al.
Publicado: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
por: Yu, Luca Jiang-Tao, et al.
Publicado: (2024)
por: Yu, Luca Jiang-Tao, et al.
Publicado: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
por: Blum'e, Ashlae
Publicado: (2025)
por: Blum'e, Ashlae
Publicado: (2025)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
por: Manukalpa, J. M. Chan Sri, et al.
Publicado: (2025)
por: Manukalpa, J. M. Chan Sri, et al.
Publicado: (2025)
Interfacing with history: Curating with audio augmented objects
por: Cliffe, Laurence
Publicado: (2024)
por: Cliffe, Laurence
Publicado: (2024)
Transhuman Ansambl - Voice Beyond Language
por: Ivsic, Lucija, et al.
Publicado: (2024)
por: Ivsic, Lucija, et al.
Publicado: (2024)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
por: Liu, Ailin, et al.
Publicado: (2024)
por: Liu, Ailin, et al.
Publicado: (2024)
Ejemplares similares
-
AI Harmonizer: Expanding Vocal Expression with a Generative Neurosymbolic Music AI System
por: Blanchard, Lancelot, et al.
Publicado: (2025) -
Acoustic Wave Modeling Using 2D FDTD: Applications in Unreal Engine For Dynamic Sound Rendering
por: Samsurya, Bilkent
Publicado: (2025) -
Two Sonification Methods for the MindCube
por: Liu, Fangzheng, et al.
Publicado: (2025) -
AudioMiXR: Spatial Audio Object Manipulation with 6DoF for Sound Design in Augmented Reality
por: Woodard, Brandon, et al.
Publicado: (2025) -
Auptimize: Optimal Placement of Spatial Audio Cues for Extended Reality
por: Cho, Hyunsung, et al.
Publicado: (2024)