A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Shuoyang, Sedó, Anna Xambó, Bryan-Kinns, Nick |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
di: Xu, Ziqing, et al.
Pubblicazione: (2025)
di: Xu, Ziqing, et al.
Pubblicazione: (2025)
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
di: Bryan-Kinns, Nick, et al.
Pubblicazione: (2024)
di: Bryan-Kinns, Nick, et al.
Pubblicazione: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
di: Blum'e, Ashlae
Pubblicazione: (2025)
di: Blum'e, Ashlae
Pubblicazione: (2025)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
di: Wang, Hongbin, et al.
Pubblicazione: (2025)
di: Wang, Hongbin, et al.
Pubblicazione: (2025)
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
di: Choi, Youngwon, et al.
Pubblicazione: (2025)
di: Choi, Youngwon, et al.
Pubblicazione: (2025)
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
di: Chhaglani, Bhawana, et al.
Pubblicazione: (2025)
di: Chhaglani, Bhawana, et al.
Pubblicazione: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
di: Li, Yinan, et al.
Pubblicazione: (2026)
di: Li, Yinan, et al.
Pubblicazione: (2026)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
di: Liu, Ailin, et al.
Pubblicazione: (2024)
di: Liu, Ailin, et al.
Pubblicazione: (2024)
Robotic Blended Sonification: Consequential Robot Sound as Creative Material for Human-Robot Interaction
di: Johansen, Stine S., et al.
Pubblicazione: (2024)
di: Johansen, Stine S., et al.
Pubblicazione: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
Collaboration Between Robots, Interfaces and Humans: Practice-Based and Audience Perspectives
di: Savery, Anna, et al.
Pubblicazione: (2024)
di: Savery, Anna, et al.
Pubblicazione: (2024)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
di: Li, Yue, et al.
Pubblicazione: (2024)
di: Li, Yue, et al.
Pubblicazione: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
BioSonix: Can Physics-Based Sonification Perceptualize Tissue Deformations From Tool Interactions?
di: Ruozzi, Veronica, et al.
Pubblicazione: (2025)
di: Ruozzi, Veronica, et al.
Pubblicazione: (2025)
Accelerating Audio Research with Robotic Dummy Heads
di: Lu, Austin, et al.
Pubblicazione: (2025)
di: Lu, Austin, et al.
Pubblicazione: (2025)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
di: Ma, Yong, et al.
Pubblicazione: (2025)
di: Ma, Yong, et al.
Pubblicazione: (2025)
Using Confidence Scores to Improve Eyes-free Detection of Speech Recognition Errors
di: Nowrin, Sadia, et al.
Pubblicazione: (2024)
di: Nowrin, Sadia, et al.
Pubblicazione: (2024)
MR-DAW: Towards Collaborative Digital Audio Workstations in Mixed Reality
di: Hopkins, Torin, et al.
Pubblicazione: (2026)
di: Hopkins, Torin, et al.
Pubblicazione: (2026)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
di: Zheng, Naijun, et al.
Pubblicazione: (2025)
di: Zheng, Naijun, et al.
Pubblicazione: (2025)
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
di: Küttner, Michael, et al.
Pubblicazione: (2025)
di: Küttner, Michael, et al.
Pubblicazione: (2025)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
di: Huh, Mina, et al.
Pubblicazione: (2026)
di: Huh, Mina, et al.
Pubblicazione: (2026)
Human Perception of Audio Deepfakes
di: Müller, Nicolas M., et al.
Pubblicazione: (2021)
di: Müller, Nicolas M., et al.
Pubblicazione: (2021)
A Framework for AI assisted Musical Devices
di: Civit, Miguel, et al.
Pubblicazione: (2024)
di: Civit, Miguel, et al.
Pubblicazione: (2024)
VoiceX: A Text-To-Speech Framework for Custom Voices
di: Mertes, Silvan, et al.
Pubblicazione: (2024)
di: Mertes, Silvan, et al.
Pubblicazione: (2024)
ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection
di: Fan, Cunhang, et al.
Pubblicazione: (2025)
di: Fan, Cunhang, et al.
Pubblicazione: (2025)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
di: Han, Hyewon, et al.
Pubblicazione: (2024)
di: Han, Hyewon, et al.
Pubblicazione: (2024)
Assessing the Viability of Wave Field Synthesis in VR-Based Cognitive Research
di: Kahl, Benjamin
Pubblicazione: (2025)
di: Kahl, Benjamin
Pubblicazione: (2025)
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
di: Mishra, Ruchik, et al.
Pubblicazione: (2024)
di: Mishra, Ruchik, et al.
Pubblicazione: (2024)
Prototype: A Keyword Spotting-Based Intelligent Audio SoC for IoT
di: Liang, Huihong, et al.
Pubblicazione: (2025)
di: Liang, Huihong, et al.
Pubblicazione: (2025)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
di: Park, Seohyun, et al.
Pubblicazione: (2025)
di: Park, Seohyun, et al.
Pubblicazione: (2025)
Interfacing with history: Curating with audio augmented objects
di: Cliffe, Laurence
Pubblicazione: (2024)
di: Cliffe, Laurence
Pubblicazione: (2024)
Transhuman Ansambl - Voice Beyond Language
di: Ivsic, Lucija, et al.
Pubblicazione: (2024)
di: Ivsic, Lucija, et al.
Pubblicazione: (2024)
Cervical Auscultation Machine Learning for Dysphagia Assessment
di: Chia, An An, et al.
Pubblicazione: (2024)
di: Chia, An An, et al.
Pubblicazione: (2024)
ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds
di: Kobayashi, Atsuya, et al.
Pubblicazione: (2020)
di: Kobayashi, Atsuya, et al.
Pubblicazione: (2020)
Documenti analoghi
-
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
di: Xu, Ziqing, et al.
Pubblicazione: (2025) -
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
di: Bryan-Kinns, Nick, et al.
Pubblicazione: (2024) -
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
di: Blum'e, Ashlae
Pubblicazione: (2025) -
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
di: Wang, Hongbin, et al.
Pubblicazione: (2025) -
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
di: Choi, Youngwon, et al.
Pubblicazione: (2025)