Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Wilson, Elizabeth, Fazekas, György, Wiggins, Geraint |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Composers' Evaluations of an AI Music Tool: Insights for Human-Centred Design
di: Row, Eleanor, et al.
Pubblicazione: (2024)
di: Row, Eleanor, et al.
Pubblicazione: (2024)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
di: Yan, Hui, et al.
Pubblicazione: (2024)
di: Yan, Hui, et al.
Pubblicazione: (2024)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
di: Tang, Jingjing, et al.
Pubblicazione: (2025)
di: Tang, Jingjing, et al.
Pubblicazione: (2025)
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
di: Benster, Tyler, et al.
Pubblicazione: (2024)
di: Benster, Tyler, et al.
Pubblicazione: (2024)
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
di: Kotowski, Błażej, et al.
Pubblicazione: (2025)
di: Kotowski, Błażej, et al.
Pubblicazione: (2025)
Open vocabulary keyword spotting through transfer learning from speech synthesis
di: V, Kesavaraj, et al.
Pubblicazione: (2024)
di: V, Kesavaraj, et al.
Pubblicazione: (2024)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
di: Fong, Hortense, et al.
Pubblicazione: (2024)
di: Fong, Hortense, et al.
Pubblicazione: (2024)
A General Close-loop Predictive Coding Framework for Auditory Working Memory
di: Yuan, Zhongju, et al.
Pubblicazione: (2025)
di: Yuan, Zhongju, et al.
Pubblicazione: (2025)
Cervical Auscultation Machine Learning for Dysphagia Assessment
di: Chia, An An, et al.
Pubblicazione: (2024)
di: Chia, An An, et al.
Pubblicazione: (2024)
Composer Style-specific Symbolic Music Generation Using Vector Quantized Discrete Diffusion Models
di: Zhang, Jincheng, et al.
Pubblicazione: (2023)
di: Zhang, Jincheng, et al.
Pubblicazione: (2023)
Mamba-Diffusion Model with Learnable Wavelet for Controllable Symbolic Music Generation
di: Zhang, Jincheng, et al.
Pubblicazione: (2025)
di: Zhang, Jincheng, et al.
Pubblicazione: (2025)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
di: Han, Hyewon, et al.
Pubblicazione: (2024)
di: Han, Hyewon, et al.
Pubblicazione: (2024)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
di: Li, Yinan, et al.
Pubblicazione: (2026)
di: Li, Yinan, et al.
Pubblicazione: (2026)
Music Generation using Human-In-The-Loop Reinforcement Learning
di: Justus, Aju Ani
Pubblicazione: (2025)
di: Justus, Aju Ani
Pubblicazione: (2025)
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
Differentiable Time-Varying Linear Prediction in the Context of End-to-End Analysis-by-Synthesis
di: Yu, Chin-Yun, et al.
Pubblicazione: (2024)
di: Yu, Chin-Yun, et al.
Pubblicazione: (2024)
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
di: Xu, Ziqing, et al.
Pubblicazione: (2025)
di: Xu, Ziqing, et al.
Pubblicazione: (2025)
Human Perception of Audio Deepfakes
di: Müller, Nicolas M., et al.
Pubblicazione: (2021)
di: Müller, Nicolas M., et al.
Pubblicazione: (2021)
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
di: Han, Runduo, et al.
Pubblicazione: (2025)
di: Han, Runduo, et al.
Pubblicazione: (2025)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
di: Brade, Stephen, et al.
Pubblicazione: (2023)
di: Brade, Stephen, et al.
Pubblicazione: (2023)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
di: Kwon, Joonwoo, et al.
Pubblicazione: (2024)
di: Kwon, Joonwoo, et al.
Pubblicazione: (2024)
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
di: Guo, Yiwei, et al.
Pubblicazione: (2023)
di: Guo, Yiwei, et al.
Pubblicazione: (2023)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
di: Wu, Liang-Yuan, et al.
Pubblicazione: (2025)
di: Wu, Liang-Yuan, et al.
Pubblicazione: (2025)
Reimagining Dance: Real-time Music Co-creation between Dancers and AI
di: Vechtomova, Olga, et al.
Pubblicazione: (2025)
di: Vechtomova, Olga, et al.
Pubblicazione: (2025)
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
di: Damacharla, Praveen, et al.
Pubblicazione: (2023)
di: Damacharla, Praveen, et al.
Pubblicazione: (2023)
MCP2OSC: Parametric Control by Natural Language
di: Fan, Yuan-Yi
Pubblicazione: (2025)
di: Fan, Yuan-Yi
Pubblicazione: (2025)
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
di: Chang, Yi, et al.
Pubblicazione: (2024)
di: Chang, Yi, et al.
Pubblicazione: (2024)
Interactive Melody Generation System for Enhancing the Creativity of Musicians
di: Hirawata, So, et al.
Pubblicazione: (2024)
di: Hirawata, So, et al.
Pubblicazione: (2024)
Tipping Points, Pulse Elasticity and Tonal Tension: An Empirical Study on What Generates Tipping Points
di: Naik, Canishk, et al.
Pubblicazione: (2024)
di: Naik, Canishk, et al.
Pubblicazione: (2024)
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
di: Sarmento, Pedro, et al.
Pubblicazione: (2024)
di: Sarmento, Pedro, et al.
Pubblicazione: (2024)
Capturing Cancer as Music: Cancer Mechanisms Expressed through Musification
di: Hnatyshyn, Rostyslav, et al.
Pubblicazione: (2024)
di: Hnatyshyn, Rostyslav, et al.
Pubblicazione: (2024)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
di: Park, Seohyun, et al.
Pubblicazione: (2025)
di: Park, Seohyun, et al.
Pubblicazione: (2025)
Interactive Sonification for Health and Energy using ChucK and Unity
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
di: Zhao, Yichun, et al.
Pubblicazione: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
di: Li, Yue, et al.
Pubblicazione: (2024)
di: Li, Yue, et al.
Pubblicazione: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
di: Yu, Luca Jiang-Tao, et al.
Pubblicazione: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
di: Blum'e, Ashlae
Pubblicazione: (2025)
di: Blum'e, Ashlae
Pubblicazione: (2025)
Documenti analoghi
-
Composers' Evaluations of an AI Music Tool: Insights for Human-Centred Design
di: Row, Eleanor, et al.
Pubblicazione: (2024) -
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
di: Yan, Hui, et al.
Pubblicazione: (2024) -
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
di: Tang, Jingjing, et al.
Pubblicazione: (2025) -
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
di: Benster, Tyler, et al.
Pubblicazione: (2024) -
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
di: Kotowski, Błażej, et al.
Pubblicazione: (2025)