HyperGANStrument: Instrument Sound Synthesis and Editing with Pitch-Invariant Hypernetworks
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Zhe, Akama, Taketo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Annotation-Free MIDI-to-Audio Synthesis via Concatenative Synthesis and Generative Refinement
di: Take, Osamu, et al.
Pubblicazione: (2024)
di: Take, Osamu, et al.
Pubblicazione: (2024)
Pitch-Conditioned Instrument Sound Synthesis From an Interactive Timbre Latent Space
di: Limberg, Christian, et al.
Pubblicazione: (2025)
di: Limberg, Christian, et al.
Pubblicazione: (2025)
A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2024)
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2024)
Naturalistic Music Decoding from EEG Data via Latent Diffusion Models
di: Postolache, Emilian, et al.
Pubblicazione: (2024)
di: Postolache, Emilian, et al.
Pubblicazione: (2024)
A Computational Analysis of Lyric Similarity Perception
di: Kim, Haven, et al.
Pubblicazione: (2024)
di: Kim, Haven, et al.
Pubblicazione: (2024)
Annotation-free Automatic Music Transcription with Scalable Synthetic Data and Adversarial Domain Confusion
di: Sato, Gakusei, et al.
Pubblicazione: (2023)
di: Sato, Gakusei, et al.
Pubblicazione: (2023)
HyperTTS: Parameter Efficient Adaptation in Text to Speech using Hypernetworks
di: Li, Yingting, et al.
Pubblicazione: (2024)
di: Li, Yingting, et al.
Pubblicazione: (2024)
PF-D2M: A Pose-free Diffusion Model for Universal Dance-to-Music Generation
di: Im, Jaekwon, et al.
Pubblicazione: (2026)
di: Im, Jaekwon, et al.
Pubblicazione: (2026)
HyperSound: Generating Implicit Neural Representations of Audio Signals with Hypernetworks
di: Szatkowski, Filip, et al.
Pubblicazione: (2022)
di: Szatkowski, Filip, et al.
Pubblicazione: (2022)
Decoding Selective Auditory Attention to Musical Elements in Ecologically Valid Music Listening
di: Akama, Taketo, et al.
Pubblicazione: (2025)
di: Akama, Taketo, et al.
Pubblicazione: (2025)
Music Proofreading with RefinPaint: Where and How to Modify Compositions given Context
di: Ramoneda, Pedro, et al.
Pubblicazione: (2024)
di: Ramoneda, Pedro, et al.
Pubblicazione: (2024)
Toward Fully Self-Supervised Multi-Pitch Estimation
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2024)
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2024)
Predicting Artificial Neural Network Representations to Learn Recognition Model for Music Identification from Brain Recordings
di: Akama, Taketo, et al.
Pubblicazione: (2024)
di: Akama, Taketo, et al.
Pubblicazione: (2024)
Pseudo-Cepstrum: Pitch Modification for Mel-Based Neural Vocoders
di: Ellinas, Nikolaos, et al.
Pubblicazione: (2025)
di: Ellinas, Nikolaos, et al.
Pubblicazione: (2025)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
Investigating an Overfitting and Degeneration Phenomenon in Self-Supervised Multi-Pitch Estimation
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2025)
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
di: Diaz-Guerra, David, et al.
Pubblicazione: (2023)
SoundMorpher: Perceptually-Uniform Sound Morphing with Diffusion Model
di: Niu, Xinlei, et al.
Pubblicazione: (2024)
di: Niu, Xinlei, et al.
Pubblicazione: (2024)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2025)
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2025)
A Lightweight Slot-Attention Framework for Multi-Instrument Multi-Pitch Estimation
di: Taenzer, Michael
Pubblicazione: (2026)
di: Taenzer, Michael
Pubblicazione: (2026)
Expressive Acoustic Guitar Sound Synthesis with an Instrument-Specific Input Representation and Diffusion Outpainting
di: Kim, Hounsu, et al.
Pubblicazione: (2024)
di: Kim, Hounsu, et al.
Pubblicazione: (2024)
SoundSculpt: Direction and Semantics Driven Ambisonic Target Sound Extraction
di: Chen, Tuochao, et al.
Pubblicazione: (2025)
di: Chen, Tuochao, et al.
Pubblicazione: (2025)
Advanced Framework for Animal Sound Classification With Features Optimization
di: Yang, Qiang, et al.
Pubblicazione: (2024)
di: Yang, Qiang, et al.
Pubblicazione: (2024)
The iNaturalist Sounds Dataset
di: Chasmai, Mustafa, et al.
Pubblicazione: (2025)
di: Chasmai, Mustafa, et al.
Pubblicazione: (2025)
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation
di: Saito, Koichi, et al.
Pubblicazione: (2024)
di: Saito, Koichi, et al.
Pubblicazione: (2024)
Automatic Equalization for Individual Instrument Tracks Using Convolutional Neural Networks
di: Mockenhaupt, Florian, et al.
Pubblicazione: (2024)
di: Mockenhaupt, Florian, et al.
Pubblicazione: (2024)
Sound event localization and classification using WASN in Outdoor Environment
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
ViolinDiff: Enhancing Expressive Violin Synthesis with Pitch Bend Conditioning
di: Kim, Daewoong, et al.
Pubblicazione: (2024)
di: Kim, Daewoong, et al.
Pubblicazione: (2024)
Lungmix: A Mixup-Based Strategy for Generalization in Respiratory Sound Classification
di: Ge, Shijia, et al.
Pubblicazione: (2024)
di: Ge, Shijia, et al.
Pubblicazione: (2024)
Audio Editing with Non-Rigid Text Prompts
di: Paissan, Francesco, et al.
Pubblicazione: (2023)
di: Paissan, Francesco, et al.
Pubblicazione: (2023)
TSE-PI: Target Sound Extraction under Reverberant Environments with Pitch Information
di: Wang, Yiwen, et al.
Pubblicazione: (2024)
di: Wang, Yiwen, et al.
Pubblicazione: (2024)
Timbre-Trap: A Low-Resource Framework for Instrument-Agnostic Music Transcription
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2023)
di: Cwitkowitz, Frank, et al.
Pubblicazione: (2023)
PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model
di: Hono, Yukiya, et al.
Pubblicazione: (2024)
di: Hono, Yukiya, et al.
Pubblicazione: (2024)
Sound Event Detection and Localization with Distance Estimation
di: Krause, Daniel Aleksander, et al.
Pubblicazione: (2024)
di: Krause, Daniel Aleksander, et al.
Pubblicazione: (2024)
Sound Tagging in Infant-centric Home Soundscapes
di: Khan, Mohammad Nur Hossain, et al.
Pubblicazione: (2024)
di: Khan, Mohammad Nur Hossain, et al.
Pubblicazione: (2024)
Focal Modulation Networks for Interpretable Sound Classification
di: Della Libera, Luca, et al.
Pubblicazione: (2024)
di: Della Libera, Luca, et al.
Pubblicazione: (2024)
Audio Geolocation: A Natural Sounds Benchmark
di: Chasmai, Mustafa, et al.
Pubblicazione: (2025)
di: Chasmai, Mustafa, et al.
Pubblicazione: (2025)
Generating Sample-Based Musical Instruments Using Neural Audio Codec Language Models
di: Nercessian, Shahan, et al.
Pubblicazione: (2024)
di: Nercessian, Shahan, et al.
Pubblicazione: (2024)
HyWA: Hypernetwork Weight Adapting Personalized Voice Activity Detection
di: Nejad, Mahsa Ghazvini, et al.
Pubblicazione: (2025)
di: Nejad, Mahsa Ghazvini, et al.
Pubblicazione: (2025)
Audio Simulation for Sound Source Localization in Virtual Evironment
di: Di Yuan, Yi, et al.
Pubblicazione: (2024)
di: Di Yuan, Yi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Annotation-Free MIDI-to-Audio Synthesis via Concatenative Synthesis and Generative Refinement
di: Take, Osamu, et al.
Pubblicazione: (2024) -
Pitch-Conditioned Instrument Sound Synthesis From an Interactive Timbre Latent Space
di: Limberg, Christian, et al.
Pubblicazione: (2025) -
A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2024) -
Naturalistic Music Decoding from EEG Data via Latent Diffusion Models
di: Postolache, Emilian, et al.
Pubblicazione: (2024) -
A Computational Analysis of Lyric Similarity Perception
di: Kim, Haven, et al.
Pubblicazione: (2024)