BemaGANv2: Discriminator Combination Strategies for GAN-based Vocoders in Long-Term Audio Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Taesoo, Jeong, Mungwi, Park, Mingyu, Kim, Narae, Kim, Junyoung, Kim, Mujung, Yoo, Jisang, Lee, Hoyun, Kim, Sanghoon, Kwon, Soonchul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
von: Kim, Minu, et al.
Veröffentlicht: (2025)
von: Kim, Minu, et al.
Veröffentlicht: (2025)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
von: Choi, Seonghwa, et al.
Veröffentlicht: (2025)
von: Choi, Seonghwa, et al.
Veröffentlicht: (2025)
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
von: Kim, Minu, et al.
Veröffentlicht: (2025)
von: Kim, Minu, et al.
Veröffentlicht: (2025)
WACA-UNet: Weakness-Aware Channel Attention for Static IR Drop Prediction in Integrated Circuit Design
von: Seo, Youngmin, et al.
Veröffentlicht: (2025)
von: Seo, Youngmin, et al.
Veröffentlicht: (2025)
Alternative Local Discriminant Bases Using Empirical Expectation and Variance Estimation
von: Fossgaard, Eirik
Veröffentlicht: (1999)
von: Fossgaard, Eirik
Veröffentlicht: (1999)
A Data Aggregation Visualization System supported by Processing-in-Memory
von: Kim, Junyoung, et al.
Veröffentlicht: (2025)
von: Kim, Junyoung, et al.
Veröffentlicht: (2025)
Detecting AI-Generated Videos with Spiking Neural Networks
von: Jang, Minsuk, et al.
Veröffentlicht: (2026)
von: Jang, Minsuk, et al.
Veröffentlicht: (2026)
A Spatio-Temporal Deep Learning Approach For High-Resolution Gridded Monsoon Prediction
von: Borah, Parashjyoti, et al.
Veröffentlicht: (2026)
von: Borah, Parashjyoti, et al.
Veröffentlicht: (2026)
Discriminative Subspace Emersion from learning feature relevances across different populations
von: Canducci, Marco, et al.
Veröffentlicht: (2025)
von: Canducci, Marco, et al.
Veröffentlicht: (2025)
Uncovering Population PK Covariates from VAE-Generated Latent Spaces
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
von: Perazzolo, Diego, et al.
Veröffentlicht: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
Persistent Multiscale Density-based Clustering
von: Bot, Daniël, et al.
Veröffentlicht: (2025)
von: Bot, Daniël, et al.
Veröffentlicht: (2025)
IMUVIE: Pickup Timeline Action Localization via Motion Movies
von: Clapham, John, et al.
Veröffentlicht: (2024)
von: Clapham, John, et al.
Veröffentlicht: (2024)
Learning from Semantic Dictionaries: Discriminative Codebook Contrastive Learning for Unified Visual Representation and Generation
von: Estepa, Imanol G., et al.
Veröffentlicht: (2026)
von: Estepa, Imanol G., et al.
Veröffentlicht: (2026)
MoXaRt: Audio-Visual Object-Guided Sound Interaction for XR
von: Xu, Tianyu, et al.
Veröffentlicht: (2026)
von: Xu, Tianyu, et al.
Veröffentlicht: (2026)
Tensor Network-Constrained Kernel Machines as Gaussian Processes
von: Wesel, Frederiek, et al.
Veröffentlicht: (2024)
von: Wesel, Frederiek, et al.
Veröffentlicht: (2024)
Quantized Fourier and Polynomial Features for more Expressive Tensor Network Models
von: Wesel, Frederiek, et al.
Veröffentlicht: (2023)
von: Wesel, Frederiek, et al.
Veröffentlicht: (2023)
A Novel Schur-Decomposition-Based Weight Projection Method for Stable State-Space Neural-Network Architectures
von: Vanegas, Sergio, et al.
Veröffentlicht: (2026)
von: Vanegas, Sergio, et al.
Veröffentlicht: (2026)
Beyond Reality: Designing Personal Experiences and Interactive Narratives in AR Theater
von: Kim, You-Jin
Veröffentlicht: (2025)
von: Kim, You-Jin
Veröffentlicht: (2025)
Pairwise Spatiotemporal Partial Trajectory Matching for Co-movement Analysis
von: Cardei, Maria, et al.
Veröffentlicht: (2024)
von: Cardei, Maria, et al.
Veröffentlicht: (2024)
Regularisation in neural networks: a survey and empirical analysis of approaches
von: Opperman, Christiaan P., et al.
Veröffentlicht: (2026)
von: Opperman, Christiaan P., et al.
Veröffentlicht: (2026)
MMM: Quantum-Chemical Molecular Representation Learning for Combinatorial Drug Recommendation
von: Kwon, Chongmyung, et al.
Veröffentlicht: (2025)
von: Kwon, Chongmyung, et al.
Veröffentlicht: (2025)
Context-dependent Causality (the Non-Nonotonic Case)
von: Billfeld, Nir, et al.
Veröffentlicht: (2024)
von: Billfeld, Nir, et al.
Veröffentlicht: (2024)
Deep Spectral Meshes: Multi-Frequency Facial Mesh Processing with Graph Neural Networks
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
Kolmogorov-Arnold Attention: Is Learnable Attention Better For Vision Transformers?
von: Maity, Subhajit, et al.
Veröffentlicht: (2025)
von: Maity, Subhajit, et al.
Veröffentlicht: (2025)
Adaptive Riemannian Graph Neural Networks
von: Wang, Xudong, et al.
Veröffentlicht: (2025)
von: Wang, Xudong, et al.
Veröffentlicht: (2025)
Learnable Kernel Density Estimation for Graphs and Its Application to Graph-Level Anomaly Detection
von: Wang, Xudong, et al.
Veröffentlicht: (2025)
von: Wang, Xudong, et al.
Veröffentlicht: (2025)
Accelerating Language Model Workflows with Prompt Choreography
von: Bai, TJ, et al.
Veröffentlicht: (2025)
von: Bai, TJ, et al.
Veröffentlicht: (2025)
Masked Contrastive Pre-Training Improves Music Audio Key Detection
von: Yonay, Ori, et al.
Veröffentlicht: (2026)
von: Yonay, Ori, et al.
Veröffentlicht: (2026)
NFCL: Simply interpretable neural networks for a short-term multivariate forecasting
von: Jo, Wonkeun, et al.
Veröffentlicht: (2024)
von: Jo, Wonkeun, et al.
Veröffentlicht: (2024)
Apictorial Jigsaw Puzzle Reconstruction Based on Curve Matching via a Corotational Beam Spline
von: Orynyak, Igor, et al.
Veröffentlicht: (2025)
von: Orynyak, Igor, et al.
Veröffentlicht: (2025)
Swish-T : Enhancing Swish Activation with Tanh Bias for Improved Neural Network Performance
von: Seo, Youngmin, et al.
Veröffentlicht: (2024)
von: Seo, Youngmin, et al.
Veröffentlicht: (2024)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
von: Aristorenas, Aris J.
Veröffentlicht: (2024)
von: Aristorenas, Aris J.
Veröffentlicht: (2024)
Revisiting GAN with Bayes-Optimal Discrimination
von: Naeini, Mohammadreza Tavasoli, et al.
Veröffentlicht: (2025)
von: Naeini, Mohammadreza Tavasoli, et al.
Veröffentlicht: (2025)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
CASE: Contrastive Activation for Saliency Estimation
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
In-situ Autoguidance: Eliciting Self-Correction in Diffusion Models
von: Gu, Enhao, et al.
Veröffentlicht: (2025)
von: Gu, Enhao, et al.
Veröffentlicht: (2025)
Impulsive pattern recognition of a myoelectric hand via Dynamic Time Warping
von: Kadilar, Mustafa Can, et al.
Veröffentlicht: (2025)
von: Kadilar, Mustafa Can, et al.
Veröffentlicht: (2025)
TACIT Benchmark: A Programmatic Visual Reasoning Benchmark for Generative and Discriminative Models
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
von: Medeiros, Daniel Nobrega
Veröffentlicht: (2026)
Ähnliche Einträge
-
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
von: Kim, Minu, et al.
Veröffentlicht: (2025) -
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
von: Choi, Seonghwa, et al.
Veröffentlicht: (2025) -
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
von: Kim, Minu, et al.
Veröffentlicht: (2025) -
WACA-UNet: Weakness-Aware Channel Attention for Static IR Drop Prediction in Integrated Circuit Design
von: Seo, Youngmin, et al.
Veröffentlicht: (2025) -
Alternative Local Discriminant Bases Using Empirical Expectation and Variance Estimation
von: Fossgaard, Eirik
Veröffentlicht: (1999)