Pseudo-Cepstrum: Pitch Modification for Mel-Based Neural Vocoders
Fuente:
arXiv
Saved in:
| Main Authors: | Ellinas, Nikolaos, Vioni, Alexandra, Kakoulidis, Panos, Vamvoukakis, Georgios, Christidou, Myrsini, Markopoulos, Konstantinos, Oh, Junkwang, Jho, Gunu, Hwang, Inchul, Chalamandaris, Aimilios, Tsiakoulis, Pirros |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
by: Kakoulidis, Panos, et al.
Published: (2024)
by: Kakoulidis, Panos, et al.
Published: (2024)
Improved Text Emotion Prediction Using Combined Valence and Arousal Ordinal Classification
by: Mitsios, Michael, et al.
Published: (2024)
by: Mitsios, Michael, et al.
Published: (2024)
MambaRate: Speech Quality Assessment Across Different Sampling Rates
by: Kakoulidis, Panos, et al.
Published: (2025)
by: Kakoulidis, Panos, et al.
Published: (2025)
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation
by: Ellinas, Nikolaos, et al.
Published: (2022)
by: Ellinas, Nikolaos, et al.
Published: (2022)
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling
by: Karapiperis, Sotirios, et al.
Published: (2024)
by: Karapiperis, Sotirios, et al.
Published: (2024)
Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models
by: Lee, Kyowoon, et al.
Published: (2025)
by: Lee, Kyowoon, et al.
Published: (2025)
FreeV: Free Lunch For Vocoders Through Pseudo Inversed Mel Filter
by: Lv, Yuanjun, et al.
Published: (2024)
by: Lv, Yuanjun, et al.
Published: (2024)
Performance of using Mel-Frequency Cepstrum Based Features in Nonlinear Classifiers for Phonocardiography Recordings
by: Ozkan, Ibrahim, et al.
Published: (2024)
by: Ozkan, Ibrahim, et al.
Published: (2024)
Is GAN Necessary for Mel-Spectrogram-based Neural Vocoder?
by: Du, Hui-Peng, et al.
Published: (2025)
by: Du, Hui-Peng, et al.
Published: (2025)
Chapter Multi-Period Attack-Aware Optical Network Planning under Demand Uncertainty
by: Manousakis, Konstantinos, et al.
Published: (2021)
by: Manousakis, Konstantinos, et al.
Published: (2021)
Real-Time Streaming Mel Vocoding with Generative Flow Matching
by: Welker, Simon, et al.
Published: (2025)
by: Welker, Simon, et al.
Published: (2025)
ESTVocoder: An Excitation-Spectral-Transformed Neural Vocoder Conditioned on Mel Spectrogram
by: Jiang, Xiao-Hang, et al.
Published: (2024)
by: Jiang, Xiao-Hang, et al.
Published: (2024)
Quantum Annealing for Robust Principal Component Analysis
by: Tomeo, Ian, et al.
Published: (2025)
by: Tomeo, Ian, et al.
Published: (2025)
Revolutionizing Heart Valve Therapy: A Translational Framework for Combining Decellularized Scaffolds With Genetic Modification
by: Nikolaos P. Tzavellas, et al.
Published: (2025)
by: Nikolaos P. Tzavellas, et al.
Published: (2025)
Protecting Cryptographic Libraries against Side-Channel and Code-Reuse Attacks
by: Tsoupidi, Rodothea Myrsini, et al.
Published: (2024)
by: Tsoupidi, Rodothea Myrsini, et al.
Published: (2024)
Convolutional Neural Network Compression via Dynamic Parameter Rank Pruning
by: Sharma, Manish, et al.
Published: (2024)
by: Sharma, Manish, et al.
Published: (2024)
PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model
by: Hono, Yukiya, et al.
Published: (2024)
by: Hono, Yukiya, et al.
Published: (2024)
A Neural Denoising Vocoder for Clean Waveform Generation from Noisy Mel-Spectrogram based on Amplitude and Phase Predictions
by: Du, Hui-Peng, et al.
Published: (2024)
by: Du, Hui-Peng, et al.
Published: (2024)
Periodic and Event-Triggering for Joint Capacity Maximization and Safe Intersection Crossing
by: Vitale, Christian, et al.
Published: (2022)
by: Vitale, Christian, et al.
Published: (2022)
Scheduling Aerial Vehicles in an Urban Air Mobility Scheme
by: Rigas, Emmanouil S., et al.
Published: (2021)
by: Rigas, Emmanouil S., et al.
Published: (2021)
Multi-Step Traffic Prediction for Multi-Period Planning in Optical Networks
by: Maryam, Hafsa, et al.
Published: (2024)
by: Maryam, Hafsa, et al.
Published: (2024)
Distributed Estimation and Control for Jamming an Aerial Target With Multiple Agents
by: Papaioannou, Savvas, et al.
Published: (2023)
by: Papaioannou, Savvas, et al.
Published: (2023)
Cepstrum-Based Texture Features for Melanoma Detection
by: Miller, Keith, et al.
Published: (2025)
by: Miller, Keith, et al.
Published: (2025)
Ultra-Low-Bitrate Mel-Spectrogram-based Neural Speech Coding with Flow-Matching-based Refinement and Vocoding-driven Reconstruction
by: Du, Hui-Peng, et al.
Published: (2026)
by: Du, Hui-Peng, et al.
Published: (2026)
Implementación de la transformada Cepstrum en hardware
by: Juliana Alejandra Arévalo Herrera
Published: (2006)
by: Juliana Alejandra Arévalo Herrera
Published: (2006)
Cepstrum-based interferometric microscopy (CIM) for quantitative phase imaging
by: Rubio-Oliver, Ricardo, et al.
Published: (2025)
by: Rubio-Oliver, Ricardo, et al.
Published: (2025)
Activity delay patterns in project networks
by: Vazquez, Alexei, et al.
Published: (2023)
by: Vazquez, Alexei, et al.
Published: (2023)
Probabilistically Robust Trajectory Planning of Multiple Aerial Agents
by: Vitale, Christian, et al.
Published: (2024)
by: Vitale, Christian, et al.
Published: (2024)
Cooperative Search and Track of Rogue Drones using Multiagent Reinforcement Learning
by: Valianti, Panayiota, et al.
Published: (2025)
by: Valianti, Panayiota, et al.
Published: (2025)
Engagement, preference and identity: Proposing a developmental engagement framework
by: Mihyun Son, et al.
Published: (2026)
by: Mihyun Son, et al.
Published: (2026)
Processing and Compositional Effects on the Stability of All‐Inorganic Metal Halide Perovskite Anodes: A Comparative Study of Dry‐ vs. Slurry‐Fabricated Electrodes
by: Kostantinos Markopoulos, et al.
Published: (2026)
by: Kostantinos Markopoulos, et al.
Published: (2026)
A Fair Federated Learning Framework for Collaborative Network Traffic Prediction and Resource Allocation
by: Panda, Saroj Kumar, et al.
Published: (2025)
by: Panda, Saroj Kumar, et al.
Published: (2025)
Comparative Analysis of Sobol and Shapley Methods for Sensitivity Analysis in Civil Engineering: Case Studies on Fibre‐Reinforced Concrete Performance
by: Nikolaos Mellios, et al.
Published: (2025)
by: Nikolaos Mellios, et al.
Published: (2025)
Entanglement on a Sphere
by: Boutivas, Konstantinos, et al.
Published: (2025)
by: Boutivas, Konstantinos, et al.
Published: (2025)
Documentation of the Urban PERIsCOPE platform and tools
by: Sabatakos, Panos, et al.
Published: (2025)
by: Sabatakos, Panos, et al.
Published: (2025)
Scalable and Personalized Oral Assessments Using Voice AI
by: Ipeirotis, Panos, et al.
Published: (2026)
by: Ipeirotis, Panos, et al.
Published: (2026)
Does AI need anything more than a single image to diagnose melanoma?
by: Aimilios Lallas, et al.
Published: (2025)
by: Aimilios Lallas, et al.
Published: (2025)
TF2AIF: Facilitating development and deployment of accelerated AI models on the cloud-edge continuum
by: Leftheriotis, Aimilios, et al.
Published: (2024)
by: Leftheriotis, Aimilios, et al.
Published: (2024)
BiVocoder: A Bidirectional Neural Vocoder Integrating Feature Extraction and Waveform Generation
by: Du, Hui-Peng, et al.
Published: (2024)
by: Du, Hui-Peng, et al.
Published: (2024)
Leveraging Multi-Step Traffic Forecasts for Multi-Period Planning Optical Networks
by: Savva, Giannis, et al.
Published: (2026)
by: Savva, Giannis, et al.
Published: (2026)
Similar Items
-
Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
by: Kakoulidis, Panos, et al.
Published: (2024) -
Improved Text Emotion Prediction Using Combined Valence and Arousal Ordinal Classification
by: Mitsios, Michael, et al.
Published: (2024) -
MambaRate: Speech Quality Assessment Across Different Sampling Rates
by: Kakoulidis, Panos, et al.
Published: (2025) -
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation
by: Ellinas, Nikolaos, et al.
Published: (2022) -
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling
by: Karapiperis, Sotirios, et al.
Published: (2024)