A Speech Enhancement Method Using Fast Fourier Transform and Convolutional Autoencoder
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kow, Pu-Yun, Kow, Pu-Zhao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Neural Network Enhanced Born Approximation for Inverse Scattering
von: Desai, Ansh, et al.
Veröffentlicht: (2025)
von: Desai, Ansh, et al.
Veröffentlicht: (2025)
A polynomial-time algorithm for deciding the Hilbert Nullstellensatz over $\mathbb{Z}_2$. A proof of $\mathbf{P}=\mathbf{NP}$ hypothesis
von: Petrov, Petar P.
Veröffentlicht: (2022)
von: Petrov, Petar P.
Veröffentlicht: (2022)
Fast and High-Quality Auto-Regressive Speech Synthesis via Speculative Decoding
von: Li, Bohan, et al.
Veröffentlicht: (2024)
von: Li, Bohan, et al.
Veröffentlicht: (2024)
Masked Modeling Duo: Towards a Universal Audio Pre-training Framework
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
Exploring Pre-trained General-purpose Audio Representations for Heart Murmur Detection
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
Multi-phase $k$-quadrature domains and applications to acoustic waves and magnetic fields
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2024)
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2024)
Perturbation Analysis and Neural Network-Based Initial Condition Estimation for the Sine-Gordon Equation
von: Ha, Junhong, et al.
Veröffentlicht: (2025)
von: Ha, Junhong, et al.
Veröffentlicht: (2025)
On the convergence of PINNs for inverse source problem in the complex Ginzburg-Landau equation
von: Cheng, Xing, et al.
Veröffentlicht: (2025)
von: Cheng, Xing, et al.
Veröffentlicht: (2025)
Gelina: Unified Speech and Gesture Synthesis via Interleaved Token Prediction
von: Guichoux, Téo, et al.
Veröffentlicht: (2025)
von: Guichoux, Téo, et al.
Veröffentlicht: (2025)
Can Sound Replace Vision in LLaVA With Token Substitution?
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
M2D-CLAP: Masked Modeling Duo Meets CLAP for Learning General-purpose Audio-Language Representation
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)
Increasing resolution and instability for linear inverse scattering problems
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2024)
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2024)
A Note on Small Percolating Sets on Hypercubes via Generative AI
von: Bérczi, Gergely, et al.
Veröffentlicht: (2024)
von: Bérczi, Gergely, et al.
Veröffentlicht: (2024)
A minimization problem with free boundary and its application to inverse scattering problems
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2023)
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2023)
Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory
von: Zhang, Xifeng, et al.
Veröffentlicht: (2026)
von: Zhang, Xifeng, et al.
Veröffentlicht: (2026)
Enhanced uncertainty quantification variational autoencoders for the solution of Bayesian inverse problems
von: Tonini, Andrea, et al.
Veröffentlicht: (2025)
von: Tonini, Andrea, et al.
Veröffentlicht: (2025)
Assessing the Utility of Audio Foundation Models for Heart and Respiratory Sound Analysis
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
Towards Pre-training an Effective Respiratory Audio Foundation Model
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
An Optimized Path Planning of Manipulator Using Spline Curves and Real Quantifier Elimination Based on Comprehensive Gröbner Systems
von: Shirato, Yusuke, et al.
Veröffentlicht: (2024)
von: Shirato, Yusuke, et al.
Veröffentlicht: (2024)
MATER: Multi-level Acoustic and Textual Emotion Representation for Interpretable Speech Emotion Recognition
von: Jon, Hyo Jin, et al.
Veröffentlicht: (2025)
von: Jon, Hyo Jin, et al.
Veröffentlicht: (2025)
Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis
von: Salehi, Pegah, et al.
Veröffentlicht: (2024)
von: Salehi, Pegah, et al.
Veröffentlicht: (2024)
SCHENO: Measuring Schema vs. Noise in Graphs
von: Hibshman, Justus Isaiah, et al.
Veröffentlicht: (2024)
von: Hibshman, Justus Isaiah, et al.
Veröffentlicht: (2024)
Sigma Flows for Image and Data Labeling and Learning Structured Prediction
von: Cassel, Jonas, et al.
Veröffentlicht: (2024)
von: Cassel, Jonas, et al.
Veröffentlicht: (2024)
Inverse problems for the spectral fractional Laplacian with inhomogeneous Dirichlet boundary data
von: Jaiswal, Ravi Shankar, et al.
Veröffentlicht: (2026)
von: Jaiswal, Ravi Shankar, et al.
Veröffentlicht: (2026)
Reconstruction of Boundary Data in the Helmholtz Equation Using Particle Swarm Optimization
von: Daoudi, Jamal, et al.
Veröffentlicht: (2025)
von: Daoudi, Jamal, et al.
Veröffentlicht: (2025)
Quality Over Quantity? LLM-Based Curation for a Data-Efficient Audio-Video Foundation Model
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
Bayesian inference for the fractional Calderón problem with a single measurement
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2025)
von: Kow, Pu-Zhao, et al.
Veröffentlicht: (2025)
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
Monaural Multi-Speaker Speech Separation Using Efficient Transformer Model
von: Rijal, S., et al.
Veröffentlicht: (2023)
von: Rijal, S., et al.
Veröffentlicht: (2023)
Inverse Electromagnetic Scattering for Doubly-Connected Cylinders using Convolutional Neural Networks
von: Mindrinos, Leonidas, et al.
Veröffentlicht: (2025)
von: Mindrinos, Leonidas, et al.
Veröffentlicht: (2025)
An Effective Trajectory Planning and an Optimized Path Planning for a 6-Degree-of-Freedom Robot Manipulator
von: Okazaki, Takumu, et al.
Veröffentlicht: (2025)
von: Okazaki, Takumu, et al.
Veröffentlicht: (2025)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
BowelRCNN: Region-based Convolutional Neural Network System for Bowel Sound Auscultation
von: Matynia, Igor, et al.
Veröffentlicht: (2025)
von: Matynia, Igor, et al.
Veröffentlicht: (2025)
Faked Speech Detection with Zero Prior Knowledge
von: Ajmi, Sahar Al, et al.
Veröffentlicht: (2022)
von: Ajmi, Sahar Al, et al.
Veröffentlicht: (2022)
UniTAF: A Modular Framework for Joint Text-to-Speech and Audio-to-Face Modeling
von: Zhou, Qiangong, et al.
Veröffentlicht: (2026)
von: Zhou, Qiangong, et al.
Veröffentlicht: (2026)
Deep Adaptive Dimension Reduction for Bayesian Inference in Inverse Problems
von: Wang, Yueyang, et al.
Veröffentlicht: (2026)
von: Wang, Yueyang, et al.
Veröffentlicht: (2026)
The learned range test method for the inverse inclusion problem
von: Sun, Shiwei, et al.
Veröffentlicht: (2024)
von: Sun, Shiwei, et al.
Veröffentlicht: (2024)
Shortest Paths in a Weighted Simplicial Complex
von: Chakraborty, Sukrit, et al.
Veröffentlicht: (2025)
von: Chakraborty, Sukrit, et al.
Veröffentlicht: (2025)
An Efficient Two-Sided Sketching Method for Large-Scale Tensor Decomposition Based on Transformed Domains
von: Cheng, Zhiguang, et al.
Veröffentlicht: (2024)
von: Cheng, Zhiguang, et al.
Veröffentlicht: (2024)
Customizing Graph Neural Networks using Path Reweighting
von: Chen, Jianpeng, et al.
Veröffentlicht: (2021)
von: Chen, Jianpeng, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
A Neural Network Enhanced Born Approximation for Inverse Scattering
von: Desai, Ansh, et al.
Veröffentlicht: (2025) -
A polynomial-time algorithm for deciding the Hilbert Nullstellensatz over $\mathbb{Z}_2$. A proof of $\mathbf{P}=\mathbf{NP}$ hypothesis
von: Petrov, Petar P.
Veröffentlicht: (2022) -
Fast and High-Quality Auto-Regressive Speech Synthesis via Speculative Decoding
von: Li, Bohan, et al.
Veröffentlicht: (2024) -
Masked Modeling Duo: Towards a Universal Audio Pre-training Framework
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024) -
Exploring Pre-trained General-purpose Audio Representations for Heart Murmur Detection
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2024)