Regularized Schrödinger Bridge: Alleviating Distortion and Exposure Bias in Solving Inverse Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yao, Qing, Gao, Lijian, Mao, Qirong, Dong, Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Audio Decoding by Inverse Problem Solving
von: T., Pedro J. Villasana, et al.
Veröffentlicht: (2024)
von: T., Pedro J. Villasana, et al.
Veröffentlicht: (2024)
A2SB: Audio-to-Audio Schrodinger Bridges
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis
von: Chen, Zehua, et al.
Veröffentlicht: (2023)
von: Chen, Zehua, et al.
Veröffentlicht: (2023)
Contrastive Regularization for Accent-Robust ASR
von: Thai, Van-Phat, et al.
Veröffentlicht: (2026)
von: Thai, Van-Phat, et al.
Veröffentlicht: (2026)
Schrödinger Bridge Mamba for One-Step Speech Enhancement
von: Yang, Jing, et al.
Veröffentlicht: (2025)
von: Yang, Jing, et al.
Veröffentlicht: (2025)
Quantifying the Corpus Bias Problem in Automatic Music Transcription Systems
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2024)
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2024)
Koopman Regularized Deep Speech Disentanglement for Speaker Verification
von: Chazaridis, Nikos, et al.
Veröffentlicht: (2026)
von: Chazaridis, Nikos, et al.
Veröffentlicht: (2026)
Bias beyond Borders: Global Inequalities in AI-Generated Music
von: Solak, Ahmet, et al.
Veröffentlicht: (2025)
von: Solak, Ahmet, et al.
Veröffentlicht: (2025)
Audio Super-Resolution with Latent Bridge Models
von: Li, Chang, et al.
Veröffentlicht: (2025)
von: Li, Chang, et al.
Veröffentlicht: (2025)
Improving Speech Emotion Recognition with Mutual Information Regularized Generative Model
von: Ahn, Chung-Soo, et al.
Veröffentlicht: (2025)
von: Ahn, Chung-Soo, et al.
Veröffentlicht: (2025)
Towards Trustworthy Audio Deepfake Detection: A Systematic Framework for Diagnosing and Mitigating Gender Bias
von: Fursule, Aishwarya, et al.
Veröffentlicht: (2026)
von: Fursule, Aishwarya, et al.
Veröffentlicht: (2026)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
von: Xu, Ji, et al.
Veröffentlicht: (2023)
von: Xu, Ji, et al.
Veröffentlicht: (2023)
CyIN: Cyclic Informative Latent Space for Bridging Complete and Incomplete Multimodal Learning
von: Lin, Ronghao, et al.
Veröffentlicht: (2026)
von: Lin, Ronghao, et al.
Veröffentlicht: (2026)
Advancing Audio Fingerprinting Accuracy Addressing Background Noise and Distortion Challenges
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
CAARMA: Class Augmentation with Adversarial Mixup Regularization
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
Acoustic Structure Inverse Design and Optimization Using Deep Learning
von: Sun, Xuecong, et al.
Veröffentlicht: (2021)
von: Sun, Xuecong, et al.
Veröffentlicht: (2021)
STEP: Detecting Audio Backdoor Attacks via Stability-based Trigger Exposure Profiling
von: Wang, Kun, et al.
Veröffentlicht: (2026)
von: Wang, Kun, et al.
Veröffentlicht: (2026)
Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis
von: Feng, Pengchao, et al.
Veröffentlicht: (2025)
von: Feng, Pengchao, et al.
Veröffentlicht: (2025)
EDSep: An Effective Diffusion-Based Method for Speech Source Separation
von: Dong, Jinwei, et al.
Veröffentlicht: (2025)
von: Dong, Jinwei, et al.
Veröffentlicht: (2025)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
von: Ding, Hanyu, et al.
Veröffentlicht: (2025)
Positive-Unlabelled Active Learning to Curate a Dataset for Orca Resident Interpretation
von: Nestor, Bret, et al.
Veröffentlicht: (2026)
von: Nestor, Bret, et al.
Veröffentlicht: (2026)
Privacy-Enhancing Infant Cry Classification with Federated Transformers and Denoising Regularization
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
Audio-Visual Continual Test-Time Adaptation without Forgetting
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2026)
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2026)
Finite Scalar Quantization Enables Redundant and Transmission-Robust Neural Audio Compression at Low Bit-rates
von: Julian, Harry, et al.
Veröffentlicht: (2025)
von: Julian, Harry, et al.
Veröffentlicht: (2025)
An Experimental Study on Joint Modeling for Sound Event Localization and Detection with Source Distance Estimation
von: Dong, Yuxuan, et al.
Veröffentlicht: (2025)
von: Dong, Yuxuan, et al.
Veröffentlicht: (2025)
SSNAPS: Audio-Visual Separation of Speech and Background Noise with Diffusion Inverse Sampling
von: Yemini, Yochai, et al.
Veröffentlicht: (2026)
von: Yemini, Yochai, et al.
Veröffentlicht: (2026)
BiSinger: Bilingual Singing Voice Synthesis
von: Zhou, Huali, et al.
Veröffentlicht: (2023)
von: Zhou, Huali, et al.
Veröffentlicht: (2023)
Mathematical Foundations of Polyphonic Music Generation via Structural Inductive Bias
von: Seo, Joonwon
Veröffentlicht: (2026)
von: Seo, Joonwon
Veröffentlicht: (2026)
Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning
von: Wu, Daiqing, et al.
Veröffentlicht: (2026)
von: Wu, Daiqing, et al.
Veröffentlicht: (2026)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
Gaussian Flow Bridges for Audio Domain Transfer with Unpaired Data
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
Mitigating Sex Bias in Audio Data-driven COPD and COVID-19 Breathing Pattern Detection Models
von: Pfeifer, Rachel, et al.
Veröffentlicht: (2024)
von: Pfeifer, Rachel, et al.
Veröffentlicht: (2024)
Bridging the Perception Gap: A Lightweight Coarse-to-Fine Architecture for Edge Audio Systems
von: Zhang, Hengfan, et al.
Veröffentlicht: (2026)
von: Zhang, Hengfan, et al.
Veröffentlicht: (2026)
Sonos Voice Control Bias Assessment Dataset: A Methodology for Demographic Bias Assessment in Voice Assistants
von: Sekkat, Chloé, et al.
Veröffentlicht: (2024)
von: Sekkat, Chloé, et al.
Veröffentlicht: (2024)
Learning to Solve Inverse Problems for Perceptual Sound Matching
von: Han, Han, et al.
Veröffentlicht: (2023)
von: Han, Han, et al.
Veröffentlicht: (2023)
An AI-enabled Bias-Free Respiratory Disease Diagnosis Model using Cough Audio: A Case Study for COVID-19
von: Saeed, Tabish, et al.
Veröffentlicht: (2024)
von: Saeed, Tabish, et al.
Veröffentlicht: (2024)
Incorporating Pre-trained Diffusion Models in Solving the Schrödinger Bridge Problem
von: Tang, Zhicong, et al.
Veröffentlicht: (2025)
von: Tang, Zhicong, et al.
Veröffentlicht: (2025)
Convolutional Neural Network Achieves Human-level Accuracy in Music Genre Classification
von: Dong, Mingwen
Veröffentlicht: (2018)
von: Dong, Mingwen
Veröffentlicht: (2018)
Bridging The Multi-Modality Gaps of Audio, Visual and Linguistic for Speech Enhancement
von: Lin, Meng-Ping, et al.
Veröffentlicht: (2025)
von: Lin, Meng-Ping, et al.
Veröffentlicht: (2025)
Regularizing Learnable Feature Extraction for Automatic Speech Recognition
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Audio Decoding by Inverse Problem Solving
von: T., Pedro J. Villasana, et al.
Veröffentlicht: (2024) -
A2SB: Audio-to-Audio Schrodinger Bridges
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025) -
Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis
von: Chen, Zehua, et al.
Veröffentlicht: (2023) -
Contrastive Regularization for Accent-Robust ASR
von: Thai, Van-Phat, et al.
Veröffentlicht: (2026) -
Schrödinger Bridge Mamba for One-Step Speech Enhancement
von: Yang, Jing, et al.
Veröffentlicht: (2025)