Conditional Flow Matching for Visually-Guided Acoustic Highlighting
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Malard, Hugo, Lan, Gael Le, Wong, Daniel, Alon, David Lou, Wu, Yi-Chiao, Parekh, Sanjeel |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
par: Lee, Dongheon, et autres
Publié: (2026)
par: Lee, Dongheon, et autres
Publié: (2026)
ArrayDPS-Refine: Generative Refinement of Discriminative Multi-Channel Speech Enhancement
par: Xu, Zhongweiyang, et autres
Publié: (2026)
par: Xu, Zhongweiyang, et autres
Publié: (2026)
High Fidelity Text-Guided Music Editing via Single-Stage Flow Matching
par: Lan, Gael Le, et autres
Publié: (2024)
par: Lan, Gael Le, et autres
Publié: (2024)
Binaural Signal Matching with Wearable Arrays for Near-Field Sources
par: Goldring, Sapir, et autres
Publié: (2025)
par: Goldring, Sapir, et autres
Publié: (2025)
Unified Diffusion Refinement for Multi-Channel Speech Enhancement and Separation
par: Xu, Zhongweiyang, et autres
Publié: (2026)
par: Xu, Zhongweiyang, et autres
Publié: (2026)
Learning to Highlight Audio by Watching Movies
par: Huang, Chao, et autres
Publié: (2025)
par: Huang, Chao, et autres
Publié: (2025)
Feasibility of iMagLS-BSM -- ILD Informed Binaural Signal Matching with Arbitrary Microphone Arrays
par: Berebi, Or, et autres
Publié: (2024)
par: Berebi, Or, et autres
Publié: (2024)
BSM-iMagLS: ILD Informed Binaural Signal Matching for Reproduction with Head-Mounted Microphone Arrays
par: Berebi, Or, et autres
Publié: (2025)
par: Berebi, Or, et autres
Publié: (2025)
FlowAVSE: Efficient Audio-Visual Speech Enhancement with Conditional Flow Matching
par: Jung, Chaeyoung, et autres
Publié: (2024)
par: Jung, Chaeyoung, et autres
Publié: (2024)
Binaural Signal Matching with Wearable Arrays for Near-Field Sources and Directional Focus
par: Goldring, Sapir, et autres
Publié: (2025)
par: Goldring, Sapir, et autres
Publié: (2025)
LP-CFM: Perceptual Invariance-Aware Conditional Flow Matching for Speech Modeling
par: Kwak, Doyeop, et autres
Publié: (2025)
par: Kwak, Doyeop, et autres
Publié: (2025)
On HRTF Notch Frequency Prediction Using Anthropometric Features and Neural Networks
par: Arbel, Lior, et autres
Publié: (2024)
par: Arbel, Lior, et autres
Publié: (2024)
CFMDCTCodec: A Low-Bitrate Neural Speech Codec with Noise-Prior-aware Conditional Flow Matching for MDCT-Spectral Enhancement
par: Jiang, Xiao-Hang, et autres
Publié: (2026)
par: Jiang, Xiao-Hang, et autres
Publié: (2026)
Improving Acoustic Scene Classification in Low-Resource Conditions
par: Chen, Zhi, et autres
Publié: (2024)
par: Chen, Zhi, et autres
Publié: (2024)
iMagLS: Interaural Level Difference with Magnitude Least-Squares Loss for Optimized First-Order Head-Related Transfer Function
par: Berebi, Or, et autres
Publié: (2023)
par: Berebi, Or, et autres
Publié: (2023)
Room Impulse Response Generation Conditioned on Acoustic Parameters
par: Arellano, Silvia, et autres
Publié: (2025)
par: Arellano, Silvia, et autres
Publié: (2025)
Assessing the Potential Impact of Direction-Dependent HRTF Selection on Sound Localization Accuracy
par: Goldring, Sapir, et autres
Publié: (2024)
par: Goldring, Sapir, et autres
Publié: (2024)
ScoreDec: A Phase-preserving High-Fidelity Audio Codec with A Generalized Score-based Diffusion Post-filter
par: Wu, Yi-Chiao, et autres
Publié: (2024)
par: Wu, Yi-Chiao, et autres
Publié: (2024)
Matching Reverberant Speech Through Learned Acoustic Embeddings and Feedback Delay Networks
par: Götz, Philipp, et autres
Publié: (2025)
par: Götz, Philipp, et autres
Publié: (2025)
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
par: Zuo, Jialong, et autres
Publié: (2025)
par: Zuo, Jialong, et autres
Publié: (2025)
MusFlow: Multimodal Music Generation via Conditional Flow Matching
par: Song, Jiahao, et autres
Publié: (2025)
par: Song, Jiahao, et autres
Publié: (2025)
SAGA-SR: Semantically and Acoustically Guided Audio Super-Resolution
par: Im, Jaekwon, et autres
Publié: (2025)
par: Im, Jaekwon, et autres
Publié: (2025)
StreamFlow: Streaming Flow Matching with Block-wise Guided Attention Mask for Speech Token Decoding
par: Guo, Dake, et autres
Publié: (2025)
par: Guo, Dake, et autres
Publié: (2025)
FlowSep: Language-Queried Sound Separation with Rectified Flow Matching
par: Yuan, Yi, et autres
Publié: (2024)
par: Yuan, Yi, et autres
Publié: (2024)
Melody-Lyrics Matching with Contrastive Alignment Loss
par: Wang, Changhong, et autres
Publié: (2025)
par: Wang, Changhong, et autres
Publié: (2025)
StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching
par: Yao, Jixun, et autres
Publié: (2024)
par: Yao, Jixun, et autres
Publié: (2024)
Blind Localization of Early Room Reflections with Arbitrary Microphone Array
par: Hadadi, Yogev, et autres
Publié: (2024)
par: Hadadi, Yogev, et autres
Publié: (2024)
Sound Event Detection with Boundary-Aware Optimization and Inference
par: Schmid, Florian, et autres
Publié: (2026)
par: Schmid, Florian, et autres
Publié: (2026)
Adaptive Deterministic Flow Matching for Target Speaker Extraction
par: Hsieh, Tsun-An, et autres
Publié: (2025)
par: Hsieh, Tsun-An, et autres
Publié: (2025)
Identifiability Conditions for Acoustic Feedback Cancellation with the Two-Channel Adaptive Feedback Canceller Algorithm
par: Roebben, Arnout, et autres
Publié: (2025)
par: Roebben, Arnout, et autres
Publié: (2025)
TACO: Training-free Sound Prompted Segmentation via Semantically Constrained Audio-visual CO-factorization
par: Malard, Hugo, et autres
Publié: (2024)
par: Malard, Hugo, et autres
Publié: (2024)
MusicFlow: Cascaded Flow Matching for Text Guided Music Generation
par: Prajwal, K R, et autres
Publié: (2024)
par: Prajwal, K R, et autres
Publié: (2024)
Flow2GAN: Hybrid Flow Matching and GAN with Multi-Resolution Network for Few-step High-Fidelity Audio Generation
par: Yao, Zengwei, et autres
Publié: (2025)
par: Yao, Zengwei, et autres
Publié: (2025)
Mixture-of-Experts Framework for Field-of-View Enhanced Signal-Dependent Binauralization of Moving Talkers
par: Mittal, Manan, et autres
Publié: (2025)
par: Mittal, Manan, et autres
Publié: (2025)
FlowSE: Flow Matching-based Speech Enhancement
par: Lee, Seonggyu, et autres
Publié: (2025)
par: Lee, Seonggyu, et autres
Publié: (2025)
Sub-band Domain Multi-Hypothesis Acoustic Echo Canceler Based Acoustic Scene Analysis
par: Southwell, Benjamin J, et autres
Publié: (2025)
par: Southwell, Benjamin J, et autres
Publié: (2025)
Summary of the NOTSOFAR-1 Challenge: Highlights and Learnings
par: Abramovski, Igor, et autres
Publié: (2025)
par: Abramovski, Igor, et autres
Publié: (2025)
FlowSE-GRPO: Training Flow Matching Speech Enhancement via Online Reinforcement Learning
par: Wang, Haoxu, et autres
Publié: (2026)
par: Wang, Haoxu, et autres
Publié: (2026)
Transfer Learning for Paediatric Sleep Apnoea Detection Using Physiology-Guided Acoustic Models
par: Niu, Chaoyue, et autres
Publié: (2025)
par: Niu, Chaoyue, et autres
Publié: (2025)
Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music Generation
par: Tal, Or, et autres
Publié: (2024)
par: Tal, Or, et autres
Publié: (2024)
Documents similaires
-
Spatial-Magnifier: Spatial upsampling for multichannel speech enhancement
par: Lee, Dongheon, et autres
Publié: (2026) -
ArrayDPS-Refine: Generative Refinement of Discriminative Multi-Channel Speech Enhancement
par: Xu, Zhongweiyang, et autres
Publié: (2026) -
High Fidelity Text-Guided Music Editing via Single-Stage Flow Matching
par: Lan, Gael Le, et autres
Publié: (2024) -
Binaural Signal Matching with Wearable Arrays for Near-Field Sources
par: Goldring, Sapir, et autres
Publié: (2025) -
Unified Diffusion Refinement for Multi-Channel Speech Enhancement and Separation
par: Xu, Zhongweiyang, et autres
Publié: (2026)