Learning to reconstruct from saturated data: audio declipping and high-dynamic range imaging
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sechaud, Victor, Jacques, Laurent, Abry, Patrice, Tachella, Julián |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
von: Sechaud, Victor, et al.
Veröffentlicht: (2024)
von: Sechaud, Victor, et al.
Veröffentlicht: (2024)
Scale-Equivariant Imaging: Self-Supervised Learning for Image Super-Resolution and Deblurring
von: Scanvic, Jérémy, et al.
Veröffentlicht: (2023)
von: Scanvic, Jérémy, et al.
Veröffentlicht: (2023)
End-to-end audio-visual learning for cochlear implant sound coding simulations in noisy environments
von: Lin, Meng-Ping, et al.
Veröffentlicht: (2025)
von: Lin, Meng-Ping, et al.
Veröffentlicht: (2025)
PGSTalker: Real-Time Audio-Driven Talking Head Generation via 3D Gaussian Splatting with Pixel-Aware Density Control
von: Zhu, Tianheng, et al.
Veröffentlicht: (2025)
von: Zhu, Tianheng, et al.
Veröffentlicht: (2025)
SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet
von: Zhong, Zhi, et al.
Veröffentlicht: (2025)
von: Zhong, Zhi, et al.
Veröffentlicht: (2025)
DensePANet: An improved generative adversarial network for photoacoustic tomography image reconstruction from sparse data
von: Hakimnejad, Hesam, et al.
Veröffentlicht: (2024)
von: Hakimnejad, Hesam, et al.
Veröffentlicht: (2024)
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models
von: Wang, Yi, et al.
Veröffentlicht: (2025)
von: Wang, Yi, et al.
Veröffentlicht: (2025)
Equivariant plug-and-play image reconstruction
von: Terris, Matthieu, et al.
Veröffentlicht: (2023)
von: Terris, Matthieu, et al.
Veröffentlicht: (2023)
Self-supervised learning for phase retrieval
von: Sechaud, Victor, et al.
Veröffentlicht: (2025)
von: Sechaud, Victor, et al.
Veröffentlicht: (2025)
Video Soundtrack Generation by Aligning Emotions and Temporal Boundaries
von: Sulun, Serkan, et al.
Veröffentlicht: (2025)
von: Sulun, Serkan, et al.
Veröffentlicht: (2025)
Normalization-equivariant Diffusion Models: Learning Posterior Samplers From Noisy And Partial Measurements
von: Levac, Brett, et al.
Veröffentlicht: (2025)
von: Levac, Brett, et al.
Veröffentlicht: (2025)
PTSD-MDNN : Fusion tardive de réseaux de neurones profonds multimodaux pour la détection du trouble de stress post-traumatique
von: Nguyen-Phuoc, Long, et al.
Veröffentlicht: (2024)
von: Nguyen-Phuoc, Long, et al.
Veröffentlicht: (2024)
Generalized Recorrupted-to-Recorrupted: Self-Supervised Learning Beyond Gaussian Noise
von: Monroy, Brayan, et al.
Veröffentlicht: (2024)
von: Monroy, Brayan, et al.
Veröffentlicht: (2024)
Deep Learning for Steganalysis of Diverse Data Types: A review of methods, taxonomy, challenges and future directions
von: Kheddar, Hamza, et al.
Veröffentlicht: (2023)
von: Kheddar, Hamza, et al.
Veröffentlicht: (2023)
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
Exploring Phonetic Context-Aware Lip-Sync For Talking Face Generation
von: Park, Se Jin, et al.
Veröffentlicht: (2023)
von: Park, Se Jin, et al.
Veröffentlicht: (2023)
MOST: MR reconstruction Optimization for multiple downStream Tasks via continual learning
von: Jeong, Hwihun, et al.
Veröffentlicht: (2024)
von: Jeong, Hwihun, et al.
Veröffentlicht: (2024)
Robust Multi-modal Task-oriented Communications with Redundancy-aware Representations
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
Whittaker-Henderson smoother for long satellite image time series interpolation
von: Fauvel, Mathieu
Veröffentlicht: (2026)
von: Fauvel, Mathieu
Veröffentlicht: (2026)
A unified deeplearning framework for contrast-phase-specific virtual monochromatic imaging
von: Jerald, Antony, et al.
Veröffentlicht: (2026)
von: Jerald, Antony, et al.
Veröffentlicht: (2026)
Extraction of 3D trajectories of mandibular condyles from 2D real-time MRI
von: Isaieva, Karyna, et al.
Veröffentlicht: (2024)
von: Isaieva, Karyna, et al.
Veröffentlicht: (2024)
Reacting like Humans: Incorporating Intrinsic Human Behaviors into NAO through Sound-Based Reactions to Fearful and Shocking Events for Enhanced Sociability
von: Ghadami, Ali, et al.
Veröffentlicht: (2023)
von: Ghadami, Ali, et al.
Veröffentlicht: (2023)
Onboard deep lossless and near-lossless predictive coding of hyperspectral images with line-based attention
von: Valsesia, Diego, et al.
Veröffentlicht: (2024)
von: Valsesia, Diego, et al.
Veröffentlicht: (2024)
Characterizing segregation in blast rock piles a deep-learning approach leveraging aerial image analysis
von: Liu, Chengeng, et al.
Veröffentlicht: (2024)
von: Liu, Chengeng, et al.
Veröffentlicht: (2024)
Deep-learning-based clustering of OCT images for biomarker discovery in age-related macular degeneration (Pinnacle study report 4)
von: Holland, Robbie, et al.
Veröffentlicht: (2024)
von: Holland, Robbie, et al.
Veröffentlicht: (2024)
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
Towards Language-Independent Face-Voice Association with Multimodal Foundation Models
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
KunquDB: An Attempt for Speaker Verification in the Chinese Opera Scenario
von: Zhou, Huali, et al.
Veröffentlicht: (2024)
von: Zhou, Huali, et al.
Veröffentlicht: (2024)
Interpretable Modeling of Articulatory Temporal Dynamics from real-time MRI for Phoneme Recognition
von: Park, Jay, et al.
Veröffentlicht: (2025)
von: Park, Jay, et al.
Veröffentlicht: (2025)
Improvement Of Audiovisual Quality Estimation Using A Nonlinear Autoregressive Exogenous Neural Network And Bitstream Parameters
von: Kossi, Koffi, et al.
Veröffentlicht: (2024)
von: Kossi, Koffi, et al.
Veröffentlicht: (2024)
Efficient Face Detection with Audio-Based Region Proposals for Human-Robot Interactions
von: Aris, William, et al.
Veröffentlicht: (2023)
von: Aris, William, et al.
Veröffentlicht: (2023)
GAN-based synthetic FDG PET images from T1 brain MRI can serve to improve performance of deep unsupervised anomaly detection models
von: Zotova, Daria, et al.
Veröffentlicht: (2025)
von: Zotova, Daria, et al.
Veröffentlicht: (2025)
Physics-informed 4D X-ray image reconstruction from ultra-sparse spatiotemporal data
von: Yao, Zisheng, et al.
Veröffentlicht: (2025)
von: Yao, Zisheng, et al.
Veröffentlicht: (2025)
Bayesian multi-exposure image fusion for robust high dynamic range ptychography
von: Kodgirwar, Shantanu, et al.
Veröffentlicht: (2024)
von: Kodgirwar, Shantanu, et al.
Veröffentlicht: (2024)
Understanding Audiovisual Deepfake Detection: Techniques, Challenges, Human Factors and Perceptual Insights
von: Hashmi, Ammarah, et al.
Veröffentlicht: (2024)
von: Hashmi, Ammarah, et al.
Veröffentlicht: (2024)
Reconstruct Anything Model: a lightweight general model for computational imaging
von: Terris, Matthieu, et al.
Veröffentlicht: (2025)
von: Terris, Matthieu, et al.
Veröffentlicht: (2025)
A Wavefield Correlation Approach to Improve Sound Speed Estimation in Ultrasound Autofocusing
von: Zhuang, Louise, et al.
Veröffentlicht: (2026)
von: Zhuang, Louise, et al.
Veröffentlicht: (2026)
AINet: Anchor Instances Learning for Regional Heterogeneity in Whole Slide Image
von: Zheng, Tingting, et al.
Veröffentlicht: (2026)
von: Zheng, Tingting, et al.
Veröffentlicht: (2026)
A speckle filter for Sentinel-1 SAR Ground Range Detected data based on Residual Convolutional Neural Networks
von: Sebastianelli, Alessandro, et al.
Veröffentlicht: (2021)
von: Sebastianelli, Alessandro, et al.
Veröffentlicht: (2021)
WaterFlow: Learning Fast & Robust Watermarks using Stable Diffusion
von: Shukla, Vinay, et al.
Veröffentlicht: (2025)
von: Shukla, Vinay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
von: Sechaud, Victor, et al.
Veröffentlicht: (2024) -
Scale-Equivariant Imaging: Self-Supervised Learning for Image Super-Resolution and Deblurring
von: Scanvic, Jérémy, et al.
Veröffentlicht: (2023) -
End-to-end audio-visual learning for cochlear implant sound coding simulations in noisy environments
von: Lin, Meng-Ping, et al.
Veröffentlicht: (2025) -
PGSTalker: Real-Time Audio-Driven Talking Head Generation via 3D Gaussian Splatting with Pixel-Aware Density Control
von: Zhu, Tianheng, et al.
Veröffentlicht: (2025) -
SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet
von: Zhong, Zhi, et al.
Veröffentlicht: (2025)