TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lu, Wanglong, Su, Lingming, Zheng, Jingjing, de Melo, Vinícius Veloso, Shoeleh, Farzaneh, Hawkin, John, Tricco, Terrence, Zhao, Hanli, Jiang, Xianta |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
di: Li, Jing, et al.
Pubblicazione: (2025)
di: Li, Jing, et al.
Pubblicazione: (2025)
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
di: Brummer, Benoit, et al.
Pubblicazione: (2025)
di: Brummer, Benoit, et al.
Pubblicazione: (2025)
Frequency-Decomposed INR for NIR-Assisted Low-Light RGB Image Denoising
di: Shi, Ligen, et al.
Pubblicazione: (2026)
di: Shi, Ligen, et al.
Pubblicazione: (2026)
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
di: Lu, Wanglong, et al.
Pubblicazione: (2022)
di: Lu, Wanglong, et al.
Pubblicazione: (2022)
Model Agnostic Defense against Adversarial Patch Attacks on Object Detection in Unmanned Aerial Vehicles
di: Pathak, Saurabh, et al.
Pubblicazione: (2024)
di: Pathak, Saurabh, et al.
Pubblicazione: (2024)
A unified Benchmark for Multi-Frame Image Restoration under Severe Refractive Warping
di: Shugaev, Maxim V., et al.
Pubblicazione: (2026)
di: Shugaev, Maxim V., et al.
Pubblicazione: (2026)
FLD+: Data-efficient Evaluation Metric for Generative Models
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Normalizing Flow-Based Metric for Image Generation
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Evaluation Metric for Quality Control and Generative Models in Histopathology Images
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Disentangling Generation and Regression in Stochastic Interpolants for Controllable Image Restoration
di: Liu, Yi, et al.
Pubblicazione: (2026)
di: Liu, Yi, et al.
Pubblicazione: (2026)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
di: Allen, M. J., et al.
Pubblicazione: (2024)
di: Allen, M. J., et al.
Pubblicazione: (2024)
Anti-ESIA: Analyzing and Mitigating Impacts of Electromagnetic Signal Injection Attacks
di: Kang, Denglin, et al.
Pubblicazione: (2024)
di: Kang, Denglin, et al.
Pubblicazione: (2024)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
di: Zhang, Jian, et al.
Pubblicazione: (2024)
di: Zhang, Jian, et al.
Pubblicazione: (2024)
Unveiling Text in Challenging Stone Inscriptions: A Character-Context-Aware Patching Strategy for Binarization
di: Jena, Pratyush, et al.
Pubblicazione: (2026)
di: Jena, Pratyush, et al.
Pubblicazione: (2026)
M3LEO: A Multi-Modal, Multi-Label Earth Observation Dataset Integrating Interferometric SAR and Multispectral Data
di: Allen, Matthew J, et al.
Pubblicazione: (2024)
di: Allen, Matthew J, et al.
Pubblicazione: (2024)
Manual Labelling Artificially Inflates Deep Learning-Based Segmentation Performance on RGB Images of Closed Canopy: Validation Using TLS
di: Allen, Matthew J., et al.
Pubblicazione: (2025)
di: Allen, Matthew J., et al.
Pubblicazione: (2025)
Image and Video Compression using Generative Sparse Representation with Fidelity Controls
di: Jiang, Wei, et al.
Pubblicazione: (2024)
di: Jiang, Wei, et al.
Pubblicazione: (2024)
Haze-Aware Attention Network for Single-Image Dehazing
di: Tong, Lihan, et al.
Pubblicazione: (2024)
di: Tong, Lihan, et al.
Pubblicazione: (2024)
Under-Canopy Terrain Reconstruction in Dense Forests Using RGB Imaging and Neural 3D Reconstruction
di: Sheffer, Refael, et al.
Pubblicazione: (2026)
di: Sheffer, Refael, et al.
Pubblicazione: (2026)
A Real-Time Diminished Reality Approach to Privacy in MR Collaboration
di: Fane, Christian
Pubblicazione: (2025)
di: Fane, Christian
Pubblicazione: (2025)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
di: Gerats, Beerend G. A., et al.
Pubblicazione: (2024)
di: Gerats, Beerend G. A., et al.
Pubblicazione: (2024)
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
di: Wu, Songhan
Pubblicazione: (2025)
di: Wu, Songhan
Pubblicazione: (2025)
GuidPaint: Class-Guided Image Inpainting with Diffusion Models
di: Wang, Qimin, et al.
Pubblicazione: (2025)
di: Wang, Qimin, et al.
Pubblicazione: (2025)
AI-assisted radiographic analysis in detecting alveolar bone-loss severity and patterns
di: Wimalasiri, Chathura, et al.
Pubblicazione: (2025)
di: Wimalasiri, Chathura, et al.
Pubblicazione: (2025)
WaveMix: A Resource-efficient Neural Network for Image Analysis
di: Jeevan, Pranav, et al.
Pubblicazione: (2022)
di: Jeevan, Pranav, et al.
Pubblicazione: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Enhancing rice leaf images: An overview of image denoising techniques
di: Chutia, Rupjyoti, et al.
Pubblicazione: (2025)
di: Chutia, Rupjyoti, et al.
Pubblicazione: (2025)
IAMAP: Unlocking Deep Learning in QGIS for non-coders and limited computing resources
di: Tresson, Paul, et al.
Pubblicazione: (2025)
di: Tresson, Paul, et al.
Pubblicazione: (2025)
Optimizing Multi-Scale Representations to Detect Effect Heterogeneity Using Earth Observation and Computer Vision: Applications to Two Anti-Poverty RCTs
di: Zhu, Fucheng Warren, et al.
Pubblicazione: (2024)
di: Zhu, Fucheng Warren, et al.
Pubblicazione: (2024)
SAR Despeckling via Log-Yeo-Johnson Transformation and Sparse Representation
di: Hu, Xuran, et al.
Pubblicazione: (2024)
di: Hu, Xuran, et al.
Pubblicazione: (2024)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
di: Duguay, Simon-Olivier, et al.
Pubblicazione: (2026)
di: Duguay, Simon-Olivier, et al.
Pubblicazione: (2026)
Event-based Solutions for Human-centered Applications: A Comprehensive Review
di: Adra, Mira, et al.
Pubblicazione: (2025)
di: Adra, Mira, et al.
Pubblicazione: (2025)
Optimizing the image correction pipeline for pedestrian detection in the thermal-infrared domain
di: Karam, Christophe, et al.
Pubblicazione: (2024)
di: Karam, Christophe, et al.
Pubblicazione: (2024)
DCMSA: Multi-Head Self-Attention Mechanism Based on Deformable Convolution For Seismic Data Denoising
di: Mingwei, Wang, et al.
Pubblicazione: (2024)
di: Mingwei, Wang, et al.
Pubblicazione: (2024)
Context-dependent Causality (the Non-Nonotonic Case)
di: Billfeld, Nir, et al.
Pubblicazione: (2024)
di: Billfeld, Nir, et al.
Pubblicazione: (2024)
CerberusDet: Unified Multi-Dataset Object Detection
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
di: Tolstykh, Irina, et al.
Pubblicazione: (2024)
Bridging Knowledge Gap Between Image Inpainting and Large-Area Visible Watermark Removal
di: Leng, Yicheng, et al.
Pubblicazione: (2025)
di: Leng, Yicheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
di: Lu, Wanglong, et al.
Pubblicazione: (2024) -
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
di: Lu, Wanglong, et al.
Pubblicazione: (2024) -
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
di: Li, Jing, et al.
Pubblicazione: (2025) -
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
di: Brummer, Benoit, et al.
Pubblicazione: (2025) -
Frequency-Decomposed INR for NIR-Assisted Low-Light RGB Image Denoising
di: Shi, Ligen, et al.
Pubblicazione: (2026)