Inverse Scene Text Removal
Fuente:
arXiv
Saved in:
| Main Authors: | Yoshimatsu, Takumi, Takezaki, Shumpei, Uchida, Seiichi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Ordinal Diffusion Model for Generating Medical Images with Different Severity Levels
by: Takezaki, Shumpei, et al.
Published: (2024)
by: Takezaki, Shumpei, et al.
Published: (2024)
Self-Relaxed Joint Training: Sample Selection for Severity Estimation with Ordinal Noisy Labels
by: Takezaki, Shumpei, et al.
Published: (2024)
by: Takezaki, Shumpei, et al.
Published: (2024)
SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images
by: Matsuzaki, Yuta, et al.
Published: (2026)
by: Matsuzaki, Yuta, et al.
Published: (2026)
Few-Part-Shot Font Generation
by: Akiba, Masaki, et al.
Published: (2025)
by: Akiba, Masaki, et al.
Published: (2025)
Font Style Interpolation with Diffusion Models
by: Kondo, Tetta, et al.
Published: (2024)
by: Kondo, Tetta, et al.
Published: (2024)
Cross-Domain Image Conversion by CycleDM
by: Shimotsumagari, Sho, et al.
Published: (2024)
by: Shimotsumagari, Sho, et al.
Published: (2024)
NoiseCutMix: A Novel Data Augmentation Approach by Mixing Estimated Noise in Diffusion Models
by: Takezaki, Shumpei, et al.
Published: (2025)
by: Takezaki, Shumpei, et al.
Published: (2025)
VesselFusion: Diffusion Models for Vessel Centerline Extraction from 3D CT Images
by: Mita, Soichi, et al.
Published: (2026)
by: Mita, Soichi, et al.
Published: (2026)
Computer-Aided Multi-Stroke Character Simplification by Stroke Removal
by: Ishiyama, Ryo, et al.
Published: (2025)
by: Ishiyama, Ryo, et al.
Published: (2025)
NoiseCollage: A Layout-Aware Text-to-Image Diffusion Model Based on Noise Cropping and Merging
by: Shirakawa, Takahiro, et al.
Published: (2024)
by: Shirakawa, Takahiro, et al.
Published: (2024)
Cell Instance Segmentation via Multi-Task Image-to-Image Schrödinger Bridge
by: Inoue, Hayato, et al.
Published: (2026)
by: Inoue, Hayato, et al.
Published: (2026)
What Text Design Characterizes Book Genres?
by: Haraguchi, Daichi, et al.
Published: (2024)
by: Haraguchi, Daichi, et al.
Published: (2024)
Guidance-base Diffusion Models for Improving Photoacoustic Image Quality
by: Eguchi, Tatsuhiro, et al.
Published: (2025)
by: Eguchi, Tatsuhiro, et al.
Published: (2025)
Typographic Text Generation with Off-the-Shelf Diffusion Model
by: Peong, KhayTze, et al.
Published: (2024)
by: Peong, KhayTze, et al.
Published: (2024)
Embedding Font Impression Word Tags Based on Co-occurrence
by: Kubota, Yugo, et al.
Published: (2025)
by: Kubota, Yugo, et al.
Published: (2025)
Learning to Kern: Set-wise Estimation of Optimal Letter Space
by: Nakatsuru, Kei, et al.
Published: (2024)
by: Nakatsuru, Kei, et al.
Published: (2024)
Enhancing Reliability of Medical Image Diagnosis through Top-rank Learning with Rejection Module
by: Ji, Xiaotong, et al.
Published: (2025)
by: Ji, Xiaotong, et al.
Published: (2025)
Impression-CLIP: Contrastive Shape-Impression Embedding for Fonts
by: Kubota, Yugo, et al.
Published: (2024)
by: Kubota, Yugo, et al.
Published: (2024)
Pseudo-label Learning with Calibrated Confidence Using an Energy-based Model
by: Toba, Masahito, et al.
Published: (2024)
by: Toba, Masahito, et al.
Published: (2024)
Hierarchical Co-Embedding of Font Shapes and Impression Tags
by: Kubota, Yugo, et al.
Published: (2026)
by: Kubota, Yugo, et al.
Published: (2026)
Font Impression Estimation in the Wild
by: Kitajima, Kazuki, et al.
Published: (2024)
by: Kitajima, Kazuki, et al.
Published: (2024)
Automatic Text Box Placement for Supporting Typographic Design
by: Muraoka, Jun, et al.
Published: (2025)
by: Muraoka, Jun, et al.
Published: (2025)
Type-R: Automatically Retouching Typos for Text-to-Image Generation
by: Shimoda, Wataru, et al.
Published: (2024)
by: Shimoda, Wataru, et al.
Published: (2024)
Ranking-Guided Semi-Supervised Domain Adaptation for Severity Classification
by: Harada, Shota, et al.
Published: (2026)
by: Harada, Shota, et al.
Published: (2026)
Total Disentanglement of Font Images into Style and Character Class Features
by: Haraguchi, Daichi, et al.
Published: (2024)
by: Haraguchi, Daichi, et al.
Published: (2024)
DiffSTR: Controlled Diffusion Models for Scene Text Removal
by: Pathak, Sanhita, et al.
Published: (2024)
by: Pathak, Sanhita, et al.
Published: (2024)
Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image Modeling
by: Wang, Zixiao, et al.
Published: (2024)
by: Wang, Zixiao, et al.
Published: (2024)
Deep Bayesian Active Learning-to-Rank with Relative Annotation for Estimation of Ulcerative Colitis Severity
by: Kadota, Takeaki, et al.
Published: (2024)
by: Kadota, Takeaki, et al.
Published: (2024)
Leveraging Vision-Language Models as Weak Annotators in Active Learning
by: Nguyen, Phuong Ngoc, et al.
Published: (2026)
by: Nguyen, Phuong Ngoc, et al.
Published: (2026)
ViTEraser: Harnessing the Power of Vision Transformers for Scene Text Removal with SegMIM Pretraining
by: Peng, Dezhi, et al.
Published: (2023)
by: Peng, Dezhi, et al.
Published: (2023)
Choose What You Need: Disentangled Representation Learning for Scene Text Recognition, Removal and Editing
by: Zhang, Boqiang, et al.
Published: (2024)
by: Zhang, Boqiang, et al.
Published: (2024)
Inverse-like Antagonistic Scene Text Spotting via Reading-Order Estimation and Dynamic Sampling
by: Zhang, Shi-Xue, et al.
Published: (2024)
by: Zhang, Shi-Xue, et al.
Published: (2024)
Instance-wise Supervision-level Optimization in Active Learning
by: Matsuo, Shinnosuke, et al.
Published: (2025)
by: Matsuo, Shinnosuke, et al.
Published: (2025)
OTR: Synthesizing Overlay Text Dataset for Text Removal
by: Zdenek, Jan, et al.
Published: (2025)
by: Zdenek, Jan, et al.
Published: (2025)
Learning from Partial Label Proportions for Whole Slide Image Segmentation
by: Matsuo, Shinnosuke, et al.
Published: (2024)
by: Matsuo, Shinnosuke, et al.
Published: (2024)
Can GPTs Evaluate Graphic Design Based on Design Principles?
by: Haraguchi, Daichi, et al.
Published: (2024)
by: Haraguchi, Daichi, et al.
Published: (2024)
Compositional Scene Understanding through Inverse Generative Modeling
by: Wang, Yanbo, et al.
Published: (2025)
by: Wang, Yanbo, et al.
Published: (2025)
ConText: Driving In-context Learning for Text Removal and Segmentation
by: Zhang, Fei, et al.
Published: (2025)
by: Zhang, Fei, et al.
Published: (2025)
Aggregated Text Transformer for Scene Text Detection
by: Zhou, Zhao, et al.
Published: (2022)
by: Zhou, Zhao, et al.
Published: (2022)
Partial Scene Text Retrieval
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Similar Items
-
An Ordinal Diffusion Model for Generating Medical Images with Different Severity Levels
by: Takezaki, Shumpei, et al.
Published: (2024) -
Self-Relaxed Joint Training: Sample Selection for Severity Estimation with Ordinal Noisy Labels
by: Takezaki, Shumpei, et al.
Published: (2024) -
SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images
by: Matsuzaki, Yuta, et al.
Published: (2026) -
Few-Part-Shot Font Generation
by: Akiba, Masaki, et al.
Published: (2025) -
Font Style Interpolation with Diffusion Models
by: Kondo, Tetta, et al.
Published: (2024)