Unveiling Text in Challenging Stone Inscriptions: A Character-Context-Aware Patching Strategy for Binarization
Fuente:
arXiv
Salvato in:
| Autori principali: | Jena, Pratyush, Joseph, Amal, Sharma, Arnav, Sarvadevabhatla, Ravi Kiran |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Frequency-Decomposed INR for NIR-Assisted Low-Light RGB Image Denoising
di: Shi, Ligen, et al.
Pubblicazione: (2026)
di: Shi, Ligen, et al.
Pubblicazione: (2026)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
di: Zhang, Jian, et al.
Pubblicazione: (2024)
di: Zhang, Jian, et al.
Pubblicazione: (2024)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
di: Lu, Wanglong, et al.
Pubblicazione: (2025)
di: Lu, Wanglong, et al.
Pubblicazione: (2025)
VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
FLD+: Data-efficient Evaluation Metric for Generative Models
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Normalizing Flow-Based Metric for Image Generation
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Mobile-Ready Automated Triage of Diabetic Retinopathy Using Digital Fundus Images
di: Joshi, Aadi, et al.
Pubblicazione: (2026)
di: Joshi, Aadi, et al.
Pubblicazione: (2026)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
di: Yasuno, Takato
Pubblicazione: (2026)
di: Yasuno, Takato
Pubblicazione: (2026)
Revisiting [CLS] and Patch Token Interaction in Vision Transformers
di: Marouani, Alexis, et al.
Pubblicazione: (2026)
di: Marouani, Alexis, et al.
Pubblicazione: (2026)
Neural Fields for 3D Tracking of Anatomy and Surgical Instruments in Monocular Laparoscopic Video Clips
di: Gerats, Beerend G. A., et al.
Pubblicazione: (2024)
di: Gerats, Beerend G. A., et al.
Pubblicazione: (2024)
A Hierarchical Self-Consistent Regularization Approach to Satellite Image Time Series Classification
di: Weikmann, Giulio, et al.
Pubblicazione: (2025)
di: Weikmann, Giulio, et al.
Pubblicazione: (2025)
Learning Joint Denoising, Demosaicing, and Compression from the Raw Natural Image Noise Dataset
di: Brummer, Benoit, et al.
Pubblicazione: (2025)
di: Brummer, Benoit, et al.
Pubblicazione: (2025)
HySparK: Hybrid Sparse Masking for Large Scale Medical Image Pre-Training
di: Tang, Fenghe, et al.
Pubblicazione: (2024)
di: Tang, Fenghe, et al.
Pubblicazione: (2024)
Quick unsupervised hyperspectral dimensionality reduction for earth observation: a comparison
di: Lupu, Daniela, et al.
Pubblicazione: (2024)
di: Lupu, Daniela, et al.
Pubblicazione: (2024)
A Novel Global Context-aware Deep Neural Network for Enhanced Brain Tumor Segmentation using Magnetic Resonance Images
di: Mukherjee, Sourjya, et al.
Pubblicazione: (2026)
di: Mukherjee, Sourjya, et al.
Pubblicazione: (2026)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
di: Zhang, Xinyi, et al.
Pubblicazione: (2026)
di: Zhang, Xinyi, et al.
Pubblicazione: (2026)
Evaluation Metric for Quality Control and Generative Models in Histopathology Images
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
di: Patel, Hitesh Laxmichand, et al.
Pubblicazione: (2025)
Fusing Structure from Motion and Simulation-Augmented Pose Regression from Optical Flow for Challenging Indoor Environments
di: Ott, Felix, et al.
Pubblicazione: (2023)
di: Ott, Felix, et al.
Pubblicazione: (2023)
The Adobe Hidden Feature and its Impact on Sensor Attribution
di: Butora, Jan, et al.
Pubblicazione: (2023)
di: Butora, Jan, et al.
Pubblicazione: (2023)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology
di: Dimasaka, Joshua, et al.
Pubblicazione: (2025)
di: Dimasaka, Joshua, et al.
Pubblicazione: (2025)
Learning to Expand Images for Efficient Visual Autoregressive Modeling
di: Yang, Ruiqing, et al.
Pubblicazione: (2025)
di: Yang, Ruiqing, et al.
Pubblicazione: (2025)
EDSNet: Efficient-DSNet for Video Summarization
di: Prasad, Ashish, et al.
Pubblicazione: (2024)
di: Prasad, Ashish, et al.
Pubblicazione: (2024)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
di: Benfeghoul, Martin, et al.
Pubblicazione: (2024)
di: Benfeghoul, Martin, et al.
Pubblicazione: (2024)
Differentiable Hierarchical Visual Tokenization
di: Aasan, Marius, et al.
Pubblicazione: (2025)
di: Aasan, Marius, et al.
Pubblicazione: (2025)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
di: Maniyar, Chintan B., et al.
Pubblicazione: (2025)
di: Maniyar, Chintan B., et al.
Pubblicazione: (2025)
Removing Motion Artifact in MRI by Using a Perceptual Loss Driven Deep Learning Framework
di: Guo, Ziheng, et al.
Pubblicazione: (2026)
di: Guo, Ziheng, et al.
Pubblicazione: (2026)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
Label Delay in Online Continual Learning
di: Csaba, Botos, et al.
Pubblicazione: (2023)
di: Csaba, Botos, et al.
Pubblicazione: (2023)
Neural Implicit Morphing of Face Images
di: Schardong, Guilherme, et al.
Pubblicazione: (2023)
di: Schardong, Guilherme, et al.
Pubblicazione: (2023)
RealHD: A High-Quality Dataset for Robust Detection of State-of-the-Art AI-Generated Images
di: Yu, Hanzhe, et al.
Pubblicazione: (2026)
di: Yu, Hanzhe, et al.
Pubblicazione: (2026)
When Style Similarity Scores Fail: Diagnosing Raw CSD Cosine in Artist-Style Evaluation
di: Frochte, Jörg
Pubblicazione: (2026)
di: Frochte, Jörg
Pubblicazione: (2026)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
Learning 3D object-centric representation through prediction
di: Day, John, et al.
Pubblicazione: (2024)
di: Day, John, et al.
Pubblicazione: (2024)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
di: Adiban, Mohammad, et al.
Pubblicazione: (2023)
di: Adiban, Mohammad, et al.
Pubblicazione: (2023)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
di: Dalal, Anurag, et al.
Pubblicazione: (2024)
di: Dalal, Anurag, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Frequency-Decomposed INR for NIR-Assisted Low-Light RGB Image Denoising
di: Shi, Ligen, et al.
Pubblicazione: (2026) -
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
di: Zhang, Jian, et al.
Pubblicazione: (2024) -
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
di: Lu, Wanglong, et al.
Pubblicazione: (2025) -
VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
di: Chen, Zhipeng, et al.
Pubblicazione: (2024) -
FLD+: Data-efficient Evaluation Metric for Generative Models
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)