ClapperText: A Benchmark for Text Recognition in Low-Resource Archival Documents
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Tingyu, Peer, Marco, Kleber, Florian, Sablatnig, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Camera Movement Classification in Historical Footage: A Comparative Study of Deep Video Models
by: Lin, Tingyu, et al.
Published: (2025)
by: Lin, Tingyu, et al.
Published: (2025)
DGME-T: Directional Grid Motion Encoding for Transformer-Based Historical Camera Movement Classification
by: Lin, Tingyu, et al.
Published: (2025)
by: Lin, Tingyu, et al.
Published: (2025)
Enhancing Historical Image Retrieval with Compositional Cues
by: Lin, Tingyu, et al.
Published: (2024)
by: Lin, Tingyu, et al.
Published: (2024)
MMM-RS: A Multi-modal, Multi-GSD, Multi-scene Remote Sensing Dataset and Benchmark for Text-to-Image Generation
by: Luo, Jialin, et al.
Published: (2024)
by: Luo, Jialin, et al.
Published: (2024)
Spatially Covariant Image Registration with Text Prompts
by: Chen, Xiang, et al.
Published: (2023)
by: Chen, Xiang, et al.
Published: (2023)
Multimodal Medical Image Binding via Shared Text Embeddings
by: Liu, Yunhao, et al.
Published: (2025)
by: Liu, Yunhao, et al.
Published: (2025)
Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts
by: Ge, Peixuan, et al.
Published: (2025)
by: Ge, Peixuan, et al.
Published: (2025)
Parametric Shadow Control for Portrait Generation in Text-to-Image Diffusion Models
by: Cai, Haoming, et al.
Published: (2025)
by: Cai, Haoming, et al.
Published: (2025)
Latent Space Synergy: Text-Guided Data Augmentation for Direct Diffusion Biomedical Segmentation
by: Aqeel, Muhammad, et al.
Published: (2025)
by: Aqeel, Muhammad, et al.
Published: (2025)
A Low-cost and Ultra-lightweight Binary Neural Network for Traffic Signal Recognition
by: Xiao, Mingke, et al.
Published: (2025)
by: Xiao, Mingke, et al.
Published: (2025)
Face-MakeUpV2: Facial Consistency Learning for Controllable Text-to-Image Generation
by: Dai, Dawei, et al.
Published: (2025)
by: Dai, Dawei, et al.
Published: (2025)
ReXGroundingCT: A 3D Chest CT Dataset for Segmentation of Findings from Free-Text Reports
by: Baharoon, Mohammed, et al.
Published: (2025)
by: Baharoon, Mohammed, et al.
Published: (2025)
Mpox Screen Lite: AI-Driven On-Device Offline Mpox Screening for Low-Resource African Mpox Emergency Response
by: Kularathne, Yudara, et al.
Published: (2024)
by: Kularathne, Yudara, et al.
Published: (2024)
Interpretable Geoscience Artificial Intelligence (XGeoS-AI): Application to Demystify Image Recognition
by: Xu, Jin-Jian, et al.
Published: (2023)
by: Xu, Jin-Jian, et al.
Published: (2023)
The Impact of Scanner Domain Shift on Deep Learning Performance in Medical Imaging: an Experimental Study
by: Guo, Brian, et al.
Published: (2024)
by: Guo, Brian, et al.
Published: (2024)
Towards the Influence of Text Quantity on Writer Retrieval
by: Peer, Marco, et al.
Published: (2025)
by: Peer, Marco, et al.
Published: (2025)
Tomato Maturity Recognition with Convolutional Transformers
by: Khan, Asim, et al.
Published: (2023)
by: Khan, Asim, et al.
Published: (2023)
Visual and Text Prompt Segmentation: A Novel Multi-Model Framework for Remote Sensing
by: Zi, Xing, et al.
Published: (2025)
by: Zi, Xing, et al.
Published: (2025)
Few-Shot Connectivity-Aware Text Line Segmentation in Historical Documents
by: Sterzinger, Rafael, et al.
Published: (2025)
by: Sterzinger, Rafael, et al.
Published: (2025)
TAVID: Text-Driven Audio-Visual Interactive Dialogue Generation
by: Kim, Ji-Hoon, et al.
Published: (2025)
by: Kim, Ji-Hoon, et al.
Published: (2025)
DRL-STNet: Unsupervised Domain Adaptation for Cross-modality Medical Image Segmentation via Disentangled Representation Learning
by: Lin, Hui, et al.
Published: (2024)
by: Lin, Hui, et al.
Published: (2024)
On the Efficacy of Text-Based Input Modalities for Action Anticipation
by: Beedu, Apoorva, et al.
Published: (2024)
by: Beedu, Apoorva, et al.
Published: (2024)
HazeMatching: Dehazing Light Microscopy Images with Guided Conditional Flow Matching
by: Ray, Anirban, et al.
Published: (2025)
by: Ray, Anirban, et al.
Published: (2025)
Fine-Grained Cat Breed Recognition with Global Context Vision Transformer
by: Hera, Mowmita Parvin, et al.
Published: (2026)
by: Hera, Mowmita Parvin, et al.
Published: (2026)
Zero-shot self-supervised learning of single breath-hold magnetic resonance cholangiopancreatography (MRCP) reconstruction
by: Kim, Jinho, et al.
Published: (2025)
by: Kim, Jinho, et al.
Published: (2025)
StreamDiT: Real-Time Streaming Text-to-Video Generation
by: Kodaira, Akio, et al.
Published: (2025)
by: Kodaira, Akio, et al.
Published: (2025)
Analysis of Transferred Pre-Trained Deep Convolution Neural Networks in Breast Masses Recognition
by: Hamad, Qusay Shihab, et al.
Published: (2024)
by: Hamad, Qusay Shihab, et al.
Published: (2024)
Foundation Artificial Intelligence Models for Health Recognition Using Face Photographs (FAHR-Face)
by: Haugg, Fridolin, et al.
Published: (2025)
by: Haugg, Fridolin, et al.
Published: (2025)
Quantization-Aware Neuromorphic Architecture for Skin Disease Classification on Resource-Constrained Devices
by: Wang, Haitian, et al.
Published: (2025)
by: Wang, Haitian, et al.
Published: (2025)
Benchmarking Self-Supervised Models for Cardiac Ultrasound View Classification
by: Megahed, Youssef, et al.
Published: (2026)
by: Megahed, Youssef, et al.
Published: (2026)
SurgWound-Bench: A Benchmark for Surgical Wound Diagnosis
by: Xu, Jiahao, et al.
Published: (2025)
by: Xu, Jiahao, et al.
Published: (2025)
FACE: Few-shot Adapter with Cross-view Fusion for Cross-subject EEG Emotion Recognition
by: Liu, Haiqi, et al.
Published: (2025)
by: Liu, Haiqi, et al.
Published: (2025)
Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition
by: Ramirez, David F., et al.
Published: (2026)
by: Ramirez, David F., et al.
Published: (2026)
Deep Learning for Surgical Instrument Recognition and Segmentation in Robotic-Assisted Surgeries: A Systematic Review
by: Ahmed, Fatimaelzahraa Ali, et al.
Published: (2024)
by: Ahmed, Fatimaelzahraa Ali, et al.
Published: (2024)
Hunting imaging biomarkers in pulmonary fibrosis: Benchmarks of the AIIB23 challenge
by: Nan, Yang, et al.
Published: (2023)
by: Nan, Yang, et al.
Published: (2023)
DM4CT: Benchmarking Diffusion Models for Computed Tomography Reconstruction
by: Shi, Jiayang, et al.
Published: (2026)
by: Shi, Jiayang, et al.
Published: (2026)
FreeTumor: Large-Scale Generative Tumor Synthesis in Computed Tomography Images for Improving Tumor Recognition
by: Wu, Linshan, et al.
Published: (2025)
by: Wu, Linshan, et al.
Published: (2025)
Internship Report: Benchmark of Deep Learning-based Imaging PPG in Automotive Domain
by: Tu, Yuqi, et al.
Published: (2024)
by: Tu, Yuqi, et al.
Published: (2024)
Towards Real-world Video Face Restoration: A New Benchmark
by: Chen, Ziyan, et al.
Published: (2024)
by: Chen, Ziyan, et al.
Published: (2024)
REHRSeg: Unleashing the Power of Self-Supervised Super-Resolution for Resource-Efficient 3D MRI Segmentation
by: Song, Zhiyun, et al.
Published: (2024)
by: Song, Zhiyun, et al.
Published: (2024)
Similar Items
-
Camera Movement Classification in Historical Footage: A Comparative Study of Deep Video Models
by: Lin, Tingyu, et al.
Published: (2025) -
DGME-T: Directional Grid Motion Encoding for Transformer-Based Historical Camera Movement Classification
by: Lin, Tingyu, et al.
Published: (2025) -
Enhancing Historical Image Retrieval with Compositional Cues
by: Lin, Tingyu, et al.
Published: (2024) -
MMM-RS: A Multi-modal, Multi-GSD, Multi-scene Remote Sensing Dataset and Benchmark for Text-to-Image Generation
by: Luo, Jialin, et al.
Published: (2024) -
Spatially Covariant Image Registration with Text Prompts
by: Chen, Xiang, et al.
Published: (2023)