OCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Peirong, Xu, Haowei, Zhang, Jiaxin, Zheng, Xuhan, Xu, Guitao, Zhang, Yuyi, Liu, Junle, Yang, Zhenhua, Zhou, Wei, Jin, Lianwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PosterVerse: A Full-Workflow Framework for Commercial-Grade Poster Generation with HTML-Based Scalable Typography
von: Liu, Junle, et al.
Veröffentlicht: (2026)
von: Liu, Junle, et al.
Veröffentlicht: (2026)
MegaHan97K: A Large-Scale Dataset for Mega-Category Chinese Character Recognition with over 97K Categories
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
Reviving Cultural Heritage: A Novel Approach for Comprehensive Historical Document Restoration
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
Online Writer Retrieval with Chinese Handwritten Phrases: A Synergistic Temporal-Frequency Representation Learning Approach
von: Zhang, Peirong, et al.
Veröffentlicht: (2024)
von: Zhang, Peirong, et al.
Veröffentlicht: (2024)
Smaller But Better: Unifying Layout Generation with Smaller Large Language Models
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
HierCode: A Lightweight Hierarchical Codebook for Zero-shot Chinese Text Recognition
von: Zhang, Yuyi, et al.
Veröffentlicht: (2024)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2024)
DocRes: A Generalist Model Toward Unifying Document Image Restoration Tasks
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2024)
Capturing More: Learning Multi-Domain Representations for Robust Online Handwriting Verification
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
UPOCR: Towards Unified Pixel-Level OCR Interface
von: Peng, Dezhi, et al.
Veröffentlicht: (2023)
von: Peng, Dezhi, et al.
Veröffentlicht: (2023)
MCCD: A Multi-Attribute Chinese Calligraphy Character Dataset Annotated with Script Styles, Dynasties, and Calligraphers
von: Zhao, Yixin, et al.
Veröffentlicht: (2025)
von: Zhao, Yixin, et al.
Veröffentlicht: (2025)
Predicting the Original Appearance of Damaged Historical Documents
von: Yang, Zhenhua, et al.
Veröffentlicht: (2024)
von: Yang, Zhenhua, et al.
Veröffentlicht: (2024)
C$^{3}$Bench: A Comprehensive Classical Chinese Understanding Benchmark for Large Language Models
von: Cao, Jiahuan, et al.
Veröffentlicht: (2024)
von: Cao, Jiahuan, et al.
Veröffentlicht: (2024)
PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research
von: Miao, Tingjia, et al.
Veröffentlicht: (2026)
von: Miao, Tingjia, et al.
Veröffentlicht: (2026)
LEGO: Self-Supervised Representation Learning for Scene Text Images
von: Ren, Yujin, et al.
Veröffentlicht: (2024)
von: Ren, Yujin, et al.
Veröffentlicht: (2024)
Privacy-Preserving Biometric Verification with Handwritten Random Digit String
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios
von: Shi, Yang, et al.
Veröffentlicht: (2025)
von: Shi, Yang, et al.
Veröffentlicht: (2025)
WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics
von: Liu, Chenxu, et al.
Veröffentlicht: (2026)
von: Liu, Chenxu, et al.
Veröffentlicht: (2026)
MemBench: Towards More Comprehensive Evaluation on the Memory of LLM-based Agents
von: Tan, Haoran, et al.
Veröffentlicht: (2025)
von: Tan, Haoran, et al.
Veröffentlicht: (2025)
CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy
von: Yang, Zhibo, et al.
Veröffentlicht: (2024)
von: Yang, Zhibo, et al.
Veröffentlicht: (2024)
AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
TCC-Bench: Benchmarking the Traditional Chinese Culture Understanding Capabilities of MLLMs
von: Xu, Pengju, et al.
Veröffentlicht: (2025)
von: Xu, Pengju, et al.
Veröffentlicht: (2025)
TongGu: Mastering Classical Chinese Understanding with Knowledge-Grounded Large Language Models
von: Cao, Jiahuan, et al.
Veröffentlicht: (2024)
von: Cao, Jiahuan, et al.
Veröffentlicht: (2024)
AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation
von: Wang, Lu, et al.
Veröffentlicht: (2025)
von: Wang, Lu, et al.
Veröffentlicht: (2025)
DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2024)
AirQualityBench: A Realistic Evaluation Benchmark for Global Air Quality Forecasting
von: Xu, Xing, et al.
Veröffentlicht: (2026)
von: Xu, Xing, et al.
Veröffentlicht: (2026)
UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
PodBench: A Comprehensive Benchmark for Instruction-Aware Audio-Oriented Podcast Script Generation
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
CausalBench: A Comprehensive Benchmark for Causal Learning Capability of LLMs
von: Zhou, Yu, et al.
Veröffentlicht: (2024)
von: Zhou, Yu, et al.
Veröffentlicht: (2024)
When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation
von: Sun, Lin, et al.
Veröffentlicht: (2026)
von: Sun, Lin, et al.
Veröffentlicht: (2026)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
WritingBench: A Comprehensive Benchmark for Generative Writing
von: Wu, Yuning, et al.
Veröffentlicht: (2025)
von: Wu, Yuning, et al.
Veröffentlicht: (2025)
OCR-Agent: Agentic OCR with Capability and Memory Reflection
von: Wen, Shimin, et al.
Veröffentlicht: (2026)
von: Wen, Shimin, et al.
Veröffentlicht: (2026)
HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models
von: Que, Haoran, et al.
Veröffentlicht: (2024)
von: Que, Haoran, et al.
Veröffentlicht: (2024)
Online Signature Verification based on the Lagrange formulation with 2D and 3D robotic models
von: Diaz, Moises, et al.
Veröffentlicht: (2025)
von: Diaz, Moises, et al.
Veröffentlicht: (2025)
OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models
von: Liu, Yuliang, et al.
Veröffentlicht: (2023)
von: Liu, Yuliang, et al.
Veröffentlicht: (2023)
KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
AI Idea Bench 2025: AI Research Idea Generation Benchmark
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PosterVerse: A Full-Workflow Framework for Commercial-Grade Poster Generation with HTML-Based Scalable Typography
von: Liu, Junle, et al.
Veröffentlicht: (2026) -
MegaHan97K: A Large-Scale Dataset for Mega-Category Chinese Character Recognition with over 97K Categories
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025) -
Reviving Cultural Heritage: A Novel Approach for Comprehensive Historical Document Restoration
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025) -
Online Writer Retrieval with Chinese Handwritten Phrases: A Synergistic Temporal-Frequency Representation Learning Approach
von: Zhang, Peirong, et al.
Veröffentlicht: (2024) -
Smaller But Better: Unifying Layout Generation with Smaller Large Language Models
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)