Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zilong, Shen, Xiaoyu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets
by: Shen, Jiyuan, et al.
Published: (2026)
by: Shen, Jiyuan, et al.
Published: (2026)
LMDX: Language Model-based Document Information Extraction and Localization
by: Perot, Vincent, et al.
Published: (2023)
by: Perot, Vincent, et al.
Published: (2023)
Team Ryu's Submission to SIGMORPHON 2024 Shared Task on Subword Tokenization
by: Li, Zilong
Published: (2024)
by: Li, Zilong
Published: (2024)
From Calculation to Adjudication: Examining LLM judges on Mathematical Reasoning Tasks
by: Stephan, Andreas, et al.
Published: (2024)
by: Stephan, Andreas, et al.
Published: (2024)
MAFA: A Multi-Agent Framework for Enterprise-Scale Annotation with Configurable Task Adaptation
by: Hegazy, Mahmood, et al.
Published: (2025)
by: Hegazy, Mahmood, et al.
Published: (2025)
Routine: A Structural Planning Framework for LLM Agent System in Enterprise
by: Zeng, Guancheng, et al.
Published: (2025)
by: Zeng, Guancheng, et al.
Published: (2025)
Case-Aware LLM-as-a-Judge Evaluation for Enterprise-Scale RAG Systems
by: Chhabra, Mukul, et al.
Published: (2026)
by: Chhabra, Mukul, et al.
Published: (2026)
On the use of Silver Standard Data for Zero-shot Classification Tasks in Information Extraction
by: Wang, Jianwei, et al.
Published: (2024)
by: Wang, Jianwei, et al.
Published: (2024)
SAIL: Sample-Centric In-Context Learning for Document Information Extraction
by: Zhang, Jinyu, et al.
Published: (2024)
by: Zhang, Jinyu, et al.
Published: (2024)
Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing
by: Liu, Ziyang
Published: (2026)
by: Liu, Ziyang
Published: (2026)
STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking
by: Chhetri, Tek Raj, et al.
Published: (2025)
by: Chhetri, Tek Raj, et al.
Published: (2025)
A Pluggable Multi-Task Learning Framework for Sentiment-Aware Financial Relation Extraction
by: Luo, Jinming, et al.
Published: (2025)
by: Luo, Jinming, et al.
Published: (2025)
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
by: Greif, Gavin, et al.
Published: (2025)
by: Greif, Gavin, et al.
Published: (2025)
Investigating OCR-Sensitive Neurons to Improve Entity Recognition in Historical Documents
by: Boros, Emanuela, et al.
Published: (2024)
by: Boros, Emanuela, et al.
Published: (2024)
Position: Avoid Overstretching LLMs for every Enterprise Task
by: Singh, Kuldeep, et al.
Published: (2026)
by: Singh, Kuldeep, et al.
Published: (2026)
KunlunBaize: LLM with Multi-Scale Convolution and Multi-Token Prediction Under TransformerX Framework
by: Li, Cheng, et al.
Published: (2025)
by: Li, Cheng, et al.
Published: (2025)
Key Coverage Matters: Semi-Structured Extraction of OCR Clinical Reports
by: Wang, Yu, et al.
Published: (2026)
by: Wang, Yu, et al.
Published: (2026)
Extract Information from Hybrid Long Documents Leveraging LLMs: A Framework and Dataset
by: Yue, Chongjian, et al.
Published: (2024)
by: Yue, Chongjian, et al.
Published: (2024)
Relation as a Prior: A Novel Paradigm for LLM-based Document-level Relation Extraction
by: Pi, Qiankun, et al.
Published: (2025)
by: Pi, Qiankun, et al.
Published: (2025)
Lightweight Spatial Modeling for Combinatorial Information Extraction From Documents
by: Dong, Yanfei, et al.
Published: (2024)
by: Dong, Yanfei, et al.
Published: (2024)
Confidence-Aware Document OCR Error Detection
by: Hemmer, Arthur, et al.
Published: (2024)
by: Hemmer, Arthur, et al.
Published: (2024)
BIM Information Extraction Through LLM-based Adaptive Exploration
by: Hellin, Sylvain, et al.
Published: (2026)
by: Hellin, Sylvain, et al.
Published: (2026)
A Positive-Unlabeled Metric Learning Framework for Document-Level Relation Extraction with Incomplete Labeling
by: Wang, Ye, et al.
Published: (2023)
by: Wang, Ye, et al.
Published: (2023)
Extract-0: A Specialized Language Model for Document Information Extraction
by: Godoy, Henrique
Published: (2025)
by: Godoy, Henrique
Published: (2025)
YAYI-UIE: A Chat-Enhanced Instruction Tuning Framework for Universal Information Extraction
by: Xiao, Xinglin, et al.
Published: (2023)
by: Xiao, Xinglin, et al.
Published: (2023)
SlovKE: A Large-Scale Dataset and LLM Evaluation for Slovak Keyphrase Extraction
by: Števaňák, David, et al.
Published: (2026)
by: Števaňák, David, et al.
Published: (2026)
Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction
by: Rashad, Mohamed
Published: (2024)
by: Rashad, Mohamed
Published: (2024)
Evaluating LLMs for Historical Document OCR: A Methodological Framework for Digital Humanities
by: Levchenko, Maria
Published: (2025)
by: Levchenko, Maria
Published: (2025)
Unifying Ontology Construction and Semantic Alignment for Deterministic Enterprise Reasoning at Scale
by: Zhu, Hongyin
Published: (2026)
by: Zhu, Hongyin
Published: (2026)
GLiNER2: An Efficient Multi-Task Information Extraction System with Schema-Driven Interface
by: Zaratiana, Urchade, et al.
Published: (2025)
by: Zaratiana, Urchade, et al.
Published: (2025)
Protect: Towards Robust Guardrailing Stack for Trustworthy Enterprise LLM Systems
by: Avinash, Karthik, et al.
Published: (2025)
by: Avinash, Karthik, et al.
Published: (2025)
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs
by: Colakoglu, Gaye, et al.
Published: (2025)
by: Colakoglu, Gaye, et al.
Published: (2025)
Exploring LLMs for Scientific Information Extraction Using The SciEx Framework
by: Li, Sha, et al.
Published: (2025)
by: Li, Sha, et al.
Published: (2025)
Enterprise Deep Research: Steerable Multi-Agent Deep Research for Enterprise Analytics
by: Prabhakar, Akshara, et al.
Published: (2025)
by: Prabhakar, Akshara, et al.
Published: (2025)
TTM-RE: Memory-Augmented Document-Level Relation Extraction
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
LLM-RadJudge: Achieving Radiologist-Level Evaluation for X-Ray Report Generation
by: Wang, Zilong, et al.
Published: (2024)
by: Wang, Zilong, et al.
Published: (2024)
Reading or Reasoning? Format Decoupled Reinforcement Learning for Document OCR
by: Zhong, Yufeng, et al.
Published: (2025)
by: Zhong, Yufeng, et al.
Published: (2025)
BURExtract-Llama: An LLM for Clinical Concept Extraction in Breast Ultrasound Reports
by: Chen, Yuxuan, et al.
Published: (2024)
by: Chen, Yuxuan, et al.
Published: (2024)
Benchmarking Complex Multimodal Document Processing Pipelines: A Unified Evaluation Framework for Enterprise AI
by: Singh, Saurabh K., et al.
Published: (2026)
by: Singh, Saurabh K., et al.
Published: (2026)
LLMs Learn Task Heuristics from Demonstrations: A Heuristic-Driven Prompting Strategy for Document-Level Event Argument Extraction
by: Zhou, Hanzhang, et al.
Published: (2023)
by: Zhou, Hanzhang, et al.
Published: (2023)
Similar Items
-
OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets
by: Shen, Jiyuan, et al.
Published: (2026) -
LMDX: Language Model-based Document Information Extraction and Localization
by: Perot, Vincent, et al.
Published: (2023) -
Team Ryu's Submission to SIGMORPHON 2024 Shared Task on Subword Tokenization
by: Li, Zilong
Published: (2024) -
From Calculation to Adjudication: Examining LLM judges on Mathematical Reasoning Tasks
by: Stephan, Andreas, et al.
Published: (2024) -
MAFA: A Multi-Agent Framework for Enterprise-Scale Annotation with Configurable Task Adaptation
by: Hegazy, Mahmood, et al.
Published: (2025)