OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets
Fuente:
arXiv
Salvato in:
| Autori principali: | Shen, Jiyuan, Yuan, Peiyue, Ghosh, Atin, Mai, Yifan, Dahlmeier, Daniel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Talk, Evaluate, Diagnose: User-aware Agent Evaluation with Automated Error Analysis
di: Chong, Penny, et al.
Pubblicazione: (2026)
di: Chong, Penny, et al.
Pubblicazione: (2026)
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026)
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026)
Document Intelligence in the Era of Large Language Models: A Survey
di: Wang, Weishi, et al.
Pubblicazione: (2025)
di: Wang, Weishi, et al.
Pubblicazione: (2025)
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
di: Wang, Zilong, et al.
Pubblicazione: (2025)
di: Wang, Zilong, et al.
Pubblicazione: (2025)
Large Language Models for Judicial Entity Extraction: A Comparative Study
di: Hussain, Atin Sakkeer, et al.
Pubblicazione: (2024)
di: Hussain, Atin Sakkeer, et al.
Pubblicazione: (2024)
GenRES: Rethinking Evaluation for Generative Relation Extraction in the Era of Large Language Models
di: Jiang, Pengcheng, et al.
Pubblicazione: (2024)
di: Jiang, Pengcheng, et al.
Pubblicazione: (2024)
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
di: Zheng, Lianmin, et al.
Pubblicazione: (2023)
di: Zheng, Lianmin, et al.
Pubblicazione: (2023)
From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities
di: Jiang, Shixin, et al.
Pubblicazione: (2024)
di: Jiang, Shixin, et al.
Pubblicazione: (2024)
Rethinking Interpretability in the Era of Large Language Models
di: Singh, Chandan, et al.
Pubblicazione: (2024)
di: Singh, Chandan, et al.
Pubblicazione: (2024)
SlovKE: A Large-Scale Dataset and LLM Evaluation for Slovak Keyphrase Extraction
di: Števaňák, David, et al.
Pubblicazione: (2026)
di: Števaňák, David, et al.
Pubblicazione: (2026)
Up to 36x Speedup: Mask-based Parallel Inference Paradigm for Key Information Extraction in MLLMs
di: Wang, Xinzhong, et al.
Pubblicazione: (2026)
di: Wang, Xinzhong, et al.
Pubblicazione: (2026)
Beyond MedQA: Towards Real-world Clinical Decision Making in the Era of LLMs
di: Xiao, Yunpeng, et al.
Pubblicazione: (2025)
di: Xiao, Yunpeng, et al.
Pubblicazione: (2025)
Preserving Knowledge Invariance: Rethinking Robustness Evaluation of Open Information Extraction
di: Qi, Ji, et al.
Pubblicazione: (2023)
di: Qi, Ji, et al.
Pubblicazione: (2023)
Span-Oriented Information Extraction -- A Unifying Perspective on Information Extraction
di: Ding, Yifan, et al.
Pubblicazione: (2024)
di: Ding, Yifan, et al.
Pubblicazione: (2024)
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
di: Greif, Gavin, et al.
Pubblicazione: (2025)
di: Greif, Gavin, et al.
Pubblicazione: (2025)
MessIRve: A Large-Scale Spanish Information Retrieval Dataset
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
di: Valentini, Francisco, et al.
Pubblicazione: (2024)
From Symbolic to Natural-Language Relations: Rethinking Knowledge Graph Construction in the Era of Large Language Models
di: Han, Kanyao, et al.
Pubblicazione: (2026)
di: Han, Kanyao, et al.
Pubblicazione: (2026)
Investigating OCR-Sensitive Neurons to Improve Entity Recognition in Historical Documents
di: Boros, Emanuela, et al.
Pubblicazione: (2024)
di: Boros, Emanuela, et al.
Pubblicazione: (2024)
RETQA: A Large-Scale Open-Domain Tabular Question Answering Dataset for Real Estate Sector
di: Wang, Zhensheng, et al.
Pubblicazione: (2024)
di: Wang, Zhensheng, et al.
Pubblicazione: (2024)
The Era of Real-World Human Interaction: RL from User Conversations
di: Jin, Chuanyang, et al.
Pubblicazione: (2025)
di: Jin, Chuanyang, et al.
Pubblicazione: (2025)
Scaling Open-Weight Large Language Models for Hydropower Regulatory Information Extraction: A Systematic Analysis
di: Yoon, Hong-Jun, et al.
Pubblicazione: (2025)
di: Yoon, Hong-Jun, et al.
Pubblicazione: (2025)
IEPile: Unearthing Large-Scale Schema-Based Information Extraction Corpus
di: Gui, Honghao, et al.
Pubblicazione: (2024)
di: Gui, Honghao, et al.
Pubblicazione: (2024)
SAIL: Sample-Centric In-Context Learning for Document Information Extraction
di: Zhang, Jinyu, et al.
Pubblicazione: (2024)
di: Zhang, Jinyu, et al.
Pubblicazione: (2024)
Lightweight Spatial Modeling for Combinatorial Information Extraction From Documents
di: Dong, Yanfei, et al.
Pubblicazione: (2024)
di: Dong, Yanfei, et al.
Pubblicazione: (2024)
Refract ICL: Rethinking Example Selection in the Era of Million-Token Models
di: Akula, Arjun R., et al.
Pubblicazione: (2025)
di: Akula, Arjun R., et al.
Pubblicazione: (2025)
AutoRE: Document-Level Relation Extraction with Large Language Models
di: Xue, Lilong, et al.
Pubblicazione: (2024)
di: Xue, Lilong, et al.
Pubblicazione: (2024)
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
di: Dong, Guanting, et al.
Pubblicazione: (2026)
di: Dong, Guanting, et al.
Pubblicazione: (2026)
Rethinking Data Selection at Scale: Random Selection is Almost All You Need
di: Xia, Tingyu, et al.
Pubblicazione: (2024)
di: Xia, Tingyu, et al.
Pubblicazione: (2024)
Key Coverage Matters: Semi-Structured Extraction of OCR Clinical Reports
di: Wang, Yu, et al.
Pubblicazione: (2026)
di: Wang, Yu, et al.
Pubblicazione: (2026)
Extract-0: A Specialized Language Model for Document Information Extraction
di: Godoy, Henrique
Pubblicazione: (2025)
di: Godoy, Henrique
Pubblicazione: (2025)
Research on Information Extraction of LCSTS Dataset Based on an Improved BERTSum-LSTM Model
di: Chen, Yiming, et al.
Pubblicazione: (2024)
di: Chen, Yiming, et al.
Pubblicazione: (2024)
MCP-SafetyBench: A Benchmark for Safety Evaluation of Large Language Models with Real-World MCP Servers
di: Zong, Xuanjun, et al.
Pubblicazione: (2025)
di: Zong, Xuanjun, et al.
Pubblicazione: (2025)
Confidence-Aware Document OCR Error Detection
di: Hemmer, Arthur, et al.
Pubblicazione: (2024)
di: Hemmer, Arthur, et al.
Pubblicazione: (2024)
MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
di: Luo, Ziyang, et al.
Pubblicazione: (2025)
di: Luo, Ziyang, et al.
Pubblicazione: (2025)
Affordance Benchmark for MLLMs
di: Wang, Junying, et al.
Pubblicazione: (2025)
di: Wang, Junying, et al.
Pubblicazione: (2025)
DocEDA: Automated Extraction and Design of Analog Circuits from Documents with Large Language Model
di: Chen, Hong Cai, et al.
Pubblicazione: (2024)
di: Chen, Hong Cai, et al.
Pubblicazione: (2024)
Evaluating Large Language Models for Real-World Engineering Tasks
di: Heesch, Rene, et al.
Pubblicazione: (2025)
di: Heesch, Rene, et al.
Pubblicazione: (2025)
Structured Extraction of Real World Medical Knowledge using LLMs for Summarization and Search
di: Kim, Edward, et al.
Pubblicazione: (2024)
di: Kim, Edward, et al.
Pubblicazione: (2024)
REALM: A Dataset of Real-World LLM Use Cases
di: Cheng, Jingwen, et al.
Pubblicazione: (2025)
di: Cheng, Jingwen, et al.
Pubblicazione: (2025)
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs
di: Colakoglu, Gaye, et al.
Pubblicazione: (2025)
di: Colakoglu, Gaye, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Talk, Evaluate, Diagnose: User-aware Agent Evaluation with Automated Error Analysis
di: Chong, Penny, et al.
Pubblicazione: (2026) -
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026) -
Document Intelligence in the Era of Large Language Models: A Survey
di: Wang, Weishi, et al.
Pubblicazione: (2025) -
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
di: Wang, Zilong, et al.
Pubblicazione: (2025) -
Large Language Models for Judicial Entity Extraction: A Comparative Study
di: Hussain, Atin Sakkeer, et al.
Pubblicazione: (2024)