LED: A Benchmark for Evaluating Layout Error Detection in Document Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Heo, Inbum, Hwang, Taewook, Jung, Jeesu, Jung, Sangkeun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LED Benchmark: Diagnosing Structural Layout Errors for Document Layout Analysis
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
von: Heo, Inbum, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models on the 2026 Korean CSAT Mathematics Exam: Measuring Mathematical Ability in a Zero-Data-Leakage Setting
von: Pyeon, Goun, et al.
Veröffentlicht: (2025)
von: Pyeon, Goun, et al.
Veröffentlicht: (2025)
ZEBRA: Leveraging Model-Behavioral Knowledge for Zero-Annotation Preference Dataset Construction
von: Jung, Jeesu, et al.
Veröffentlicht: (2025)
von: Jung, Jeesu, et al.
Veröffentlicht: (2025)
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
Benchmarking Graph Neural Networks for Document Layout Analysis in Public Affairs
von: Lopez-Duran, Miguel, et al.
Veröffentlicht: (2025)
von: Lopez-Duran, Miguel, et al.
Veröffentlicht: (2025)
LayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding
von: Luo, Chuwei, et al.
Veröffentlicht: (2024)
von: Luo, Chuwei, et al.
Veröffentlicht: (2024)
What Drives Paper Acceptance? A Process-Centric Analysis of Modern Peer Review
von: Jung, Sangkeun, et al.
Veröffentlicht: (2025)
von: Jung, Sangkeun, et al.
Veröffentlicht: (2025)
KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
Fooling the LVLM Judges: Visual Biases in LVLM-Based Evaluation
von: Hwang, Yerin, et al.
Veröffentlicht: (2025)
von: Hwang, Yerin, et al.
Veröffentlicht: (2025)
Exploring Domain Robust Lightweight Reward Models based on Router Mechanism
von: Namgoong, Hyuk, et al.
Veröffentlicht: (2024)
von: Namgoong, Hyuk, et al.
Veröffentlicht: (2024)
Visually Guided Generative Text-Layout Pre-training for Document Intelligence
von: Mao, Zhiming, et al.
Veröffentlicht: (2024)
von: Mao, Zhiming, et al.
Veröffentlicht: (2024)
LAPDoc: Layout-Aware Prompting for Documents
von: Lamott, Marcel, et al.
Veröffentlicht: (2024)
von: Lamott, Marcel, et al.
Veröffentlicht: (2024)
A Spectrum Evaluation Benchmark for Medical Multi-Modal Large Language Models
von: Liu, Jie, et al.
Veröffentlicht: (2024)
von: Liu, Jie, et al.
Veröffentlicht: (2024)
From Codicology to Code: A Comparative Study of Transformer and YOLO-based Detectors for Layout Analysis in Historical Documents
von: Aguilar, Sergio Torres
Veröffentlicht: (2025)
von: Aguilar, Sergio Torres
Veröffentlicht: (2025)
Confidence-Aware Document OCR Error Detection
von: Hemmer, Arthur, et al.
Veröffentlicht: (2024)
von: Hemmer, Arthur, et al.
Veröffentlicht: (2024)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
KH-FUNSD: A Hierarchical and Fine-Grained Layout Analysis Dataset for Low-Resource Khmer Business Document
von: Thuon, Nimol, et al.
Veröffentlicht: (2025)
von: Thuon, Nimol, et al.
Veröffentlicht: (2025)
Co-Layout: LLM-driven Co-optimization for Interior Layout
von: Xiang, Chucheng, et al.
Veröffentlicht: (2025)
von: Xiang, Chucheng, et al.
Veröffentlicht: (2025)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
von: Zhu, Wanrong, et al.
Veröffentlicht: (2024)
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
von: Lee, Segyu, et al.
Veröffentlicht: (2026)
von: Lee, Segyu, et al.
Veröffentlicht: (2026)
Show, don't tell -- Providing Visual Error Feedback for Handwritten Documents
von: Yasin, Said, et al.
Veröffentlicht: (2026)
von: Yasin, Said, et al.
Veröffentlicht: (2026)
READoc: A Unified Benchmark for Realistic Document Structured Extraction
von: Li, Zichao, et al.
Veröffentlicht: (2024)
von: Li, Zichao, et al.
Veröffentlicht: (2024)
Recovering Policy-Induced Errors: Benchmarking and Trajectory Synthesis for Robust GUI Agents
von: Bu, Tianpeng, et al.
Veröffentlicht: (2026)
von: Bu, Tianpeng, et al.
Veröffentlicht: (2026)
Factorized Diffusion: Perceptual Illusions by Noise Decomposition
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models
von: Geng, Daniel, et al.
Veröffentlicht: (2023)
von: Geng, Daniel, et al.
Veröffentlicht: (2023)
Human-Aligned MLLM Judges for Fine-Grained Image Editing Evaluation: A Benchmark, Framework, and Analysis
von: Liu, Runzhou, et al.
Veröffentlicht: (2026)
von: Liu, Runzhou, et al.
Veröffentlicht: (2026)
LLMs Behind the Scenes: Enabling Narrative Scene Illustration
von: Roemmele, Melissa, et al.
Veröffentlicht: (2025)
von: Roemmele, Melissa, et al.
Veröffentlicht: (2025)
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts
von: Hwang, Taebaek, et al.
Veröffentlicht: (2025)
von: Hwang, Taebaek, et al.
Veröffentlicht: (2025)
UNIDOC-BENCH: A Unified Benchmark for Document-Centric Multimodal RAG
von: Peng, Xiangyu, et al.
Veröffentlicht: (2025)
von: Peng, Xiangyu, et al.
Veröffentlicht: (2025)
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
Actions and Objects Pathways for Domain Adaptation in Video Question Answering
von: Mohamud, Safaa Abdullahi Moallim, et al.
Veröffentlicht: (2024)
von: Mohamud, Safaa Abdullahi Moallim, et al.
Veröffentlicht: (2024)
GDI-Bench: A Benchmark for General Document Intelligence with Vision and Reasoning Decoupling
von: Li, Siqi, et al.
Veröffentlicht: (2025)
von: Li, Siqi, et al.
Veröffentlicht: (2025)
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
A Hybrid Approach for Document Layout Analysis in Document images
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
Guidance-Based Prompt Data Augmentation in Specialized Domains for Named Entity Recognition
von: Kang, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Kang, Hyeonseok, et al.
Veröffentlicht: (2024)
$λ$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
Towards Khmer Scene Document Layout Detection
von: Kong, Marry, et al.
Veröffentlicht: (2026)
von: Kong, Marry, et al.
Veröffentlicht: (2026)
ROAP: A Reading-Order and Attention-Prior Pipeline for Optimizing Layout Transformers in Key Information Extraction
von: Xie, Tingwei, et al.
Veröffentlicht: (2026)
von: Xie, Tingwei, et al.
Veröffentlicht: (2026)
DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding
von: Feng, Xiang, et al.
Veröffentlicht: (2026)
von: Feng, Xiang, et al.
Veröffentlicht: (2026)
MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LED Benchmark: Diagnosing Structural Layout Errors for Document Layout Analysis
von: Heo, Inbum, et al.
Veröffentlicht: (2025) -
Evaluating Large Language Models on the 2026 Korean CSAT Mathematics Exam: Measuring Mathematical Ability in a Zero-Data-Leakage Setting
von: Pyeon, Goun, et al.
Veröffentlicht: (2025) -
ZEBRA: Leveraging Model-Behavioral Knowledge for Zero-Annotation Preference Dataset Construction
von: Jung, Jeesu, et al.
Veröffentlicht: (2025) -
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
von: Park, ChaeHun, et al.
Veröffentlicht: (2024) -
Benchmarking Graph Neural Networks for Document Layout Analysis in Public Affairs
von: Lopez-Duran, Miguel, et al.
Veröffentlicht: (2025)