KIEval: Evaluation Metric for Document Key Information Extraction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khang, Minsoo, Jung, Sang Chul, Park, Sungrae, Hong, Teakgyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CReSt: A Comprehensive Benchmark for Retrieval-Augmented Generation with Complex Reasoning over Structured Documents
von: Khang, Minsoo, et al.
Veröffentlicht: (2025)
von: Khang, Minsoo, et al.
Veröffentlicht: (2025)
TFLOP: Table Structure Recognition Framework with Layout Pointer Mechanism
von: Khang, Minsoo, et al.
Veröffentlicht: (2025)
von: Khang, Minsoo, et al.
Veröffentlicht: (2025)
ZERA: Zero-init Instruction Evolving Refinement Agent -- From Zero Instructions to Structured Prompts via Principle-based Optimization
von: Yi, Seungyoun, et al.
Veröffentlicht: (2025)
von: Yi, Seungyoun, et al.
Veröffentlicht: (2025)
System Message Generation for User Preferences using Open-Source Models
von: Jeong, Minbyul, et al.
Veröffentlicht: (2025)
von: Jeong, Minbyul, et al.
Veröffentlicht: (2025)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
KIEval: A Knowledge-grounded Interactive Evaluation Framework for Large Language Models
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
von: Yu, Zhuohao, et al.
Veröffentlicht: (2024)
User-Oriented Multi-Turn Dialogue Generation with Tool Use at scale
von: Cho, Jungho, et al.
Veröffentlicht: (2026)
von: Cho, Jungho, et al.
Veröffentlicht: (2026)
PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses
von: Hong, Minki, et al.
Veröffentlicht: (2026)
von: Hong, Minki, et al.
Veröffentlicht: (2026)
UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
von: Ji, Yifan, et al.
Veröffentlicht: (2026)
von: Ji, Yifan, et al.
Veröffentlicht: (2026)
Bounded Hyperbolic Tangent: A Stable and Efficient Alternative to Pre-Layer Normalization in Large Language Models
von: Byun, Hoyoon, et al.
Veröffentlicht: (2025)
von: Byun, Hoyoon, et al.
Veröffentlicht: (2025)
Examining the Metrics for Document-Level Claim Extraction in Czech and Slovak
von: Makaiova, Lucia, et al.
Veröffentlicht: (2025)
von: Makaiova, Lucia, et al.
Veröffentlicht: (2025)
Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs
von: Tan, Nelvin, et al.
Veröffentlicht: (2026)
von: Tan, Nelvin, et al.
Veröffentlicht: (2026)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
Neurosymbolic Information Extraction from Transactional Documents
von: Hemmer, Arthur, et al.
Veröffentlicht: (2025)
von: Hemmer, Arthur, et al.
Veröffentlicht: (2025)
SLIDE: Sliding Localized Information for Document Extraction
von: Singh, Divyansh, et al.
Veröffentlicht: (2025)
von: Singh, Divyansh, et al.
Veröffentlicht: (2025)
Quantitative Information Extraction from Humanitarian Documents
von: Liberatore, Daniele, et al.
Veröffentlicht: (2024)
von: Liberatore, Daniele, et al.
Veröffentlicht: (2024)
LongKey: Keyphrase Extraction for Long Documents
von: Alves, Jeovane Honorio, et al.
Veröffentlicht: (2024)
von: Alves, Jeovane Honorio, et al.
Veröffentlicht: (2024)
Improving Conversational Abilities of Quantized Large Language Models via Direct Preference Alignment
von: Lee, Janghwan, et al.
Veröffentlicht: (2024)
von: Lee, Janghwan, et al.
Veröffentlicht: (2024)
Spatial ModernBERT: Spatial-Aware Transformer for Table and Key-Value Extraction in Financial Documents at Scale
von: Javis AI Team, et al.
Veröffentlicht: (2025)
von: Javis AI Team, et al.
Veröffentlicht: (2025)
Evaluating the Utility of Grounding Documents with Reference-Free LLM-based Metrics
von: Hua, Yilun, et al.
Veröffentlicht: (2026)
von: Hua, Yilun, et al.
Veröffentlicht: (2026)
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
Up to 36x Speedup: Mask-based Parallel Inference Paradigm for Key Information Extraction in MLLMs
von: Wang, Xinzhong, et al.
Veröffentlicht: (2026)
von: Wang, Xinzhong, et al.
Veröffentlicht: (2026)
SAGE:Specification-Aware Grammar Extraction for Automated Test Case Generation with LLMs
von: Aditi, et al.
Veröffentlicht: (2025)
von: Aditi, et al.
Veröffentlicht: (2025)
Dual-Scale World Models for LLM Agents Towards Hard-Exploration Problems
von: Kim, Minsoo, et al.
Veröffentlicht: (2025)
von: Kim, Minsoo, et al.
Veröffentlicht: (2025)
ViBERTgrid BiLSTM-CRF: Multimodal Key Information Extraction from Unstructured Financial Documents
von: Pala, Furkan, et al.
Veröffentlicht: (2024)
von: Pala, Furkan, et al.
Veröffentlicht: (2024)
A Positive-Unlabeled Metric Learning Framework for Document-Level Relation Extraction with Incomplete Labeling
von: Wang, Ye, et al.
Veröffentlicht: (2023)
von: Wang, Ye, et al.
Veröffentlicht: (2023)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
Document-Level Zero-Shot Relation Extraction with Entity Side Information
von: Chanthran, Mohan Raj, et al.
Veröffentlicht: (2026)
von: Chanthran, Mohan Raj, et al.
Veröffentlicht: (2026)
"What is the value of {templates}?" Rethinking Document Information Extraction Datasets for LLMs
von: Zmigrod, Ran, et al.
Veröffentlicht: (2024)
von: Zmigrod, Ran, et al.
Veröffentlicht: (2024)
NovAScore: A New Automated Metric for Evaluating Document Level Novelty
von: Ai, Lin, et al.
Veröffentlicht: (2024)
von: Ai, Lin, et al.
Veröffentlicht: (2024)
Enhancing Argument Summarization: Prioritizing Exhaustiveness in Key Point Generation and Introducing an Automatic Coverage Evaluation Metric
von: Khosravani, Mohammad, et al.
Veröffentlicht: (2024)
von: Khosravani, Mohammad, et al.
Veröffentlicht: (2024)
Information Extraction From Fiscal Documents Using LLMs
von: Aggarwal, Vikram, et al.
Veröffentlicht: (2025)
von: Aggarwal, Vikram, et al.
Veröffentlicht: (2025)
PACE-RAG: Patient-Aware Contextual and Evidence-based Policy RAG for Clinical Drug Recommendation
von: Huh, Chaeyoung, et al.
Veröffentlicht: (2026)
von: Huh, Chaeyoung, et al.
Veröffentlicht: (2026)
Unveiling Key Aspects of Fine-Tuning in Sentence Embeddings: A Representation Rank Analysis
von: Jung, Euna, et al.
Veröffentlicht: (2024)
von: Jung, Euna, et al.
Veröffentlicht: (2024)
AgenticIE: An Adaptive Agent for Information Extraction from Complex Regulatory Documents
von: Colakoglu, Gaye, et al.
Veröffentlicht: (2025)
von: Colakoglu, Gaye, et al.
Veröffentlicht: (2025)
Claim Extraction for Fact-Checking: Data, Models, and Automated Metrics
von: Ullrich, Herbert, et al.
Veröffentlicht: (2025)
von: Ullrich, Herbert, et al.
Veröffentlicht: (2025)
BuDDIE: A Business Document Dataset for Multi-task Information Extraction
von: Zmigrod, Ran, et al.
Veröffentlicht: (2024)
von: Zmigrod, Ran, et al.
Veröffentlicht: (2024)
Do not be greedy, Think Twice: Sampling and Selection for Document-level Information Extraction
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2026)
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2026)
Assisted Data Annotation for Business Process Information Extraction from Textual Documents
von: Neuberger, Julian, et al.
Veröffentlicht: (2024)
von: Neuberger, Julian, et al.
Veröffentlicht: (2024)
TeXBLEU: Automatic Metric for Evaluate LaTeX Format
von: Jung, Kyudan, et al.
Veröffentlicht: (2024)
von: Jung, Kyudan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CReSt: A Comprehensive Benchmark for Retrieval-Augmented Generation with Complex Reasoning over Structured Documents
von: Khang, Minsoo, et al.
Veröffentlicht: (2025) -
TFLOP: Table Structure Recognition Framework with Layout Pointer Mechanism
von: Khang, Minsoo, et al.
Veröffentlicht: (2025) -
ZERA: Zero-init Instruction Evolving Refinement Agent -- From Zero Instructions to Structured Prompts via Principle-based Optimization
von: Yi, Seungyoun, et al.
Veröffentlicht: (2025) -
System Message Generation for User Preferences using Open-Source Models
von: Jeong, Minbyul, et al.
Veröffentlicht: (2025) -
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)