LLMStructBench: Benchmarking Large Language Model Structured Data Extraction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tenckhoff, Sönke, Koddenbrock, Mario, Rodner, Erik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026)
von: Park, Sungho, et al.
Veröffentlicht: (2026)
Digital Twin sensors in cultural heritage applications
von: Niccolucci, Franco, et al.
Veröffentlicht: (2024)
von: Niccolucci, Franco, et al.
Veröffentlicht: (2024)
Active Data
von: Arthur, Richard, et al.
Veröffentlicht: (2026)
von: Arthur, Richard, et al.
Veröffentlicht: (2026)
Adapting PromptORE for Modern History: Information Extraction from Hispanic Monarchy Documents of the XVIth Century
von: Hidalgo, Hèctor Loopez, et al.
Veröffentlicht: (2024)
von: Hidalgo, Hèctor Loopez, et al.
Veröffentlicht: (2024)
Detection of Personal Data in Structured Datasets Using a Large Language Model
von: Ntwali, Albert Agisha, et al.
Veröffentlicht: (2025)
von: Ntwali, Albert Agisha, et al.
Veröffentlicht: (2025)
Bi-View Embedding Fusion: A Hybrid Learning Approach for Knowledge Graph's Nodes Classification Addressing Problems with Limited Data
von: Napoli, Rosario, et al.
Veröffentlicht: (2025)
von: Napoli, Rosario, et al.
Veröffentlicht: (2025)
Tensor Manifold-Based Graph-Vector Fusion for AI-Native Academic Literature Retrieval
von: Wei, Xing, et al.
Veröffentlicht: (2026)
von: Wei, Xing, et al.
Veröffentlicht: (2026)
Learned Compression of Nonlinear Time Series With Random Access
von: Guerra, Andrea, et al.
Veröffentlicht: (2024)
von: Guerra, Andrea, et al.
Veröffentlicht: (2024)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
von: Whittaker, Edward, et al.
Veröffentlicht: (2024)
von: Whittaker, Edward, et al.
Veröffentlicht: (2024)
LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
Big Help or Big Brother? Auditing Tracking, Profiling, and Personalization in Generative AI Assistants
von: Vekaria, Yash, et al.
Veröffentlicht: (2025)
von: Vekaria, Yash, et al.
Veröffentlicht: (2025)
ArcheType: A Novel Framework for Open-Source Column Type Annotation using Large Language Models
von: Feuer, Benjamin, et al.
Veröffentlicht: (2023)
von: Feuer, Benjamin, et al.
Veröffentlicht: (2023)
When Large Language Models are More PersuasiveThan Incentivized Humans, and Why
von: Schoenegger, Philipp, et al.
Veröffentlicht: (2025)
von: Schoenegger, Philipp, et al.
Veröffentlicht: (2025)
Assisted morbidity coding: the SISCO.web use case for identifying the main diagnosis in Hospital Discharge Records
von: Cardillo, Elena, et al.
Veröffentlicht: (2024)
von: Cardillo, Elena, et al.
Veröffentlicht: (2024)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
von: Budigi, Venkata Krishna Prasanth, et al.
Veröffentlicht: (2026)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
Domain Specific Data Distillation and Multi-modal Embedding Generation
von: Peddiraju, Sharadind, et al.
Veröffentlicht: (2024)
von: Peddiraju, Sharadind, et al.
Veröffentlicht: (2024)
3D Data Long-Term Preservation in Cultural Heritage
von: Amico, Nicola, et al.
Veröffentlicht: (2024)
von: Amico, Nicola, et al.
Veröffentlicht: (2024)
compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data
von: Termignon, Lucie, et al.
Veröffentlicht: (2026)
von: Termignon, Lucie, et al.
Veröffentlicht: (2026)
The Cultural Gene of Large Language Models: A Study on the Impact of Cross-Corpus Training on Model Values and Biases
von: Fenech-Borg, Emanuel Z., et al.
Veröffentlicht: (2025)
von: Fenech-Borg, Emanuel Z., et al.
Veröffentlicht: (2025)
PAKT: Perspectivized Argumentation Knowledge Graph and Tool for Deliberation Analysis (with Supplementary Materials)
von: Plenz, Moritz, et al.
Veröffentlicht: (2024)
von: Plenz, Moritz, et al.
Veröffentlicht: (2024)
Data accounting and error counting
von: Gajda, Michał J.
Veröffentlicht: (2023)
von: Gajda, Michał J.
Veröffentlicht: (2023)
AuthorityBench: Benchmarking LLM Authority Perception for Reliable Retrieval-Augmented Generation
von: Yao, Zhihui, et al.
Veröffentlicht: (2026)
von: Yao, Zhihui, et al.
Veröffentlicht: (2026)
Partial Adaptive Indexing for Approximate Query Answering
von: Maroulis, Stavros, et al.
Veröffentlicht: (2024)
von: Maroulis, Stavros, et al.
Veröffentlicht: (2024)
Siren Federate: Bridging document, relational, and graph models for exploratory graph analysis
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
von: Bordea, Georgeta, et al.
Veröffentlicht: (2025)
A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
Utilizing Large Language Models to Synthesize Product Desirability Datasets
von: Hastings, John D., et al.
Veröffentlicht: (2024)
von: Hastings, John D., et al.
Veröffentlicht: (2024)
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
Representation Fidelity:Auditing Algorithmic Decisions About Humans Using Self-Descriptions
von: Elstner, Theresa, et al.
Veröffentlicht: (2026)
von: Elstner, Theresa, et al.
Veröffentlicht: (2026)
Challenging the Validity of Personality Tests for Large Language Models
von: Sühr, Tom, et al.
Veröffentlicht: (2023)
von: Sühr, Tom, et al.
Veröffentlicht: (2023)
GraphAr: An Efficient Storage Scheme for Graph Data in Data Lakes
von: Li, Xue, et al.
Veröffentlicht: (2023)
von: Li, Xue, et al.
Veröffentlicht: (2023)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
ALIGNS: Unlocking nomological networks in psychological measurement through a large language model
von: Larsen, Kai R., et al.
Veröffentlicht: (2025)
von: Larsen, Kai R., et al.
Veröffentlicht: (2025)
OnPair: Short Strings Compression for Fast Random Access
von: Gargiulo, Francesco, et al.
Veröffentlicht: (2025)
von: Gargiulo, Francesco, et al.
Veröffentlicht: (2025)
From Native Memes to Global Moderation: Cross-Cultural Evaluation of Vision-Language Models for Hateful Meme Detection
von: Wang, Mo, et al.
Veröffentlicht: (2026)
von: Wang, Mo, et al.
Veröffentlicht: (2026)
LiveVectorLake: A Real-Time Versioned Knowledge Base Architecture for Streaming Vector Updates and Temporal Retrieval
von: Prajapati, Tarun
Veröffentlicht: (2025)
von: Prajapati, Tarun
Veröffentlicht: (2025)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
From Code to Compliance: Assessing ChatGPT's Utility in Designing an Accessible Webpage -- A Case Study
von: Ahmed, Ammar, et al.
Veröffentlicht: (2025)
von: Ahmed, Ammar, et al.
Veröffentlicht: (2025)
Gender and Race Bias in Consumer Product Recommendations by Large Language Models
von: Xu, Ke, et al.
Veröffentlicht: (2026)
von: Xu, Ke, et al.
Veröffentlicht: (2026)
An Analysis of XML Compression Efficiency
von: Augeri, Christopher James, et al.
Veröffentlicht: (2024)
von: Augeri, Christopher James, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026) -
Digital Twin sensors in cultural heritage applications
von: Niccolucci, Franco, et al.
Veröffentlicht: (2024) -
Active Data
von: Arthur, Richard, et al.
Veröffentlicht: (2026) -
Adapting PromptORE for Modern History: Information Extraction from Hispanic Monarchy Documents of the XVIth Century
von: Hidalgo, Hèctor Loopez, et al.
Veröffentlicht: (2024) -
Detection of Personal Data in Structured Datasets Using a Large Language Model
von: Ntwali, Albert Agisha, et al.
Veröffentlicht: (2025)