LLM-KG-Bench 3.0: A Compass for SemanticTechnology Capabilities in the Ocean of LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Meyer, Lars-Peter, Frey, Johannes, Heim, Desiree, Brei, Felix, Stadler, Claus, Junghanns, Kurt, Martin, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-assisted Knowledge Graph Engineering: Experiments with ChatGPT
by: Meyer, Lars-Peter, et al.
Published: (2023)
by: Meyer, Lars-Peter, et al.
Published: (2023)
Assessing SPARQL capabilities of Large Language Models
by: Meyer, Lars-Peter, et al.
Published: (2024)
by: Meyer, Lars-Peter, et al.
Published: (2024)
ARUQULA -- An LLM based Text2SPARQL Approach using ReAct and Knowledge Graph Exploration Utilities
by: Brei, Felix, et al.
Published: (2025)
by: Brei, Felix, et al.
Published: (2025)
Leveraging small language models for Text2SPARQL tasks to improve the resilience of AI assistance
by: Brei, Felix, et al.
Published: (2024)
by: Brei, Felix, et al.
Published: (2024)
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
by: Kassem, Aly M., et al.
Published: (2025)
by: Kassem, Aly M., et al.
Published: (2025)
The Grammar of FAIR: A Granular Architecture of Semantic Units for FAIR Semantics, Inspired by Biology and Linguistics
by: Vogt, Lars, et al.
Published: (2025)
by: Vogt, Lars, et al.
Published: (2025)
FAIR 2.0: Extending the FAIR Guiding Principles to Address Semantic Interoperability
by: Vogt, Lars, et al.
Published: (2024)
by: Vogt, Lars, et al.
Published: (2024)
LLM+KG@VLDB'24 Workshop Summary
by: Khan, Arijit, et al.
Published: (2024)
by: Khan, Arijit, et al.
Published: (2024)
How do Scaling Laws Apply to Knowledge Graph Engineering Tasks? The Impact of Model Size on Large Language Model Performance
by: Heim, Desiree, et al.
Published: (2025)
by: Heim, Desiree, et al.
Published: (2025)
ELT-Bench-Verified: Benchmark Quality Issues Underestimate AI Agent Capabilities
by: Zanoli, Christopher, et al.
Published: (2026)
by: Zanoli, Christopher, et al.
Published: (2026)
Rethinking OWL Expressivity: Semantic Units for FAIR and Cognitively Interoperable Knowledge Graphs Why OWLs don't have to understand everything they say
by: Vogt, Lars
Published: (2024)
by: Vogt, Lars
Published: (2024)
Evaluating Data Quality Tools: Measurement Capabilities and LLM Integration
by: Rehberger, Tobias, et al.
Published: (2026)
by: Rehberger, Tobias, et al.
Published: (2026)
RNA-KG v2.0: An RNA-centered Knowledge Graph with Properties
by: Cavalleri, Emanuele, et al.
Published: (2025)
by: Cavalleri, Emanuele, et al.
Published: (2025)
CompassDB: Pioneering High-Performance Key-Value Store with Perfect Hash
by: Jiang, Jin, et al.
Published: (2024)
by: Jiang, Jin, et al.
Published: (2024)
MetaboKG: An Analysis-centric Knowledge Graph Framework for Untargeted Metabolomics
by: Féraud, Matthieu, et al.
Published: (2026)
by: Féraud, Matthieu, et al.
Published: (2026)
SemBench: A Benchmark for Semantic Query Processing Engines
by: Lao, Jiale, et al.
Published: (2025)
by: Lao, Jiale, et al.
Published: (2025)
The Semantic Ladder: A Framework for Progressive Formalization of Natural Language Content for Knowledge Graphs and AI Systems
by: Vogt, Lars
Published: (2026)
by: Vogt, Lars
Published: (2026)
Compass: General Filtered Search across Vector and Structured Data
by: Ye, Chunxiao, et al.
Published: (2025)
by: Ye, Chunxiao, et al.
Published: (2025)
DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning
by: Ahmed, Ahmed G. A. H, et al.
Published: (2026)
by: Ahmed, Ahmed G. A. H, et al.
Published: (2026)
Designing and Comparing RPQ Semantics
by: Marsault, Victor, et al.
Published: (2026)
by: Marsault, Victor, et al.
Published: (2026)
Compass: SLO-aware Query Planner for Compound AI Serving at Scale
by: Liu, Banruo, et al.
Published: (2025)
by: Liu, Banruo, et al.
Published: (2025)
VectraFlow: Long-Horizon Semantic Processing over Data and Event Streams with LLMs
by: Chen, Shu, et al.
Published: (2026)
by: Chen, Shu, et al.
Published: (2026)
Think2SQL: Reinforce LLM Reasoning Capabilities for Text2SQL
by: Papicchio, Simone, et al.
Published: (2025)
by: Papicchio, Simone, et al.
Published: (2025)
Sema: A High-performance System for LLM-based Semantic Query Processing
by: Qi, Kangkang, et al.
Published: (2026)
by: Qi, Kangkang, et al.
Published: (2026)
ClinDet-Bench: Beyond Abstention, Evaluating Judgment Determinability of LLMs in Clinical Decision-Making
by: Watanabe, Yusuke, et al.
Published: (2026)
by: Watanabe, Yusuke, et al.
Published: (2026)
PM4Py.LLM: a Comprehensive Module for Implementing PM on LLMs
by: Berti, Alessandro
Published: (2024)
by: Berti, Alessandro
Published: (2024)
TabSQLify: Enhancing Reasoning Capabilities of LLMs Through Table Decomposition
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
Towards Evolution Capabilities in Data Pipelines
by: Kramer, Kevin M.
Published: (2023)
by: Kramer, Kevin M.
Published: (2023)
A Framework for FAIR and CLEAR Ecological Data and Knowledge: Semantic Units for Synthesis and Causal Modelling
by: Vogt, Lars, et al.
Published: (2025)
by: Vogt, Lars, et al.
Published: (2025)
Toward Real-World Table Agents: Capabilities, Workflows, and Design Principles for LLM-based Table Intelligence
by: Tian, Jiaming, et al.
Published: (2025)
by: Tian, Jiaming, et al.
Published: (2025)
LLM-assisted Labeling Function Generation for Semantic Type Detection
by: Li, Chenjie, et al.
Published: (2024)
by: Li, Chenjie, et al.
Published: (2024)
An LLM Agent-Based Complex Semantic Table Annotation Approach
by: Geng, Yilin, et al.
Published: (2025)
by: Geng, Yilin, et al.
Published: (2025)
Actionable Understanding: Action Units for Bridging the Knowledge-Action Gap in Post-FAIR Knowledge Infrastructures
by: Vogt, Lars
Published: (2026)
by: Vogt, Lars
Published: (2026)
The FAIREr Guiding Principles: Organizing data and metadata into semantically meaningful types of FAIR Digital Objects to increase their human explorability and cognitive interoperability
by: Vogt, Lars
Published: (2023)
by: Vogt, Lars
Published: (2023)
LST-Bench: Benchmarking Log-Structured Tables in the Cloud
by: Camacho-Rodríguez, Jesús, et al.
Published: (2023)
by: Camacho-Rodríguez, Jesús, et al.
Published: (2023)
Is Your Learned Query Optimizer Behaving As You Expect? A Machine Learning Perspective
by: Lehmann, Claude, et al.
Published: (2023)
by: Lehmann, Claude, et al.
Published: (2023)
LLM-SQL-Solver: Can LLMs Determine SQL Equivalence?
by: Zhao, Fuheng, et al.
Published: (2023)
by: Zhao, Fuheng, et al.
Published: (2023)
A Semantic Approach for Big Data Exploration in Industry 4.0
by: Berges, Idoia, et al.
Published: (2024)
by: Berges, Idoia, et al.
Published: (2024)
ResBench: A Comprehensive Framework for Evaluating Database Resilience
by: Hu, Puyun, et al.
Published: (2025)
by: Hu, Puyun, et al.
Published: (2025)
The KG-ER Conceptual Schema Language
by: Franconi, Enrico, et al.
Published: (2025)
by: Franconi, Enrico, et al.
Published: (2025)
Similar Items
-
LLM-assisted Knowledge Graph Engineering: Experiments with ChatGPT
by: Meyer, Lars-Peter, et al.
Published: (2023) -
Assessing SPARQL capabilities of Large Language Models
by: Meyer, Lars-Peter, et al.
Published: (2024) -
ARUQULA -- An LLM based Text2SPARQL Approach using ReAct and Knowledge Graph Exploration Utilities
by: Brei, Felix, et al.
Published: (2025) -
Leveraging small language models for Text2SPARQL tasks to improve the resilience of AI assistance
by: Brei, Felix, et al.
Published: (2024) -
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
by: Kassem, Aly M., et al.
Published: (2025)