Human Evaluation of Procedural Knowledge Graph Extraction from Text with Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Carriero, Valentina Anita, Azzini, Antonia, Baroni, Ilaria, Scrocca, Mario, Celino, Irene |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Conversational Approach to Well-being Awareness Creation and Behavioural Intention
von: Azzini, Antonia, et al.
Veröffentlicht: (2025)
von: Azzini, Antonia, et al.
Veröffentlicht: (2025)
Procedural Knowledge Ontology (PKO)
von: Carriero, Valentina Anita, et al.
Veröffentlicht: (2025)
von: Carriero, Valentina Anita, et al.
Veröffentlicht: (2025)
Knowledge Affordances for Hybrid Human-AI Information Seeking
von: Celino, Irene
Veröffentlicht: (2026)
von: Celino, Irene
Veröffentlicht: (2026)
VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs
von: Xiang, Yurui, et al.
Veröffentlicht: (2026)
von: Xiang, Yurui, et al.
Veröffentlicht: (2026)
Are Humans as Brittle as Large Language Models?
von: Li, Jiahui, et al.
Veröffentlicht: (2025)
von: Li, Jiahui, et al.
Veröffentlicht: (2025)
HealthGenie: Empowering Users with Healthy Dietary Guidance through Knowledge Graph and Large Language Models
von: Gao, Fan, et al.
Veröffentlicht: (2025)
von: Gao, Fan, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Identifying Knowledge Components
von: Wang, Canwen, et al.
Veröffentlicht: (2025)
von: Wang, Canwen, et al.
Veröffentlicht: (2025)
Large Language Models for Virtual Human Gesture Selection
von: Torshizi, Parisa Ghanad, et al.
Veröffentlicht: (2025)
von: Torshizi, Parisa Ghanad, et al.
Veröffentlicht: (2025)
Large Language Models for Depression Recognition in Spoken Language Integrating Psychological Knowledge
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
SciDaSynth: Interactive Structured Data Extraction from Scientific Literature with Large Language Model
von: Wang, Xingbo, et al.
Veröffentlicht: (2024)
von: Wang, Xingbo, et al.
Veröffentlicht: (2024)
Hallucinations and Key Information Extraction in Medical Texts: A Comprehensive Assessment of Open-Source Large Language Models
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2025)
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2025)
Large Language Model-based Human-Agent Collaboration for Complex Task Solving
von: Feng, Xueyang, et al.
Veröffentlicht: (2024)
von: Feng, Xueyang, et al.
Veröffentlicht: (2024)
Evaluating the Usage of African-American Vernacular English in Large Language Models
von: Dunlap, Deja, et al.
Veröffentlicht: (2026)
von: Dunlap, Deja, et al.
Veröffentlicht: (2026)
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
von: Si, Chenglei, et al.
Veröffentlicht: (2023)
von: Si, Chenglei, et al.
Veröffentlicht: (2023)
The Effectiveness of Style Vectors for Steering Large Language Models: A Human Evaluation
von: Diallo, Diaoulé, et al.
Veröffentlicht: (2026)
von: Diallo, Diaoulé, et al.
Veröffentlicht: (2026)
DTS-SQL: Decomposed Text-to-SQL with Small Large Language Models
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
An Actor-Critic Approach to Boosting Text-to-SQL Large Language Model
von: Zheng, Ziyang, et al.
Veröffentlicht: (2024)
von: Zheng, Ziyang, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Learning Complex Legal Concepts through Storytelling
von: Jiang, Hang, et al.
Veröffentlicht: (2024)
von: Jiang, Hang, et al.
Veröffentlicht: (2024)
Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering Systems
von: Martin-Boyle, Anna, et al.
Veröffentlicht: (2026)
von: Martin-Boyle, Anna, et al.
Veröffentlicht: (2026)
Toward Automated Qualitative Analysis: Leveraging Large Language Models for Tutoring Dialogue Evaluation
von: Gu, Megan, et al.
Veröffentlicht: (2025)
von: Gu, Megan, et al.
Veröffentlicht: (2025)
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
von: Mena, Omar, et al.
Veröffentlicht: (2025)
von: Mena, Omar, et al.
Veröffentlicht: (2025)
Evaluating Telugu Proficiency in Large Language Models_ A Comparative Analysis of ChatGPT and Gemini
von: Kishore, Katikela Sreeharsha, et al.
Veröffentlicht: (2024)
von: Kishore, Katikela Sreeharsha, et al.
Veröffentlicht: (2024)
LalaEval: A Holistic Human Evaluation Framework for Domain-Specific Large Language Models
von: Sun, Chongyan, et al.
Veröffentlicht: (2024)
von: Sun, Chongyan, et al.
Veröffentlicht: (2024)
rapid-triples: Adaptive Forms for Semi-automatic Knowledge Collection in RDF
von: Scrocca, Mario, et al.
Veröffentlicht: (2025)
von: Scrocca, Mario, et al.
Veröffentlicht: (2025)
Guiding Generative Storytelling with Knowledge Graphs
von: Pan, Zhijun, et al.
Veröffentlicht: (2025)
von: Pan, Zhijun, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models in Theory of Mind Tasks
von: Kosinski, Michal
Veröffentlicht: (2023)
von: Kosinski, Michal
Veröffentlicht: (2023)
Evaluation Of P300 Speller Performance Using Large Language Models Along With Cross-Subject Training
von: Parthasarathy, Nithin, et al.
Veröffentlicht: (2024)
von: Parthasarathy, Nithin, et al.
Veröffentlicht: (2024)
Automated Novelty Evaluation of Academic Paper: A Collaborative Approach Integrating Human and Large Language Model Knowledge
von: Wu, Wenqing, et al.
Veröffentlicht: (2025)
von: Wu, Wenqing, et al.
Veröffentlicht: (2025)
Practicing with Language Models Cultivates Human Empathic Communication
von: Kumar, Aakriti, et al.
Veröffentlicht: (2026)
von: Kumar, Aakriti, et al.
Veröffentlicht: (2026)
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
von: Ibrahim, Lujain, et al.
Veröffentlicht: (2025)
von: Ibrahim, Lujain, et al.
Veröffentlicht: (2025)
An Evaluation of Estimative Uncertainty in Large Language Models
von: Tang, Zhisheng, et al.
Veröffentlicht: (2024)
von: Tang, Zhisheng, et al.
Veröffentlicht: (2024)
Evaluating the Prompt Steerability of Large Language Models
von: Miehling, Erik, et al.
Veröffentlicht: (2024)
von: Miehling, Erik, et al.
Veröffentlicht: (2024)
Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
von: Chen, Chaoran, et al.
Veröffentlicht: (2025)
von: Chen, Chaoran, et al.
Veröffentlicht: (2025)
Phraselette: A Poet's Procedural Palette
von: Calderwood, Alex, et al.
Veröffentlicht: (2025)
von: Calderwood, Alex, et al.
Veröffentlicht: (2025)
Large Language Models Pass the Turing Test
von: Jones, Cameron R., et al.
Veröffentlicht: (2025)
von: Jones, Cameron R., et al.
Veröffentlicht: (2025)
Social Skill Training with Large Language Models
von: Yang, Diyi, et al.
Veröffentlicht: (2024)
von: Yang, Diyi, et al.
Veröffentlicht: (2024)
Towards Valid Student Simulation with Large Language Models
von: Yuan, Zhihao, et al.
Veröffentlicht: (2026)
von: Yuan, Zhihao, et al.
Veröffentlicht: (2026)
Evaluating Large Language Models in Analysing Classroom Dialogue
von: Long, Yun, et al.
Veröffentlicht: (2024)
von: Long, Yun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Conversational Approach to Well-being Awareness Creation and Behavioural Intention
von: Azzini, Antonia, et al.
Veröffentlicht: (2025) -
Procedural Knowledge Ontology (PKO)
von: Carriero, Valentina Anita, et al.
Veröffentlicht: (2025) -
Knowledge Affordances for Hybrid Human-AI Information Seeking
von: Celino, Irene
Veröffentlicht: (2026) -
VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs
von: Xiang, Yurui, et al.
Veröffentlicht: (2026) -
Are Humans as Brittle as Large Language Models?
von: Li, Jiahui, et al.
Veröffentlicht: (2025)