Human-Centered Evaluation of an LLM-Based Process Modeling Copilot: A Mixed-Methods Study with Domain Experts
Fuente:
arXiv
Guardado en:
| Autores principales: | Lauer, Chantale, Pfeiffer, Peter, Mehdiyev, Nijat |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Assessing the Business Process Modeling Competences of Large Language Models
por: Lauer, Chantale, et al.
Publicado: (2026)
por: Lauer, Chantale, et al.
Publicado: (2026)
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
por: Steigerwald, Philipp, et al.
Publicado: (2026)
por: Steigerwald, Philipp, et al.
Publicado: (2026)
From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering
por: Dong, Tao, et al.
Publicado: (2025)
por: Dong, Tao, et al.
Publicado: (2025)
The RealHumanEval: Evaluating Large Language Models' Abilities to Support Programmers
por: Mozannar, Hussein, et al.
Publicado: (2024)
por: Mozannar, Hussein, et al.
Publicado: (2024)
Qualitative Evaluation of LLM-Designed GUI
por: Sawicki, Bartosz, et al.
Publicado: (2026)
por: Sawicki, Bartosz, et al.
Publicado: (2026)
Single Conversation Methodology: A Human-Centered Protocol for AI-Assisted Software Development
por: Escobedo, Salvador D.
Publicado: (2025)
por: Escobedo, Salvador D.
Publicado: (2025)
From Prompt to Product: A Human-Centered Benchmark of Agentic App Generation Systems
por: Ortiz, Marcos, et al.
Publicado: (2025)
por: Ortiz, Marcos, et al.
Publicado: (2025)
On the Utility of Domain Modeling Assistance with Large Language Models
por: Chaaben, Meriem Ben, et al.
Publicado: (2024)
por: Chaaben, Meriem Ben, et al.
Publicado: (2024)
Exploring Human-AI Collaboration in Agile: Customised LLM Meeting Assistants
por: Cabrero-Daniel, Beatriz, et al.
Publicado: (2024)
por: Cabrero-Daniel, Beatriz, et al.
Publicado: (2024)
GIS Copilot: Towards an Autonomous GIS Agent for Spatial Analysis
por: Akinboyewa, Temitope, et al.
Publicado: (2024)
por: Akinboyewa, Temitope, et al.
Publicado: (2024)
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
por: van der Maden, Willem, et al.
Publicado: (2026)
por: van der Maden, Willem, et al.
Publicado: (2026)
From Theory to Practice: Real-World Use Cases on Trustworthy LLM-Driven Process Modeling, Prediction and Automation
por: Pfeiffer, Peter, et al.
Publicado: (2025)
por: Pfeiffer, Peter, et al.
Publicado: (2025)
From Prompts to Propositions: A Logic-Based Lens on Student-LLM Interactions
por: Alfageeh, Ali, et al.
Publicado: (2025)
por: Alfageeh, Ali, et al.
Publicado: (2025)
Improving Energy Efficiency in Manufacturing: A Novel Expert System Shell
por: Ioshchikhes, Borys, et al.
Publicado: (2024)
por: Ioshchikhes, Borys, et al.
Publicado: (2024)
The Impact of LLM-Assistants on Software Developer Productivity: A Systematic Review and Mapping Study
por: Mohamed, Amr, et al.
Publicado: (2025)
por: Mohamed, Amr, et al.
Publicado: (2025)
EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention
por: Zhang, Yifan, et al.
Publicado: (2025)
por: Zhang, Yifan, et al.
Publicado: (2025)
Thoughtful Things: Building Human-Centric Smart Devices with Small Language Models
por: King, Evan, et al.
Publicado: (2024)
por: King, Evan, et al.
Publicado: (2024)
Investigating Multimodal Large Language Models to Support Usability Evaluation
por: Lubos, Sebastian, et al.
Publicado: (2025)
por: Lubos, Sebastian, et al.
Publicado: (2025)
A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development
por: Geyer, Werner, et al.
Publicado: (2025)
por: Geyer, Werner, et al.
Publicado: (2025)
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers
por: Fan, Aysa Xuemo, et al.
Publicado: (2024)
por: Fan, Aysa Xuemo, et al.
Publicado: (2024)
Using an LLM to Help With Code Understanding
por: Nam, Daye, et al.
Publicado: (2023)
por: Nam, Daye, et al.
Publicado: (2023)
CentaurEval: Benchmarking Human-in-the-Loop Value in Agentic Coding
por: Luo, Hanjun, et al.
Publicado: (2025)
por: Luo, Hanjun, et al.
Publicado: (2025)
Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security
por: Li, Yuanchun, et al.
Publicado: (2024)
por: Li, Yuanchun, et al.
Publicado: (2024)
EyeTrans: Merging Human and Machine Attention for Neural Code Summarization
por: Zhang, Yifan, et al.
Publicado: (2024)
por: Zhang, Yifan, et al.
Publicado: (2024)
LLM Interactive Optimization of Open Source Python Libraries -- Case Studies and Generalization
por: Florath, Andreas
Publicado: (2023)
por: Florath, Andreas
Publicado: (2023)
Toward Epistemic Stability: Engineering Consistent Procedures for Industrial LLM Hallucination Reduction
por: Freeman, Brian, et al.
Publicado: (2026)
por: Freeman, Brian, et al.
Publicado: (2026)
Context Branching for LLM Conversations: A Version Control Approach to Exploratory Programming
por: Nanjundappa, Bhargav Chickmagalur, et al.
Publicado: (2025)
por: Nanjundappa, Bhargav Chickmagalur, et al.
Publicado: (2025)
Prompt-with-Me: in-IDE Structured Prompt Management for LLM-Driven Software Engineering
por: Li, Ziyou, et al.
Publicado: (2025)
por: Li, Ziyou, et al.
Publicado: (2025)
Understanding and supporting how developers prompt for LLM-powered code editing in practice
por: Nam, Daye, et al.
Publicado: (2025)
por: Nam, Daye, et al.
Publicado: (2025)
Dynamic Framework for Collaborative Learning: Leveraging Advanced LLM with Adaptive Feedback Mechanisms
por: Tahir, Hassam, et al.
Publicado: (2026)
por: Tahir, Hassam, et al.
Publicado: (2026)
Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety
por: Uddin, S M Jamil
Publicado: (2026)
por: Uddin, S M Jamil
Publicado: (2026)
Optimizing LLM Code Suggestions: Feedback-Driven Timing with Lightweight State Bounds
por: Awad, Mohammad Nour Al, et al.
Publicado: (2025)
por: Awad, Mohammad Nour Al, et al.
Publicado: (2025)
Human-AI Experience in Integrated Development Environments: A Systematic Literature Review
por: Sergeyuk, Agnia, et al.
Publicado: (2025)
por: Sergeyuk, Agnia, et al.
Publicado: (2025)
How Developers Interact with AI: A Taxonomy of Human-AI Collaboration in Software Engineering
por: Treude, Christoph, et al.
Publicado: (2025)
por: Treude, Christoph, et al.
Publicado: (2025)
Pre-Filtering Code Suggestions using Developer Behavioral Telemetry to Optimize LLM-Assisted Programming
por: Awad, Mohammad Nour Al, et al.
Publicado: (2025)
por: Awad, Mohammad Nour Al, et al.
Publicado: (2025)
LLMs' Reshaping of People, Processes, Products, and Society in Software Development: A Comprehensive Exploration with Early Adopters
por: Tabarsi, Benyamin, et al.
Publicado: (2025)
por: Tabarsi, Benyamin, et al.
Publicado: (2025)
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
por: Pelechano, Vicente, et al.
Publicado: (2026)
por: Pelechano, Vicente, et al.
Publicado: (2026)
Agentic Metacognition: Designing a "Self-Aware" Low-Code Agent for Failure Prediction and Human Handoff
por: Xu, Jiexi
Publicado: (2025)
por: Xu, Jiexi
Publicado: (2025)
The Impact of Generative AI on Collaborative Open-Source Software Development: Evidence from GitHub Copilot
por: Song, Fangchen, et al.
Publicado: (2024)
por: Song, Fangchen, et al.
Publicado: (2024)
Beyond the Prompt: An Empirical Study of Cursor Rules
por: Jiang, Shaokang, et al.
Publicado: (2025)
por: Jiang, Shaokang, et al.
Publicado: (2025)
Ejemplares similares
-
Assessing the Business Process Modeling Competences of Large Language Models
por: Lauer, Chantale, et al.
Publicado: (2026) -
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
por: Steigerwald, Philipp, et al.
Publicado: (2026) -
From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering
por: Dong, Tao, et al.
Publicado: (2025) -
The RealHumanEval: Evaluating Large Language Models' Abilities to Support Programmers
por: Mozannar, Hussein, et al.
Publicado: (2024) -
Qualitative Evaluation of LLM-Designed GUI
por: Sawicki, Bartosz, et al.
Publicado: (2026)