The Expert Validation Framework (EVF): Enabling Domain Expert Control in AI Engineering
Fuente:
arXiv
Saved in:
| Main Authors: | Gren, Lucas, Dobslaw, Felix |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparing Open-Source and Commercial LLMs for Domain-Specific Analysis and Reporting: Software Engineering Challenges and Design Trade-offs
by: Koraag, Theo, et al.
Published: (2025)
by: Koraag, Theo, et al.
Published: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
by: Dobslaw, Felix, et al.
Published: (2025)
by: Dobslaw, Felix, et al.
Published: (2025)
AI-Assisted Requirements Engineering: An Empirical Evaluation Relative to Expert Judgment
by: Levy, Oz, et al.
Published: (2026)
by: Levy, Oz, et al.
Published: (2026)
Analysis of LLMs vs Human Experts in Requirements Engineering
by: Hymel, Cory, et al.
Published: (2025)
by: Hymel, Cory, et al.
Published: (2025)
Cross-Functional AI Task Forces (X-FAITs) for AI Transformation of Software Organizations
by: Gren, Lucas, et al.
Published: (2025)
by: Gren, Lucas, et al.
Published: (2025)
Mining Software Repositories for Expert Recommendation
by: Marshall, Chad, et al.
Published: (2025)
by: Marshall, Chad, et al.
Published: (2025)
From Group Psychology to Software Engineering Research to Automotive R&D: Measuring Team Development at Volvo Cars
by: Gren, Lucas, et al.
Published: (2024)
by: Gren, Lucas, et al.
Published: (2024)
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
by: Steigerwald, Philipp, et al.
Published: (2026)
by: Steigerwald, Philipp, et al.
Published: (2026)
GEMS: Generative Expert Metric System through Iterative Prompt Priming
by: Cheng, Ti-Chung, et al.
Published: (2024)
by: Cheng, Ti-Chung, et al.
Published: (2024)
Automatic techniques for issue report classification: A systematic mapping study
by: Laiq, Muhammad, et al.
Published: (2025)
by: Laiq, Muhammad, et al.
Published: (2025)
LUK: Empowering Log Understanding with Expert Knowledge from Large Language Models
by: Ma, Lipeng, et al.
Published: (2024)
by: Ma, Lipeng, et al.
Published: (2024)
LogReasoner: Empowering LLMs with Expert-like Coarse-to-Fine Reasoning for Automated Log Analysis
by: Ma, Lipeng, et al.
Published: (2025)
by: Ma, Lipeng, et al.
Published: (2025)
The Cognitive Circuit Breaker: A Systems Engineering Framework for Intrinsic AI Reliability
by: Pan, Jonathan
Published: (2026)
by: Pan, Jonathan
Published: (2026)
Automated Validation of LLM-based Evaluators for Software Engineering Artifacts
by: Fandina, Ora Nova, et al.
Published: (2025)
by: Fandina, Ora Nova, et al.
Published: (2025)
The Social Psychology of Software Security (Psycurity)
by: Gren, Lucas, et al.
Published: (2024)
by: Gren, Lucas, et al.
Published: (2024)
Prototypical Leadership in Agile Software Development
by: Dawood, Jina, et al.
Published: (2024)
by: Dawood, Jina, et al.
Published: (2024)
Investigating team maturity in an agile automotive reorganization
by: Gren, Lucas, et al.
Published: (2024)
by: Gren, Lucas, et al.
Published: (2024)
Unified Software Engineering Agent as AI Software Engineer
by: Applis, Leonhard, et al.
Published: (2025)
by: Applis, Leonhard, et al.
Published: (2025)
Engineering AI Judge Systems
by: Lin, Jiahuei, et al.
Published: (2024)
by: Lin, Jiahuei, et al.
Published: (2024)
Naming the Pain in Machine Learning-Enabled Systems Engineering
by: Kalinowski, Marcos, et al.
Published: (2024)
by: Kalinowski, Marcos, et al.
Published: (2024)
Shift-Up: A Framework for Software Engineering Guardrails in AI-native Software Development -- Initial Findings
by: Lipsanen, Petrus, et al.
Published: (2026)
by: Lipsanen, Petrus, et al.
Published: (2026)
Reversa: A Reverse Documentation Engineering Framework for Converting Legacy Software into Operational Specifications for AI Agents
by: de Macedo, Sanderson Oliveira, et al.
Published: (2026)
by: de Macedo, Sanderson Oliveira, et al.
Published: (2026)
AI-Tutoring in Software Engineering Education
by: Frankford, Eduard, et al.
Published: (2024)
by: Frankford, Eduard, et al.
Published: (2024)
Human-Centered Evaluation of an LLM-Based Process Modeling Copilot: A Mixed-Methods Study with Domain Experts
by: Lauer, Chantale, et al.
Published: (2026)
by: Lauer, Chantale, et al.
Published: (2026)
Agentic AI Software Engineers: Programming with Trust
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
An Agent-Based Framework for the Automatic Validation of Mathematical Optimization Models
by: Zadorojniy, Alexander, et al.
Published: (2025)
by: Zadorojniy, Alexander, et al.
Published: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
by: Baqar, Mohammad, et al.
Published: (2024)
by: Baqar, Mohammad, et al.
Published: (2024)
Understanding on the Edge: LLM-generated Boundary Test Explanations
by: Akbarova, Sabinakhon, et al.
Published: (2026)
by: Akbarova, Sabinakhon, et al.
Published: (2026)
SPDZCoder: Combining Expert Knowledge with LLMs for Generating Privacy-Computing Code
by: Dong, Xiaoning, et al.
Published: (2024)
by: Dong, Xiaoning, et al.
Published: (2024)
Towards an Appropriate Level of Reliance on AI: A Preliminary Reliance-Control Framework for AI in Software Engineering
by: Ferino, Samuel, et al.
Published: (2026)
by: Ferino, Samuel, et al.
Published: (2026)
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
by: Aggarwal, Pooja, et al.
Published: (2024)
by: Aggarwal, Pooja, et al.
Published: (2024)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
by: Li, Jingyue, et al.
Published: (2026)
by: Li, Jingyue, et al.
Published: (2026)
Generative AI for Requirements Engineering: A Systematic Literature Review
by: Cheng, Haowei, et al.
Published: (2024)
by: Cheng, Haowei, et al.
Published: (2024)
Will AI replace Software Engineers? Do not hold your breath
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
GenAI for Simulation Model in Model-Based Systems Engineering
by: Zhang, Lin, et al.
Published: (2025)
by: Zhang, Lin, et al.
Published: (2025)
Generative AI and Empirical Software Engineering: A Paradigm Shift
by: Treude, Christoph, et al.
Published: (2025)
by: Treude, Christoph, et al.
Published: (2025)
When Prompt Engineering Meets Software Engineering: CNL-P as Natural and Robust "APIs'' for Human-AI Interaction
by: Xing, Zhenchang, et al.
Published: (2025)
by: Xing, Zhenchang, et al.
Published: (2025)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
by: Alonso, Silvio, et al.
Published: (2025)
by: Alonso, Silvio, et al.
Published: (2025)
Greening AI-enabled Systems with Software Engineering: A Research Agenda for Environmentally Sustainable AI Practices
by: Cruz, Luís, et al.
Published: (2025)
by: Cruz, Luís, et al.
Published: (2025)
Software Reuse in the Generative AI Era: From Cargo Cult Towards AI Native Software Engineering
by: Mikkonen, Tommi, et al.
Published: (2025)
by: Mikkonen, Tommi, et al.
Published: (2025)
Similar Items
-
Comparing Open-Source and Commercial LLMs for Domain-Specific Analysis and Reporting: Software Engineering Challenges and Design Trade-offs
by: Koraag, Theo, et al.
Published: (2025) -
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
by: Dobslaw, Felix, et al.
Published: (2025) -
AI-Assisted Requirements Engineering: An Empirical Evaluation Relative to Expert Judgment
by: Levy, Oz, et al.
Published: (2026) -
Analysis of LLMs vs Human Experts in Requirements Engineering
by: Hymel, Cory, et al.
Published: (2025) -
Cross-Functional AI Task Forces (X-FAITs) for AI Transformation of Software Organizations
by: Gren, Lucas, et al.
Published: (2025)