Extended Empirical Validation of the Explainability Solution Space
Fuente:
arXiv
Saved in:
| Main Authors: | Mestre, Antoni, Albert, Manoli, Gil, Miriam, Pelechano, Vicente |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
by: Pelechano, Vicente, et al.
Published: (2026)
by: Pelechano, Vicente, et al.
Published: (2026)
Empirical Computation
by: Tang, Eric, et al.
Published: (2025)
by: Tang, Eric, et al.
Published: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
by: Yu, Jiongchi, et al.
Published: (2025)
by: Yu, Jiongchi, et al.
Published: (2025)
Agentic Frameworks for Reasoning Tasks: An Empirical Study
by: Rasheed, Zeeshan, et al.
Published: (2026)
by: Rasheed, Zeeshan, et al.
Published: (2026)
An Empirical Study of AI Techniques in Mobile Applications
by: Li, Yinghua, et al.
Published: (2022)
by: Li, Yinghua, et al.
Published: (2022)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
by: Li, Jingyue, et al.
Published: (2026)
by: Li, Jingyue, et al.
Published: (2026)
XSearch: Explainable Code Search via Concept-to-Code Alignment
by: Liu, Yiming, et al.
Published: (2026)
by: Liu, Yiming, et al.
Published: (2026)
Verification and Validation of Autonomous Systems
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
by: Shetiya, Sneha Sudhir, et al.
Published: (2024)
An Empirical Study of the Imbalance Issue in Software Vulnerability Detection
by: Guo, Yuejun, et al.
Published: (2026)
by: Guo, Yuejun, et al.
Published: (2026)
Using LLMs in Software Requirements Specifications: An Empirical Evaluation
by: Krishna, Madhava, et al.
Published: (2024)
by: Krishna, Madhava, et al.
Published: (2024)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
by: Wang, Ruiqi, et al.
Published: (2025)
by: Wang, Ruiqi, et al.
Published: (2025)
An Empirical Study of Challenges in Machine Learning Asset Management
by: Zhao, Zhimin, et al.
Published: (2024)
by: Zhao, Zhimin, et al.
Published: (2024)
Advanced Detection of Source Code Clones via an Ensemble of Unsupervised Similarity Measures
by: Martinez-Gil, Jorge
Published: (2024)
by: Martinez-Gil, Jorge
Published: (2024)
Automated Validation of COBOL to Java Transformation
by: Kumar, Atul, et al.
Published: (2025)
by: Kumar, Atul, et al.
Published: (2025)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
by: Ebrahimi, Amir M., et al.
Published: (2026)
by: Ebrahimi, Amir M., et al.
Published: (2026)
LLM-Based Robustness Testing of Microservice Applications: An Empirical Study
by: Tigulla, Hrushitha Goud, et al.
Published: (2026)
by: Tigulla, Hrushitha Goud, et al.
Published: (2026)
How Do Agents Perform Code Optimization? An Empirical Study
by: Peng, Huiyun, et al.
Published: (2025)
by: Peng, Huiyun, et al.
Published: (2025)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
by: Meng, Xiangxin, et al.
Published: (2024)
by: Meng, Xiangxin, et al.
Published: (2024)
An Empirical Study of Agent Developer Practices in AI Agent Frameworks
by: Wang, Yanlin, et al.
Published: (2025)
by: Wang, Yanlin, et al.
Published: (2025)
Generative AI and Empirical Software Engineering: A Paradigm Shift
by: Treude, Christoph, et al.
Published: (2025)
by: Treude, Christoph, et al.
Published: (2025)
Can GPT-4 Replicate Empirical Software Engineering Research?
by: Liang, Jenny T., et al.
Published: (2023)
by: Liang, Jenny T., et al.
Published: (2023)
An Empirical Study of OpenAI API Discussions on Stack Overflow
by: Chen, Xiang, et al.
Published: (2025)
by: Chen, Xiang, et al.
Published: (2025)
An Empirical Framework for Evaluating Semantic Preservation Using Hugging Face
by: Jia, Nan, et al.
Published: (2025)
by: Jia, Nan, et al.
Published: (2025)
Bugs in Large Language Models Generated Code: An Empirical Study
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Towards Explainable Stakeholder-Aware Requirements Prioritisation in Aged-Care Digital Health
by: Xiao, Yuqing, et al.
Published: (2026)
by: Xiao, Yuqing, et al.
Published: (2026)
From Ranking to Reasoning: Explainable Web API Recommendation via Semantic Reasoning
by: Xu, Zishuo, et al.
Published: (2025)
by: Xu, Zishuo, et al.
Published: (2025)
AI-Assisted Requirements Engineering: An Empirical Evaluation Relative to Expert Judgment
by: Levy, Oz, et al.
Published: (2026)
by: Levy, Oz, et al.
Published: (2026)
Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
by: Nirujan, Hinduja, et al.
Published: (2026)
by: Nirujan, Hinduja, et al.
Published: (2026)
An Empirical Study of Proactive Coding Assistants in Real-World Software Development
by: Li, Lehui, et al.
Published: (2026)
by: Li, Lehui, et al.
Published: (2026)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
by: Majgaonkar, Oorja, et al.
Published: (2025)
by: Majgaonkar, Oorja, et al.
Published: (2025)
Open the Oyster: Empirical Evaluation and Improvement of Code Reasoning Confidence in LLMs
by: Wang, Shufan, et al.
Published: (2025)
by: Wang, Shufan, et al.
Published: (2025)
Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering Studies
by: Angermeir, Florian, et al.
Published: (2025)
by: Angermeir, Florian, et al.
Published: (2025)
Analyzing Prompt Influence on Automated Method Generation: An Empirical Study with Copilot
by: Fagadau, Ionut Daniel, et al.
Published: (2024)
by: Fagadau, Ionut Daniel, et al.
Published: (2024)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
by: Nguyen, Thu-Trang, et al.
Published: (2024)
by: Nguyen, Thu-Trang, et al.
Published: (2024)
Enhancing Automated Program Repair with Solution Design
by: Zhao, Jiuang, et al.
Published: (2024)
by: Zhao, Jiuang, et al.
Published: (2024)
Large Language Models for Validating Network Protocol Parsers
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
by: Wang, Xiaoyin, et al.
Published: (2024)
by: Wang, Xiaoyin, et al.
Published: (2024)
Multimodal Auto Validation For Self-Refinement in Web Agents
by: Azam, Ruhana, et al.
Published: (2024)
by: Azam, Ruhana, et al.
Published: (2024)
TENET: Leveraging Tests Beyond Validation for Code Generation
by: Hu, Yiran, et al.
Published: (2025)
by: Hu, Yiran, et al.
Published: (2025)
Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study
by: Bansal, Kaushal
Published: (2026)
by: Bansal, Kaushal
Published: (2026)
Similar Items
-
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
by: Pelechano, Vicente, et al.
Published: (2026) -
Empirical Computation
by: Tang, Eric, et al.
Published: (2025) -
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
by: Yu, Jiongchi, et al.
Published: (2025) -
Agentic Frameworks for Reasoning Tasks: An Empirical Study
by: Rasheed, Zeeshan, et al.
Published: (2026) -
An Empirical Study of AI Techniques in Mobile Applications
by: Li, Yinghua, et al.
Published: (2022)