Extended Empirical Validation of the Explainability Solution Space
Fuente:
arXiv
Guardado en:
| Autores principales: | Mestre, Antoni, Albert, Manoli, Gil, Miriam, Pelechano, Vicente |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
por: Pelechano, Vicente, et al.
Publicado: (2026)
por: Pelechano, Vicente, et al.
Publicado: (2026)
Empirical Computation
por: Tang, Eric, et al.
Publicado: (2025)
por: Tang, Eric, et al.
Publicado: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
por: Yu, Jiongchi, et al.
Publicado: (2025)
por: Yu, Jiongchi, et al.
Publicado: (2025)
Agentic Frameworks for Reasoning Tasks: An Empirical Study
por: Rasheed, Zeeshan, et al.
Publicado: (2026)
por: Rasheed, Zeeshan, et al.
Publicado: (2026)
An Empirical Study of AI Techniques in Mobile Applications
por: Li, Yinghua, et al.
Publicado: (2022)
por: Li, Yinghua, et al.
Publicado: (2022)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
por: Li, Jingyue, et al.
Publicado: (2026)
por: Li, Jingyue, et al.
Publicado: (2026)
XSearch: Explainable Code Search via Concept-to-Code Alignment
por: Liu, Yiming, et al.
Publicado: (2026)
por: Liu, Yiming, et al.
Publicado: (2026)
Verification and Validation of Autonomous Systems
por: Shetiya, Sneha Sudhir, et al.
Publicado: (2024)
por: Shetiya, Sneha Sudhir, et al.
Publicado: (2024)
An Empirical Study of the Imbalance Issue in Software Vulnerability Detection
por: Guo, Yuejun, et al.
Publicado: (2026)
por: Guo, Yuejun, et al.
Publicado: (2026)
Using LLMs in Software Requirements Specifications: An Empirical Evaluation
por: Krishna, Madhava, et al.
Publicado: (2024)
por: Krishna, Madhava, et al.
Publicado: (2024)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
por: Wang, Ruiqi, et al.
Publicado: (2025)
por: Wang, Ruiqi, et al.
Publicado: (2025)
An Empirical Study of Challenges in Machine Learning Asset Management
por: Zhao, Zhimin, et al.
Publicado: (2024)
por: Zhao, Zhimin, et al.
Publicado: (2024)
Advanced Detection of Source Code Clones via an Ensemble of Unsupervised Similarity Measures
por: Martinez-Gil, Jorge
Publicado: (2024)
por: Martinez-Gil, Jorge
Publicado: (2024)
Automated Validation of COBOL to Java Transformation
por: Kumar, Atul, et al.
Publicado: (2025)
por: Kumar, Atul, et al.
Publicado: (2025)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
LLM-Based Robustness Testing of Microservice Applications: An Empirical Study
por: Tigulla, Hrushitha Goud, et al.
Publicado: (2026)
por: Tigulla, Hrushitha Goud, et al.
Publicado: (2026)
How Do Agents Perform Code Optimization? An Empirical Study
por: Peng, Huiyun, et al.
Publicado: (2025)
por: Peng, Huiyun, et al.
Publicado: (2025)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
por: Meng, Xiangxin, et al.
Publicado: (2024)
por: Meng, Xiangxin, et al.
Publicado: (2024)
An Empirical Study of Agent Developer Practices in AI Agent Frameworks
por: Wang, Yanlin, et al.
Publicado: (2025)
por: Wang, Yanlin, et al.
Publicado: (2025)
Generative AI and Empirical Software Engineering: A Paradigm Shift
por: Treude, Christoph, et al.
Publicado: (2025)
por: Treude, Christoph, et al.
Publicado: (2025)
Can GPT-4 Replicate Empirical Software Engineering Research?
por: Liang, Jenny T., et al.
Publicado: (2023)
por: Liang, Jenny T., et al.
Publicado: (2023)
An Empirical Study of OpenAI API Discussions on Stack Overflow
por: Chen, Xiang, et al.
Publicado: (2025)
por: Chen, Xiang, et al.
Publicado: (2025)
An Empirical Framework for Evaluating Semantic Preservation Using Hugging Face
por: Jia, Nan, et al.
Publicado: (2025)
por: Jia, Nan, et al.
Publicado: (2025)
Bugs in Large Language Models Generated Code: An Empirical Study
por: Tambon, Florian, et al.
Publicado: (2024)
por: Tambon, Florian, et al.
Publicado: (2024)
Towards Explainable Stakeholder-Aware Requirements Prioritisation in Aged-Care Digital Health
por: Xiao, Yuqing, et al.
Publicado: (2026)
por: Xiao, Yuqing, et al.
Publicado: (2026)
From Ranking to Reasoning: Explainable Web API Recommendation via Semantic Reasoning
por: Xu, Zishuo, et al.
Publicado: (2025)
por: Xu, Zishuo, et al.
Publicado: (2025)
AI-Assisted Requirements Engineering: An Empirical Evaluation Relative to Expert Judgment
por: Levy, Oz, et al.
Publicado: (2026)
por: Levy, Oz, et al.
Publicado: (2026)
Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
por: Nirujan, Hinduja, et al.
Publicado: (2026)
por: Nirujan, Hinduja, et al.
Publicado: (2026)
An Empirical Study of Proactive Coding Assistants in Real-World Software Development
por: Li, Lehui, et al.
Publicado: (2026)
por: Li, Lehui, et al.
Publicado: (2026)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
por: Majgaonkar, Oorja, et al.
Publicado: (2025)
por: Majgaonkar, Oorja, et al.
Publicado: (2025)
Open the Oyster: Empirical Evaluation and Improvement of Code Reasoning Confidence in LLMs
por: Wang, Shufan, et al.
Publicado: (2025)
por: Wang, Shufan, et al.
Publicado: (2025)
Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering Studies
por: Angermeir, Florian, et al.
Publicado: (2025)
por: Angermeir, Florian, et al.
Publicado: (2025)
Analyzing Prompt Influence on Automated Method Generation: An Empirical Study with Copilot
por: Fagadau, Ionut Daniel, et al.
Publicado: (2024)
por: Fagadau, Ionut Daniel, et al.
Publicado: (2024)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
por: Nguyen, Thu-Trang, et al.
Publicado: (2024)
por: Nguyen, Thu-Trang, et al.
Publicado: (2024)
Enhancing Automated Program Repair with Solution Design
por: Zhao, Jiuang, et al.
Publicado: (2024)
por: Zhao, Jiuang, et al.
Publicado: (2024)
Large Language Models for Validating Network Protocol Parsers
por: Zheng, Mingwei, et al.
Publicado: (2025)
por: Zheng, Mingwei, et al.
Publicado: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
por: Wang, Xiaoyin, et al.
Publicado: (2024)
por: Wang, Xiaoyin, et al.
Publicado: (2024)
Multimodal Auto Validation For Self-Refinement in Web Agents
por: Azam, Ruhana, et al.
Publicado: (2024)
por: Azam, Ruhana, et al.
Publicado: (2024)
TENET: Leveraging Tests Beyond Validation for Code Generation
por: Hu, Yiran, et al.
Publicado: (2025)
por: Hu, Yiran, et al.
Publicado: (2025)
Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study
por: Bansal, Kaushal
Publicado: (2026)
por: Bansal, Kaushal
Publicado: (2026)
Ejemplares similares
-
HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems
por: Pelechano, Vicente, et al.
Publicado: (2026) -
Empirical Computation
por: Tang, Eric, et al.
Publicado: (2025) -
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
por: Yu, Jiongchi, et al.
Publicado: (2025) -
Agentic Frameworks for Reasoning Tasks: An Empirical Study
por: Rasheed, Zeeshan, et al.
Publicado: (2026) -
An Empirical Study of AI Techniques in Mobile Applications
por: Li, Yinghua, et al.
Publicado: (2022)