Statistical Confidence in Functional Correctness: An Approach for AI Product Functional Correctness Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Albertini, Wallace, Araújo, Marina Condé, Araújo, Júlia Condé, Alves, Antonio Pedro Santos, Kalinowski, Marcos |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Applying a Requirements-Focused Agile Management Approach for Machine Learning-Enabled Systems
by: Romao, Lucas, et al.
Published: (2026)
by: Romao, Lucas, et al.
Published: (2026)
Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape
by: Trinkenreich, Bianca, et al.
Published: (2026)
by: Trinkenreich, Bianca, et al.
Published: (2026)
Domain Knowledge in Requirements Engineering: A Systematic Mapping Study
by: Araújo, Marina, et al.
Published: (2025)
by: Araújo, Marina, et al.
Published: (2025)
LLM-Assisted Thematic Analysis: Opportunities, Limitations, and Recommendations
by: Ornelas, Tatiane, et al.
Published: (2025)
by: Ornelas, Tatiane, et al.
Published: (2025)
Impacts of Generative AI on Agile Teams' Productivity: A Multi-Case Longitudinal Study
by: Tomaz, Rafael, et al.
Published: (2026)
by: Tomaz, Rafael, et al.
Published: (2026)
Understanding and measuring software engineer behavior: What can we learn from the behavioral sciences?
by: Araújo, Allysson Allex, et al.
Published: (2024)
by: Araújo, Allysson Allex, et al.
Published: (2024)
Teaching Survey Research in Software Engineering
by: Kalinowski, Marcos, et al.
Published: (2024)
by: Kalinowski, Marcos, et al.
Published: (2024)
Towards Effective Collaboration between Software Engineers and Data Scientists developing Machine Learning-Enabled Systems
by: Busquim, Gabriel, et al.
Published: (2024)
by: Busquim, Gabriel, et al.
Published: (2024)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
by: Alonso, Silvio, et al.
Published: (2025)
by: Alonso, Silvio, et al.
Published: (2025)
Embracing Experiential Learning: Hackathons as an Educational Strategy for Shaping Soft Skills in Software Engineering
by: Araújo, Allysson Allex, et al.
Published: (2025)
by: Araújo, Allysson Allex, et al.
Published: (2025)
Professional Insights into Benefits and Limitations of Implementing MLOps Principles
by: Araujo, Gabriel, et al.
Published: (2024)
by: Araujo, Gabriel, et al.
Published: (2024)
Can participation in a hackathon impact the motivation of software engineering students? A preliminary case study analysis
by: Araújo, Allysson Allex, et al.
Published: (2024)
by: Araújo, Allysson Allex, et al.
Published: (2024)
Investigating Issues that Lead to Code Technical Debt in Machine Learning Systems
by: Ximenes, Rodrigo, et al.
Published: (2025)
by: Ximenes, Rodrigo, et al.
Published: (2025)
Investigating Benefits and Limitations of Migrating to a Micro-Frontends Architecture
by: Antunes, Fabio, et al.
Published: (2024)
by: Antunes, Fabio, et al.
Published: (2024)
Towards Emotionally Intelligent Software Engineers: Understanding Students' Self-Perceptions After a Cooperative Learning Experience
by: Araújo, Allysson Allex, et al.
Published: (2025)
by: Araújo, Allysson Allex, et al.
Published: (2025)
Investigating the Online Recruitment and Selection Journey of Novice Software Engineers: Anti-patterns and Recommendations
by: Setúbal, Miguel, et al.
Published: (2024)
by: Setúbal, Miguel, et al.
Published: (2024)
Teaching Practically Relevant Research Problem Formulation in Software Engineering with Lean Research Inception
by: Pereira, Anrafel Fernandes, et al.
Published: (2026)
by: Pereira, Anrafel Fernandes, et al.
Published: (2026)
Towards a Framework for Operationalizing the Specification of Trustworthy AI Requirements
by: Villamizar, Hugo, et al.
Published: (2025)
by: Villamizar, Hugo, et al.
Published: (2025)
Industrial Practices of Requirements Engineering for ML-Enabled Systems in Brazil
by: Alves, Antonio Pedro Santos, et al.
Published: (2024)
by: Alves, Antonio Pedro Santos, et al.
Published: (2024)
A Quasi-Experimental Evaluation of Coaching to Mitigate the Impostor Phenomenon in Early-Career Software Engineers
by: Guenes, Paloma, et al.
Published: (2026)
by: Guenes, Paloma, et al.
Published: (2026)
An Empirical Study of Generative AI Adoption in Software Engineering
by: Giray, Görkem, et al.
Published: (2025)
by: Giray, Görkem, et al.
Published: (2025)
CorrectBench: Automatic Testbench Generation with Functional Self-Correction using LLMs for HDL Design
by: Qiu, Ruidi, et al.
Published: (2024)
by: Qiu, Ruidi, et al.
Published: (2024)
Correctness Witnesses with Function Contracts
by: Heizmann, Matthias, et al.
Published: (2025)
by: Heizmann, Matthias, et al.
Published: (2025)
POLARIS: A framework to guide the development of Trustworthy AI systems
by: Baldassarre, Maria Teresa, et al.
Published: (2024)
by: Baldassarre, Maria Teresa, et al.
Published: (2024)
On the Interaction between Software Engineers and Data Scientists when building Machine Learning-Enabled Systems
by: Busquim, Gabriel, et al.
Published: (2024)
by: Busquim, Gabriel, et al.
Published: (2024)
Constructive Patterns for Human-Centered Tech Hiring
by: Araújo, Allysson Allex, et al.
Published: (2026)
by: Araújo, Allysson Allex, et al.
Published: (2026)
Towards AI Agents Supported Research Problem Formulation
by: Pereira, Anrafel Fernandes, et al.
Published: (2025)
by: Pereira, Anrafel Fernandes, et al.
Published: (2025)
Adoption of Large Language Models in Scrum Management: Insights from Brazilian Practitioners
by: Perkusich, Mirko, et al.
Published: (2026)
by: Perkusich, Mirko, et al.
Published: (2026)
Agile Minds, Innovative Solutions, and Industry-Academia Collaboration: Lean R&D Meets Problem-Based Learning in Software Engineering Education
by: Romao, Lucas, et al.
Published: (2024)
by: Romao, Lucas, et al.
Published: (2024)
Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects
by: Kashif, Syed Mohammad, et al.
Published: (2026)
by: Kashif, Syed Mohammad, et al.
Published: (2026)
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
by: Peng, Yibo, et al.
Published: (2025)
by: Peng, Yibo, et al.
Published: (2025)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
Software Engineering Podcasts: An Empirical Study of Their Potential as a Research Resource
by: Wyrich, Marvin, et al.
Published: (2026)
by: Wyrich, Marvin, et al.
Published: (2026)
A Multivocal Literature Review on the Benefits and Limitations of Automated Machine Learning Tools
by: Azevedo, Kelly, et al.
Published: (2024)
by: Azevedo, Kelly, et al.
Published: (2024)
Investigating the Use of LLMs for Evidence Briefings Generation in Software Engineering
by: Marcelino, Mauro, et al.
Published: (2025)
by: Marcelino, Mauro, et al.
Published: (2025)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
by: Liu, Fang, et al.
Published: (2024)
by: Liu, Fang, et al.
Published: (2024)
Ensuring Functional Correctness of Large Code Models with Selective Generation
by: Jeong, Jaewoo, et al.
Published: (2025)
by: Jeong, Jaewoo, et al.
Published: (2025)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
Teaching Empirical Research Methods in Software Engineering: An Editorial Introduction
by: Mendez, Daniel, et al.
Published: (2025)
by: Mendez, Daniel, et al.
Published: (2025)
Designing a Syllabus for a Course on Empirical Software Engineering
by: Avgeriou, Paris, et al.
Published: (2025)
by: Avgeriou, Paris, et al.
Published: (2025)
Similar Items
-
Applying a Requirements-Focused Agile Management Approach for Machine Learning-Enabled Systems
by: Romao, Lucas, et al.
Published: (2026) -
Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape
by: Trinkenreich, Bianca, et al.
Published: (2026) -
Domain Knowledge in Requirements Engineering: A Systematic Mapping Study
by: Araújo, Marina, et al.
Published: (2025) -
LLM-Assisted Thematic Analysis: Opportunities, Limitations, and Recommendations
by: Ornelas, Tatiane, et al.
Published: (2025) -
Impacts of Generative AI on Agile Teams' Productivity: A Multi-Case Longitudinal Study
by: Tomaz, Rafael, et al.
Published: (2026)