Statistical Confidence in Functional Correctness: An Approach for AI Product Functional Correctness Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Albertini, Wallace, Araújo, Marina Condé, Araújo, Júlia Condé, Alves, Antonio Pedro Santos, Kalinowski, Marcos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Applying a Requirements-Focused Agile Management Approach for Machine Learning-Enabled Systems
di: Romao, Lucas, et al.
Pubblicazione: (2026)
di: Romao, Lucas, et al.
Pubblicazione: (2026)
Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape
di: Trinkenreich, Bianca, et al.
Pubblicazione: (2026)
di: Trinkenreich, Bianca, et al.
Pubblicazione: (2026)
Domain Knowledge in Requirements Engineering: A Systematic Mapping Study
di: Araújo, Marina, et al.
Pubblicazione: (2025)
di: Araújo, Marina, et al.
Pubblicazione: (2025)
LLM-Assisted Thematic Analysis: Opportunities, Limitations, and Recommendations
di: Ornelas, Tatiane, et al.
Pubblicazione: (2025)
di: Ornelas, Tatiane, et al.
Pubblicazione: (2025)
Impacts of Generative AI on Agile Teams' Productivity: A Multi-Case Longitudinal Study
di: Tomaz, Rafael, et al.
Pubblicazione: (2026)
di: Tomaz, Rafael, et al.
Pubblicazione: (2026)
Understanding and measuring software engineer behavior: What can we learn from the behavioral sciences?
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2024)
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2024)
Teaching Survey Research in Software Engineering
di: Kalinowski, Marcos, et al.
Pubblicazione: (2024)
di: Kalinowski, Marcos, et al.
Pubblicazione: (2024)
Towards Effective Collaboration between Software Engineers and Data Scientists developing Machine Learning-Enabled Systems
di: Busquim, Gabriel, et al.
Pubblicazione: (2024)
di: Busquim, Gabriel, et al.
Pubblicazione: (2024)
Define-ML: An Approach to Ideate Machine Learning-Enabled Systems
di: Alonso, Silvio, et al.
Pubblicazione: (2025)
di: Alonso, Silvio, et al.
Pubblicazione: (2025)
Embracing Experiential Learning: Hackathons as an Educational Strategy for Shaping Soft Skills in Software Engineering
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2025)
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2025)
Professional Insights into Benefits and Limitations of Implementing MLOps Principles
di: Araujo, Gabriel, et al.
Pubblicazione: (2024)
di: Araujo, Gabriel, et al.
Pubblicazione: (2024)
Can participation in a hackathon impact the motivation of software engineering students? A preliminary case study analysis
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2024)
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2024)
Investigating Issues that Lead to Code Technical Debt in Machine Learning Systems
di: Ximenes, Rodrigo, et al.
Pubblicazione: (2025)
di: Ximenes, Rodrigo, et al.
Pubblicazione: (2025)
Investigating Benefits and Limitations of Migrating to a Micro-Frontends Architecture
di: Antunes, Fabio, et al.
Pubblicazione: (2024)
di: Antunes, Fabio, et al.
Pubblicazione: (2024)
Towards Emotionally Intelligent Software Engineers: Understanding Students' Self-Perceptions After a Cooperative Learning Experience
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2025)
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2025)
Investigating the Online Recruitment and Selection Journey of Novice Software Engineers: Anti-patterns and Recommendations
di: Setúbal, Miguel, et al.
Pubblicazione: (2024)
di: Setúbal, Miguel, et al.
Pubblicazione: (2024)
Teaching Practically Relevant Research Problem Formulation in Software Engineering with Lean Research Inception
di: Pereira, Anrafel Fernandes, et al.
Pubblicazione: (2026)
di: Pereira, Anrafel Fernandes, et al.
Pubblicazione: (2026)
Towards a Framework for Operationalizing the Specification of Trustworthy AI Requirements
di: Villamizar, Hugo, et al.
Pubblicazione: (2025)
di: Villamizar, Hugo, et al.
Pubblicazione: (2025)
Industrial Practices of Requirements Engineering for ML-Enabled Systems in Brazil
di: Alves, Antonio Pedro Santos, et al.
Pubblicazione: (2024)
di: Alves, Antonio Pedro Santos, et al.
Pubblicazione: (2024)
A Quasi-Experimental Evaluation of Coaching to Mitigate the Impostor Phenomenon in Early-Career Software Engineers
di: Guenes, Paloma, et al.
Pubblicazione: (2026)
di: Guenes, Paloma, et al.
Pubblicazione: (2026)
An Empirical Study of Generative AI Adoption in Software Engineering
di: Giray, Görkem, et al.
Pubblicazione: (2025)
di: Giray, Görkem, et al.
Pubblicazione: (2025)
CorrectBench: Automatic Testbench Generation with Functional Self-Correction using LLMs for HDL Design
di: Qiu, Ruidi, et al.
Pubblicazione: (2024)
di: Qiu, Ruidi, et al.
Pubblicazione: (2024)
Correctness Witnesses with Function Contracts
di: Heizmann, Matthias, et al.
Pubblicazione: (2025)
di: Heizmann, Matthias, et al.
Pubblicazione: (2025)
POLARIS: A framework to guide the development of Trustworthy AI systems
di: Baldassarre, Maria Teresa, et al.
Pubblicazione: (2024)
di: Baldassarre, Maria Teresa, et al.
Pubblicazione: (2024)
On the Interaction between Software Engineers and Data Scientists when building Machine Learning-Enabled Systems
di: Busquim, Gabriel, et al.
Pubblicazione: (2024)
di: Busquim, Gabriel, et al.
Pubblicazione: (2024)
Constructive Patterns for Human-Centered Tech Hiring
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2026)
di: Araújo, Allysson Allex, et al.
Pubblicazione: (2026)
Towards AI Agents Supported Research Problem Formulation
di: Pereira, Anrafel Fernandes, et al.
Pubblicazione: (2025)
di: Pereira, Anrafel Fernandes, et al.
Pubblicazione: (2025)
Adoption of Large Language Models in Scrum Management: Insights from Brazilian Practitioners
di: Perkusich, Mirko, et al.
Pubblicazione: (2026)
di: Perkusich, Mirko, et al.
Pubblicazione: (2026)
Agile Minds, Innovative Solutions, and Industry-Academia Collaboration: Lean R&D Meets Problem-Based Learning in Software Engineering Education
di: Romao, Lucas, et al.
Pubblicazione: (2024)
di: Romao, Lucas, et al.
Pubblicazione: (2024)
Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects
di: Kashif, Syed Mohammad, et al.
Pubblicazione: (2026)
di: Kashif, Syed Mohammad, et al.
Pubblicazione: (2026)
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
di: Peng, Yibo, et al.
Pubblicazione: (2025)
di: Peng, Yibo, et al.
Pubblicazione: (2025)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
di: Yang, Guang, et al.
Pubblicazione: (2024)
di: Yang, Guang, et al.
Pubblicazione: (2024)
Software Engineering Podcasts: An Empirical Study of Their Potential as a Research Resource
di: Wyrich, Marvin, et al.
Pubblicazione: (2026)
di: Wyrich, Marvin, et al.
Pubblicazione: (2026)
A Multivocal Literature Review on the Benefits and Limitations of Automated Machine Learning Tools
di: Azevedo, Kelly, et al.
Pubblicazione: (2024)
di: Azevedo, Kelly, et al.
Pubblicazione: (2024)
Investigating the Use of LLMs for Evidence Briefings Generation in Software Engineering
di: Marcelino, Mauro, et al.
Pubblicazione: (2025)
di: Marcelino, Mauro, et al.
Pubblicazione: (2025)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
di: Liu, Fang, et al.
Pubblicazione: (2024)
di: Liu, Fang, et al.
Pubblicazione: (2024)
Ensuring Functional Correctness of Large Code Models with Selective Generation
di: Jeong, Jaewoo, et al.
Pubblicazione: (2025)
di: Jeong, Jaewoo, et al.
Pubblicazione: (2025)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
di: Santana Jr, E. G., et al.
Pubblicazione: (2025)
di: Santana Jr, E. G., et al.
Pubblicazione: (2025)
Teaching Empirical Research Methods in Software Engineering: An Editorial Introduction
di: Mendez, Daniel, et al.
Pubblicazione: (2025)
di: Mendez, Daniel, et al.
Pubblicazione: (2025)
Designing a Syllabus for a Course on Empirical Software Engineering
di: Avgeriou, Paris, et al.
Pubblicazione: (2025)
di: Avgeriou, Paris, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Applying a Requirements-Focused Agile Management Approach for Machine Learning-Enabled Systems
di: Romao, Lucas, et al.
Pubblicazione: (2026) -
Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape
di: Trinkenreich, Bianca, et al.
Pubblicazione: (2026) -
Domain Knowledge in Requirements Engineering: A Systematic Mapping Study
di: Araújo, Marina, et al.
Pubblicazione: (2025) -
LLM-Assisted Thematic Analysis: Opportunities, Limitations, and Recommendations
di: Ornelas, Tatiane, et al.
Pubblicazione: (2025) -
Impacts of Generative AI on Agile Teams' Productivity: A Multi-Case Longitudinal Study
di: Tomaz, Rafael, et al.
Pubblicazione: (2026)