Salvato in:
| Autori principali: | Dristi, Simantika Bhattacharjee, Dwyer, Matthew B. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2509.15397 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Differential Fuzzing-Based Evaluation of Functional Equivalence in LLM-Generated Code Refactorings
di: Dristi, Simantika Bhattacharjee, et al.
Pubblicazione: (2026)
di: Dristi, Simantika Bhattacharjee, et al.
Pubblicazione: (2026)
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks
di: Siddiq, Mohammed Latif, et al.
Pubblicazione: (2024)
di: Siddiq, Mohammed Latif, et al.
Pubblicazione: (2024)
Bias Testing and Mitigation in LLM-based Code Generation
di: Huang, Dong, et al.
Pubblicazione: (2023)
di: Huang, Dong, et al.
Pubblicazione: (2023)
TOGLL: Correct and Strong Test Oracle Generation with LLMs
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024)
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024)
Doc2OracLL: Investigating the Impact of Documentation on LLM-based Test Oracle Generation
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024)
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024)
Generating Realistic, Diverse, and Fault-Revealing Inputs with Latent Space Interpolation for Testing Deep Neural Networks
di: Duan, Bin, et al.
Pubblicazione: (2025)
di: Duan, Bin, et al.
Pubblicazione: (2025)
STADA: Specification-based Testing for Autonomous Driving Agents
di: Saha, Joy, et al.
Pubblicazione: (2026)
di: Saha, Joy, et al.
Pubblicazione: (2026)
Bias Unveiled: Investigating Social Bias in LLM-Generated Code
di: Ling, Lin, et al.
Pubblicazione: (2024)
di: Ling, Lin, et al.
Pubblicazione: (2024)
Mitigating Omitted Variable Bias in Empirical Software Engineering
di: Furia, Carlo A., et al.
Pubblicazione: (2025)
di: Furia, Carlo A., et al.
Pubblicazione: (2025)
Social Bias in LLM-Generated Code: Benchmark and Mitigation
di: Rabbi, Fazle, et al.
Pubblicazione: (2026)
di: Rabbi, Fazle, et al.
Pubblicazione: (2026)
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
di: Chen, QiHong, et al.
Pubblicazione: (2025)
di: Chen, QiHong, et al.
Pubblicazione: (2025)
Can Code Evaluation Metrics Detect Code Plagiarism?
di: Ebrahim, Fahad, et al.
Pubblicazione: (2026)
di: Ebrahim, Fahad, et al.
Pubblicazione: (2026)
Static Code Analyzer Recommendation via Preference Mining
di: Ge, Xiuting, et al.
Pubblicazione: (2024)
di: Ge, Xiuting, et al.
Pubblicazione: (2024)
Benchmarks and Metrics for Evaluations of Code Generation: A Critical Review
di: Paul, Debalina Ghosh, et al.
Pubblicazione: (2024)
di: Paul, Debalina Ghosh, et al.
Pubblicazione: (2024)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
di: Du, Yongkang, et al.
Pubblicazione: (2025)
di: Du, Yongkang, et al.
Pubblicazione: (2025)
Exploring Multi-Lingual Bias of Large Code Models in Code Generation
di: Wang, Chaozheng, et al.
Pubblicazione: (2024)
di: Wang, Chaozheng, et al.
Pubblicazione: (2024)
Analyzing and Mitigating (with LLMs) the Security Misconfigurations of Helm Charts from Artifact Hub
di: Minna, Francesco, et al.
Pubblicazione: (2024)
di: Minna, Francesco, et al.
Pubblicazione: (2024)
Generating Maximal Configurations and Their Variants Using Code Metrics
di: Yavuz, Tuba, et al.
Pubblicazione: (2024)
di: Yavuz, Tuba, et al.
Pubblicazione: (2024)
Mitigating Gender Bias in Code Large Language Models via Model Editing
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
di: Qin, Zhanyue, et al.
Pubblicazione: (2024)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
di: Yang, Guang, et al.
Pubblicazione: (2024)
di: Yang, Guang, et al.
Pubblicazione: (2024)
Analyzing and Evaluating the Behavior of Git Diff and Merge
di: Glodny, Niels
Pubblicazione: (2025)
di: Glodny, Niels
Pubblicazione: (2025)
ChatGPT for Code Refactoring: Analyzing Topics, Interaction, and Effective Prompts
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2025)
di: AlOmar, Eman Abdullah, et al.
Pubblicazione: (2025)
Analyzing Dependency Distribution Changes Arising from Code Smell Interactions
di: Zhang, Zushuai, et al.
Pubblicazione: (2025)
di: Zhang, Zushuai, et al.
Pubblicazione: (2025)
Code-Survey: An LLM-Driven Methodology for Analyzing Large-Scale Codebases
di: Zheng, Yusheng, et al.
Pubblicazione: (2024)
di: Zheng, Yusheng, et al.
Pubblicazione: (2024)
Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models
di: Xiao, Yisong, et al.
Pubblicazione: (2025)
di: Xiao, Yisong, et al.
Pubblicazione: (2025)
How Reliable Are FOSS Popularity Metrics? Analyzing the Effort Required for Spoofing Common Software Popularity Metrics
di: Swierzy, Ben, et al.
Pubblicazione: (2025)
di: Swierzy, Ben, et al.
Pubblicazione: (2025)
Hallucinations in Code Change to Natural Language Generation: Prevalence and Evaluation of Detection Metrics
di: Liu, Chunhua, et al.
Pubblicazione: (2025)
di: Liu, Chunhua, et al.
Pubblicazione: (2025)
CODE-DITING: A Reasoning-Based Metric for Functional Alignment in Code Evaluation
di: Yang, Guang, et al.
Pubblicazione: (2025)
di: Yang, Guang, et al.
Pubblicazione: (2025)
NRevisit: A Cognitive Behavioral Metric for Code Understandability Assessment
di: Hao, Gao, et al.
Pubblicazione: (2025)
di: Hao, Gao, et al.
Pubblicazione: (2025)
Software Code Quality Measurement: Implications from Metric Distributions
di: Jin, Siyuan, et al.
Pubblicazione: (2023)
di: Jin, Siyuan, et al.
Pubblicazione: (2023)
Integrating Code Metrics into Automated Documentation Generation for Computational Notebooks
di: Ghahfarokhi, Mojtaba Mostafavi, et al.
Pubblicazione: (2026)
di: Ghahfarokhi, Mojtaba Mostafavi, et al.
Pubblicazione: (2026)
Towards Understanding the Impact of Code Modifications on Software Quality Metrics
di: Karanikiotis, Thomas, et al.
Pubblicazione: (2024)
di: Karanikiotis, Thomas, et al.
Pubblicazione: (2024)
Who Introduces and Who Fixes? Analyzing Code Quality in Collaborative Student's Projects
di: Ferrao, Rafael Corsi, et al.
Pubblicazione: (2025)
di: Ferrao, Rafael Corsi, et al.
Pubblicazione: (2025)
Exploring Large Language Models for Analyzing and Improving Method Names in Scientific Code
di: Larsen, Gunnar, et al.
Pubblicazione: (2025)
di: Larsen, Gunnar, et al.
Pubblicazione: (2025)
PyGress: Tool for Analyzing the Progression of Code Proficiency in Python OSS Projects
di: Charatvaraphan, Rujiphart, et al.
Pubblicazione: (2025)
di: Charatvaraphan, Rujiphart, et al.
Pubblicazione: (2025)
Assessing Quality Metrics for Neural Reality Gap Input Mitigation in Autonomous Driving Testing
di: Lambertenghi, Stefano Carlo, et al.
Pubblicazione: (2024)
di: Lambertenghi, Stefano Carlo, et al.
Pubblicazione: (2024)
A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics
di: Katzy, Jonathan, et al.
Pubblicazione: (2025)
di: Katzy, Jonathan, et al.
Pubblicazione: (2025)
FAIL: Analyzing Software Failures from the News Using LLMs
di: Anandayuvaraj, Dharun, et al.
Pubblicazione: (2024)
di: Anandayuvaraj, Dharun, et al.
Pubblicazione: (2024)
Analyzing Prominent LLMs: An Empirical Study of Performance and Complexity in Solving LeetCode Problems
di: Guimaraes, Everton, et al.
Pubblicazione: (2025)
di: Guimaraes, Everton, et al.
Pubblicazione: (2025)
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
di: Ghammam, Anwar, et al.
Pubblicazione: (2026)
di: Ghammam, Anwar, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Differential Fuzzing-Based Evaluation of Functional Equivalence in LLM-Generated Code Refactorings
di: Dristi, Simantika Bhattacharjee, et al.
Pubblicazione: (2026) -
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks
di: Siddiq, Mohammed Latif, et al.
Pubblicazione: (2024) -
Bias Testing and Mitigation in LLM-based Code Generation
di: Huang, Dong, et al.
Pubblicazione: (2023) -
TOGLL: Correct and Strong Test Oracle Generation with LLMs
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024) -
Doc2OracLL: Investigating the Impact of Documentation on LLM-based Test Oracle Generation
di: Hossain, Soneya Binta, et al.
Pubblicazione: (2024)