Assessing Correctness in LLM-Based Code Generation via Uncertainty Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Arindam, David, Cristina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using Semantic Distance to Estimate Uncertainty in LLM-Based Code Generation
by: He, Weilin, et al.
Published: (2026)
by: He, Weilin, et al.
Published: (2026)
Ensemble-Based Uncertainty Estimation for Code Correctness Estimation
by: Wei, Yunxiang, et al.
Published: (2026)
by: Wei, Yunxiang, et al.
Published: (2026)
Beyond Code Generation: Assessing Code LLM Maturity with Postconditions
by: He, Fusen, et al.
Published: (2024)
by: He, Fusen, et al.
Published: (2024)
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation
by: Sarker, Laboni, et al.
Published: (2024)
by: Sarker, Laboni, et al.
Published: (2024)
In-Context Learning as an Effective Estimator of Functional Correctness of LLM-Generated Code
by: Das, Susmita, et al.
Published: (2025)
by: Das, Susmita, et al.
Published: (2025)
Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
by: Vu, Thanh Trong, et al.
Published: (2025)
by: Vu, Thanh Trong, et al.
Published: (2025)
Structured Safety Auditing for Balancing Code Correctness and Content Safety in LLM-Generated Code
by: Tan, Honghao, et al.
Published: (2026)
by: Tan, Honghao, et al.
Published: (2026)
Uncertainty Quantification for LLM-based Code Generation
by: Xu, Senrong, et al.
Published: (2026)
by: Xu, Senrong, et al.
Published: (2026)
Assessing the Impact of Requirement Ambiguity on LLM-based Function-Level Code Generation
by: Yang, Di, et al.
Published: (2026)
by: Yang, Di, et al.
Published: (2026)
CodeCircuit: Toward Inferring LLM-Generated Code Correctness via Attribution Graphs
by: He, Yicheng, et al.
Published: (2026)
by: He, Yicheng, et al.
Published: (2026)
Assessing AI-Based Code Assistants in Method Generation Tasks
by: Corso, Vincenzo, et al.
Published: (2024)
by: Corso, Vincenzo, et al.
Published: (2024)
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026)
by: AKLI, Amal, et al.
Published: (2026)
Validated Code Translation for Projects with External Libraries
by: Zhang, Hanliang, et al.
Published: (2026)
by: Zhang, Hanliang, et al.
Published: (2026)
CodeScore-R: An Automated Robustness Metric for Assessing the FunctionalCorrectness of Code Synthesis
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
SemGuard: Real-Time Semantic Evaluator for Correcting LLM-Generated Code
by: Wang, Qinglin, et al.
Published: (2025)
by: Wang, Qinglin, et al.
Published: (2025)
Assessing Code Generation with Intermediate Languages
by: Deng, Xun, et al.
Published: (2024)
by: Deng, Xun, et al.
Published: (2024)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
by: Khati, Dipin, et al.
Published: (2026)
by: Khati, Dipin, et al.
Published: (2026)
Static Analysis as a Feedback Loop: Enhancing LLM-Generated Code Beyond Correctness
by: Blyth, Scott, et al.
Published: (2025)
by: Blyth, Scott, et al.
Published: (2025)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
Prompt Optimization for LLM Code Generation via Reinforcement Learning
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
by: Liu, Fang, et al.
Published: (2024)
by: Liu, Fang, et al.
Published: (2024)
CodeCoR: An LLM-Based Self-Reflective Multi-Agent Framework for Code Generation
by: Pan, Ruwei, et al.
Published: (2025)
by: Pan, Ruwei, et al.
Published: (2025)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
by: Zhu, Yuqi, et al.
Published: (2025)
by: Zhu, Yuqi, et al.
Published: (2025)
Automated Repair of Ambiguous Problem Descriptions for LLM-Based Code Generation
by: Jia, Haoxiang, et al.
Published: (2025)
by: Jia, Haoxiang, et al.
Published: (2025)
Quantum-Guided Test Case Minimization for LLM-Based Code Generation
by: Zhang, Huixiang, et al.
Published: (2025)
by: Zhang, Huixiang, et al.
Published: (2025)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
Evaluating Efficiency and Novelty of LLM-Generated Code for Graph Analysis
by: Nia, Atieh Barati, et al.
Published: (2025)
by: Nia, Atieh Barati, et al.
Published: (2025)
Demystifying Faulty Code with LLM: Step-by-Step Reasoning for Explainable Fault Localization
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
Correctness isnt Efficiency: Runtime Memory Divergence in LLM-Generated Code
by: Rajput, Prateek, et al.
Published: (2026)
by: Rajput, Prateek, et al.
Published: (2026)
Enhancing LLM Code Generation with Ensembles: A Similarity-Based Selection Approach
by: Mahmud, Tarek, et al.
Published: (2025)
by: Mahmud, Tarek, et al.
Published: (2025)
SemOpt: LLM-Driven Code Optimization via Rule-Based Analysis
by: Zhao, Yuwei, et al.
Published: (2025)
by: Zhao, Yuwei, et al.
Published: (2025)
When is Generated Code Difficult to Comprehend? Assessing AI Agent Python Code Proficiency in the Wild
by: Temkulkiat, Nanthit, et al.
Published: (2026)
by: Temkulkiat, Nanthit, et al.
Published: (2026)
AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length
by: Cheng, Junhang, et al.
Published: (2025)
by: Cheng, Junhang, et al.
Published: (2025)
Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks
by: Hasan, Md Mahade, et al.
Published: (2025)
by: Hasan, Md Mahade, et al.
Published: (2025)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
by: Du, Mingzhe, et al.
Published: (2025)
by: Du, Mingzhe, et al.
Published: (2025)
A Differential Fuzzing-Based Evaluation of Functional Equivalence in LLM-Generated Code Refactorings
by: Dristi, Simantika Bhattacharjee, et al.
Published: (2026)
by: Dristi, Simantika Bhattacharjee, et al.
Published: (2026)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
by: Fakhoury, Sarah, et al.
Published: (2024)
by: Fakhoury, Sarah, et al.
Published: (2024)
Smoke and Mirrors: Jailbreaking LLM-based Code Generation via Implicit Malicious Prompts
by: Ouyang, Sheng, et al.
Published: (2025)
by: Ouyang, Sheng, et al.
Published: (2025)
De-Hallucinator: Mitigating LLM Hallucinations in Code Generation Tasks via Iterative Grounding
by: Eghbali, Aryaz, et al.
Published: (2024)
by: Eghbali, Aryaz, et al.
Published: (2024)
AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search
by: Li, Qingyao, et al.
Published: (2026)
by: Li, Qingyao, et al.
Published: (2026)
Similar Items
-
Using Semantic Distance to Estimate Uncertainty in LLM-Based Code Generation
by: He, Weilin, et al.
Published: (2026) -
Ensemble-Based Uncertainty Estimation for Code Correctness Estimation
by: Wei, Yunxiang, et al.
Published: (2026) -
Beyond Code Generation: Assessing Code LLM Maturity with Postconditions
by: He, Fusen, et al.
Published: (2024) -
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation
by: Sarker, Laboni, et al.
Published: (2024) -
In-Context Learning as an Effective Estimator of Functional Correctness of LLM-Generated Code
by: Das, Susmita, et al.
Published: (2025)