Correctness Assessment of Code Generated by Large Language Models Using Internal Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bui, Tuan-Dung, Vu, Thanh Trong, Nguyen, Thu-Trang, Nguyen, Son, Vo, Hieu Dinh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2025)
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2025)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
Automated Description Generation for Software Patches
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML
von: Le, Dat, et al.
Veröffentlicht: (2026)
von: Le, Dat, et al.
Veröffentlicht: (2026)
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
Generating Critical Scenarios for Testing Automated Driving Systems
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
CABENCH: Benchmarking Composable AI for Solving Complex Tasks through Composing Ready-to-Use Models
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
Impact of Code Transformation on Detection of Smart Contract Vulnerabilities
von: Manh, Cuong Tran, et al.
Veröffentlicht: (2024)
von: Manh, Cuong Tran, et al.
Veröffentlicht: (2024)
COBOL-Coder: Domain-Adapted Large Language Models for COBOL Code Generation and Translation
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
On the Impacts of Contexts on Repository-Level Code Generation
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
Towards Test Generation from Task Description for Mobile Testing with Multi-modal Reasoning
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
VisualCoder: Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
Segment-Based Test Case Prioritization: A Multi-objective Approach
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
von: Le, Huy, et al.
Veröffentlicht: (2025)
von: Le, Huy, et al.
Veröffentlicht: (2025)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
Larger Is Not Always Better: Leveraging Structured Code Diffs for Comment Inconsistency Detection
von: Nguyen, Phong, et al.
Veröffentlicht: (2025)
von: Nguyen, Phong, et al.
Veröffentlicht: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
LEGION: Harnessing Pre-trained Language Models for GitHub Topic Recommendations with Distribution-Balance Loss
von: Dang, Yen-Trang, et al.
Veröffentlicht: (2024)
von: Dang, Yen-Trang, et al.
Veröffentlicht: (2024)
An Empirical Study on Self-correcting Large Language Models for Data Science Code Generation
von: Quoc, Thai Tang, et al.
Veröffentlicht: (2024)
von: Quoc, Thai Tang, et al.
Veröffentlicht: (2024)
KAT: Dependency-aware Automated API Testing with Large Language Models
von: Le, Tri, et al.
Veröffentlicht: (2024)
von: Le, Tri, et al.
Veröffentlicht: (2024)
CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
When Retriever Meets Generator: A Joint Model for Code Comment Generation
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
Verifying DNN-based Semantic Communication Against Generative Adversarial Noise
von: Le, Thanh, et al.
Veröffentlicht: (2026)
von: Le, Thanh, et al.
Veröffentlicht: (2026)
Do Not Treat Code as Natural Language: Implications for Repository-Level Code Generation and Beyond
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
On LLMs' Internal Representation of Code Correctness
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
Inferring Properties of Graph Neural Networks
von: Nguyen, Dat, et al.
Veröffentlicht: (2024)
von: Nguyen, Dat, et al.
Veröffentlicht: (2024)
TestWeaver: Execution-aware, Feedback-driven Regression Testing Generation with Large Language Models
von: Le, Cuong Chi, et al.
Veröffentlicht: (2025)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2025)
Toward Generation of Test Cases from Task Descriptions via History-aware Planning
von: Cao, Duy, et al.
Veröffentlicht: (2025)
von: Cao, Duy, et al.
Veröffentlicht: (2025)
DocChecker: Bootstrapping Code Large Language Model for Detecting and Resolving Code-Comment Inconsistencies
von: Dau, Anh T. V., et al.
Veröffentlicht: (2023)
von: Dau, Anh T. V., et al.
Veröffentlicht: (2023)
CRACI: A Cloud-Native Reference Architecture for the Industrial Compute Continuum
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
Encoding Version History Context for Better Code Representation
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
Natural Is The Best: Model-Agnostic Code Simplification for Pre-trained Large Language Models
von: Wang, Yan, et al.
Veröffentlicht: (2024)
von: Wang, Yan, et al.
Veröffentlicht: (2024)
Exploring the Reasoning Depth of Small Language Models in Software Architecture: A Multidimensional Evaluation Framework Towards Software Engineering 2.0
von: Vo, Ha, et al.
Veröffentlicht: (2026)
von: Vo, Ha, et al.
Veröffentlicht: (2026)
Adversarial Attacks on Code Models with Discriminative Graph Patterns
von: Nguyen, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Nguyen, Thanh-Dat, et al.
Veröffentlicht: (2023)
DesignCoder: Hierarchy-Aware and Self-Correcting UI Code Generation with Large Language Models
von: Chen, Yunnong, et al.
Veröffentlicht: (2025)
von: Chen, Yunnong, et al.
Veröffentlicht: (2025)
On the Effectiveness of Code Representation in Deep Learning-Based Automated Patch Correctness Assessment
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2025) -
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024) -
Automated Description Generation for Software Patches
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024) -
iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML
von: Le, Dat, et al.
Veröffentlicht: (2026) -
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)