Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vu, Thanh Trong, Bui, Tuan-Dung, Nguyen, Thu-Trang, Nguyen, Son, Vo, Hieu Dinh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Correctness Assessment of Code Generated by Large Language Models Using Internal Representations
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2025)
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2025)
Automated Description Generation for Software Patches
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024)
iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML
von: Le, Dat, et al.
Veröffentlicht: (2026)
von: Le, Dat, et al.
Veröffentlicht: (2026)
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)
Generating Critical Scenarios for Testing Automated Driving Systems
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
von: Nguyen, Trung-Hieu, et al.
Veröffentlicht: (2024)
CABENCH: Benchmarking Composable AI for Solving Complex Tasks through Composing Ready-to-Use Models
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
von: Pham, Tung-Thuy, et al.
Veröffentlicht: (2025)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
von: Nguyen, Tien-Quang, et al.
Veröffentlicht: (2024)
Impact of Code Transformation on Detection of Smart Contract Vulnerabilities
von: Manh, Cuong Tran, et al.
Veröffentlicht: (2024)
von: Manh, Cuong Tran, et al.
Veröffentlicht: (2024)
On the Impacts of Contexts on Repository-Level Code Generation
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
Towards Test Generation from Task Description for Mobile Testing with Multi-modal Reasoning
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
Segment-Based Test Case Prioritization: A Multi-objective Approach
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
von: Huynh, Hieu, et al.
Veröffentlicht: (2024)
COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
von: Huynh, Hieu, et al.
Veröffentlicht: (2025)
Larger Is Not Always Better: Leveraging Structured Code Diffs for Comment Inconsistency Detection
von: Nguyen, Phong, et al.
Veröffentlicht: (2025)
von: Nguyen, Phong, et al.
Veröffentlicht: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
von: Manh, Dung Nguyen, et al.
Veröffentlicht: (2024)
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
von: Le, Huy, et al.
Veröffentlicht: (2025)
von: Le, Huy, et al.
Veröffentlicht: (2025)
Verifying DNN-based Semantic Communication Against Generative Adversarial Noise
von: Le, Thanh, et al.
Veröffentlicht: (2026)
von: Le, Thanh, et al.
Veröffentlicht: (2026)
On LLMs' Internal Representation of Code Correctness
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
COBOL-Coder: Domain-Adapted Large Language Models for COBOL Code Generation and Translation
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
von: Dau, Anh T. V., et al.
Veröffentlicht: (2026)
Inferring Properties of Graph Neural Networks
von: Nguyen, Dat, et al.
Veröffentlicht: (2024)
von: Nguyen, Dat, et al.
Veröffentlicht: (2024)
Toward Generation of Test Cases from Task Descriptions via History-aware Planning
von: Cao, Duy, et al.
Veröffentlicht: (2025)
von: Cao, Duy, et al.
Veröffentlicht: (2025)
When Retriever Meets Generator: A Joint Model for Code Comment Generation
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
von: Le, Tien P. T., et al.
Veröffentlicht: (2025)
CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
CRACI: A Cloud-Native Reference Architecture for the Industrial Compute Continuum
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
LEGION: Harnessing Pre-trained Language Models for GitHub Topic Recommendations with Distribution-Balance Loss
von: Dang, Yen-Trang, et al.
Veröffentlicht: (2024)
von: Dang, Yen-Trang, et al.
Veröffentlicht: (2024)
CodeFlow: Program Behavior Prediction with Dynamic Dependencies Learning
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
Encoding Version History Context for Better Code Representation
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Huy, et al.
Veröffentlicht: (2024)
Natural Is The Best: Model-Agnostic Code Simplification for Pre-trained Large Language Models
von: Wang, Yan, et al.
Veröffentlicht: (2024)
von: Wang, Yan, et al.
Veröffentlicht: (2024)
VisualCoder: Guiding Large Language Models in Code Execution with Fine-grained Multimodal Chain-of-Thought Reasoning
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
von: Le, Cuong Chi, et al.
Veröffentlicht: (2024)
On the Effectiveness of Code Representation in Deep Learning-Based Automated Patch Correctness Assessment
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
von: Zhang, Quanjun, et al.
Veröffentlicht: (2026)
Do Not Treat Code as Natural Language: Implications for Repository-Level Code Generation and Beyond
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
von: Le-Anh, Minh, et al.
Veröffentlicht: (2026)
An LLM-based multi-agent framework for agile effort estimation
von: Bui, Thanh-Long, et al.
Veröffentlicht: (2025)
von: Bui, Thanh-Long, et al.
Veröffentlicht: (2025)
CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation
von: Doan, Duy Tung, et al.
Veröffentlicht: (2026)
von: Doan, Duy Tung, et al.
Veröffentlicht: (2026)
Functional Overlap Reranking for Neural Code Generation
von: To, Hung Quoc, et al.
Veröffentlicht: (2023)
von: To, Hung Quoc, et al.
Veröffentlicht: (2023)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Detection of Technical Debt in Java Source Code
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
von: Hai, Nam Le, et al.
Veröffentlicht: (2024)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Correctness Assessment of Code Generated by Large Language Models Using Internal Representations
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2025) -
Automated Description Generation for Software Patches
von: Vu, Thanh Trong, et al.
Veröffentlicht: (2024) -
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
von: Nguyen, Thu-Trang, et al.
Veröffentlicht: (2024) -
iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML
von: Le, Dat, et al.
Veröffentlicht: (2026) -
RAMBO: Enhancing RAG-based Repository-Level Method Body Completion
von: Bui, Tuan-Dung, et al.
Veröffentlicht: (2024)