Larger Is Not Always Better: Leveraging Structured Code Diffs for Comment Inconsistency Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Phong, Bui, Anh M. T., Nguyen, Phuong T. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DocChecker: Bootstrapping Code Large Language Model for Detecting and Resolving Code-Comment Inconsistencies
by: Dau, Anh T. V., et al.
Published: (2023)
by: Dau, Anh T. V., et al.
Published: (2023)
When Retriever Meets Generator: A Joint Model for Code Comment Generation
by: Le, Tien P. T., et al.
Published: (2025)
by: Le, Tien P. T., et al.
Published: (2025)
Detection of Technical Debt in Java Source Code
by: Hai, Nam Le, et al.
Published: (2024)
by: Hai, Nam Le, et al.
Published: (2024)
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
by: Le, Huy, et al.
Published: (2025)
by: Le, Huy, et al.
Published: (2025)
Good things come in three: Generating SO Post Titles with Pre-Trained Models, Self Improvement and Post Ranking
by: Le, Duc Anh, et al.
Published: (2024)
by: Le, Duc Anh, et al.
Published: (2024)
EnseSmells: Deep ensemble and programming language models for automated code smells detection
by: Ho, Anh, et al.
Published: (2025)
by: Ho, Anh, et al.
Published: (2025)
Larger Is Not Always Better: Exploring Small Open-source Language Models in Logging Statement Generation
by: Zhong, Renyi, et al.
Published: (2025)
by: Zhong, Renyi, et al.
Published: (2025)
Bake Two Cakes with One Oven: RL for Defusing Popularity Bias and Cold-start in Third-Party Library Recommendations
by: Vuong, Minh Hoang, et al.
Published: (2025)
by: Vuong, Minh Hoang, et al.
Published: (2025)
HyperAgent: Generalist Software Engineering Agents to Solve Coding Tasks at Scale
by: Phan, Huy Nhat, et al.
Published: (2024)
by: Phan, Huy Nhat, et al.
Published: (2024)
CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases
by: Hoang, Anh Nguyen, et al.
Published: (2025)
by: Hoang, Anh Nguyen, et al.
Published: (2025)
CCISolver: End-to-End Detection and Repair of Method-Level Code-Comment Inconsistency
by: Zhong, Renyi, et al.
Published: (2025)
by: Zhong, Renyi, et al.
Published: (2025)
Investigating the Impact of Code Comment Inconsistency on Bug Introducing
by: Radmanesh, Shiva, et al.
Published: (2024)
by: Radmanesh, Shiva, et al.
Published: (2024)
Qualitative Coding Analysis through Open-Source Large Language Models: A User Study and Design Recommendations
by: Ngo, Tung T., et al.
Published: (2026)
by: Ngo, Tung T., et al.
Published: (2026)
Encoding Version History Context for Better Code Representation
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
AgileCoder: Dynamic Collaborative Agents for Software Development based on Agile Methodology
by: Nguyen, Minh Huynh, et al.
Published: (2024)
by: Nguyen, Minh Huynh, et al.
Published: (2024)
Detecting Malicious Source Code in PyPI Packages with LLMs: Does RAG Come in Handy?
by: Ibiyo, Motunrayo, et al.
Published: (2025)
by: Ibiyo, Motunrayo, et al.
Published: (2025)
Documentation-Guided Agentic Codebase Migration from C to Rust
by: Le-Anh, Minh, et al.
Published: (2026)
by: Le-Anh, Minh, et al.
Published: (2026)
The Effect of Code Obfuscation on Human Program Comprehension
by: Nguyen, Anh H. N., et al.
Published: (2026)
by: Nguyen, Anh H. N., et al.
Published: (2026)
iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML
by: Le, Dat, et al.
Published: (2026)
by: Le, Dat, et al.
Published: (2026)
COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation
by: Dau, Anh T. V., et al.
Published: (2026)
by: Dau, Anh T. V., et al.
Published: (2026)
Do Not Treat Code as Natural Language: Implications for Repository-Level Code Generation and Beyond
by: Le-Anh, Minh, et al.
Published: (2026)
by: Le-Anh, Minh, et al.
Published: (2026)
COBOL-Coder: Domain-Adapted Large Language Models for COBOL Code Generation and Translation
by: Dau, Anh T. V., et al.
Published: (2026)
by: Dau, Anh T. V., et al.
Published: (2026)
Envisioning the Next-Generation AI Coding Assistants: Insights & Proposals
by: Nghiem, Khanh, et al.
Published: (2024)
by: Nghiem, Khanh, et al.
Published: (2024)
When simplicity meets effectiveness: Detecting code comments coherence with word embeddings and LSTM
by: Igbomezie, Michael Dubem, et al.
Published: (2024)
by: Igbomezie, Michael Dubem, et al.
Published: (2024)
Model-Agnostic Correctness Assessment for LLM-Generated Code via Dynamic Internal Representation Selection
by: Vu, Thanh Trong, et al.
Published: (2025)
by: Vu, Thanh Trong, et al.
Published: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
by: Manh, Dung Nguyen, et al.
Published: (2024)
by: Manh, Dung Nguyen, et al.
Published: (2024)
Simplicity by Obfuscation: Evaluating LLM-Driven Code Transformation with Semantic Elasticity
by: De Tomasi, Lorenzo, et al.
Published: (2025)
by: De Tomasi, Lorenzo, et al.
Published: (2025)
An Empirical Study of Multi-Agent RAG for Real-World University Admissions Counseling
by: Nguyen-Duc, Anh, et al.
Published: (2025)
by: Nguyen-Duc, Anh, et al.
Published: (2025)
RBCTest: Leveraging LLMs to Mine and Verify Oracles of API Response Bodies for RESTful API Testing
by: Huynh, Hieu, et al.
Published: (2025)
by: Huynh, Hieu, et al.
Published: (2025)
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
by: Hassid, Michael, et al.
Published: (2024)
by: Hassid, Michael, et al.
Published: (2024)
Leveraging Design-Aware Context in Large Language Models for Code Comment Generation
by: Mitra, Aritra, et al.
Published: (2025)
by: Mitra, Aritra, et al.
Published: (2025)
LEGION: Harnessing Pre-trained Language Models for GitHub Topic Recommendations with Distribution-Balance Loss
by: Dang, Yen-Trang, et al.
Published: (2024)
by: Dang, Yen-Trang, et al.
Published: (2024)
On the Impacts of Contexts on Repository-Level Code Generation
by: Hai, Nam Le, et al.
Published: (2024)
by: Hai, Nam Le, et al.
Published: (2024)
Correctness Assessment of Code Generated by Large Language Models Using Internal Representations
by: Bui, Tuan-Dung, et al.
Published: (2025)
by: Bui, Tuan-Dung, et al.
Published: (2025)
Leveraging Reviewer Experience in Code Review Comment Generation
by: Lin, Hong Yi, et al.
Published: (2024)
by: Lin, Hong Yi, et al.
Published: (2024)
Towards Automated Detection of Inline Code Comment Smells
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
Novice Developers Produce Larger Review Overhead for Project Maintainers while Vibe Coding
by: Asdaque, Syed Ammar, et al.
Published: (2026)
by: Asdaque, Syed Ammar, et al.
Published: (2026)
ROSE: Transformer-Based Refactoring Recommendation for Architectural Smells
by: Nursapa, Samal, et al.
Published: (2025)
by: Nursapa, Samal, et al.
Published: (2025)
XBIDetective: Leveraging Vision Language Models for Identifying Cross-Browser Visual Inconsistencies
by: Grewal, Balreet, et al.
Published: (2025)
by: Grewal, Balreet, et al.
Published: (2025)
CodeFlow: Program Behavior Prediction with Dynamic Dependencies Learning
by: Le, Cuong Chi, et al.
Published: (2024)
by: Le, Cuong Chi, et al.
Published: (2024)
Similar Items
-
DocChecker: Bootstrapping Code Large Language Model for Detecting and Resolving Code-Comment Inconsistencies
by: Dau, Anh T. V., et al.
Published: (2023) -
When Retriever Meets Generator: A Joint Model for Code Comment Generation
by: Le, Tien P. T., et al.
Published: (2025) -
Detection of Technical Debt in Java Source Code
by: Hai, Nam Le, et al.
Published: (2024) -
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
by: Le, Huy, et al.
Published: (2025) -
Good things come in three: Generating SO Post Titles with Pre-Trained Models, Self Improvement and Post Ranking
by: Le, Duc Anh, et al.
Published: (2024)