Artificial or Just Artful? Do LLMs Bend the Rules in Programming?
Fuente:
arXiv
Saved in:
| Main Authors: | Sghaier, Oussama Ben, Delcourt, Kevin, Sahraoui, Houari |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving the Learning of Code Review Successive Tasks with Cross-Task Knowledge Distillation
by: Sghaier, Oussama Ben, et al.
Published: (2024)
by: Sghaier, Oussama Ben, et al.
Published: (2024)
Harnessing Large Language Models for Curated Code Reviews
by: Sghaier, Oussama Ben, et al.
Published: (2025)
by: Sghaier, Oussama Ben, et al.
Published: (2025)
Combining Large Language Models with Static Analyzers for Code Review Generation
by: Jaoua, Imen, et al.
Published: (2025)
by: Jaoua, Imen, et al.
Published: (2025)
Leveraging Reward Models for Guiding Code Review Comment Generation
by: Sghaier, Oussama Ben, et al.
Published: (2025)
by: Sghaier, Oussama Ben, et al.
Published: (2025)
Automation in Model-Driven Engineering: A look back, and ahead
by: Burgueño, Lola, et al.
Published: (2024)
by: Burgueño, Lola, et al.
Published: (2024)
MONO2REST: Identifying and Exposing Microservices: a Reusable RESTification Approach
by: Lecrivain, Matthéo, et al.
Published: (2025)
by: Lecrivain, Matthéo, et al.
Published: (2025)
On the synchronization between Hugging Face pre-trained language models and their upstream GitHub repository
by: Ajibode, Adekunle, et al.
Published: (2025)
by: Ajibode, Adekunle, et al.
Published: (2025)
On the Utility of Domain Modeling Assistance with Large Language Models
by: Chaaben, Meriem Ben, et al.
Published: (2024)
by: Chaaben, Meriem Ben, et al.
Published: (2024)
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences
by: Weyssow, Martin, et al.
Published: (2024)
by: Weyssow, Martin, et al.
Published: (2024)
On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code
by: Weyssow, Martin, et al.
Published: (2023)
by: Weyssow, Martin, et al.
Published: (2023)
Modeling Sampling Workflows for Code Repositories
by: Lefeuvre, Romain, et al.
Published: (2026)
by: Lefeuvre, Romain, et al.
Published: (2026)
Exploring Parameter-Efficient Fine-Tuning Techniques for Code Generation with Large Language Models
by: Weyssow, Martin, et al.
Published: (2023)
by: Weyssow, Martin, et al.
Published: (2023)
RuleFlow : Generating Reusable Program Optimizations with LLMs
by: Singh, Avaljot, et al.
Published: (2026)
by: Singh, Avaljot, et al.
Published: (2026)
An Exploratory Study on Just-in-Time Multi-Programming-Language Bug Prediction
by: Li, Zengyang, et al.
Published: (2024)
by: Li, Zengyang, et al.
Published: (2024)
Do Code LLMs Do Static Analysis?
by: Su, Chia-Yi, et al.
Published: (2025)
by: Su, Chia-Yi, et al.
Published: (2025)
Beyond Rules: LLM-Powered Linting for Quantum Programs
by: Cassieri, Pietro, et al.
Published: (2026)
by: Cassieri, Pietro, et al.
Published: (2026)
Can LLMs Solve Science or Just Write Code? Evaluating Quantum Solver Generation
by: Baresi, Luciano, et al.
Published: (2026)
by: Baresi, Luciano, et al.
Published: (2026)
Test Plan Generation for Live Testing of Cloud Services
by: Jebbar, Oussama, et al.
Published: (2025)
by: Jebbar, Oussama, et al.
Published: (2025)
When LLMs Meet API Documentation: Can Retrieval Augmentation Aid Code Generation Just as It Helps Developers?
by: Chen, Jingyi, et al.
Published: (2025)
by: Chen, Jingyi, et al.
Published: (2025)
Combining Example-Based and Rule-Based Program Transformations to Resolve Build Conflicts
by: Towqir, Sheikh Shadab, et al.
Published: (2025)
by: Towqir, Sheikh Shadab, et al.
Published: (2025)
Experimenting a New Programming Practice with LLMs
by: Zhang, Simiao, et al.
Published: (2024)
by: Zhang, Simiao, et al.
Published: (2024)
Galapagos: Automated N-Version Programming with LLMs
by: Ron, Javier, et al.
Published: (2024)
by: Ron, Javier, et al.
Published: (2024)
Toward Realistic Evaluations of Just-In-Time Vulnerability Prediction
by: Nguyen, Duong, et al.
Published: (2025)
by: Nguyen, Duong, et al.
Published: (2025)
Do RESTful API Design Rules Have an Impact on the Understandability of Web APIs? A Web-Based Experiment with API Descriptions
by: Bogner, Justus, et al.
Published: (2023)
by: Bogner, Justus, et al.
Published: (2023)
Integrating Symbolic Execution with LLMs for Automated Generation of Program Specifications
by: Yang, Fanpeng, et al.
Published: (2025)
by: Yang, Fanpeng, et al.
Published: (2025)
StepGrade: Grading Programming Assignments with Context-Aware LLMs
by: Akyash, Mohammad, et al.
Published: (2025)
by: Akyash, Mohammad, et al.
Published: (2025)
Rethinking Kernel Program Repair: Benchmarking and Enhancing LLMs with RGym
by: Shehada, Kareem, et al.
Published: (2025)
by: Shehada, Kareem, et al.
Published: (2025)
Do LLMs generate test oracles that capture the actual or the expected program behaviour?
by: Konstantinou, Michael, et al.
Published: (2024)
by: Konstantinou, Michael, et al.
Published: (2024)
From LLMs to Agents in Programming: The Impact of Providing an LLM with a Compiler
by: Kjellberg, Viktor, et al.
Published: (2026)
by: Kjellberg, Viktor, et al.
Published: (2026)
DePro: Understanding the Role of LLMs in Debugging Competitive Programming Code
by: Parvez, Nabiha, et al.
Published: (2026)
by: Parvez, Nabiha, et al.
Published: (2026)
Enhancing LLMs in Long Code Translation through Instrumentation and Program State Alignment
by: Xin-Ye, Li, et al.
Published: (2025)
by: Xin-Ye, Li, et al.
Published: (2025)
Can LLMs Recover Program Semantics? A Systematic Evaluation with Symbolic Execution
by: Feng, Rong, et al.
Published: (2025)
by: Feng, Rong, et al.
Published: (2025)
PALM: Synergizing Program Analysis and LLMs to Enhance Rust Unit Test Coverage
by: Chu, Bei, et al.
Published: (2025)
by: Chu, Bei, et al.
Published: (2025)
Quantum Program Linting with LLMs: Emerging Results from a Comparative Study
by: Shin, Seung Yeob, et al.
Published: (2025)
by: Shin, Seung Yeob, et al.
Published: (2025)
Programming Language Confusion: When Code LLMs Can't Keep their Languages Straight
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
ScratchEval : A Multimodal Evaluation Framework for LLMs in Block-Based Programming
by: Si, Yuan, et al.
Published: (2026)
by: Si, Yuan, et al.
Published: (2026)
Codellm-Devkit: A Framework for Contextualizing Code LLMs with Program Analysis Insights
by: Krishna, Rahul, et al.
Published: (2024)
by: Krishna, Rahul, et al.
Published: (2024)
ReDef: Do Code Language Models Truly Understand Code Changes for Just-in-Time Software Defect Prediction?
by: Nam, Doha, et al.
Published: (2025)
by: Nam, Doha, et al.
Published: (2025)
Disproving Program Equivalence with LLMs
by: Allamanis, Miltiadis, et al.
Published: (2025)
by: Allamanis, Miltiadis, et al.
Published: (2025)
SmartNote: An LLM-Powered, Personalised Release Note Generator That Just Works
by: Daneshyan, Farbod, et al.
Published: (2025)
by: Daneshyan, Farbod, et al.
Published: (2025)
Similar Items
-
Improving the Learning of Code Review Successive Tasks with Cross-Task Knowledge Distillation
by: Sghaier, Oussama Ben, et al.
Published: (2024) -
Harnessing Large Language Models for Curated Code Reviews
by: Sghaier, Oussama Ben, et al.
Published: (2025) -
Combining Large Language Models with Static Analyzers for Code Review Generation
by: Jaoua, Imen, et al.
Published: (2025) -
Leveraging Reward Models for Guiding Code Review Comment Generation
by: Sghaier, Oussama Ben, et al.
Published: (2025) -
Automation in Model-Driven Engineering: A look back, and ahead
by: Burgueño, Lola, et al.
Published: (2024)