Understanding Robustness of Model Editing in Code LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Chhetri, Vinaik, Fereidouni, Moghis, Siddique, A. B, Farooq, Umar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Analyzing the Evolution and Maintenance of Quantum Software Repositories
by: Upadhyay, Krishna, et al.
Published: (2025)
by: Upadhyay, Krishna, et al.
Published: (2025)
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
by: Chhetri, Vinaik, et al.
Published: (2025)
by: Chhetri, Vinaik, et al.
Published: (2025)
A Framework for Generating Conversational Recommendation Datasets from Behavioral Interactions
by: Chhetri, Vinaik, et al.
Published: (2025)
by: Chhetri, Vinaik, et al.
Published: (2025)
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
by: Liu, Daniel, et al.
Published: (2026)
by: Liu, Daniel, et al.
Published: (2026)
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
by: Fakorede, Moshood A., et al.
Published: (2026)
by: Fakorede, Moshood A., et al.
Published: (2026)
MobileConvRec: A Conversational Dataset for Mobile Apps Recommendations
by: Maji, Srijata, et al.
Published: (2024)
by: Maji, Srijata, et al.
Published: (2024)
Looking into Black Box Code Language Models
by: Haider, Muhammad Umair, et al.
Published: (2024)
by: Haider, Muhammad Umair, et al.
Published: (2024)
INTERPOS: Interaction Rhythm Guided Positional Morphing for Mobile App Recommender Systems
by: Maqbool, M. H., et al.
Published: (2025)
by: Maqbool, M. H., et al.
Published: (2025)
Operational Robustness of LLMs on Code Generation
by: Paul, Debalina Ghosh, et al.
Published: (2026)
by: Paul, Debalina Ghosh, et al.
Published: (2026)
Applying the Chinese Wall Reverse Engineering Technique to Large Language Model Code Editing
by: Hanmongkolchai, Manatsawin
Published: (2025)
by: Hanmongkolchai, Manatsawin
Published: (2025)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
by: Wang, Yan, et al.
Published: (2025)
by: Wang, Yan, et al.
Published: (2025)
Teaching Code Refactoring Using LLMs
by: Khairnar, Anshul, et al.
Published: (2025)
by: Khairnar, Anshul, et al.
Published: (2025)
Towards Verified Code Reasoning by LLMs
by: Sistla, Meghana, et al.
Published: (2025)
by: Sistla, Meghana, et al.
Published: (2025)
How Robustly do LLMs Understand Execution Semantics?
by: Spiess, Claudio, et al.
Published: (2026)
by: Spiess, Claudio, et al.
Published: (2026)
OSS-Bench: Benchmark Generator for Coding LLMs
by: Jiang, Yuancheng, et al.
Published: (2025)
by: Jiang, Yuancheng, et al.
Published: (2025)
Robust Learning of Diverse Code Edits
by: Aggarwal, Tushar, et al.
Published: (2025)
by: Aggarwal, Tushar, et al.
Published: (2025)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
by: Xia, Yunhui, et al.
Published: (2025)
by: Xia, Yunhui, et al.
Published: (2025)
Unsupervised Evaluation of Code LLMs with Round-Trip Correctness
by: Allamanis, Miltiadis, et al.
Published: (2024)
by: Allamanis, Miltiadis, et al.
Published: (2024)
Themis: Training Robust Multilingual Code Reward Models for Flexible Multi-Criteria Scoring
by: Paul, Indraneil, et al.
Published: (2026)
by: Paul, Indraneil, et al.
Published: (2026)
EditLord: Learning Code Transformation Rules for Code Editing
by: Li, Weichen, et al.
Published: (2025)
by: Li, Weichen, et al.
Published: (2025)
Mechanistic Interpretability of Code Correctness in LLMs via Sparse Autoencoders
by: Tahimic, Kriz, et al.
Published: (2025)
by: Tahimic, Kriz, et al.
Published: (2025)
PromSec: Prompt Optimization for Secure Generation of Functional Source Code with Large Language Models (LLMs)
by: Nazzal, Mahmoud, et al.
Published: (2024)
by: Nazzal, Mahmoud, et al.
Published: (2024)
TritonRL: Training LLMs to Think and Code Triton Without Cheating
by: Woo, Jiin, et al.
Published: (2025)
by: Woo, Jiin, et al.
Published: (2025)
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs
by: Haque, Mirazul, et al.
Published: (2025)
by: Haque, Mirazul, et al.
Published: (2025)
K-ASTRO: Structure-Aware Adaptation of LLMs for Code Vulnerability Detection
by: Zhang, Yifan, et al.
Published: (2022)
by: Zhang, Yifan, et al.
Published: (2022)
Where Do LLMs Still Struggle? An In-Depth Analysis of Code Generation Benchmarks
by: Sharifloo, Amir Molzam, et al.
Published: (2025)
by: Sharifloo, Amir Molzam, et al.
Published: (2025)
Leveraging LLMs for Legacy Code Modernization: Challenges and Opportunities for LLM-Generated Documentation
by: Diggs, Colin, et al.
Published: (2024)
by: Diggs, Colin, et al.
Published: (2024)
Towards Understanding What Code Language Models Learned
by: Ahmed, Toufique, et al.
Published: (2023)
by: Ahmed, Toufique, et al.
Published: (2023)
CoCoNUT: Structural Code Understanding does not fall out of a tree
by: Beger, Claas, et al.
Published: (2025)
by: Beger, Claas, et al.
Published: (2025)
On The Importance of Reasoning for Context Retrieval in Repository-Level Code Editing
by: Kovrigin, Alexander, et al.
Published: (2024)
by: Kovrigin, Alexander, et al.
Published: (2024)
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
by: Khati, Dipin, et al.
Published: (2025)
by: Khati, Dipin, et al.
Published: (2025)
Understanding and Detecting Platform-Specific Violations in Android Auto Apps
by: Fakorede, Moshood, et al.
Published: (2025)
by: Fakorede, Moshood, et al.
Published: (2025)
Can LLMs Find Bugs in Code? An Evaluation from Beginner Errors to Security Vulnerabilities in Python and C++
by: Mhatre, Akshay, et al.
Published: (2025)
by: Mhatre, Akshay, et al.
Published: (2025)
Renaissance of Literate Programming in the Era of LLMs: Enhancing LLM-Based Code Generation in Large-Scale Projects
by: Zhang, Wuyang, et al.
Published: (2024)
by: Zhang, Wuyang, et al.
Published: (2024)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
by: Yan, Kaiwen, et al.
Published: (2025)
by: Yan, Kaiwen, et al.
Published: (2025)
Calibration and Correctness of Language Models for Code
by: Spiess, Claudio, et al.
Published: (2024)
by: Spiess, Claudio, et al.
Published: (2024)
CodeEditorBench: Evaluating Code Editing Capability of Large Language Models
by: Guo, Jiawei, et al.
Published: (2024)
by: Guo, Jiawei, et al.
Published: (2024)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
Assessing and Enhancing Quantum Readiness in Mobile Apps
by: Strauss, Joseph, et al.
Published: (2025)
by: Strauss, Joseph, et al.
Published: (2025)
On LLMs' Internal Representation of Code Correctness
by: Ribeiro, Francisco, et al.
Published: (2025)
by: Ribeiro, Francisco, et al.
Published: (2025)
Similar Items
-
Analyzing the Evolution and Maintenance of Quantum Software Repositories
by: Upadhyay, Krishna, et al.
Published: (2025) -
What Users Value and Critique: Large-Scale Analysis of User Feedback on AI-Powered Mobile Apps
by: Chhetri, Vinaik, et al.
Published: (2025) -
A Framework for Generating Conversational Recommendation Datasets from Behavioral Interactions
by: Chhetri, Vinaik, et al.
Published: (2025) -
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
by: Liu, Daniel, et al.
Published: (2026) -
MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development
by: Fakorede, Moshood A., et al.
Published: (2026)