Gespeichert in:
| Hauptverfasser: | Rodriguez-Cardenas, Daniel, Velasco, Alejandro, Poshyvanyk, Denys |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.07046 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards More Trustworthy and Interpretable LLMs for Code through Syntax-Grounded Explanations
von: Palacio, David N., et al.
Veröffentlicht: (2024)
von: Palacio, David N., et al.
Veröffentlicht: (2024)
Toward a Theory of Causation for Interpreting Neural Code Models
von: Palacio, David N., et al.
Veröffentlicht: (2023)
von: Palacio, David N., et al.
Veröffentlicht: (2023)
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
von: Khati, Dipin, et al.
Veröffentlicht: (2026)
von: Khati, Dipin, et al.
Veröffentlicht: (2026)
Tricky$^2$: Towards a Benchmark for Evaluating Human and LLM Error Interactions
von: Granger, Cole, et al.
Veröffentlicht: (2026)
von: Granger, Cole, et al.
Veröffentlicht: (2026)
Understanding Privacy Risks in Code Models Through Training Dynamics: A Causal Approach
von: Yang, Hua, et al.
Veröffentlicht: (2025)
von: Yang, Hua, et al.
Veröffentlicht: (2025)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
von: Yang, Hua, et al.
Veröffentlicht: (2025)
von: Yang, Hua, et al.
Veröffentlicht: (2025)
On Interpreting the Effectiveness of Unsupervised Software Traceability with Information Theory
von: Palacio, David N., et al.
Veröffentlicht: (2024)
von: Palacio, David N., et al.
Veröffentlicht: (2024)
Which Syntactic Capabilities Are Statistically Learned by Masked Language Models for Code?
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024)
Toward Explaining Large Language Models in Software Engineering Tasks
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
Toward Neurosymbolic Program Comprehension
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
von: Rodriguez-Cardenas, Daniel, et al.
Veröffentlicht: (2026)
von: Rodriguez-Cardenas, Daniel, et al.
Veröffentlicht: (2026)
On The Importance of Reasoning for Context Retrieval in Repository-Level Code Editing
von: Kovrigin, Alexander, et al.
Veröffentlicht: (2024)
von: Kovrigin, Alexander, et al.
Veröffentlicht: (2024)
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We?
von: O'Brien, Conor, et al.
Veröffentlicht: (2024)
von: O'Brien, Conor, et al.
Veröffentlicht: (2024)
Evaluating the Use of LLMs for Documentation to Code Traceability
von: Alor, Ebube, et al.
Veröffentlicht: (2025)
von: Alor, Ebube, et al.
Veröffentlicht: (2025)
Mapping the Trust Terrain: LLMs in Software Engineering -- Insights and Perspectives
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
von: Khati, Dipin, et al.
Veröffentlicht: (2025)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
von: Rahman, Imranur, et al.
Veröffentlicht: (2025)
A Causal Perspective on Measuring, Explaining and Mitigating Smells in LLM-Generated Code
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
von: Velasco, Alejandro, et al.
Veröffentlicht: (2025)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
von: Vulićević, Jelena Ilić
Veröffentlicht: (2026)
von: Vulićević, Jelena Ilić
Veröffentlicht: (2026)
LiCoEval: Evaluating LLMs on License Compliance in Code Generation
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
A Comprehensive Framework for Evaluating API-oriented Code Generation in Large Language Models
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
On LLMs' Internal Representation of Code Correctness
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
Operational Robustness of LLMs on Code Generation
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
von: Thillen, Alex, et al.
Veröffentlicht: (2026)
von: Thillen, Alex, et al.
Veröffentlicht: (2026)
Automating Code Adaptation for MLOps -- A Benchmarking Study on LLMs
von: Patel, Harsh, et al.
Veröffentlicht: (2024)
von: Patel, Harsh, et al.
Veröffentlicht: (2024)
Drawing Pandas: A Benchmark for LLMs in Generating Plotting Code
von: Galimzyanov, Timur, et al.
Veröffentlicht: (2024)
von: Galimzyanov, Timur, et al.
Veröffentlicht: (2024)
"Don't Be Afraid, Just Learn": Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI
von: Otten, Daniel, et al.
Veröffentlicht: (2026)
von: Otten, Daniel, et al.
Veröffentlicht: (2026)
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
GREPO: A Benchmark for Graph Neural Networks on Repository-Level Bug Localization
von: Wang, Juntong, et al.
Veröffentlicht: (2026)
von: Wang, Juntong, et al.
Veröffentlicht: (2026)
DRAGON: Robust Classification for Very Large Collections of Software Repositories
von: Balla, Stefano, et al.
Veröffentlicht: (2026)
von: Balla, Stefano, et al.
Veröffentlicht: (2026)
Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering
von: Abedu, Samuel, et al.
Veröffentlicht: (2024)
von: Abedu, Samuel, et al.
Veröffentlicht: (2024)
Free and Customizable Code Documentation with LLMs: A Fine-Tuning Approach
von: Chakrabarty, Sayak, et al.
Veröffentlicht: (2024)
von: Chakrabarty, Sayak, et al.
Veröffentlicht: (2024)
LLMs in Coding and their Impact on the Commercial Software Engineering Landscape
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
von: Belozerov, Vladislav, et al.
Veröffentlicht: (2025)
Protocode: Prototype-Driven Interpretability for Code Generation in LLMs
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
The Struggles of LLMs in Cross-lingual Code Clone Detection
von: Moumoula, Micheline Bénédicte, et al.
Veröffentlicht: (2024)
von: Moumoula, Micheline Bénédicte, et al.
Veröffentlicht: (2024)
Evaluation of LLMs on Syntax-Aware Code Fill-in-the-Middle Tasks
von: Gong, Linyuan, et al.
Veröffentlicht: (2024)
von: Gong, Linyuan, et al.
Veröffentlicht: (2024)
Lessons Learned: A Multi-Agent Framework for Code LLMs to Learn and Improve
von: Liu, Yuanzhe, et al.
Veröffentlicht: (2025)
von: Liu, Yuanzhe, et al.
Veröffentlicht: (2025)
LLM-based Content Classification Approach for GitHub Repositories by the README Files
von: Mehmood, Malik Uzair, et al.
Veröffentlicht: (2025)
von: Mehmood, Malik Uzair, et al.
Veröffentlicht: (2025)
CSR-Bench: Benchmarking LLM Agents in Deployment of Computer Science Research Repositories
von: Xiao, Yijia, et al.
Veröffentlicht: (2025)
von: Xiao, Yijia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards More Trustworthy and Interpretable LLMs for Code through Syntax-Grounded Explanations
von: Palacio, David N., et al.
Veröffentlicht: (2024) -
Toward a Theory of Causation for Interpreting Neural Code Models
von: Palacio, David N., et al.
Veröffentlicht: (2023) -
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
von: Khati, Dipin, et al.
Veröffentlicht: (2025) -
How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
von: Velasco, Alejandro, et al.
Veröffentlicht: (2024) -
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
von: Khati, Dipin, et al.
Veröffentlicht: (2026)