Patched RTC: evaluating LLMs for diverse software development tasks
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Sharma, Asankhaya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Patched MOA: optimizing inference for diverse software development tasks
von: Sharma, Asankhaya
Veröffentlicht: (2024)
von: Sharma, Asankhaya
Veröffentlicht: (2024)
PELLI: Framework to effectively integrate LLMs for quality software generation
von: Krebs, Rasmus, et al.
Veröffentlicht: (2026)
von: Krebs, Rasmus, et al.
Veröffentlicht: (2026)
Exploring the extent of similarities in software failures across industries using LLMs
von: Detloff, Martin
Veröffentlicht: (2024)
von: Detloff, Martin
Veröffentlicht: (2024)
Supporting architecture evaluation for ATAM scenarios with LLMs
von: Capilla, Rafael, et al.
Veröffentlicht: (2025)
von: Capilla, Rafael, et al.
Veröffentlicht: (2025)
Empirical evaluation of LLMs in predicting fixes of Configuration bugs in Smart Home System
von: Monisha, Sheikh Moonwara Anjum, et al.
Veröffentlicht: (2025)
von: Monisha, Sheikh Moonwara Anjum, et al.
Veröffentlicht: (2025)
Evaluating LLMs for One-Shot Patching of Real and Artificial Vulnerabilities
von: Garg, Aayush, et al.
Veröffentlicht: (2025)
von: Garg, Aayush, et al.
Veröffentlicht: (2025)
PatchFinder: A Two-Phase Approach to Security Patch Tracing for Disclosed Vulnerabilities in Open-Source Software
von: Li, Kaixuan, et al.
Veröffentlicht: (2024)
von: Li, Kaixuan, et al.
Veröffentlicht: (2024)
RiskBridge: Turning CVEs into Business-Aligned Patch Priorities
von: Sheikh, Yelena Mujibur, et al.
Veröffentlicht: (2026)
von: Sheikh, Yelena Mujibur, et al.
Veröffentlicht: (2026)
Repository-Aware File Path Retrieval via Fine-Tuned LLMs
von: Yanuganti, Vasudha, et al.
Veröffentlicht: (2025)
von: Yanuganti, Vasudha, et al.
Veröffentlicht: (2025)
From Patches to Trajectories: Privileged Process Supervision for Software-Engineering Agents
von: Ma, Murong, et al.
Veröffentlicht: (2026)
von: Ma, Murong, et al.
Veröffentlicht: (2026)
Learning From Developers: Towards Reliable Patch Validation at Scale for Linux
von: Lin, Chih-En, et al.
Veröffentlicht: (2026)
von: Lin, Chih-En, et al.
Veröffentlicht: (2026)
Algorithmic algorithm development with LLMs: A Case Study on LLM-Usage for Contraction Order Optimization in Tensor Networks
von: Hoppe, Fabian, et al.
Veröffentlicht: (2026)
von: Hoppe, Fabian, et al.
Veröffentlicht: (2026)
Agentic AI-assisted coding offers a unique opportunity to instill epistemic grounding during software development
von: Palmblad, Magnus, et al.
Veröffentlicht: (2026)
von: Palmblad, Magnus, et al.
Veröffentlicht: (2026)
Assessing LLM code generation quality through path planning tasks
von: Chen, Wanyi, et al.
Veröffentlicht: (2025)
von: Chen, Wanyi, et al.
Veröffentlicht: (2025)
Repeton: Structured Bug Repair with ReAct-Guided Patch-and-Test Cycles
von: Vinh, Nguyen Phu, et al.
Veröffentlicht: (2025)
von: Vinh, Nguyen Phu, et al.
Veröffentlicht: (2025)
PyBench: Evaluating LLM Agent on various real-world coding tasks
von: Zhang, Yaolun, et al.
Veröffentlicht: (2024)
von: Zhang, Yaolun, et al.
Veröffentlicht: (2024)
The importance of visual modelling languages in generative software engineering
von: Rossi, Roberto
Veröffentlicht: (2024)
von: Rossi, Roberto
Veröffentlicht: (2024)
Memory-Efficient Large Language Models for Program Repair with Semantic-Guided Patch Generation
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2024)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2024)
InfCode: Adversarial Iterative Refinement of Tests and Patches for Reliable Software Issue Resolution
von: Li, KeFan, et al.
Veröffentlicht: (2025)
von: Li, KeFan, et al.
Veröffentlicht: (2025)
VulnLLMEval: A Framework for Evaluating Large Language Models in Software Vulnerability Detection and Patching
von: Zibaeirad, Arastoo, et al.
Veröffentlicht: (2024)
von: Zibaeirad, Arastoo, et al.
Veröffentlicht: (2024)
Automating Patch Set Generation from Code Review Comments Using Large Language Models
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval
von: Liu, Ye, et al.
Veröffentlicht: (2024)
von: Liu, Ye, et al.
Veröffentlicht: (2024)
RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment
von: Fuster-Pena, Marcos, et al.
Veröffentlicht: (2025)
von: Fuster-Pena, Marcos, et al.
Veröffentlicht: (2025)
LLMs: A Game-Changer for Software Engineers?
von: Haque, Md Asraful
Veröffentlicht: (2024)
von: Haque, Md Asraful
Veröffentlicht: (2024)
CoCo-Bench: A Comprehensive Code Benchmark For Multi-task Large Language Model Evaluation
von: Yin, Wenjing, et al.
Veröffentlicht: (2025)
von: Yin, Wenjing, et al.
Veröffentlicht: (2025)
Artificial intelligence for context-aware visual change detection in software test automation
von: Moradi, Milad, et al.
Veröffentlicht: (2024)
von: Moradi, Milad, et al.
Veröffentlicht: (2024)
Document Retrieval Augmented Fine-Tuning (DRAFT) for safety-critical software assessments
von: Bolton, Regan, et al.
Veröffentlicht: (2025)
von: Bolton, Regan, et al.
Veröffentlicht: (2025)
Evaluating Software Development Agents: Patch Patterns, Code Quality, and Issue Complexity in Real-World GitHub Scenarios
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
Co-PatcheR: Collaborative Software Patching with Component(s)-specific Small Reasoning Models
von: Tang, Yuheng, et al.
Veröffentlicht: (2025)
von: Tang, Yuheng, et al.
Veröffentlicht: (2025)
Private GPTs for LLM-driven testing in software development and machine learning
von: Jagielski, Jakub, et al.
Veröffentlicht: (2025)
von: Jagielski, Jakub, et al.
Veröffentlicht: (2025)
CMSA algorithm for solving the prioritized pairwise test data generation problem in software product lines
von: Ferrer, Javier, et al.
Veröffentlicht: (2024)
von: Ferrer, Javier, et al.
Veröffentlicht: (2024)
Sentiment analysis for software engineering: How far can zero-shot learning (ZSL) go?
von: Alfayez, Reem, et al.
Veröffentlicht: (2026)
von: Alfayez, Reem, et al.
Veröffentlicht: (2026)
ContextCov: Deriving and Enforcing Executable Constraints from Agent Instruction Files
von: Sharma, Reshabh K
Veröffentlicht: (2026)
von: Sharma, Reshabh K
Veröffentlicht: (2026)
Evaluating LLMs for Visualization Tasks
von: Khan, Saadiq Rauf, et al.
Veröffentlicht: (2025)
von: Khan, Saadiq Rauf, et al.
Veröffentlicht: (2025)
An explainable hybrid deep learning-enabled intelligent fault detection and diagnosis approach for automotive software systems validation
von: Abboush, Mohammad, et al.
Veröffentlicht: (2026)
von: Abboush, Mohammad, et al.
Veröffentlicht: (2026)
Machine Learning Experiences: A story of learning AI for use in enterprise software testing that can be used by anyone
von: Cohoon, Michael, et al.
Veröffentlicht: (2025)
von: Cohoon, Michael, et al.
Veröffentlicht: (2025)
On the Effectiveness of LLMs for Manual Test Verifications
von: Peixoto, Myron David Lucena Campos, et al.
Veröffentlicht: (2024)
von: Peixoto, Myron David Lucena Campos, et al.
Veröffentlicht: (2024)
Are LLMs Correctly Integrated into Software Systems?
von: Shao, Yuchen, et al.
Veröffentlicht: (2024)
von: Shao, Yuchen, et al.
Veröffentlicht: (2024)
Generating Energy-efficient code with LLMs
von: Cappendijk, Tom, et al.
Veröffentlicht: (2024)
von: Cappendijk, Tom, et al.
Veröffentlicht: (2024)
Analysis on LLMs Performance for Code Summarization
von: Akib, Md. Ahnaf, et al.
Veröffentlicht: (2024)
von: Akib, Md. Ahnaf, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Patched MOA: optimizing inference for diverse software development tasks
von: Sharma, Asankhaya
Veröffentlicht: (2024) -
PELLI: Framework to effectively integrate LLMs for quality software generation
von: Krebs, Rasmus, et al.
Veröffentlicht: (2026) -
Exploring the extent of similarities in software failures across industries using LLMs
von: Detloff, Martin
Veröffentlicht: (2024) -
Supporting architecture evaluation for ATAM scenarios with LLMs
von: Capilla, Rafael, et al.
Veröffentlicht: (2025) -
Empirical evaluation of LLMs in predicting fixes of Configuration bugs in Smart Home System
von: Monisha, Sheikh Moonwara Anjum, et al.
Veröffentlicht: (2025)