From Helpful to Trustworthy: LLM Agents for Pair Programming
Fuente:
arXiv
Saved in:
| Main Author: | Ayon, Ragib Shahariar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpecPylot: Python Specification Generation using Large Language Models
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
AutoReSpec: A Framework for Generating Specification using Large Language Models
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
by: Bouzenia, Islem, et al.
Published: (2024)
by: Bouzenia, Islem, et al.
Published: (2024)
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
by: Chakroborti, Apu Kumar, et al.
Published: (2025)
by: Chakroborti, Apu Kumar, et al.
Published: (2025)
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
by: Islam, Niful, et al.
Published: (2026)
by: Islam, Niful, et al.
Published: (2026)
Facilitating Trustworthy Human-Agent Collaboration in LLM-based Multi-Agent System oriented Software Engineering
by: Ronanki, Krishna
Published: (2025)
by: Ronanki, Krishna
Published: (2025)
From User Interface to Agent Interface: Efficiency Optimization of UI Representations for LLM Agents
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
Identifying Helpful Context for LLM-based Vulnerability Repair: A Preliminary Study
by: Antal, Gábor, et al.
Published: (2025)
by: Antal, Gábor, et al.
Published: (2025)
Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?
by: Gloaguen, Thibaud, et al.
Published: (2026)
by: Gloaguen, Thibaud, et al.
Published: (2026)
Reflection-Driven Control for Trustworthy Code Agents
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?
by: Garg, Spandan, et al.
Published: (2026)
by: Garg, Spandan, et al.
Published: (2026)
SWE-Skills-Bench: Do Agent Skills Actually Help in Real-World Software Engineering?
by: Han, Tingxu, et al.
Published: (2026)
by: Han, Tingxu, et al.
Published: (2026)
HAFixAgent: History-Aware Program Repair Agent
by: Shi, Yu, et al.
Published: (2025)
by: Shi, Yu, et al.
Published: (2025)
Does Code Cleanliness Affect Coding Agents? A Controlled Minimal-Pair Study
by: Trivedi, Priyansh, et al.
Published: (2026)
by: Trivedi, Priyansh, et al.
Published: (2026)
A Pair Programming Framework for Code Generation via Multi-Plan Exploration and Feedback-Driven Refinement
by: Zhang, Huan, et al.
Published: (2024)
by: Zhang, Huan, et al.
Published: (2024)
Help Without Being Asked: A Deployed Proactive Agent System for On-Call Support with Continuous Self-Improvement
by: Liu, Fengrui, et al.
Published: (2026)
by: Liu, Fengrui, et al.
Published: (2026)
From Flat Logs to Causal Graphs: Hierarchical Failure Attribution for LLM-based Multi-Agent Systems
by: Wang, Yawen, et al.
Published: (2026)
by: Wang, Yawen, et al.
Published: (2026)
Collaborative Agents for Automated Program Repair in Ruby
by: Akbarpour, Nikta, et al.
Published: (2025)
by: Akbarpour, Nikta, et al.
Published: (2025)
Evaluating Agent-based Program Repair at Google
by: Rondon, Pat, et al.
Published: (2025)
by: Rondon, Pat, et al.
Published: (2025)
ProgramBench: Can Language Models Rebuild Programs From Scratch?
by: Yang, John, et al.
Published: (2026)
by: Yang, John, et al.
Published: (2026)
STELP: Secure Transpilation and Execution of LLM-Generated Programs
by: Shinde, Swapnil, et al.
Published: (2026)
by: Shinde, Swapnil, et al.
Published: (2026)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
by: Wang, Xiaoyin, et al.
Published: (2024)
by: Wang, Xiaoyin, et al.
Published: (2024)
Reducing Cost of LLM Agents with Trajectory Reduction
by: Xiao, Yuan-An, et al.
Published: (2025)
by: Xiao, Yuan-An, et al.
Published: (2025)
LLM Collaboration With Multi-Agent Reinforcement Learning
by: Liu, Shuo, et al.
Published: (2025)
by: Liu, Shuo, et al.
Published: (2025)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
by: Chen, Zhi, et al.
Published: (2026)
by: Chen, Zhi, et al.
Published: (2026)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
by: Trae Research Team, et al.
Published: (2025)
by: Trae Research Team, et al.
Published: (2025)
On the Role of Fault Localization Context for LLM-Based Program Repair
by: Sepidband, Melika, et al.
Published: (2026)
by: Sepidband, Melika, et al.
Published: (2026)
LLM Test Generation via Iterative Hybrid Program Analysis
by: Gu, Sijia, et al.
Published: (2025)
by: Gu, Sijia, et al.
Published: (2025)
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming
by: Agarwal, Anisha, et al.
Published: (2024)
by: Agarwal, Anisha, et al.
Published: (2024)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
by: Gao, Shuzheng, et al.
Published: (2025)
by: Gao, Shuzheng, et al.
Published: (2025)
Using an LLM to Help With Code Understanding
by: Nam, Daye, et al.
Published: (2023)
by: Nam, Daye, et al.
Published: (2023)
AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering
by: Kumar, Rajesh, et al.
Published: (2026)
by: Kumar, Rajesh, et al.
Published: (2026)
Stop Comparing LLM Agents Without Disclosing the Harness
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
Evolving Excellence: Automated Optimization of LLM-based Agents
by: Brookes, Paul, et al.
Published: (2025)
by: Brookes, Paul, et al.
Published: (2025)
Static Program Analysis Guided LLM Based Unit Test Generation
by: Roychowdhury, Sujoy, et al.
Published: (2025)
by: Roychowdhury, Sujoy, et al.
Published: (2025)
Towards Secure Program Partitioning for Smart Contracts with LLM's In-Context Learning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
Test It Before You Trust It: Applying Software Testing for Trustworthy In-context Learning
by: Racharak, Teeradaj, et al.
Published: (2025)
by: Racharak, Teeradaj, et al.
Published: (2025)
LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering
by: Qiu, Jielin, et al.
Published: (2025)
by: Qiu, Jielin, et al.
Published: (2025)
Constraint Decay: The Fragility of LLM Agents in Backend Code Generation
by: Dente, Francesco, et al.
Published: (2026)
by: Dente, Francesco, et al.
Published: (2026)
An Empirical Study on LLM-based Agents for Automated Bug Fixing
by: Meng, Xiangxin, et al.
Published: (2024)
by: Meng, Xiangxin, et al.
Published: (2024)
Similar Items
-
SpecPylot: Python Specification Generation using Large Language Models
by: Ayon, Ragib Shahariar, et al.
Published: (2026) -
AutoReSpec: A Framework for Generating Specification using Large Language Models
by: Ayon, Ragib Shahariar, et al.
Published: (2026) -
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
by: Bouzenia, Islem, et al.
Published: (2024) -
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
by: Chakroborti, Apu Kumar, et al.
Published: (2025) -
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
by: Islam, Niful, et al.
Published: (2026)