A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Tabassum, Anika, Hossain, Md Sifat, Arefin, Md. Fahim, Islam, Tariqul, Zaman, Tarannum Shaila |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-ProS: Analyzing Large Language Models' Performance in Competitive Problem Solving
by: Hossain, Md Sifat, et al.
Published: (2025)
by: Hossain, Md Sifat, et al.
Published: (2025)
OLAF: Towards Robust LLM-Based Annotation Framework in Empirical Software Engineering
by: Imran, Mia Mohammad, et al.
Published: (2025)
by: Imran, Mia Mohammad, et al.
Published: (2025)
SmartShift: A Secure and Efficient Approach to Smart Contract Migration
by: Hossain, Tahrim, et al.
Published: (2025)
by: Hossain, Tahrim, et al.
Published: (2025)
DePro: Understanding the Role of LLMs in Debugging Competitive Programming Code
by: Parvez, Nabiha, et al.
Published: (2026)
by: Parvez, Nabiha, et al.
Published: (2026)
SysPro: Reproducing System-level Concurrency Bugs from Bug Reports
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
SECite: Analyzing and Summarizing Citations in Software Engineering Literature
by: Pyreddy, Shireesh Reddy, et al.
Published: (2026)
by: Pyreddy, Shireesh Reddy, et al.
Published: (2026)
Gender Dynamics in Software Engineering: Insights from Research on Concurrency Bug Reproduction
by: Zaman, Tarannum Shaila, et al.
Published: (2025)
by: Zaman, Tarannum Shaila, et al.
Published: (2025)
LLPut: Investigating Large Language Models for Bug Report-Based Input Generation
by: Hasan, Alif Al, et al.
Published: (2025)
by: Hasan, Alif Al, et al.
Published: (2025)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
PyGen: A Collaborative Human-AI Approach to Python Package Creation
by: Barua, Saikat, et al.
Published: (2024)
by: Barua, Saikat, et al.
Published: (2024)
A Combined Feature Embedding Tools for Multi-Class Software Defect and Identification
by: Sultan, Md. Fahim, et al.
Published: (2024)
by: Sultan, Md. Fahim, et al.
Published: (2024)
CODESIM: Multi-Agent Code Generation and Problem Solving through Simulation-Driven Planning and Debugging
by: Islam, Md. Ashraful, et al.
Published: (2025)
by: Islam, Md. Ashraful, et al.
Published: (2025)
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
by: Taherkhani, Hamed, et al.
Published: (2024)
by: Taherkhani, Hamed, et al.
Published: (2024)
Enhanced LLM-Based Framework for Predicting Null Pointer Dereference in Source Code
by: Sultan, Md. Fahim, et al.
Published: (2024)
by: Sultan, Md. Fahim, et al.
Published: (2024)
VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
Towards Reliable Evaluation of Neural Program Repair with Natural Robustness Testing
by: Le-Cong, Thanh, et al.
Published: (2024)
by: Le-Cong, Thanh, et al.
Published: (2024)
LLMs: A Game-Changer for Software Engineers?
by: Haque, Md Asraful
Published: (2024)
by: Haque, Md Asraful
Published: (2024)
Swiss Cheese Model for AI Safety: A Taxonomy and Reference Architecture for Multi-Layered Guardrails of Foundation Model Based Agents
by: Shamsujjoha, Md, et al.
Published: (2024)
by: Shamsujjoha, Md, et al.
Published: (2024)
Error Understanding in Program Code With LLM-DL for Multi-label Classification
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
Programming with AI: Evaluating ChatGPT, Gemini, AlphaCode, and GitHub Copilot for Programmers
by: Siam, Md Kamrul, et al.
Published: (2024)
by: Siam, Md Kamrul, et al.
Published: (2024)
CIgrate: Automating CI Service Migration with Large Language Models
by: Hossain, Md Nazmul, et al.
Published: (2025)
by: Hossain, Md Nazmul, et al.
Published: (2025)
A Large-Scale Empirical Study of COVID-19 Contact Tracing Mobile App Reviews
by: Parisa, Sifat Ishmam, et al.
Published: (2024)
by: Parisa, Sifat Ishmam, et al.
Published: (2024)
Analysis on LLMs Performance for Code Summarization
by: Akib, Md. Ahnaf, et al.
Published: (2024)
by: Akib, Md. Ahnaf, et al.
Published: (2024)
Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development
by: Basu, Srijita, et al.
Published: (2026)
by: Basu, Srijita, et al.
Published: (2026)
Transforming Software Development: Evaluating the Efficiency and Challenges of GitHub Copilot in Real-World Projects
by: Pandey, Ruchika, et al.
Published: (2024)
by: Pandey, Ruchika, et al.
Published: (2024)
EnStack: An Ensemble Stacking Framework of Large Language Models for Enhanced Vulnerability Detection in Source Code
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
Software Self-Extension with SelfEvolve: an Agentic Architecture for Runtime Code Generation
by: Fahim, Md Asif Iqbal, et al.
Published: (2026)
by: Fahim, Md Asif Iqbal, et al.
Published: (2026)
Cybernaut: Towards Reliable Web Automation
by: Tomar, Ankur, et al.
Published: (2025)
by: Tomar, Ankur, et al.
Published: (2025)
A Pair Programming Framework for Code Generation via Multi-Plan Exploration and Feedback-Driven Refinement
by: Zhang, Huan, et al.
Published: (2024)
by: Zhang, Huan, et al.
Published: (2024)
A Self-Healing Framework for Reliable LLM-Based Autonomous Agents
by: Jeong, Cheonsu, et al.
Published: (2026)
by: Jeong, Cheonsu, et al.
Published: (2026)
ChatGPT vs. DeepSeek: A Comparative Study on AI-Based Code Generation
by: Manik, Md Motaleb Hossen
Published: (2025)
by: Manik, Md Motaleb Hossen
Published: (2025)
Flexible Control Flow Graph Alignment for Delivering Data-Driven Feedback to Novice Programming Learners
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
by: Chowdhury, Md Towhidul Absar, et al.
Published: (2024)
Code Comprehension with GitHub Copilot: Performance Gains, Comprehension Trade-offs, and Behavioral Predictors in Brownfield Programming
by: Qiao, Yunhan, et al.
Published: (2025)
by: Qiao, Yunhan, et al.
Published: (2025)
AutoCodeRover: Autonomous Program Improvement
by: Zhang, Yuntong, et al.
Published: (2024)
by: Zhang, Yuntong, et al.
Published: (2024)
Toward Reliable Design of LLM-Enabled Agentic Workflows: Optimizing Latency-Reliability-Cost Tradeoffs
by: Yang, Ya-Ting, et al.
Published: (2026)
by: Yang, Ya-Ting, et al.
Published: (2026)
Best Practices for Facing the Security Challenges of Internet of Things Devices Focusing on Software Development Life Cycle
by: Islam, Md Rafid, et al.
Published: (2024)
by: Islam, Md Rafid, et al.
Published: (2024)
Benchmarking ChatGPT, Codeium, and GitHub Copilot: A Comparative Study of AI-Driven Programming and Debugging Assistants
by: Ovi, Md Sultanul Islam, et al.
Published: (2024)
by: Ovi, Md Sultanul Islam, et al.
Published: (2024)
Evaluating Software Process Models for Multi-Agent Class-Level Code Generation
by: Shafin, Wasique Islam, et al.
Published: (2025)
by: Shafin, Wasique Islam, et al.
Published: (2025)
The Effects of GitHub Copilot on Computing Students' Programming Effectiveness, Efficiency, and Processes in Brownfield Programming Tasks
by: Shihab, Md Istiak Hossain, et al.
Published: (2025)
by: Shihab, Md Istiak Hossain, et al.
Published: (2025)
RAILS: Retrieval-Augmented Intelligence for Learning Software Development
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
by: Abdullah, Wali Mohammad, et al.
Published: (2025)
Similar Items
-
LLM-ProS: Analyzing Large Language Models' Performance in Competitive Problem Solving
by: Hossain, Md Sifat, et al.
Published: (2025) -
OLAF: Towards Robust LLM-Based Annotation Framework in Empirical Software Engineering
by: Imran, Mia Mohammad, et al.
Published: (2025) -
SmartShift: A Secure and Efficient Approach to Smart Contract Migration
by: Hossain, Tahrim, et al.
Published: (2025) -
DePro: Understanding the Role of LLMs in Debugging Competitive Programming Code
by: Parvez, Nabiha, et al.
Published: (2026) -
SysPro: Reproducing System-level Concurrency Bugs from Bug Reports
by: Zaman, Tarannum Shaila, et al.
Published: (2026)