Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
Fuente:
arXiv
Salvato in:
| Autore principale: | Iscan, Mehmet |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
di: Palacios, Diego Cabezas
Pubblicazione: (2026)
di: Palacios, Diego Cabezas
Pubblicazione: (2026)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
di: Zolduoarrati, Elijah, et al.
Pubblicazione: (2025)
di: Zolduoarrati, Elijah, et al.
Pubblicazione: (2025)
PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents
di: Iscan, Mehmet
Pubblicazione: (2026)
di: Iscan, Mehmet
Pubblicazione: (2026)
Automated Bug Triaging using Instruction-Tuned Large Language Models
di: Kiashemshaki, Kiana, et al.
Pubblicazione: (2025)
di: Kiashemshaki, Kiana, et al.
Pubblicazione: (2025)
Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents
di: Iscan, Mehmet
Pubblicazione: (2026)
di: Iscan, Mehmet
Pubblicazione: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
Generative AI and the Transformation of Software Development Practices
di: Acharya, Vivek
Pubblicazione: (2025)
di: Acharya, Vivek
Pubblicazione: (2025)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
di: Zhou, Yue, et al.
Pubblicazione: (2025)
di: Zhou, Yue, et al.
Pubblicazione: (2025)
OODEval: Evaluating Large Language Models on Object-Oriented Design
di: Xiao, Bingxu, et al.
Pubblicazione: (2026)
di: Xiao, Bingxu, et al.
Pubblicazione: (2026)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
di: Alshaikh, Moaath, et al.
Pubblicazione: (2026)
di: Alshaikh, Moaath, et al.
Pubblicazione: (2026)
Natural Language Summarization Enables Multi-Repository Bug Localization by LLMs in Microservice Architectures
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval
di: Ashrafi, Nazmus
Pubblicazione: (2026)
di: Ashrafi, Nazmus
Pubblicazione: (2026)
Context Engineering for Multi-Agent LLM Code Assistants Using Elicit, NotebookLM, ChatGPT, and Claude Code
di: Haseeb, Muhammad
Pubblicazione: (2025)
di: Haseeb, Muhammad
Pubblicazione: (2025)
LLMCup: Ranking-Enhanced Comment Updating with LLMs
di: Ge, Hua, et al.
Pubblicazione: (2025)
di: Ge, Hua, et al.
Pubblicazione: (2025)
MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization
di: Tanjim, Md Mehrab, et al.
Pubblicazione: (2026)
di: Tanjim, Md Mehrab, et al.
Pubblicazione: (2026)
GALA: Multimodal Graph Alignment for Bug Localization in Automated Program Repair
di: Liu, Zhuoyao, et al.
Pubblicazione: (2026)
di: Liu, Zhuoyao, et al.
Pubblicazione: (2026)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
di: Nguyen, Quang-Dung, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Dung, et al.
Pubblicazione: (2025)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
di: Wang, Yuchen, et al.
Pubblicazione: (2026)
di: Wang, Yuchen, et al.
Pubblicazione: (2026)
Mining Subscenario Refactoring Opportunities in Behaviour-Driven Software Test Suites: ML Classifiers and LLM-Judge Baselines
di: Mughal, Ali Hassaan, et al.
Pubblicazione: (2026)
di: Mughal, Ali Hassaan, et al.
Pubblicazione: (2026)
Can AI Assist in Olympiad Coding
di: Ren, Samuel
Pubblicazione: (2025)
di: Ren, Samuel
Pubblicazione: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
di: Souza, Débora, et al.
Pubblicazione: (2026)
di: Souza, Débora, et al.
Pubblicazione: (2026)
RelRepair: Enhancing Automated Program Repair by Retrieving Relevant Code
di: Liu, Shunyu, et al.
Pubblicazione: (2025)
di: Liu, Shunyu, et al.
Pubblicazione: (2025)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2025)
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2025)
AgentModernize: Preserving Business Logic in Legacy Modernization with Multi-Agent LLMs and Behavioral Specification Graphs
di: Ahmed, Sheikh Nazib, et al.
Pubblicazione: (2026)
di: Ahmed, Sheikh Nazib, et al.
Pubblicazione: (2026)
DRS-OSS: Practical Diff Risk Scoring with LLMs
di: Sayedsalehi, Ali, et al.
Pubblicazione: (2025)
di: Sayedsalehi, Ali, et al.
Pubblicazione: (2025)
Reinforcement Learning for Dynamic Workflow Optimization in CI/CD Pipelines
di: Soni, Aniket Abhishek, et al.
Pubblicazione: (2026)
di: Soni, Aniket Abhishek, et al.
Pubblicazione: (2026)
Mind the Metrics: Patterns for Telemetry-Aware In-IDE AI Application Development using the Model Context Protocol (MCP)
di: Koc, Vincent, et al.
Pubblicazione: (2025)
di: Koc, Vincent, et al.
Pubblicazione: (2025)
L2MAC: Large Language Model Automatic Computer for Extensive Code Generation
di: Holt, Samuel, et al.
Pubblicazione: (2023)
di: Holt, Samuel, et al.
Pubblicazione: (2023)
Scattered Forest Search: Smarter Code Space Exploration with LLMs
di: Light, Jonathan, et al.
Pubblicazione: (2024)
di: Light, Jonathan, et al.
Pubblicazione: (2024)
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
di: Jana, Prithwish, et al.
Pubblicazione: (2026)
di: Jana, Prithwish, et al.
Pubblicazione: (2026)
VulScribeR: Exploring RAG-based Vulnerability Augmentation with LLMs
di: Daneshvar, Seyed Shayan, et al.
Pubblicazione: (2024)
di: Daneshvar, Seyed Shayan, et al.
Pubblicazione: (2024)
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
di: Zhang, Jiahao, et al.
Pubblicazione: (2026)
A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification
di: Odmark, Joshua, et al.
Pubblicazione: (2026)
di: Odmark, Joshua, et al.
Pubblicazione: (2026)
CoTran: An LLM-based Code Translator using Reinforcement Learning with Feedback from Compiler and Symbolic Execution
di: Jana, Prithwish, et al.
Pubblicazione: (2023)
di: Jana, Prithwish, et al.
Pubblicazione: (2023)
Using LLMs to Establish Implicit User Sentiment of Software Desirability
di: Weitl-Harms, Sherri, et al.
Pubblicazione: (2024)
di: Weitl-Harms, Sherri, et al.
Pubblicazione: (2024)
CIFE: Code Instruction-Following Evaluation
di: Gunnu, Sravani, et al.
Pubblicazione: (2025)
di: Gunnu, Sravani, et al.
Pubblicazione: (2025)
SLEAN: Simple Lightweight Ensemble Analysis Network for Multi-Provider LLM Coordination: Design, Implementation, and Vibe Coding Bug Investigation Case Study
di: Vargas, Matheus J. T.
Pubblicazione: (2025)
di: Vargas, Matheus J. T.
Pubblicazione: (2025)
Understanding and Detecting Flaky Builds in GitHub Actions
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
Survey Transfer Learning: Recycling Data with Silicon Responses
di: Amini, Ali
Pubblicazione: (2025)
di: Amini, Ali
Pubblicazione: (2025)
CodeTracer: Towards Traceable Agent States
di: Li, Han, et al.
Pubblicazione: (2026)
di: Li, Han, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
di: Palacios, Diego Cabezas
Pubblicazione: (2026) -
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
di: Zolduoarrati, Elijah, et al.
Pubblicazione: (2025) -
PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents
di: Iscan, Mehmet
Pubblicazione: (2026) -
Automated Bug Triaging using Instruction-Tuned Large Language Models
di: Kiashemshaki, Kiana, et al.
Pubblicazione: (2025) -
Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents
di: Iscan, Mehmet
Pubblicazione: (2026)