GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wu, Jie JW, Herlihy, Ayanda Patrick, Mirza, Ahmad Saleem, Afoud, Ali, Fard, Fatemeh |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent
par: Wu, Jie JW, et autres
Publié: (2024)
par: Wu, Jie JW, et autres
Publié: (2024)
MergeRepair: An Exploratory Study on Merging Task-Specific Adapters in Code LLMs for Automated Program Repair
par: Dehghan, Meghdad, et autres
Publié: (2024)
par: Dehghan, Meghdad, et autres
Publié: (2024)
A Survey on LLM-based Code Generation for Low-Resource and Domain-Specific Programming Languages
par: Joel, Sathvik, et autres
Publié: (2024)
par: Joel, Sathvik, et autres
Publié: (2024)
AutoOffAB: Toward Automated Offline A/B Testing for Data-Driven Requirement Engineering
par: Wu, Jie JW
Publié: (2023)
par: Wu, Jie JW
Publié: (2023)
MANTRA: a Framework for Multi-stage Adaptive Noise TReAtment During Training
par: Zhao, Zixiao, et autres
Publié: (2025)
par: Zhao, Zixiao, et autres
Publié: (2025)
Unveiling Ruby: Insights from Stack Overflow and Developer Survey
par: Akbarpour, Nikta, et autres
Publié: (2025)
par: Akbarpour, Nikta, et autres
Publié: (2025)
Can Code Language Models Learn Clarification-Seeking Behaviors?
par: Wu, Jie JW, et autres
Publié: (2025)
par: Wu, Jie JW, et autres
Publié: (2025)
Investigating the Efficacy of Large Language Models for Code Clone Detection
par: Khajezade, Mohamad, et autres
Publié: (2024)
par: Khajezade, Mohamad, et autres
Publié: (2024)
Collaborative Agents for Automated Program Repair in Ruby
par: Akbarpour, Nikta, et autres
Publié: (2025)
par: Akbarpour, Nikta, et autres
Publié: (2025)
Context-Augmented Code Generation Using Programming Knowledge Graphs
par: Saberi, Iman, et autres
Publié: (2024)
par: Saberi, Iman, et autres
Publié: (2024)
Large Language Models Should Ask Clarifying Questions to Increase Confidence in Generated Code
par: Wu, Jie JW
Publié: (2023)
par: Wu, Jie JW
Publié: (2023)
Do Current Language Models Support Code Intelligence for R Programming Language?
par: Zhao, ZiXiao, et autres
Publié: (2024)
par: Zhao, ZiXiao, et autres
Publié: (2024)
Bias in the Loop: Auditing LLM-as-a-Judge for Software Engineering
par: Zhao, Zixiao, et autres
Publié: (2026)
par: Zhao, Zixiao, et autres
Publié: (2026)
Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing
par: Peng, Jiaren, et autres
Publié: (2026)
par: Peng, Jiaren, et autres
Publié: (2026)
CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions
par: Shi, Jingwei, et autres
Publié: (2026)
par: Shi, Jingwei, et autres
Publié: (2026)
An Exploratory Study of V-Model in Building ML-Enabled Software: A Systems Engineering Perspective
par: Wu, Jie JW
Publié: (2023)
par: Wu, Jie JW
Publié: (2023)
Studying Vulnerable Code Entities in R
par: Zhao, Zixiao, et autres
Publié: (2024)
par: Zhao, Zixiao, et autres
Publié: (2024)
Empirical Studies of Parameter Efficient Methods for Large Language Models of Code and Knowledge Transfer to R
par: Esmaeili, Amirreza, et autres
Publié: (2024)
par: Esmaeili, Amirreza, et autres
Publié: (2024)
ARISE: A Repository-level Graph Representation and Toolset for Agentic Fault Localization and Program Repair
par: Seddik, Shahd, et autres
Publié: (2026)
par: Seddik, Shahd, et autres
Publié: (2026)
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
par: Wang, Zimu, et autres
Publié: (2026)
par: Wang, Zimu, et autres
Publié: (2026)
Evaluating LLM Agents on Automated Software Analysis Tasks
par: Bouzenia, Islem, et autres
Publié: (2026)
par: Bouzenia, Islem, et autres
Publié: (2026)
Evolving Excellence: Automated Optimization of LLM-based Agents
par: Brookes, Paul, et autres
Publié: (2025)
par: Brookes, Paul, et autres
Publié: (2025)
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
par: Khajezade, Mohamad, et autres
Publié: (2026)
par: Khajezade, Mohamad, et autres
Publié: (2026)
Automated Code Editing with Search-Generate-Modify
par: Liu, Changshu, et autres
Publié: (2023)
par: Liu, Changshu, et autres
Publié: (2023)
Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation
par: Pysklo, Hubert M., et autres
Publié: (2026)
par: Pysklo, Hubert M., et autres
Publié: (2026)
AgenticTCAD: A LLM-based Multi-Agent Framework for Automated TCAD Code Generation and Device Optimization
par: Fan, Guangxi, et autres
Publié: (2025)
par: Fan, Guangxi, et autres
Publié: (2025)
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
par: Jahromi, Ali Soltanian Fard, et autres
Publié: (2026)
par: Jahromi, Ali Soltanian Fard, et autres
Publié: (2026)
Prompt Optimization for LLM Code Generation via Reinforcement Learning
par: Esfahani, Ali Mohammadi, et autres
Publié: (2026)
par: Esfahani, Ali Mohammadi, et autres
Publié: (2026)
Context-Augmented Code Generation Using Programming Knowledge Graphs
par: Seddik, Shahd, et autres
Publié: (2026)
par: Seddik, Shahd, et autres
Publié: (2026)
Analysis of AdvFusion: Adapter-based Multilingual Learning for Code Large Language Models
par: Esmaeili, Amirreza, et autres
Publié: (2025)
par: Esmaeili, Amirreza, et autres
Publié: (2025)
Wired for Reuse: Automating Context-Aware Code Adaptation in IDEs via LLM-Based Agent
par: Wang, Taiming, et autres
Publié: (2025)
par: Wang, Taiming, et autres
Publié: (2025)
Beyond pip install: Evaluating LLM Agents for the Automated Installation of Python Projects
par: Milliken, Louis, et autres
Publié: (2024)
par: Milliken, Louis, et autres
Publié: (2024)
AdvFusion: Adapter-based Knowledge Transfer for Code Summarization on Code Language Models
par: Saberi, Iman, et autres
Publié: (2023)
par: Saberi, Iman, et autres
Publié: (2023)
LLM Agents for Automated Dependency Upgrades
par: Tawosi, Vali, et autres
Publié: (2025)
par: Tawosi, Vali, et autres
Publié: (2025)
Industrial LLM-based Code Optimization under Regulation: A Mixture-of-Agents Approach
par: Ashiga, Mari, et autres
Publié: (2025)
par: Ashiga, Mari, et autres
Publié: (2025)
Code Hallucination
par: Rahman, Mirza Masfiqur, et autres
Publié: (2024)
par: Rahman, Mirza Masfiqur, et autres
Publié: (2024)
MarsCode Agent: AI-native Automated Bug Fixing
par: Liu, Yizhou, et autres
Publié: (2024)
par: Liu, Yizhou, et autres
Publié: (2024)
ProcCtrlBench: Evaluating Process-Level Defects and Control Preservation in LLM Coding Agents
par: He, Jiawei, et autres
Publié: (2026)
par: He, Jiawei, et autres
Publié: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
par: Gong, Zhihao, et autres
Publié: (2026)
par: Gong, Zhihao, et autres
Publié: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
par: Gong, Zhihao, et autres
Publié: (2025)
par: Gong, Zhihao, et autres
Publié: (2025)
Documents similaires
-
HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent
par: Wu, Jie JW, et autres
Publié: (2024) -
MergeRepair: An Exploratory Study on Merging Task-Specific Adapters in Code LLMs for Automated Program Repair
par: Dehghan, Meghdad, et autres
Publié: (2024) -
A Survey on LLM-based Code Generation for Low-Resource and Domain-Specific Programming Languages
par: Joel, Sathvik, et autres
Publié: (2024) -
AutoOffAB: Toward Automated Offline A/B Testing for Data-Driven Requirement Engineering
par: Wu, Jie JW
Publié: (2023) -
MANTRA: a Framework for Multi-stage Adaptive Noise TReAtment During Training
par: Zhao, Zixiao, et autres
Publié: (2025)