AI-assisted Code Authoring at Scale: Fine-tuning, deploying, and mixed methods evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Murali, Vijayaraghavan, Maddila, Chandra, Ahmad, Imad, Bolin, Michael, Cheng, Daniel, Ghorbani, Negar, Fernandez, Renuka, Nagappan, Nachiappan, Rigby, Peter C. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-line AI-assisted Code Authoring
by: Dunay, Omer, et al.
Published: (2024)
by: Dunay, Omer, et al.
Published: (2024)
AI-Assisted Fixes to Code Review Comments at Scale
by: Maddila, Chandra, et al.
Published: (2025)
by: Maddila, Chandra, et al.
Published: (2025)
AI-Assisted SQL Authoring at Industry Scale
by: Maddila, Chandra, et al.
Published: (2024)
by: Maddila, Chandra, et al.
Published: (2024)
Moving Faster and Reducing Risk: Using LLMs in Release Deployment
by: Abreu, Rui, et al.
Published: (2024)
by: Abreu, Rui, et al.
Published: (2024)
Improving Code Reviewer Recommendation: Accuracy, Latency, Workload, and Bystanders
by: Rigby, Peter C., et al.
Published: (2023)
by: Rigby, Peter C., et al.
Published: (2023)
REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage
by: Jha, Smriti, et al.
Published: (2026)
by: Jha, Smriti, et al.
Published: (2026)
Agentic Program Repair from Test Failures at Scale: A Neuro-symbolic approach with static analysis and test execution feedback
by: Maddila, Chandra, et al.
Published: (2025)
by: Maddila, Chandra, et al.
Published: (2025)
Whodunit: Classifying Code as Human Authored or GPT-4 Generated -- A case study on CodeChef problems
by: Idialu, Oseremen Joy, et al.
Published: (2024)
by: Idialu, Oseremen Joy, et al.
Published: (2024)
Test-Driven Development for Code Generation
by: Mathews, Noble Saji, et al.
Published: (2024)
by: Mathews, Noble Saji, et al.
Published: (2024)
Wink: Recovering from Misbehaviors in Coding Agents
by: Nanda, Rahul, et al.
Published: (2026)
by: Nanda, Rahul, et al.
Published: (2026)
Is GitHub's Copilot as Bad as Humans at Introducing Vulnerabilities in Code?
by: Asare, Owura, et al.
Published: (2022)
by: Asare, Owura, et al.
Published: (2022)
Examining LLMs Ability to Summarize Code Through Mutation-Analysis
by: Khatib, Lara, et al.
Published: (2026)
by: Khatib, Lara, et al.
Published: (2026)
Code Improvement Practices at Meta
by: Mockus, Audris, et al.
Published: (2025)
by: Mockus, Audris, et al.
Published: (2025)
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
by: Zhu, Yuecai, et al.
Published: (2026)
by: Zhu, Yuecai, et al.
Published: (2026)
FuzzSlice: Pruning False Positives in Static Analysis Warnings Through Function-Level Fuzzing
by: Murali, Aniruddhan, et al.
Published: (2024)
by: Murali, Aniruddhan, et al.
Published: (2024)
Measuring the Runtime Performance of C++ Code Written by Humans using GitHub Copilot
by: Erhabor, Daniel, et al.
Published: (2023)
by: Erhabor, Daniel, et al.
Published: (2023)
Improving Code Search with Hard Negative Sampling Based on Fine-tuning
by: Dong, Hande, et al.
Published: (2023)
by: Dong, Hande, et al.
Published: (2023)
Context-Aware CodeLLM Eviction for AI-assisted Coding
by: Thangarajah, Kishanthan, et al.
Published: (2025)
by: Thangarajah, Kishanthan, et al.
Published: (2025)
Codexity: Secure AI-assisted Code Generation
by: Kim, Sung Yong, et al.
Published: (2024)
by: Kim, Sung Yong, et al.
Published: (2024)
MORepair: Teaching LLMs to Repair Code via Multi-Objective Fine-tuning
by: Yang, Boyang, et al.
Published: (2024)
by: Yang, Boyang, et al.
Published: (2024)
Devstral: Fine-tuning Language Models for Coding Agent Applications
by: Rastogi, Abhinav, et al.
Published: (2025)
by: Rastogi, Abhinav, et al.
Published: (2025)
Is Your Automated Software Engineer Trustworthy?
by: Mathews, Noble Saji, et al.
Published: (2025)
by: Mathews, Noble Saji, et al.
Published: (2025)
RAG or Fine-tuning? A Comparative Study on LCMs-based Code Completion in Industry
by: Wang, Chaozheng, et al.
Published: (2025)
by: Wang, Chaozheng, et al.
Published: (2025)
Checker Bug Detection and Repair in Deep Learning Libraries
by: Harzevili, Nima Shiri, et al.
Published: (2024)
by: Harzevili, Nima Shiri, et al.
Published: (2024)
Code Less, Align More: Efficient LLM Fine-tuning for Code Generation with Data Pruning
by: Tsai, Yun-Da, et al.
Published: (2024)
by: Tsai, Yun-Da, et al.
Published: (2024)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Understanding the Human-LLM Dynamic: A Literature Survey of LLM Use in Programming Tasks
by: Etsenake, Deborah, et al.
Published: (2024)
by: Etsenake, Deborah, et al.
Published: (2024)
Structure-aware Fine-tuning for Code Pre-trained Models
by: Wu, Jiayi, et al.
Published: (2024)
by: Wu, Jiayi, et al.
Published: (2024)
AddressWatcher: Sanitizer-Based Localization of Memory Leak Fixes
by: Murali, Aniruddhan, et al.
Published: (2024)
by: Murali, Aniruddhan, et al.
Published: (2024)
Analyzing Message-Code Inconsistency in AI Coding Agent-Authored Pull Requests
by: Gong, Jingzhi, et al.
Published: (2026)
by: Gong, Jingzhi, et al.
Published: (2026)
Extending Behavioral Software Engineering: Decision-Making and Collaboration in Human-AI Teams for Responsible Software Engineering
by: Rani, Lekshmi Murali
Published: (2025)
by: Rani, Lekshmi Murali
Published: (2025)
The development and deployment of formal methods in the UK
by: Jones, Cliff B., et al.
Published: (2020)
by: Jones, Cliff B., et al.
Published: (2020)
Design choices made by LLM-based test generators prevent them from finding bugs
by: Mathews, Noble Saji, et al.
Published: (2024)
by: Mathews, Noble Saji, et al.
Published: (2024)
AssertFlip: Reproducing Bugs via Inversion of LLM-Generated Passing Tests
by: Khatib, Lara, et al.
Published: (2025)
by: Khatib, Lara, et al.
Published: (2025)
Does SWE-Bench-Verified Test Agent Ability or Model Memory?
by: Prathifkumar, Thanosan, et al.
Published: (2025)
by: Prathifkumar, Thanosan, et al.
Published: (2025)
UniASM: Binary Code Similarity Detection without Fine-tuning
by: Gu, Yeming, et al.
Published: (2022)
by: Gu, Yeming, et al.
Published: (2022)
Repoformer: Selective Retrieval for Repository-Level Code Completion
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Does the Order of Fine-tuning Matter and Why?
by: Chen, Qihong, et al.
Published: (2024)
by: Chen, Qihong, et al.
Published: (2024)
RubberDuckBench: A Benchmark for AI Coding Assistants
by: Mohammed, Ferida, et al.
Published: (2026)
by: Mohammed, Ferida, et al.
Published: (2026)
A User-centered Security Evaluation of Copilot
by: Asare, Owura, et al.
Published: (2023)
by: Asare, Owura, et al.
Published: (2023)
Similar Items
-
Multi-line AI-assisted Code Authoring
by: Dunay, Omer, et al.
Published: (2024) -
AI-Assisted Fixes to Code Review Comments at Scale
by: Maddila, Chandra, et al.
Published: (2025) -
AI-Assisted SQL Authoring at Industry Scale
by: Maddila, Chandra, et al.
Published: (2024) -
Moving Faster and Reducing Risk: Using LLMs in Release Deployment
by: Abreu, Rui, et al.
Published: (2024) -
Improving Code Reviewer Recommendation: Accuracy, Latency, Workload, and Bystanders
by: Rigby, Peter C., et al.
Published: (2023)