Adaptive Request Scheduling for CodeLLM Serving with SLA Guarantees
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Shi, Chen, Boyuan, Thangarajah, Kishanthan, Lutfiyya, Hanan, Hassan, Ahmed E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Aware CodeLLM Eviction for AI-assisted Coding
by: Thangarajah, Kishanthan, et al.
Published: (2025)
by: Thangarajah, Kishanthan, et al.
Published: (2025)
SynConfRoute: Syntax-Aware Routing for Efficient Code Completion with Small CodeLLMs
by: Thangarajah, Kishanthan, et al.
Published: (2026)
by: Thangarajah, Kishanthan, et al.
Published: (2026)
SLA-Awareness for AI-assisted coding
by: Thangarajah, Kishanthan, et al.
Published: (2025)
by: Thangarajah, Kishanthan, et al.
Published: (2025)
Software Performance Engineering for Foundation Model-Powered Software
by: Zhang, Haoxiang, et al.
Published: (2024)
by: Zhang, Haoxiang, et al.
Published: (2024)
Mastering the Craft of Data Synthesis for CodeLLMs
by: Chen, Meng, et al.
Published: (2024)
by: Chen, Meng, et al.
Published: (2024)
On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
by: Watanabe, Miku, et al.
Published: (2025)
by: Watanabe, Miku, et al.
Published: (2025)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
by: Manh, Dung Nguyen, et al.
Published: (2024)
by: Manh, Dung Nguyen, et al.
Published: (2024)
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
by: Hassan, Ahmed E., et al.
Published: (2024)
by: Hassan, Ahmed E., et al.
Published: (2024)
Collaborator or Assistant? How AI Coding Agents Partition Work Across Pull Request Lifecycles
by: Jo, Young, et al.
Published: (2026)
by: Jo, Young, et al.
Published: (2026)
Empirical Analysis of Pull Requests for Google Summer of Code
by: Popoola, Saheed
Published: (2024)
by: Popoola, Saheed
Published: (2024)
An Empirical Study on Developers Shared Conversations with ChatGPT in GitHub Pull Requests and Issues
by: Hao, Huizi, et al.
Published: (2024)
by: Hao, Huizi, et al.
Published: (2024)
Can Large Language Models Serve as Evaluators for Code Summarization?
by: Wu, Yang, et al.
Published: (2024)
by: Wu, Yang, et al.
Published: (2024)
On the Effect of Token Merging on Pre-trained Models for Code
by: Saad, Mootez, et al.
Published: (2025)
by: Saad, Mootez, et al.
Published: (2025)
How Do Developers Use Code Suggestions in Pull Request Reviews?
by: Bouraffa, Abir, et al.
Published: (2025)
by: Bouraffa, Abir, et al.
Published: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
by: Vasilevski, Kirill, et al.
Published: (2025)
by: Vasilevski, Kirill, et al.
Published: (2025)
Do AI Coding Agents Log Like Humans? An Empirical Study
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
Comparing AI Coding Agents: A Task-Stratified Analysis of Pull Request Acceptance
by: Pinna, Giovanni, et al.
Published: (2026)
by: Pinna, Giovanni, et al.
Published: (2026)
Quality Gatekeepers: Investigating the Effects ofCode Review Bots on Pull Request Activities
by: Wessel, Mairieli, et al.
Published: (2021)
by: Wessel, Mairieli, et al.
Published: (2021)
On Unified Prompt Tuning for Request Quality Assurance in Public Code Review
by: Chen, Xinyu, et al.
Published: (2024)
by: Chen, Xinyu, et al.
Published: (2024)
Rethinking Software Engineering in the Foundation Model Era: From Task-Driven AI Copilots to Goal-Driven AI Pair Programmers
by: Hassan, Ahmed E., et al.
Published: (2024)
by: Hassan, Ahmed E., et al.
Published: (2024)
Towards AI-Native Software Engineering (SE 3.0): A Vision and a Challenge Roadmap
by: Hassan, Ahmed E., et al.
Published: (2024)
by: Hassan, Ahmed E., et al.
Published: (2024)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
Do Autonomous Agents Contribute Test Code? A Study of Tests in Agentic Pull Requests
by: Haque, Sabrina, et al.
Published: (2026)
by: Haque, Sabrina, et al.
Published: (2026)
Analyzing Message-Code Inconsistency in AI Coding Agent-Authored Pull Requests
by: Gong, Jingzhi, et al.
Published: (2026)
by: Gong, Jingzhi, et al.
Published: (2026)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
Knowledge-Guided Prompt Learning for Request Quality Assurance in Public Code Review
by: Li, Lin, et al.
Published: (2024)
by: Li, Lin, et al.
Published: (2024)
Demystifying Feature Requests: Leveraging LLMs to Refine Feature Requests in Open-Source Software
by: KC, Pragyan, et al.
Published: (2025)
by: KC, Pragyan, et al.
Published: (2025)
Characterizing Multi-Hunk Patches: Divergence, Proximity, and LLM Repair Challenges
by: Nashid, Noor, et al.
Published: (2025)
by: Nashid, Noor, et al.
Published: (2025)
Code Change Characteristics and Description Alignment: A Comparative Study of Agentic versus Human Pull Requests
by: Pham, Dung, et al.
Published: (2026)
by: Pham, Dung, et al.
Published: (2026)
Let's Make Every Pull Request Meaningful: An Empirical Analysis of Developer and Agentic Pull Requests
by: Yoshioka, Haruhiko, et al.
Published: (2026)
by: Yoshioka, Haruhiko, et al.
Published: (2026)
AgenticSZZ: Temporal Knowledge Graph-Guided Agentic Bug-Inducing Commit Identification
by: Shi, Yu, et al.
Published: (2026)
by: Shi, Yu, et al.
Published: (2026)
When Elo Lies: Hidden Biases in Codeforces-Based Evaluation of Large Language Models
by: Zheng, Shenyu, et al.
Published: (2026)
by: Zheng, Shenyu, et al.
Published: (2026)
Assessing and Improving the Representativeness of Code Generation Benchmarks Using Knowledge Units (KUs) of Programming Languages -- An Empirical Study
by: Ahasanuzzaman, Md, et al.
Published: (2026)
by: Ahasanuzzaman, Md, et al.
Published: (2026)
Beyond Bug Fixes: An Empirical Investigation of Post-Merge Code Quality Issues in Agent-Generated Pull Requests
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
by: Cynthia, Shamse Tasnim, et al.
Published: (2026)
Human or LLM? A Comparative Study on Accessible Code Generation Capability
by: Suh, Hyunjae, et al.
Published: (2025)
by: Suh, Hyunjae, et al.
Published: (2025)
LLM-Redactor: An Empirical Evaluation of Eight Techniques for Privacy-Preserving LLM Requests
by: Agyemang, Justice Owusu, et al.
Published: (2026)
by: Agyemang, Justice Owusu, et al.
Published: (2026)
The Value of Effective Pull Request Description
by: Pirouzkhah, Shirin, et al.
Published: (2026)
by: Pirouzkhah, Shirin, et al.
Published: (2026)
Sphinx: Benchmarking and Modeling for LLM-Driven Pull Request Review
by: Zhang, Daoan, et al.
Published: (2026)
by: Zhang, Daoan, et al.
Published: (2026)
AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length
by: Cheng, Junhang, et al.
Published: (2025)
by: Cheng, Junhang, et al.
Published: (2025)
Similar Items
-
Context-Aware CodeLLM Eviction for AI-assisted Coding
by: Thangarajah, Kishanthan, et al.
Published: (2025) -
SynConfRoute: Syntax-Aware Routing for Efficient Code Completion with Small CodeLLMs
by: Thangarajah, Kishanthan, et al.
Published: (2026) -
SLA-Awareness for AI-assisted coding
by: Thangarajah, Kishanthan, et al.
Published: (2025) -
Software Performance Engineering for Foundation Model-Powered Software
by: Zhang, Haoxiang, et al.
Published: (2024) -
Mastering the Craft of Data Synthesis for CodeLLMs
by: Chen, Meng, et al.
Published: (2024)