A First Look at Bugs in LLM Inference Engines
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Mugeng, Zhong, Siqi, Bi, Weichen, Zhang, Yixuan, Chen, Zhiyang, Chen, Zhenpeng, Liu, Xuanzhe, Ma, Yun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Research on WebAssembly Runtimes: A Survey
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
Research Artifacts in Software Engineering Publications: Status and Trends
by: Liu, Mugeng, et al.
Published: (2024)
by: Liu, Mugeng, et al.
Published: (2024)
LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
by: Liu, Kaibo, et al.
Published: (2024)
by: Liu, Kaibo, et al.
Published: (2024)
Promptware Engineering: Software Engineering for Prompt-Enabled Systems
by: Chen, Zhenpeng, et al.
Published: (2025)
by: Chen, Zhenpeng, et al.
Published: (2025)
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
An Empirical Study of Refactoring Engine Bugs
by: Wang, Haibo, et al.
Published: (2024)
by: Wang, Haibo, et al.
Published: (2024)
SCOPE: Performance Testing for Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2023)
by: Wen, Jinfeng, et al.
Published: (2023)
Testing Refactoring Engine via Historical Bug Report driven LLM
by: Wang, Haibo, et al.
Published: (2025)
by: Wang, Haibo, et al.
Published: (2025)
Finding Cross-rule Optimization Bugs in Datalog Engines
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
Personality-Guided Code Generation Using Large Language Models
by: Guo, Yaoqi, et al.
Published: (2024)
by: Guo, Yaoqi, et al.
Published: (2024)
Large Language Model-Based Agents for Software Engineering: A Survey
by: Liu, Junwei, et al.
Published: (2024)
by: Liu, Junwei, et al.
Published: (2024)
BugScope: Learn to Find Bugs Like Human
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
An Empirical Study of Bugs in Modern LLM Agent Frameworks
by: Zhu, Xinxue, et al.
Published: (2026)
by: Zhu, Xinxue, et al.
Published: (2026)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
by: Guo, Liwei, et al.
Published: (2025)
by: Guo, Liwei, et al.
Published: (2025)
LLM-Based Misconfiguration Detection for AWS Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2024)
by: Wen, Jinfeng, et al.
Published: (2024)
Towards Understanding the Bugs in Solidity Compiler
by: Ma, Haoyang, et al.
Published: (2024)
by: Ma, Haoyang, et al.
Published: (2024)
Characterizing Bugs in Login Processes of Android Applications: An Empirical Study
by: Zhou, Zixu, et al.
Published: (2025)
by: Zhou, Zixu, et al.
Published: (2025)
GUIPilot: A Consistency-based Mobile GUI Testing Approach for Detecting Application-specific Bugs
by: Liu, Ruofan, et al.
Published: (2025)
by: Liu, Ruofan, et al.
Published: (2025)
Bias Behind the Wheel: Fairness Testing of Autonomous Driving Systems
by: Li, Xinyue, et al.
Published: (2023)
by: Li, Xinyue, et al.
Published: (2023)
Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
Understanding Bugs in Template Engine-Based Applications: Symptoms, Root Causes, and Fix Patterns
by: Gao, Kai, et al.
Published: (2026)
by: Gao, Kai, et al.
Published: (2026)
STALL+: Boosting LLM-based Repository-level Code Completion with Static Analysis
by: Liu, Junwei, et al.
Published: (2024)
by: Liu, Junwei, et al.
Published: (2024)
DeepFWI: Identifying Bug-Sensitive Warnings with Multi-Modal Code-Warning Semantics
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
A Dual-Loop Agent Framework for Automated Vulnerability Reproduction
by: Liu, Bin, et al.
Published: (2026)
by: Liu, Bin, et al.
Published: (2026)
TreeMind: Automatically Reproducing Android Bug Reports via LLM-empowered Monte Carlo Tree Search
by: Chen, Zhengyu, et al.
Published: (2025)
by: Chen, Zhengyu, et al.
Published: (2025)
Towards Trustworthy LLMs for Code: A Data-Centric Synergistic Auditing Framework
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
LLM-Powered Silent Bug Fuzzing in Deep Learning Libraries via Versatile and Controlled Bug Transfer
by: Zhang, Kunpeng, et al.
Published: (2026)
by: Zhang, Kunpeng, et al.
Published: (2026)
Understanding Bug-Reproducing Tests: A First Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
Causal Inference for the Effect of Code Coverage on Bug Introduction
by: Schulte, Lukas, et al.
Published: (2026)
by: Schulte, Lukas, et al.
Published: (2026)
BugForge: Constructing and Utilizing DBMS Bug Repository to Enhance DBMS Testing
by: Li, Dawei, et al.
Published: (2026)
by: Li, Dawei, et al.
Published: (2026)
A Comprehensive Study of Bugs in Modern Distributed Deep Learning Systems
by: Ma, Xiaoxue, et al.
Published: (2025)
by: Ma, Xiaoxue, et al.
Published: (2025)
An Empirical Study on the Characteristics of Database Access Bugs in Java Applications
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
Characterising Bugs in Jupyter Platform
by: Tang, Yutian, et al.
Published: (2025)
by: Tang, Yutian, et al.
Published: (2025)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
SysPro: Reproducing System-level Concurrency Bugs from Bug Reports
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
by: Zaman, Tarannum Shaila, et al.
Published: (2026)
Improving Compiler Bug Isolation by Leveraging Large Language Models
by: Qi, Yixian, et al.
Published: (2025)
by: Qi, Yixian, et al.
Published: (2025)
PanicFI: An Infrastructure for Fixing Panic Bugs in Real-World Rust Programs
by: Ni, Yunbo, et al.
Published: (2024)
by: Ni, Yunbo, et al.
Published: (2024)
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
by: Lyu, Yunbo, et al.
Published: (2026)
by: Lyu, Yunbo, et al.
Published: (2026)
Bug Priority Change Prediction: An Exploratory Study on Apache Software
by: Cai, Guangzong, et al.
Published: (2025)
by: Cai, Guangzong, et al.
Published: (2025)
Large Language Models Based JSON Parser Fuzzing for Bug Discovery and Behavioral Analysis
by: Zhong, Zhiyuan, et al.
Published: (2024)
by: Zhong, Zhiyuan, et al.
Published: (2024)
Similar Items
-
Research on WebAssembly Runtimes: A Survey
by: Zhang, Yixuan, et al.
Published: (2024) -
Research Artifacts in Software Engineering Publications: Status and Trends
by: Liu, Mugeng, et al.
Published: (2024) -
LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
by: Liu, Kaibo, et al.
Published: (2024) -
Promptware Engineering: Software Engineering for Prompt-Enabled Systems
by: Chen, Zhenpeng, et al.
Published: (2025) -
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
by: Li, Xinyue, et al.
Published: (2026)