Saved in:
| Main Authors: | Herter, Patrick, Ahlrichs, Vincent, Açilan, Ridvan, Horsch, Julian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.01609 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stack Trace-Based Crash Deduplication with Transformer Adaptation
by: Mamun, Md Afif Al, et al.
Published: (2025)
by: Mamun, Md Afif Al, et al.
Published: (2025)
Fuzzing BusyBox: Leveraging LLM and Crash Reuse for Embedded Bug Unearthing
by: Asmita, et al.
Published: (2024)
by: Asmita, et al.
Published: (2024)
Finding the Needle in the Crash Stack: Industrial-Scale Crash Root Cause Localization with AutoCrashFL
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
CrashJS: A NodeJS Benchmark for Automated Crash Reproduction
by: Oliver, Philip, et al.
Published: (2024)
by: Oliver, Philip, et al.
Published: (2024)
Outrunning LLM Cutoffs: A Live Kernel Crash Resolution Benchmark for All
by: Huang, Chenxi, et al.
Published: (2026)
by: Huang, Chenxi, et al.
Published: (2026)
BugLens: Leveraging Bisection for Lightweight Compiler Bug Deduplication
by: Zhou, Xintong, et al.
Published: (2025)
by: Zhou, Xintong, et al.
Published: (2025)
Crash-free Deductive Verifiers
by: Nauta, Wander, et al.
Published: (2026)
by: Nauta, Wander, et al.
Published: (2026)
SAFE: Harnessing LLM for Scenario-Driven ADS Testing from Multimodal Crash Data
by: Luo, Siwei, et al.
Published: (2025)
by: Luo, Siwei, et al.
Published: (2025)
Deduplicating and Ranking Solution Programs for Suggesting Reference Solutions
by: Shirafuji, Atsushi, et al.
Published: (2023)
by: Shirafuji, Atsushi, et al.
Published: (2023)
Predicting the Impact of Crashes Across Release Channels
by: Mujahid, Suhaib, et al.
Published: (2024)
by: Mujahid, Suhaib, et al.
Published: (2024)
A Study of Using Multimodal LLMs for Non-Crash Functional Bug Detection in Android Apps
by: Ju, Bangyan, et al.
Published: (2024)
by: Ju, Bangyan, et al.
Published: (2024)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
by: Lam, Man Ho, et al.
Published: (2025)
by: Lam, Man Ho, et al.
Published: (2025)
Crash Report Enhancement with Large Language Models: An Empirical Study
by: Fahim, S M Farah Al, et al.
Published: (2025)
by: Fahim, S M Farah Al, et al.
Published: (2025)
DaiFu: In-Situ Crash Recovery for Deep Learning Systems
by: He, Zilong, et al.
Published: (2025)
by: He, Zilong, et al.
Published: (2025)
Beyond Crash-to-Patch: Patch Evolution for Linux Kernel Repair
by: Bai, Luyao, et al.
Published: (2026)
by: Bai, Luyao, et al.
Published: (2026)
Runtime-Augmented LLMs for Crash Detection and Diagnosis in ML Notebooks
by: Wang, Yiran, et al.
Published: (2026)
by: Wang, Yiran, et al.
Published: (2026)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
by: Shibaev, Egor, et al.
Published: (2024)
by: Shibaev, Egor, et al.
Published: (2024)
Better Debugging: Combining Static Analysis and LLMs for Explainable Crashing Fault Localization
by: Yan, Jiwei, et al.
Published: (2024)
by: Yan, Jiwei, et al.
Published: (2024)
Crash Report Accumulation During Continuous Fuzzing
by: Yegorov, Ilya, et al.
Published: (2024)
by: Yegorov, Ilya, et al.
Published: (2024)
JunoBench: A Benchmark Dataset of Crashes in Python Machine Learning Jupyter Notebooks
by: Wang, Yiran, et al.
Published: (2025)
by: Wang, Yiran, et al.
Published: (2025)
Why do Machine Learning Notebooks Crash? An Empirical Study on Public Python Jupyter Notebooks
by: Wang, Yiran, et al.
Published: (2024)
by: Wang, Yiran, et al.
Published: (2024)
KGym: A Platform and Dataset to Benchmark Large Language Models on Linux Kernel Crash Resolution
by: Mathai, Alex, et al.
Published: (2024)
by: Mathai, Alex, et al.
Published: (2024)
APT-ClaritySet: A Large-Scale, High-Fidelity Labeled Dataset for APT Malware with Alias Normalization and Graph-Based Deduplication
by: Yin, Zhenhao, et al.
Published: (2025)
by: Yin, Zhenhao, et al.
Published: (2025)
Agents in the Sandbox: End-to-End Crash Bug Reproduction for Minecraft
by: Yapağcı, Eray, et al.
Published: (2025)
by: Yapağcı, Eray, et al.
Published: (2025)
Effective, Platform-Independent GUI Testing via Image Embedding and Reinforcement Learning
by: Yu, Shengcheng, et al.
Published: (2022)
by: Yu, Shengcheng, et al.
Published: (2022)
On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization
by: Crupi, Giuseppe, et al.
Published: (2025)
by: Crupi, Giuseppe, et al.
Published: (2025)
Refining Fuzzed Crashing Inputs for Better Fault Diagnosis
by: Kim, Kieun, et al.
Published: (2025)
by: Kim, Kieun, et al.
Published: (2025)
Studying and Understanding the Effectiveness and Failures of Conversational LLM-Based Repair
by: Chen, Aolin, et al.
Published: (2025)
by: Chen, Aolin, et al.
Published: (2025)
An Effective Approach to Embedding Source Code by Combining Large Language and Sentence Embedding Models
by: Xian, Zixiang, et al.
Published: (2024)
by: Xian, Zixiang, et al.
Published: (2024)
Fixing 7,400 Bugs for 1$: Cheap Crash-Site Program Repair
by: Zheng, Han, et al.
Published: (2025)
by: Zheng, Han, et al.
Published: (2025)
The Impact Of Bug Localization Based on Crash Report Mining: A Developers' Perspective
by: Medeiros, Marcos, et al.
Published: (2024)
by: Medeiros, Marcos, et al.
Published: (2024)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
by: Kuang, Shiqi, et al.
Published: (2025)
by: Kuang, Shiqi, et al.
Published: (2025)
Green LLM Techniques in Action: How Effective Are Existing Techniques for Improving the Energy Efficiency of LLM-Based Applications in Industry?
by: Kuran, Pelin Rabia, et al.
Published: (2026)
by: Kuran, Pelin Rabia, et al.
Published: (2026)
KTester: Leveraging Domain and Testing Knowledge for More Effective LLM-based Test Generation
by: Li, Anji, et al.
Published: (2025)
by: Li, Anji, et al.
Published: (2025)
CrashFixer: A crash resolution agent for the Linux kernel
by: Mathai, Alex, et al.
Published: (2025)
by: Mathai, Alex, et al.
Published: (2025)
LLM4Perf: Large Language Models Are Effective Samplers for Multi-Objective Performance Modeling
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
FalseCrashReducer: Mitigating False Positive Crashes in OSS-Fuzz-Gen Using Agentic AI
by: Amusuo, Paschal C., et al.
Published: (2025)
by: Amusuo, Paschal C., et al.
Published: (2025)
Using an LLM to Help With Code Understanding
by: Nam, Daye, et al.
Published: (2023)
by: Nam, Daye, et al.
Published: (2023)
EcoScratch: Cost-Effective Multimodal Repair for Scratch Using Execution Feedback
by: Si, Yuan, et al.
Published: (2026)
by: Si, Yuan, et al.
Published: (2026)
EmbC-Test: How to Speed Up Embedded Software Testing Using LLMs and RAG
by: Harnot, Maximilian, et al.
Published: (2026)
by: Harnot, Maximilian, et al.
Published: (2026)
Similar Items
-
Stack Trace-Based Crash Deduplication with Transformer Adaptation
by: Mamun, Md Afif Al, et al.
Published: (2025) -
Fuzzing BusyBox: Leveraging LLM and Crash Reuse for Embedded Bug Unearthing
by: Asmita, et al.
Published: (2024) -
Finding the Needle in the Crash Stack: Industrial-Scale Crash Root Cause Localization with AutoCrashFL
by: Kang, Sungmin, et al.
Published: (2025) -
CrashJS: A NodeJS Benchmark for Automated Crash Reproduction
by: Oliver, Philip, et al.
Published: (2024) -
Outrunning LLM Cutoffs: A Live Kernel Crash Resolution Benchmark for All
by: Huang, Chenxi, et al.
Published: (2026)