Revisiting Code Debloating with Ground Truth-based Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Bilal, Muhammad, Ali, Moiz, Kumar, Mohit, Zaffar, Fareed, Shaon, Fahad, Gehani, Ashish, Rahaman, Sazzadur |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SoK: Software Debloating Landscape and Future Directions
by: Alhanahnah, Mohannad, et al.
Published: (2024)
by: Alhanahnah, Mohannad, et al.
Published: (2024)
An Investigation into Protestware
by: Finken, Tanner, et al.
Published: (2024)
by: Finken, Tanner, et al.
Published: (2024)
GraphQLify: Automated and Type Safety-Preserving GraphQL API Adoption
by: Amareen, Saleh, et al.
Published: (2026)
by: Amareen, Saleh, et al.
Published: (2026)
AI-Generated Code Is Not Reproducible (Yet): An Empirical Study of Dependency Gaps in LLM-Based Coding Agents
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
The Cure is in the Cause: A Filesystem for Container Debloating
by: Zhang, Huaifeng, et al.
Published: (2023)
by: Zhang, Huaifeng, et al.
Published: (2023)
On the Contents and Utility of IoT Cybersecurity Guidelines
by: Chen, Jesse, et al.
Published: (2023)
by: Chen, Jesse, et al.
Published: (2023)
Improving Program Debloating with 1-DU Chain Minimality
by: Kim, Myeongsoo, et al.
Published: (2024)
by: Kim, Myeongsoo, et al.
Published: (2024)
A Soundness and Precision Benchmark for Java Debloating Tools
by: Klauke, Jonas, et al.
Published: (2025)
by: Klauke, Jonas, et al.
Published: (2025)
A Broad Comparative Evaluation of Software Debloating Tools
by: Brown, Michael D., et al.
Published: (2023)
by: Brown, Michael D., et al.
Published: (2023)
Large Language Models-Aided Program Debloating
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
RVDebloater: Mode-based Adaptive Firmware Debloating for Robotic Vehicles
by: Salehi, Mohsen, et al.
Published: (2026)
by: Salehi, Mohsen, et al.
Published: (2026)
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision?
by: Fan, Lishui, et al.
Published: (2026)
by: Fan, Lishui, et al.
Published: (2026)
Detecting Call Graph Unsoundness without Ground Truth
by: Zhong, Fangtian, et al.
Published: (2026)
by: Zhong, Fangtian, et al.
Published: (2026)
Can Code Evaluation Metrics Detect Code Plagiarism?
by: Ebrahim, Fahad, et al.
Published: (2026)
by: Ebrahim, Fahad, et al.
Published: (2026)
A Ground-Truth-Based Evaluation of Vulnerability Detection Across Multiple Ecosystems
by: Mandl, Peter, et al.
Published: (2026)
by: Mandl, Peter, et al.
Published: (2026)
Enhanced LLM-Based Framework for Predicting Null Pointer Dereference in Source Code
by: Sultan, Md. Fahim, et al.
Published: (2024)
by: Sultan, Md. Fahim, et al.
Published: (2024)
Impeding LLM-assisted Cheating in Introductory Programming Assignments via Adversarial Perturbation
by: Salim, Saiful Islam, et al.
Published: (2024)
by: Salim, Saiful Islam, et al.
Published: (2024)
MutaGReP: Execution-Free Repository-Grounded Plan Search for Code-Use
by: Khan, Zaid, et al.
Published: (2025)
by: Khan, Zaid, et al.
Published: (2025)
LeanBin: Harnessing Lifting and Recompilation to Debloat Binaries
by: Wodiany, Igor, et al.
Published: (2024)
by: Wodiany, Igor, et al.
Published: (2024)
Consistency Meets Verification: Enhancing Test Generation Quality in Large Language Models Without Ground-Truth Solutions
by: Taherkhani, Hamed, et al.
Published: (2026)
by: Taherkhani, Hamed, et al.
Published: (2026)
Revisiting Code Search in a Two-Stage Paradigm
by: Hu, Fan, et al.
Published: (2022)
by: Hu, Fan, et al.
Published: (2022)
GAN-enhanced Simulation-driven DNN Testing in Absence of Ground Truth
by: Attaoui, Mohammed, et al.
Published: (2025)
by: Attaoui, Mohammed, et al.
Published: (2025)
On the Robustness of Fairness Practices: A Causal Framework for Systematic Evaluation
by: Monjezi, Verya, et al.
Published: (2026)
by: Monjezi, Verya, et al.
Published: (2026)
Low-Code Paradox in DevOps: Security and Governance Insights from Practitioners
by: Akbar, Muhammad Azeem, et al.
Published: (2026)
by: Akbar, Muhammad Azeem, et al.
Published: (2026)
Code Hallucination
by: Rahman, Mirza Masfiqur, et al.
Published: (2024)
by: Rahman, Mirza Masfiqur, et al.
Published: (2024)
Revisiting Evolutionary Program Repair via Code Language Model
by: Wang, Yunan, et al.
Published: (2024)
by: Wang, Yunan, et al.
Published: (2024)
EnStack: An Ensemble Stacking Framework of Large Language Models for Enhanced Vulnerability Detection in Source Code
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
by: Ridoy, Shahriyar Zaman, et al.
Published: (2024)
Benchmarking and Revisiting Code Generation Assessment: A Mutation-Based Approach
by: Wang, Longtian, et al.
Published: (2025)
by: Wang, Longtian, et al.
Published: (2025)
CodeScore: Evaluating Code Generation by Learning Code Execution
by: Dong, Yihong, et al.
Published: (2023)
by: Dong, Yihong, et al.
Published: (2023)
UniCode: Augmenting Evaluation for Code Reasoning
by: Zheng, Xinyue, et al.
Published: (2025)
by: Zheng, Xinyue, et al.
Published: (2025)
SGCR: A Specification-Grounded Framework for Trustworthy LLM Code Review
by: Wang, Kai, et al.
Published: (2025)
by: Wang, Kai, et al.
Published: (2025)
On the Adoption of AI Coding Agents in Open-source Android and iOS Development
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
Revisiting the Role of Natural Language Code Comments in Code Translation
by: Gupta, Monika, et al.
Published: (2026)
by: Gupta, Monika, et al.
Published: (2026)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
by: Zhang, Binquan, et al.
Published: (2025)
by: Zhang, Binquan, et al.
Published: (2025)
Sustainable Code Generation Using Large Language Models: A Systematic Literature Review
by: Ali, Sabiya Banu Masthan, et al.
Published: (2026)
by: Ali, Sabiya Banu Masthan, et al.
Published: (2026)
InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation
by: Chen, Qiaosheng, et al.
Published: (2025)
by: Chen, Qiaosheng, et al.
Published: (2025)
Novice Developers Produce Larger Review Overhead for Project Maintainers while Vibe Coding
by: Asdaque, Syed Ammar, et al.
Published: (2026)
by: Asdaque, Syed Ammar, et al.
Published: (2026)
LogSD: Detecting Anomalies from System Logs through Self-supervised Learning and Frequency-based Masking
by: Xie, Yongzheng, et al.
Published: (2024)
by: Xie, Yongzheng, et al.
Published: (2024)
Evaluating Small-Scale Code Models for Code Clone Detection
by: Martinez-Gil, Jorge
Published: (2025)
by: Martinez-Gil, Jorge
Published: (2025)
Exploring the Impact of Code Style in Identifying Good Programmers
by: Yasir, Rafed Muhammad, et al.
Published: (2022)
by: Yasir, Rafed Muhammad, et al.
Published: (2022)
Similar Items
-
SoK: Software Debloating Landscape and Future Directions
by: Alhanahnah, Mohannad, et al.
Published: (2024) -
An Investigation into Protestware
by: Finken, Tanner, et al.
Published: (2024) -
GraphQLify: Automated and Type Safety-Preserving GraphQL API Adoption
by: Amareen, Saleh, et al.
Published: (2026) -
AI-Generated Code Is Not Reproducible (Yet): An Empirical Study of Dependency Gaps in LLM-Based Coding Agents
by: Vangala, Bhanu Prakash, et al.
Published: (2025) -
The Cure is in the Cause: A Filesystem for Container Debloating
by: Zhang, Huaifeng, et al.
Published: (2023)