From Code Changes to Quality Gains: An Empirical Study in Python ML Systems with PyQu
Fuente:
arXiv
Saved in:
| Main Authors: | Almukhtar, Mohamed, Ghammam, Anwar, Kessentini, Marouane, Ming, Hua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
by: Ghammam, Anwar, et al.
Published: (2026)
by: Ghammam, Anwar, et al.
Published: (2026)
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
by: Almukhtar, Mohamed, et al.
Published: (2026)
by: Almukhtar, Mohamed, et al.
Published: (2026)
Build Code Needs Maintenance Too: A Study on Refactoring and Technical Debt in Build Systems
by: Ghammam, Anwar, et al.
Published: (2025)
by: Ghammam, Anwar, et al.
Published: (2025)
From REST to MCP: An Empirical Study of API Wrapping and Automated Server Generation for LLM Agents
by: Mastouri, Meriem, et al.
Published: (2025)
by: Mastouri, Meriem, et al.
Published: (2025)
Less is More? An Empirical Study on Configuration Issues in Python PyPI Ecosystem
by: Peng, Yun, et al.
Published: (2023)
by: Peng, Yun, et al.
Published: (2023)
Refactoring for Dockerfile Quality: A Dive into Developer Practices and Automation Potential
by: Ksontini, Emna, et al.
Published: (2025)
by: Ksontini, Emna, et al.
Published: (2025)
Optimizing Code Embeddings and ML Classifiers for Python Source Code Vulnerability Detection
by: Farasat, Talaya, et al.
Published: (2025)
by: Farasat, Talaya, et al.
Published: (2025)
An Empirical Study on the Impact of Gender Diversity on Code Quality in AI Systems
by: Cynthia, Shamse Tasnim, et al.
Published: (2025)
by: Cynthia, Shamse Tasnim, et al.
Published: (2025)
PyGress: Tool for Analyzing the Progression of Code Proficiency in Python OSS Projects
by: Charatvaraphan, Rujiphart, et al.
Published: (2025)
by: Charatvaraphan, Rujiphart, et al.
Published: (2025)
Performance Smells in ML and Non-ML Python Projects: A Comparative Study
by: Belias, François, et al.
Published: (2025)
by: Belias, François, et al.
Published: (2025)
Comparing ML-Specific and General Python Code Smells Across Project Characteristics
by: Agh, Halimeh, et al.
Published: (2026)
by: Agh, Halimeh, et al.
Published: (2026)
An Empirical Study of Fault Localization in Python Programs
by: Rezaalipour, Mohammad, et al.
Published: (2023)
by: Rezaalipour, Mohammad, et al.
Published: (2023)
An Empirical Study on the Performance and Energy Usage of Compiled Python Code
by: Stoico, Vincenzo, et al.
Published: (2025)
by: Stoico, Vincenzo, et al.
Published: (2025)
PyTy: Repairing Static Type Errors in Python
by: Chow, Yiu Wai, et al.
Published: (2024)
by: Chow, Yiu Wai, et al.
Published: (2024)
FauxPy: A Fault Localization Tool for Python
by: Rezaalipour, Mohammad, et al.
Published: (2024)
by: Rezaalipour, Mohammad, et al.
Published: (2024)
An Empirical Study on Package-Level Deprecation in Python Ecosystem
by: Zhong, Zhiqing, et al.
Published: (2024)
by: Zhong, Zhiqing, et al.
Published: (2024)
PyTracer: Automatically profiling numerical instabilities in Python
by: Chatelain, Yohan, et al.
Published: (2021)
by: Chatelain, Yohan, et al.
Published: (2021)
DyPyBench: A Benchmark of Executable Python Software
by: Bouzenia, Islem, et al.
Published: (2024)
by: Bouzenia, Islem, et al.
Published: (2024)
An Empirical Study on the Impact of Code Duplication-aware Refactoring Practices on Quality Metrics
by: AlOmar, Eman Abdullah
Published: (2025)
by: AlOmar, Eman Abdullah
Published: (2025)
Delving into Parameter-Efficient Fine-Tuning in Code Change Learning: An Empirical Study
by: Liu, Shuo, et al.
Published: (2024)
by: Liu, Shuo, et al.
Published: (2024)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
BugsInPy: A Database of Existing Bugs in Python Programs to Enable Controlled Testing and Debugging Studies
by: Widyasari, Ratnadira, et al.
Published: (2024)
by: Widyasari, Ratnadira, et al.
Published: (2024)
PyTrim: A Practical Tool for Reducing Python Dependency Bloat
by: Karakatsanis, Konstantinos, et al.
Published: (2025)
by: Karakatsanis, Konstantinos, et al.
Published: (2025)
HaPy-Bug -- Human Annotated Python Bug Resolution Dataset
by: Przymus, Piotr, et al.
Published: (2025)
by: Przymus, Piotr, et al.
Published: (2025)
An Empirical Study of Python Library Migration Using Large Language Models
by: Islam, Md Mohayeminul, et al.
Published: (2025)
by: Islam, Md Mohayeminul, et al.
Published: (2025)
CoQuIR: A Comprehensive Benchmark for Code Quality-Aware Information Retrieval
by: Geng, Jiahui, et al.
Published: (2025)
by: Geng, Jiahui, et al.
Published: (2025)
Sentiment Analysis of ML Projects: Bridging Emotional Intelligence and Code Quality
by: Ahmed, Md Shoaib, et al.
Published: (2024)
by: Ahmed, Md Shoaib, et al.
Published: (2024)
HQPEF-Py: Metrics, Python Patterns, and Guidance for Evaluating Hybrid Quantum Programs
by: Osei, Michael Adjei, et al.
Published: (2025)
by: Osei, Michael Adjei, et al.
Published: (2025)
PyExamine A Comprehensive, UnOpinionated Smell Detection Tool for Python
by: Shivashankar, Karthik, et al.
Published: (2025)
by: Shivashankar, Karthik, et al.
Published: (2025)
TypeEvalPy: A Micro-benchmarking Framework for Python Type Inference Tools
by: Venkatesh, Ashwin Prasad Shivarpatna, et al.
Published: (2023)
by: Venkatesh, Ashwin Prasad Shivarpatna, et al.
Published: (2023)
PyPulse: A Python Library for Biosignal Imputation
by: Gao, Kevin, et al.
Published: (2024)
by: Gao, Kevin, et al.
Published: (2024)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
by: Cheng, Zaiyu, et al.
Published: (2026)
by: Cheng, Zaiyu, et al.
Published: (2026)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
A Comprehensive Multi-Vocal Empirical Study of ML Cloud Service Misuses
by: Amor, Hadil Ben, et al.
Published: (2025)
by: Amor, Hadil Ben, et al.
Published: (2025)
The Product Beyond the Model -- An Empirical Study of Repositories of Open-Source ML Products
by: Nahar, Nadia, et al.
Published: (2023)
by: Nahar, Nadia, et al.
Published: (2023)
Description and Comparative Analysis of QuRE: A New Industrial Requirements Quality Dataset
by: Femmer, Henning, et al.
Published: (2025)
by: Femmer, Henning, et al.
Published: (2025)
An ML-based Approach to Predicting Software Change Dependencies: Insights from an Empirical Study on OpenStack
by: Arabat, Ali, et al.
Published: (2025)
by: Arabat, Ali, et al.
Published: (2025)
Why do Machine Learning Notebooks Crash? An Empirical Study on Public Python Jupyter Notebooks
by: Wang, Yiran, et al.
Published: (2024)
by: Wang, Yiran, et al.
Published: (2024)
An Empirical Study of Large Language Models for Type and Call Graph Analysis in Python and JavaScript
by: Venkatesh, Ashwin Prasad Shivarpatna, et al.
Published: (2024)
by: Venkatesh, Ashwin Prasad Shivarpatna, et al.
Published: (2024)
Precision or Peril: A PoC of Python Code Quality from Quantized Large Language Models
by: Melin, Eric L., et al.
Published: (2024)
by: Melin, Eric L., et al.
Published: (2024)
Similar Items
-
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
by: Ghammam, Anwar, et al.
Published: (2026) -
Quality and Security Signals in AI-Generated Python Refactoring Pull Requests
by: Almukhtar, Mohamed, et al.
Published: (2026) -
Build Code Needs Maintenance Too: A Study on Refactoring and Technical Debt in Build Systems
by: Ghammam, Anwar, et al.
Published: (2025) -
From REST to MCP: An Empirical Study of API Wrapping and Automated Server Generation for LLM Agents
by: Mastouri, Meriem, et al.
Published: (2025) -
Less is More? An Empirical Study on Configuration Issues in Python PyPI Ecosystem
by: Peng, Yun, et al.
Published: (2023)