Cluster-guided LLM-Based Anonymization of Software Analytics Data: Studying Privacy-Utility Trade-offs in JIT Defect Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Khan, Maaz, Khan, Gul Sher, Raza, Ahsan, Ullah, Pir Sami, Bangash, Abdul Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IRJIT: A Simple, Online, Information Retrieval Approach for Just-In-Time Software Defect Prediction
by: Sahar, Hareem, et al.
Published: (2022)
by: Sahar, Hareem, et al.
Published: (2022)
Towards MLOps: A DevOps Tools Recommender System for Machine Learning System
by: Shah, Pir Sami Ullah, et al.
Published: (2024)
by: Shah, Pir Sami Ullah, et al.
Published: (2024)
On the Adoption of AI Coding Agents in Open-source Android and iOS Development
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
by: Khan, Muhammad Ahmad, et al.
Published: (2026)
The State of Documentation Practices of Third-party Machine Learning Models and Datasets
by: Oreamuno, Ernesto Lang, et al.
Published: (2023)
by: Oreamuno, Ernesto Lang, et al.
Published: (2023)
Protecting Privacy in Software Logs: What Should Be Anonymized?
by: Aghili, Roozbeh, et al.
Published: (2024)
by: Aghili, Roozbeh, et al.
Published: (2024)
An Empirical Study on JIT Defect Prediction Based on BERT-style Model
by: Guo, Yuxiang, et al.
Published: (2024)
by: Guo, Yuxiang, et al.
Published: (2024)
Agentic AI in 6G Software Businesses: A Layered Maturity Model
by: Zohaib, Muhammad, et al.
Published: (2025)
by: Zohaib, Muhammad, et al.
Published: (2025)
Enhancing Collaboration for Software Engineers through Matching
by: Azim, Nayaab, et al.
Published: (2025)
by: Azim, Nayaab, et al.
Published: (2025)
An Improved Quantum Software Challenges Classification Approach using Transfer Learning and Explainable AI
by: Khan, Nek Dil, et al.
Published: (2025)
by: Khan, Nek Dil, et al.
Published: (2025)
Evaluating LLM-Based Test Generation Under Software Evolution
by: Haroon, Sabaat, et al.
Published: (2026)
by: Haroon, Sabaat, et al.
Published: (2026)
Java JIT Testing with Template Extraction
by: Zang, Zhiqiang, et al.
Published: (2024)
by: Zang, Zhiqiang, et al.
Published: (2024)
Novice Developers Produce Larger Review Overhead for Project Maintainers while Vibe Coding
by: Asdaque, Syed Ammar, et al.
Published: (2026)
by: Asdaque, Syed Ammar, et al.
Published: (2026)
HAFix: History-Augmented Large Language Models for Bug Fixing
by: Shi, Yu, et al.
Published: (2025)
by: Shi, Yu, et al.
Published: (2025)
Understanding and Finding JIT Compiler Performance Bugs
by: Yi, Zijian, et al.
Published: (2026)
by: Yi, Zijian, et al.
Published: (2026)
Mining Q&A Platforms for Empirical Evidence on Quantum Software Programming
by: Khan, Arif Ali, et al.
Published: (2025)
by: Khan, Arif Ali, et al.
Published: (2025)
On the Footprints of Reviewer Bots Feedback on Agentic Pull Requests in OSS GitHub Repositories
by: Fatima, Syeda Kaneez, et al.
Published: (2026)
by: Fatima, Syeda Kaneez, et al.
Published: (2026)
Contrasting the Hyperparameter Tuning Impact Across Software Defect Prediction Scenarios
by: Rakha, Mohamed Sami, et al.
Published: (2025)
by: Rakha, Mohamed Sami, et al.
Published: (2025)
The State of the SBOM Tool Ecosystems: A Comparative Analysis of SPDX and CycloneDX
by: Bangash, Abdul Ali, et al.
Published: (2025)
by: Bangash, Abdul Ali, et al.
Published: (2025)
On the synchronization between Hugging Face pre-trained language models and their upstream GitHub repository
by: Ajibode, Adekunle, et al.
Published: (2025)
by: Ajibode, Adekunle, et al.
Published: (2025)
Towards Semantic Versioning of Open Pre-trained Language Model Releases on Hugging Face
by: Ajibode, Adekunle, et al.
Published: (2024)
by: Ajibode, Adekunle, et al.
Published: (2024)
An Empirical Study of Challenges in Machine Learning Asset Management
by: Zhao, Zhimin, et al.
Published: (2024)
by: Zhao, Zhimin, et al.
Published: (2024)
Advancing Quantum Software Engineering: A Vision of Hybrid Full-Stack Iterative Model
by: Khan, Arif Ali, et al.
Published: (2024)
by: Khan, Arif Ali, et al.
Published: (2024)
Investigating Adversarial Attacks in Software Analytics via Machine Learning Explainability
by: Awal, MD Abdul, et al.
Published: (2024)
by: Awal, MD Abdul, et al.
Published: (2024)
On the Workflows and Smells of Leaderboard Operations (LBOps): An Exploratory Study of Foundation Model Leaderboards
by: Zhao, Zhimin, et al.
Published: (2024)
by: Zhao, Zhimin, et al.
Published: (2024)
Pattern-Based Peephole Optimizations with Java JIT Tests
by: Zang, Zhiqiang, et al.
Published: (2024)
by: Zang, Zhiqiang, et al.
Published: (2024)
Large Language Models as Robust Data Generators in Software Analytics: Are We There Yet?
by: Awal, Md. Abdul, et al.
Published: (2024)
by: Awal, Md. Abdul, et al.
Published: (2024)
Advancing Mobile UI Testing by Learning Screen Usage Semantics
by: Khan, Safwat Ali
Published: (2025)
by: Khan, Safwat Ali
Published: (2025)
A Dataset of Low-Rated Applications from the Amazon Appstore for User Feedback Analysis
by: Khan, Nek Dil, et al.
Published: (2026)
by: Khan, Nek Dil, et al.
Published: (2026)
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
by: Zhao, Zhimin, et al.
Published: (2026)
by: Zhao, Zhimin, et al.
Published: (2026)
Prioritizing Software Requirements Using Large Language Models
by: Sami, Malik Abdul, et al.
Published: (2024)
by: Sami, Malik Abdul, et al.
Published: (2024)
EvaluateXAI: A Framework to Evaluate the Reliability and Consistency of Rule-based XAI Techniques for Software Analytics Tasks
by: Awal, Md Abdul, et al.
Published: (2024)
by: Awal, Md Abdul, et al.
Published: (2024)
The Role of Generative AI in Strengthening Secure Software Coding Practices: A Systematic Perspective
by: Alwageed, Hathal S., et al.
Published: (2025)
by: Alwageed, Hathal S., et al.
Published: (2025)
Reliability of AI Bots Footprints in GitHub Actions CI/CD Workflows
by: Shah, Syed Muhammad Ashhar, et al.
Published: (2026)
by: Shah, Syed Muhammad Ashhar, et al.
Published: (2026)
Strategic Motivators for Ethical AI System Development: An Empirical and Holistic Model
by: Akbar, Muhammad Azeem, et al.
Published: (2025)
by: Akbar, Muhammad Azeem, et al.
Published: (2025)
Experimenting with Multi-Agent Software Development: Towards a Unified Platform
by: Sami, Malik Abdul, et al.
Published: (2024)
by: Sami, Malik Abdul, et al.
Published: (2024)
Vibe Coding on Trial: Operating Characteristics of Unanimous LLM Juries
by: Ullah, Muhammad Aziz, et al.
Published: (2026)
by: Ullah, Muhammad Aziz, et al.
Published: (2026)
Understanding Prompt Management in GitHub Repositories: A Call for Best Practices
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Practical Feasibility of Sustainable Software Engineering Tools and Techniques
by: Ghanta, Satwik, et al.
Published: (2026)
by: Ghanta, Satwik, et al.
Published: (2026)
TimeLess: A Vision for the Next Generation of Software Development
by: Rasheed, Zeeshan, et al.
Published: (2024)
by: Rasheed, Zeeshan, et al.
Published: (2024)
BioDefect: The First Dataset for Defect Detection in Bioinformatics Software
by: Xu, Tianxiang, et al.
Published: (2026)
by: Xu, Tianxiang, et al.
Published: (2026)
Similar Items
-
IRJIT: A Simple, Online, Information Retrieval Approach for Just-In-Time Software Defect Prediction
by: Sahar, Hareem, et al.
Published: (2022) -
Towards MLOps: A DevOps Tools Recommender System for Machine Learning System
by: Shah, Pir Sami Ullah, et al.
Published: (2024) -
On the Adoption of AI Coding Agents in Open-source Android and iOS Development
by: Khan, Muhammad Ahmad, et al.
Published: (2026) -
The State of Documentation Practices of Third-party Machine Learning Models and Datasets
by: Oreamuno, Ernesto Lang, et al.
Published: (2023) -
Protecting Privacy in Software Logs: What Should Be Anonymized?
by: Aghili, Roozbeh, et al.
Published: (2024)