Harden and Catch for Just-in-Time Assured LLM-Based Software Testing: Open Research Challenges
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Harman, Mark, O'Hearn, Peter, Sengupta, Shubho |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Assured LLM-Based Software Engineering
par: Alshahwan, Nadia, et autres
Publié: (2024)
par: Alshahwan, Nadia, et autres
Publié: (2024)
Just-in-Time Catching Test Generation at Meta
par: Becker, Matthew, et autres
Publié: (2026)
par: Becker, Matthew, et autres
Publié: (2026)
Mutation-Guided LLM-based Test Generation at Meta
par: Foster, Christopher, et autres
Publié: (2025)
par: Foster, Christopher, et autres
Publié: (2025)
WybeCoder: Verified Imperative Code Generation
par: Gloeckle, Fabian, et autres
Publié: (2026)
par: Gloeckle, Fabian, et autres
Publié: (2026)
Generative AI for Testing of Autonomous Driving Systems: A Survey
par: Song, Qunying, et autres
Publié: (2025)
par: Song, Qunying, et autres
Publié: (2025)
LLM-Based Agentic Systems for Software Engineering: Challenges and Opportunities
par: Tang, Yongjian, et autres
Publié: (2026)
par: Tang, Yongjian, et autres
Publié: (2026)
Evaluating LLM-Based Test Generation Under Software Evolution
par: Haroon, Sabaat, et autres
Publié: (2026)
par: Haroon, Sabaat, et autres
Publié: (2026)
Multimodal Learning for Just-In-Time Software Defect Prediction in Autonomous Driving Systems
par: Mohammad, Faisal, et autres
Publié: (2025)
par: Mohammad, Faisal, et autres
Publié: (2025)
Just-In-Time Software Defect Prediction via Bi-modal Change Representation Learning
par: Jiang, Yuze, et autres
Publié: (2024)
par: Jiang, Yuze, et autres
Publié: (2024)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
par: Dobslaw, Felix, et autres
Publié: (2025)
par: Dobslaw, Felix, et autres
Publié: (2025)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
par: Chen, Zhi, et autres
Publié: (2026)
par: Chen, Zhi, et autres
Publié: (2026)
Reasoning-Based Software Testing
par: Giamattei, Luca, et autres
Publié: (2023)
par: Giamattei, Luca, et autres
Publié: (2023)
Can LLM Generate Regression Tests for Software Commits?
par: Liu, Jing, et autres
Publié: (2025)
par: Liu, Jing, et autres
Publié: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
par: Yu, Jiongchi, et autres
Publié: (2025)
par: Yu, Jiongchi, et autres
Publié: (2025)
ReDef: Do Code Language Models Truly Understand Code Changes for Just-in-Time Software Defect Prediction?
par: Nam, Doha, et autres
Publié: (2025)
par: Nam, Doha, et autres
Publié: (2025)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
par: Lin, Zhihao, et autres
Publié: (2024)
par: Lin, Zhihao, et autres
Publié: (2024)
Large Language Models for Software Testing: A Research Roadmap
par: Augusto, Cristian, et autres
Publié: (2025)
par: Augusto, Cristian, et autres
Publié: (2025)
Automated Unit Test Improvement using Large Language Models at Meta
par: Alshahwan, Nadia, et autres
Publié: (2024)
par: Alshahwan, Nadia, et autres
Publié: (2024)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
par: Trae Research Team, et autres
Publié: (2025)
par: Trae Research Team, et autres
Publié: (2025)
Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report
par: Sengupta, Satadru, et autres
Publié: (2026)
par: Sengupta, Satadru, et autres
Publié: (2026)
Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead
par: Rao, Hongzhou, et autres
Publié: (2025)
par: Rao, Hongzhou, et autres
Publié: (2025)
Rethinking Testing for LLM Applications: Characteristics, Challenges, and a Lightweight Interaction Protocol
par: Ma, Wei, et autres
Publié: (2025)
par: Ma, Wei, et autres
Publié: (2025)
Towards Specification-Driven LLM-Based Generation of Embedded Automotive Software
par: Patil, Minal Suresh, et autres
Publié: (2024)
par: Patil, Minal Suresh, et autres
Publié: (2024)
Fairness Is Not Just Ethical: Performance Trade-Off via Data Correlation Tuning to Mitigate Bias in ML Software
par: Xiao, Ying, et autres
Publié: (2025)
par: Xiao, Ying, et autres
Publié: (2025)
"Don't Be Afraid, Just Learn": Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI
par: Otten, Daniel, et autres
Publié: (2026)
par: Otten, Daniel, et autres
Publié: (2026)
Thinking Longer, Not Larger: Enhancing Software Engineering Agents via Scaling Test-Time Compute
par: Ma, Yingwei, et autres
Publié: (2025)
par: Ma, Yingwei, et autres
Publié: (2025)
Fuzzy Inference System for Test Case Prioritization in Software Testing
par: Karatayev, Aron, et autres
Publié: (2024)
par: Karatayev, Aron, et autres
Publié: (2024)
FVSpec: Real-World Property-Based Tests as Lean Challenges
par: Dougherty, Quinn, et autres
Publié: (2026)
par: Dougherty, Quinn, et autres
Publié: (2026)
Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering Studies
par: Angermeir, Florian, et autres
Publié: (2025)
par: Angermeir, Florian, et autres
Publié: (2025)
OLAF: Towards Robust LLM-Based Annotation Framework in Empirical Software Engineering
par: Imran, Mia Mohammad, et autres
Publié: (2025)
par: Imran, Mia Mohammad, et autres
Publié: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
par: Baqar, Mohammad, et autres
Publié: (2024)
par: Baqar, Mohammad, et autres
Publié: (2024)
A Defect Classification Framework for AI-Based Software Systems (AI-ODC)
par: Alannsary, Mohammed O.
Publié: (2025)
par: Alannsary, Mohammed O.
Publié: (2025)
LLM-Based Robustness Testing of Microservice Applications: An Empirical Study
par: Tigulla, Hrushitha Goud, et autres
Publié: (2026)
par: Tigulla, Hrushitha Goud, et autres
Publié: (2026)
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
par: Ziftci, Celal, et autres
Publié: (2026)
par: Ziftci, Celal, et autres
Publié: (2026)
Test It Before You Trust It: Applying Software Testing for Trustworthy In-context Learning
par: Racharak, Teeradaj, et autres
Publié: (2025)
par: Racharak, Teeradaj, et autres
Publié: (2025)
Multi-Agent LLM Committees for Autonomous Software Beta Testing
par: Karanam, Sumanth Bharadwaj Hachalli, et autres
Publié: (2025)
par: Karanam, Sumanth Bharadwaj Hachalli, et autres
Publié: (2025)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
par: Liu, Shu, et autres
Publié: (2026)
par: Liu, Shu, et autres
Publié: (2026)
The Role of Artificial Intelligence and Machine Learning in Software Testing
par: Ramadan, Ahmed, et autres
Publié: (2024)
par: Ramadan, Ahmed, et autres
Publié: (2024)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
par: Ran, Dezhi, et autres
Publié: (2024)
par: Ran, Dezhi, et autres
Publié: (2024)
Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios
par: Hu, Ruida, et autres
Publié: (2026)
par: Hu, Ruida, et autres
Publié: (2026)
Documents similaires
-
Assured LLM-Based Software Engineering
par: Alshahwan, Nadia, et autres
Publié: (2024) -
Just-in-Time Catching Test Generation at Meta
par: Becker, Matthew, et autres
Publié: (2026) -
Mutation-Guided LLM-based Test Generation at Meta
par: Foster, Christopher, et autres
Publié: (2025) -
WybeCoder: Verified Imperative Code Generation
par: Gloeckle, Fabian, et autres
Publié: (2026) -
Generative AI for Testing of Autonomous Driving Systems: A Survey
par: Song, Qunying, et autres
Publié: (2025)