Test Case Prioritization: A Snowballing Literature Review and TCPFramework with Approach Combinators
Fuente:
arXiv
Guardado en:
| Autores principales: | Chojnacki, Tomasz, Madeyski, Lech |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Triage: Routing Software Engineering Tasks to Cost-Effective LLM Tiers via Code Quality Signals
por: Madeyski, Lech
Publicado: (2026)
por: Madeyski, Lech
Publicado: (2026)
LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews
por: Madeyski, Lech, et al.
Publicado: (2025)
por: Madeyski, Lech, et al.
Publicado: (2025)
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
por: Chang, Hung-Fu, et al.
Publicado: (2025)
por: Chang, Hung-Fu, et al.
Publicado: (2025)
Learning Software Bug Reports: A Systematic Literature Review
por: Long, Guoming, et al.
Publicado: (2025)
por: Long, Guoming, et al.
Publicado: (2025)
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
por: Ramírez, Aurora, et al.
Publicado: (2024)
por: Ramírez, Aurora, et al.
Publicado: (2024)
Towards Continuous Assurance with Formal Verification and Assurance Cases
por: Abeywickrama, Dhaminda B., et al.
Publicado: (2025)
por: Abeywickrama, Dhaminda B., et al.
Publicado: (2025)
Fine-Tuning LLMs to Analyze Multiple Dimensions of Code Review: A Maximum Entropy Regulated Long Chain-of-Thought Approach
por: Yu, Yongda, et al.
Publicado: (2025)
por: Yu, Yongda, et al.
Publicado: (2025)
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
por: Li, Meiziniu, et al.
Publicado: (2024)
por: Li, Meiziniu, et al.
Publicado: (2024)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
por: Bradbury, Jeremy S., et al.
Publicado: (2024)
por: Bradbury, Jeremy S., et al.
Publicado: (2024)
Leveraging Large Language Models for Use Case Model Generation from Software Requirements
por: Eisenreich, Tobias, et al.
Publicado: (2025)
por: Eisenreich, Tobias, et al.
Publicado: (2025)
CCCI: Code Completion with Contextual Information for Complex Data Transfer Tasks Using Large Language Models
por: Jin, Hangzhan, et al.
Publicado: (2025)
por: Jin, Hangzhan, et al.
Publicado: (2025)
Adversarial Agent Collaboration for Correctness Improvements of C to Safe Rust Translation
por: Li, Tianyu, et al.
Publicado: (2025)
por: Li, Tianyu, et al.
Publicado: (2025)
The Rise and Fall(?) of Software Engineering
por: Mastropaolo, Antonio, et al.
Publicado: (2024)
por: Mastropaolo, Antonio, et al.
Publicado: (2024)
Assessing LLMs for Front-end Software Architecture Knowledge
por: Guerra, L. P. Franciscatto, et al.
Publicado: (2025)
por: Guerra, L. P. Franciscatto, et al.
Publicado: (2025)
You Don't Need Public Tests to Generate Correct Code
por: Silva, Kaushitha, et al.
Publicado: (2026)
por: Silva, Kaushitha, et al.
Publicado: (2026)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
por: Kessel, Marcus
Publicado: (2024)
por: Kessel, Marcus
Publicado: (2024)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
por: Terragni, Valerio
Publicado: (2026)
por: Terragni, Valerio
Publicado: (2026)
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
por: Li, Meiziniu, et al.
Publicado: (2022)
por: Li, Meiziniu, et al.
Publicado: (2022)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
por: Rehan, Tzafrir
Publicado: (2026)
por: Rehan, Tzafrir
Publicado: (2026)
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
por: More, Riddhi, et al.
Publicado: (2025)
por: More, Riddhi, et al.
Publicado: (2025)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
por: Nguyen, Quang-Dung, et al.
Publicado: (2025)
por: Nguyen, Quang-Dung, et al.
Publicado: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
por: Palacios, Diego Cabezas
Publicado: (2026)
por: Palacios, Diego Cabezas
Publicado: (2026)
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
por: Bhardwaj, Varun Pratap
Publicado: (2026)
por: Bhardwaj, Varun Pratap
Publicado: (2026)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
por: More, Riddhi, et al.
Publicado: (2025)
por: More, Riddhi, et al.
Publicado: (2025)
Distilling Desired Comments for Enhanced Code Review with Large Language Models
por: Yu, Yongda, et al.
Publicado: (2024)
por: Yu, Yongda, et al.
Publicado: (2024)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
por: Cambronero, José, et al.
Publicado: (2025)
por: Cambronero, José, et al.
Publicado: (2025)
ATLAS: A Layered Constraint-Guided Framework for Structured Artifact Generation in LLM-Assisted MDE
por: Ma, Tong, et al.
Publicado: (2025)
por: Ma, Tong, et al.
Publicado: (2025)
Social, Legal, Ethical, Empathetic and Cultural Norm Operationalisation for AI Agents
por: Calinescu, Radu, et al.
Publicado: (2026)
por: Calinescu, Radu, et al.
Publicado: (2026)
CodeCompass: Navigating the Navigation Paradox in Agentic Code Intelligence
por: Paipuru, Tarakanath
Publicado: (2026)
por: Paipuru, Tarakanath
Publicado: (2026)
Automating Domain-Driven Design: Experience with a Prompting Framework
por: Eisenreich, Tobias, et al.
Publicado: (2026)
por: Eisenreich, Tobias, et al.
Publicado: (2026)
Quantitative Analysis of Technical Debt and Pattern Violation in Large Language Model Architectures
por: Slater, Tyler
Publicado: (2025)
por: Slater, Tyler
Publicado: (2025)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
por: Ravi, Ravin, et al.
Publicado: (2026)
por: Ravi, Ravin, et al.
Publicado: (2026)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
por: Trooskens, Geert, et al.
Publicado: (2026)
por: Trooskens, Geert, et al.
Publicado: (2026)
Experience with GitHub Copilot for Developer Productivity at Zoominfo
por: Bakal, Gal, et al.
Publicado: (2025)
por: Bakal, Gal, et al.
Publicado: (2025)
LLM-Assisted Translation of Legacy FORTRAN Codes to C++: A Cross-Platform Study
por: Ranasinghe, Nishath Rajiv, et al.
Publicado: (2025)
por: Ranasinghe, Nishath Rajiv, et al.
Publicado: (2025)
EvoGraph: Hybrid Directed Graph Evolution toward Software 3.0
por: Costa, Igor, et al.
Publicado: (2025)
por: Costa, Igor, et al.
Publicado: (2025)
A Pattern Language for Resilient Visual Agents
por: Gidey, Habtom Kahsay, et al.
Publicado: (2026)
por: Gidey, Habtom Kahsay, et al.
Publicado: (2026)
SLEGO: A Collaborative Data Analytics System with LLM Recommender for Diverse Users
por: Ng, Siu Lung, et al.
Publicado: (2024)
por: Ng, Siu Lung, et al.
Publicado: (2024)
Rocks Coding, Not Development--A Human-Centric, Experimental Evaluation of LLM-Supported SE Tasks
por: Wang, Wei, et al.
Publicado: (2024)
por: Wang, Wei, et al.
Publicado: (2024)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
Ejemplares similares
-
Triage: Routing Software Engineering Tasks to Cost-Effective LLM Tiers via Code Quality Signals
por: Madeyski, Lech
Publicado: (2026) -
LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews
por: Madeyski, Lech, et al.
Publicado: (2025) -
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
por: Chang, Hung-Fu, et al.
Publicado: (2025) -
Learning Software Bug Reports: A Systematic Literature Review
por: Long, Guoming, et al.
Publicado: (2025) -
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
por: Ramírez, Aurora, et al.
Publicado: (2024)