MUCOCO: Automated Consistency Testing of Code LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Chou, Chua Jin, Lwin, Khant That, Soremekun, Ezekiel |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Directed Grammar-Based Test Generation
par: Kirschner, Lukas, et autres
Publié: (2025)
par: Kirschner, Lukas, et autres
Publié: (2025)
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
par: Sohn, Jeongju, et autres
Publié: (2025)
par: Sohn, Jeongju, et autres
Publié: (2025)
Malicious ML Model Detection by Learning Dynamic Behaviors
par: Nambiar, Sarang, et autres
Publié: (2026)
par: Nambiar, Sarang, et autres
Publié: (2026)
End-user Comprehension of Transfer Risks in Smart Contracts
par: Panicker, Yustynn, et autres
Publié: (2024)
par: Panicker, Yustynn, et autres
Publié: (2024)
Software Fairness: An Analysis and Survey
par: Soremekun, Ezekiel, et autres
Publié: (2022)
par: Soremekun, Ezekiel, et autres
Publié: (2022)
Distribution-aware Fairness Test Generation
par: Rajan, Sai Sathiesh, et autres
Publié: (2023)
par: Rajan, Sai Sathiesh, et autres
Publié: (2023)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
par: Khant, Kyi Shin, et autres
Publié: (2025)
par: Khant, Kyi Shin, et autres
Publié: (2025)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
par: Huang, Huihui, et autres
Publié: (2026)
par: Huang, Huihui, et autres
Publié: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
par: Li, Ziyu, et autres
Publié: (2024)
par: Li, Ziyu, et autres
Publié: (2024)
Write Your Own CodeChecker: An Automated Test-Driven Checker Development Approach with LLMs
par: Liu, Jun, et autres
Publié: (2024)
par: Liu, Jun, et autres
Publié: (2024)
Test Oracle Automation in the era of LLMs
par: Molina, Facundo, et autres
Publié: (2024)
par: Molina, Facundo, et autres
Publié: (2024)
Automated Harmfulness Testing for Code Large Language Models
par: Tan, Honghao, et autres
Publié: (2025)
par: Tan, Honghao, et autres
Publié: (2025)
Consistent or Sensitive? Automated Code Revision Tools Against Semantics-Preserving Perturbations
par: Pirouzkhah, Shirin, et autres
Publié: (2026)
par: Pirouzkhah, Shirin, et autres
Publié: (2026)
Unprecedented Code Change Automation: The Fusion of LLMs and Transformation by Example
par: Dilhara, Malinda, et autres
Publié: (2024)
par: Dilhara, Malinda, et autres
Publié: (2024)
Automated and Context-Aware Code Documentation Leveraging Advanced LLMs
par: Sarker, Swapnil Sharma, et autres
Publié: (2025)
par: Sarker, Swapnil Sharma, et autres
Publié: (2025)
SimCT: A Simple Consistency Test Protocol in LLMs Development Lifecycle
par: Zhao, Fufangchen, et autres
Publié: (2024)
par: Zhao, Fufangchen, et autres
Publié: (2024)
Extracting Formal Specifications from Documents Using LLMs for Automated Testing
par: Li, Hui, et autres
Publié: (2025)
par: Li, Hui, et autres
Publié: (2025)
Automated Refactoring of Non-Idiomatic Python Code: A Differentiated Replication with LLMs
par: Midolo, Alessandro, et autres
Publié: (2025)
par: Midolo, Alessandro, et autres
Publié: (2025)
Automating Code Generation for Semiconductor Equipment Control from Developer Utterances with LLMs
par: Kim, Youngkyoung, et autres
Publié: (2025)
par: Kim, Youngkyoung, et autres
Publié: (2025)
Figma2Code: Automating Multimodal Design to Code in the Wild
par: Gui, Yi, et autres
Publié: (2026)
par: Gui, Yi, et autres
Publié: (2026)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
par: Rao, Nikitha, et autres
Publié: (2024)
par: Rao, Nikitha, et autres
Publié: (2024)
Automated Soap Opera Testing Directed by LLMs and Scenario Knowledge: Feasibility, Challenges, and Road Ahead
par: Su, Yanqi, et autres
Publié: (2024)
par: Su, Yanqi, et autres
Publié: (2024)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
par: Cheng, Yiran, et autres
Publié: (2025)
par: Cheng, Yiran, et autres
Publié: (2025)
ConAIR:Consistency-Augmented Iterative Interaction Framework to Enhance the Reliability of Code Generation
par: Dong, Jinhao, et autres
Publié: (2024)
par: Dong, Jinhao, et autres
Publié: (2024)
Leveraging LLMs for Automated Translation of Legacy Code: A Case Study on PL/SQL to Java Transformation
par: Solovyeva, Lola, et autres
Publié: (2025)
par: Solovyeva, Lola, et autres
Publié: (2025)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
par: Li, Junjie, et autres
Publié: (2025)
par: Li, Junjie, et autres
Publié: (2025)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
par: Zhu, Yuqi, et autres
Publié: (2025)
par: Zhu, Yuqi, et autres
Publié: (2025)
LLMs for Automated Unit Test Generation and Assessment in Java: The AgoneTest Framework
par: Lops, Andrea, et autres
Publié: (2025)
par: Lops, Andrea, et autres
Publié: (2025)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
par: Lwin, Naing Oo, et autres
Publié: (2026)
par: Lwin, Naing Oo, et autres
Publié: (2026)
REACCEPT: Automated Co-evolution of Production and Test Code Based on Dynamic Validation and Large Language Models
par: Chi, Jianlei, et autres
Publié: (2024)
par: Chi, Jianlei, et autres
Publié: (2024)
Automated Code Review In Practice
par: Cihan, Umut, et autres
Publié: (2024)
par: Cihan, Umut, et autres
Publié: (2024)
CA2: Code-Aware Agent for Automated Game Testing
par: Adaikkappan, Valliappan Chidambaram, et autres
Publié: (2026)
par: Adaikkappan, Valliappan Chidambaram, et autres
Publié: (2026)
Hybrid Privacy Policy-Code Consistency Check using Knowledge Graphs and LLMs
par: Mao, Zhenyu, et autres
Publié: (2025)
par: Mao, Zhenyu, et autres
Publié: (2025)
TAM-Eval: Evaluating LLMs for Automated Unit Test Maintenance
par: Bruches, Elena, et autres
Publié: (2026)
par: Bruches, Elena, et autres
Publié: (2026)
PrediQL: Automated Testing of GraphQL APIs with LLMs
par: Liu, Shaolun, et autres
Publié: (2025)
par: Liu, Shaolun, et autres
Publié: (2025)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
par: Sherifi, Betim, et autres
Publié: (2024)
par: Sherifi, Betim, et autres
Publié: (2024)
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
par: Jin, Shuo, et autres
Publié: (2025)
par: Jin, Shuo, et autres
Publié: (2025)
Invariant-Driven Automated Testing
par: Ribeiro, Ana Catarina
Publié: (2026)
par: Ribeiro, Ana Catarina
Publié: (2026)
Automated Unit Test Refactoring
par: Gao, Yi, et autres
Publié: (2024)
par: Gao, Yi, et autres
Publié: (2024)
Enabling Cost-Effective UI Automation Testing with Retrieval-Based LLMs: A Case Study in WeChat
par: Feng, Sidong, et autres
Publié: (2024)
par: Feng, Sidong, et autres
Publié: (2024)
Documents similaires
-
Directed Grammar-Based Test Generation
par: Kirschner, Lukas, et autres
Publié: (2025) -
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
par: Sohn, Jeongju, et autres
Publié: (2025) -
Malicious ML Model Detection by Learning Dynamic Behaviors
par: Nambiar, Sarang, et autres
Publié: (2026) -
End-user Comprehension of Transfer Risks in Smart Contracts
par: Panicker, Yustynn, et autres
Publié: (2024) -
Software Fairness: An Analysis and Survey
par: Soremekun, Ezekiel, et autres
Publié: (2022)