MUCOCO: Automated Consistency Testing of Code LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chou, Chua Jin, Lwin, Khant That, Soremekun, Ezekiel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Directed Grammar-Based Test Generation
von: Kirschner, Lukas, et al.
Veröffentlicht: (2025)
von: Kirschner, Lukas, et al.
Veröffentlicht: (2025)
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
von: Sohn, Jeongju, et al.
Veröffentlicht: (2025)
von: Sohn, Jeongju, et al.
Veröffentlicht: (2025)
Malicious ML Model Detection by Learning Dynamic Behaviors
von: Nambiar, Sarang, et al.
Veröffentlicht: (2026)
von: Nambiar, Sarang, et al.
Veröffentlicht: (2026)
End-user Comprehension of Transfer Risks in Smart Contracts
von: Panicker, Yustynn, et al.
Veröffentlicht: (2024)
von: Panicker, Yustynn, et al.
Veröffentlicht: (2024)
Software Fairness: An Analysis and Survey
von: Soremekun, Ezekiel, et al.
Veröffentlicht: (2022)
von: Soremekun, Ezekiel, et al.
Veröffentlicht: (2022)
Distribution-aware Fairness Test Generation
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2023)
von: Rajan, Sai Sathiesh, et al.
Veröffentlicht: (2023)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
von: Khant, Kyi Shin, et al.
Veröffentlicht: (2025)
von: Khant, Kyi Shin, et al.
Veröffentlicht: (2025)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
von: Huang, Huihui, et al.
Veröffentlicht: (2026)
von: Huang, Huihui, et al.
Veröffentlicht: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
von: Li, Ziyu, et al.
Veröffentlicht: (2024)
von: Li, Ziyu, et al.
Veröffentlicht: (2024)
Write Your Own CodeChecker: An Automated Test-Driven Checker Development Approach with LLMs
von: Liu, Jun, et al.
Veröffentlicht: (2024)
von: Liu, Jun, et al.
Veröffentlicht: (2024)
Test Oracle Automation in the era of LLMs
von: Molina, Facundo, et al.
Veröffentlicht: (2024)
von: Molina, Facundo, et al.
Veröffentlicht: (2024)
Automated Harmfulness Testing for Code Large Language Models
von: Tan, Honghao, et al.
Veröffentlicht: (2025)
von: Tan, Honghao, et al.
Veröffentlicht: (2025)
Consistent or Sensitive? Automated Code Revision Tools Against Semantics-Preserving Perturbations
von: Pirouzkhah, Shirin, et al.
Veröffentlicht: (2026)
von: Pirouzkhah, Shirin, et al.
Veröffentlicht: (2026)
Unprecedented Code Change Automation: The Fusion of LLMs and Transformation by Example
von: Dilhara, Malinda, et al.
Veröffentlicht: (2024)
von: Dilhara, Malinda, et al.
Veröffentlicht: (2024)
Automated and Context-Aware Code Documentation Leveraging Advanced LLMs
von: Sarker, Swapnil Sharma, et al.
Veröffentlicht: (2025)
von: Sarker, Swapnil Sharma, et al.
Veröffentlicht: (2025)
SimCT: A Simple Consistency Test Protocol in LLMs Development Lifecycle
von: Zhao, Fufangchen, et al.
Veröffentlicht: (2024)
von: Zhao, Fufangchen, et al.
Veröffentlicht: (2024)
Extracting Formal Specifications from Documents Using LLMs for Automated Testing
von: Li, Hui, et al.
Veröffentlicht: (2025)
von: Li, Hui, et al.
Veröffentlicht: (2025)
Automated Refactoring of Non-Idiomatic Python Code: A Differentiated Replication with LLMs
von: Midolo, Alessandro, et al.
Veröffentlicht: (2025)
von: Midolo, Alessandro, et al.
Veröffentlicht: (2025)
Automating Code Generation for Semiconductor Equipment Control from Developer Utterances with LLMs
von: Kim, Youngkyoung, et al.
Veröffentlicht: (2025)
von: Kim, Youngkyoung, et al.
Veröffentlicht: (2025)
Figma2Code: Automating Multimodal Design to Code in the Wild
von: Gui, Yi, et al.
Veröffentlicht: (2026)
von: Gui, Yi, et al.
Veröffentlicht: (2026)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
von: Rao, Nikitha, et al.
Veröffentlicht: (2024)
von: Rao, Nikitha, et al.
Veröffentlicht: (2024)
Automated Soap Opera Testing Directed by LLMs and Scenario Knowledge: Feasibility, Challenges, and Road Ahead
von: Su, Yanqi, et al.
Veröffentlicht: (2024)
von: Su, Yanqi, et al.
Veröffentlicht: (2024)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
von: Cheng, Yiran, et al.
Veröffentlicht: (2025)
ConAIR:Consistency-Augmented Iterative Interaction Framework to Enhance the Reliability of Code Generation
von: Dong, Jinhao, et al.
Veröffentlicht: (2024)
von: Dong, Jinhao, et al.
Veröffentlicht: (2024)
Leveraging LLMs for Automated Translation of Legacy Code: A Case Study on PL/SQL to Java Transformation
von: Solovyeva, Lola, et al.
Veröffentlicht: (2025)
von: Solovyeva, Lola, et al.
Veröffentlicht: (2025)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
LLMs for Automated Unit Test Generation and Assessment in Java: The AgoneTest Framework
von: Lops, Andrea, et al.
Veröffentlicht: (2025)
von: Lops, Andrea, et al.
Veröffentlicht: (2025)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
von: Lwin, Naing Oo, et al.
Veröffentlicht: (2026)
von: Lwin, Naing Oo, et al.
Veröffentlicht: (2026)
REACCEPT: Automated Co-evolution of Production and Test Code Based on Dynamic Validation and Large Language Models
von: Chi, Jianlei, et al.
Veröffentlicht: (2024)
von: Chi, Jianlei, et al.
Veröffentlicht: (2024)
Automated Code Review In Practice
von: Cihan, Umut, et al.
Veröffentlicht: (2024)
von: Cihan, Umut, et al.
Veröffentlicht: (2024)
CA2: Code-Aware Agent for Automated Game Testing
von: Adaikkappan, Valliappan Chidambaram, et al.
Veröffentlicht: (2026)
von: Adaikkappan, Valliappan Chidambaram, et al.
Veröffentlicht: (2026)
Hybrid Privacy Policy-Code Consistency Check using Knowledge Graphs and LLMs
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
TAM-Eval: Evaluating LLMs for Automated Unit Test Maintenance
von: Bruches, Elena, et al.
Veröffentlicht: (2026)
von: Bruches, Elena, et al.
Veröffentlicht: (2026)
PrediQL: Automated Testing of GraphQL APIs with LLMs
von: Liu, Shaolun, et al.
Veröffentlicht: (2025)
von: Liu, Shaolun, et al.
Veröffentlicht: (2025)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
von: Sherifi, Betim, et al.
Veröffentlicht: (2024)
von: Sherifi, Betim, et al.
Veröffentlicht: (2024)
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
von: Jin, Shuo, et al.
Veröffentlicht: (2025)
von: Jin, Shuo, et al.
Veröffentlicht: (2025)
Invariant-Driven Automated Testing
von: Ribeiro, Ana Catarina
Veröffentlicht: (2026)
von: Ribeiro, Ana Catarina
Veröffentlicht: (2026)
Automated Unit Test Refactoring
von: Gao, Yi, et al.
Veröffentlicht: (2024)
von: Gao, Yi, et al.
Veröffentlicht: (2024)
Enabling Cost-Effective UI Automation Testing with Retrieval-Based LLMs: A Case Study in WeChat
von: Feng, Sidong, et al.
Veröffentlicht: (2024)
von: Feng, Sidong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Directed Grammar-Based Test Generation
von: Kirschner, Lukas, et al.
Veröffentlicht: (2025) -
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
von: Sohn, Jeongju, et al.
Veröffentlicht: (2025) -
Malicious ML Model Detection by Learning Dynamic Behaviors
von: Nambiar, Sarang, et al.
Veröffentlicht: (2026) -
End-user Comprehension of Transfer Risks in Smart Contracts
von: Panicker, Yustynn, et al.
Veröffentlicht: (2024) -
Software Fairness: An Analysis and Survey
von: Soremekun, Ezekiel, et al.
Veröffentlicht: (2022)