MUCOCO: Automated Consistency Testing of Code LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Chou, Chua Jin, Lwin, Khant That, Soremekun, Ezekiel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Directed Grammar-Based Test Generation
por: Kirschner, Lukas, et al.
Publicado: (2025)
por: Kirschner, Lukas, et al.
Publicado: (2025)
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
por: Sohn, Jeongju, et al.
Publicado: (2025)
por: Sohn, Jeongju, et al.
Publicado: (2025)
Malicious ML Model Detection by Learning Dynamic Behaviors
por: Nambiar, Sarang, et al.
Publicado: (2026)
por: Nambiar, Sarang, et al.
Publicado: (2026)
End-user Comprehension of Transfer Risks in Smart Contracts
por: Panicker, Yustynn, et al.
Publicado: (2024)
por: Panicker, Yustynn, et al.
Publicado: (2024)
Software Fairness: An Analysis and Survey
por: Soremekun, Ezekiel, et al.
Publicado: (2022)
por: Soremekun, Ezekiel, et al.
Publicado: (2022)
Distribution-aware Fairness Test Generation
por: Rajan, Sai Sathiesh, et al.
Publicado: (2023)
por: Rajan, Sai Sathiesh, et al.
Publicado: (2023)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
por: Khant, Kyi Shin, et al.
Publicado: (2025)
por: Khant, Kyi Shin, et al.
Publicado: (2025)
PenForge: On-the-Fly Expert Agent Construction for Automated Penetration Testing
por: Huang, Huihui, et al.
Publicado: (2026)
por: Huang, Huihui, et al.
Publicado: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
por: Li, Ziyu, et al.
Publicado: (2024)
por: Li, Ziyu, et al.
Publicado: (2024)
Write Your Own CodeChecker: An Automated Test-Driven Checker Development Approach with LLMs
por: Liu, Jun, et al.
Publicado: (2024)
por: Liu, Jun, et al.
Publicado: (2024)
Test Oracle Automation in the era of LLMs
por: Molina, Facundo, et al.
Publicado: (2024)
por: Molina, Facundo, et al.
Publicado: (2024)
Automated Harmfulness Testing for Code Large Language Models
por: Tan, Honghao, et al.
Publicado: (2025)
por: Tan, Honghao, et al.
Publicado: (2025)
Consistent or Sensitive? Automated Code Revision Tools Against Semantics-Preserving Perturbations
por: Pirouzkhah, Shirin, et al.
Publicado: (2026)
por: Pirouzkhah, Shirin, et al.
Publicado: (2026)
Unprecedented Code Change Automation: The Fusion of LLMs and Transformation by Example
por: Dilhara, Malinda, et al.
Publicado: (2024)
por: Dilhara, Malinda, et al.
Publicado: (2024)
Automated and Context-Aware Code Documentation Leveraging Advanced LLMs
por: Sarker, Swapnil Sharma, et al.
Publicado: (2025)
por: Sarker, Swapnil Sharma, et al.
Publicado: (2025)
SimCT: A Simple Consistency Test Protocol in LLMs Development Lifecycle
por: Zhao, Fufangchen, et al.
Publicado: (2024)
por: Zhao, Fufangchen, et al.
Publicado: (2024)
Extracting Formal Specifications from Documents Using LLMs for Automated Testing
por: Li, Hui, et al.
Publicado: (2025)
por: Li, Hui, et al.
Publicado: (2025)
Automated Refactoring of Non-Idiomatic Python Code: A Differentiated Replication with LLMs
por: Midolo, Alessandro, et al.
Publicado: (2025)
por: Midolo, Alessandro, et al.
Publicado: (2025)
Automating Code Generation for Semiconductor Equipment Control from Developer Utterances with LLMs
por: Kim, Youngkyoung, et al.
Publicado: (2025)
por: Kim, Youngkyoung, et al.
Publicado: (2025)
Figma2Code: Automating Multimodal Design to Code in the Wild
por: Gui, Yi, et al.
Publicado: (2026)
por: Gui, Yi, et al.
Publicado: (2026)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
por: Rao, Nikitha, et al.
Publicado: (2024)
por: Rao, Nikitha, et al.
Publicado: (2024)
Automated Soap Opera Testing Directed by LLMs and Scenario Knowledge: Feasibility, Challenges, and Road Ahead
por: Su, Yanqi, et al.
Publicado: (2024)
por: Su, Yanqi, et al.
Publicado: (2024)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
por: Cheng, Yiran, et al.
Publicado: (2025)
por: Cheng, Yiran, et al.
Publicado: (2025)
ConAIR:Consistency-Augmented Iterative Interaction Framework to Enhance the Reliability of Code Generation
por: Dong, Jinhao, et al.
Publicado: (2024)
por: Dong, Jinhao, et al.
Publicado: (2024)
Leveraging LLMs for Automated Translation of Legacy Code: A Case Study on PL/SQL to Java Transformation
por: Solovyeva, Lola, et al.
Publicado: (2025)
por: Solovyeva, Lola, et al.
Publicado: (2025)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
por: Li, Junjie, et al.
Publicado: (2025)
por: Li, Junjie, et al.
Publicado: (2025)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
por: Zhu, Yuqi, et al.
Publicado: (2025)
por: Zhu, Yuqi, et al.
Publicado: (2025)
LLMs for Automated Unit Test Generation and Assessment in Java: The AgoneTest Framework
por: Lops, Andrea, et al.
Publicado: (2025)
por: Lops, Andrea, et al.
Publicado: (2025)
Deterministic vs. LLM-Controlled Orchestration for COBOL-to-Python Modernization
por: Lwin, Naing Oo, et al.
Publicado: (2026)
por: Lwin, Naing Oo, et al.
Publicado: (2026)
REACCEPT: Automated Co-evolution of Production and Test Code Based on Dynamic Validation and Large Language Models
por: Chi, Jianlei, et al.
Publicado: (2024)
por: Chi, Jianlei, et al.
Publicado: (2024)
Automated Code Review In Practice
por: Cihan, Umut, et al.
Publicado: (2024)
por: Cihan, Umut, et al.
Publicado: (2024)
CA2: Code-Aware Agent for Automated Game Testing
por: Adaikkappan, Valliappan Chidambaram, et al.
Publicado: (2026)
por: Adaikkappan, Valliappan Chidambaram, et al.
Publicado: (2026)
Hybrid Privacy Policy-Code Consistency Check using Knowledge Graphs and LLMs
por: Mao, Zhenyu, et al.
Publicado: (2025)
por: Mao, Zhenyu, et al.
Publicado: (2025)
TAM-Eval: Evaluating LLMs for Automated Unit Test Maintenance
por: Bruches, Elena, et al.
Publicado: (2026)
por: Bruches, Elena, et al.
Publicado: (2026)
PrediQL: Automated Testing of GraphQL APIs with LLMs
por: Liu, Shaolun, et al.
Publicado: (2025)
por: Liu, Shaolun, et al.
Publicado: (2025)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
por: Sherifi, Betim, et al.
Publicado: (2024)
por: Sherifi, Betim, et al.
Publicado: (2024)
RulER: Automated Rule-Based Semantic Error Localization and Repair for Code Translation
por: Jin, Shuo, et al.
Publicado: (2025)
por: Jin, Shuo, et al.
Publicado: (2025)
Invariant-Driven Automated Testing
por: Ribeiro, Ana Catarina
Publicado: (2026)
por: Ribeiro, Ana Catarina
Publicado: (2026)
Automated Unit Test Refactoring
por: Gao, Yi, et al.
Publicado: (2024)
por: Gao, Yi, et al.
Publicado: (2024)
Enabling Cost-Effective UI Automation Testing with Retrieval-Based LLMs: A Case Study in WeChat
por: Feng, Sidong, et al.
Publicado: (2024)
por: Feng, Sidong, et al.
Publicado: (2024)
Ejemplares similares
-
Directed Grammar-Based Test Generation
por: Kirschner, Lukas, et al.
Publicado: (2025) -
Latent Mutants: A large-scale study on the Interplay between mutation testing and software evolution
por: Sohn, Jeongju, et al.
Publicado: (2025) -
Malicious ML Model Detection by Learning Dynamic Behaviors
por: Nambiar, Sarang, et al.
Publicado: (2026) -
End-user Comprehension of Transfer Risks in Smart Contracts
por: Panicker, Yustynn, et al.
Publicado: (2024) -
Software Fairness: An Analysis and Survey
por: Soremekun, Ezekiel, et al.
Publicado: (2022)