LLM Assisted Coding with Metamorphic Specification Mutation Agent
Fuente:
arXiv
Saved in:
| Main Authors: | Akhond, Mostafijur Rahman, Uddin, Gias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM For Loop Invariant Generation and Fixing: How Far Are We?
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
Applications and Challenges of Fairness APIs in Machine Learning Software
by: Das, Ajoy, et al.
Published: (2025)
by: Das, Ajoy, et al.
Published: (2025)
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
by: Salimian, Sina, et al.
Published: (2025)
by: Salimian, Sina, et al.
Published: (2025)
PAGENT: Learning to Patch Software Engineering Agents
by: Xue, Haoran, et al.
Published: (2025)
by: Xue, Haoran, et al.
Published: (2025)
Optimized Log Parsing with Syntactic Modifications
by: Enan, Nafid, et al.
Published: (2025)
by: Enan, Nafid, et al.
Published: (2025)
ABTest: Behavior-Driven Testing for AI Coding Agents
by: Dai, Wuyang, et al.
Published: (2026)
by: Dai, Wuyang, et al.
Published: (2026)
Assessing the Influence of Toxic and Gender Discriminatory Communication on Perceptible Diversity in OSS Projects
by: Sultana, Sayma, et al.
Published: (2024)
by: Sultana, Sayma, et al.
Published: (2024)
Evaluating the Environmental Impact of using SLMs and Prompt Engineering for Code Generation
by: Mamun, Md Afif Al, et al.
Published: (2026)
by: Mamun, Md Afif Al, et al.
Published: (2026)
An Empirical Study on Bug Severity Estimation using Source Code Metrics and Static Analysis
by: Mashhadi, Ehsan, et al.
Published: (2022)
by: Mashhadi, Ehsan, et al.
Published: (2022)
Secret Leak Detection in Software Issue Reports using LLMs: A Comprehensive Evaluation
by: Ahmed, Sadif, et al.
Published: (2024)
by: Ahmed, Sadif, et al.
Published: (2024)
Engineering Pitfalls in AI Coding Tools: An Empirical Study of Bugs in Claude Code, Codex, and Gemini CLI
by: Zhang, Ruixin, et al.
Published: (2026)
by: Zhang, Ruixin, et al.
Published: (2026)
SWE-Bench+: Enhanced Coding Benchmark for LLMs
by: Aleithan, Reem, et al.
Published: (2024)
by: Aleithan, Reem, et al.
Published: (2024)
Reputation Gaming in Stack Overflow
by: Mazloomzadeh, Iren, et al.
Published: (2021)
by: Mazloomzadeh, Iren, et al.
Published: (2021)
BLAgent: Agentic RAG for File-Level Bug Localization
by: Mamun, Md Afif Al, et al.
Published: (2026)
by: Mamun, Md Afif Al, et al.
Published: (2026)
IssueGuard: Real-Time Secret Leak Prevention Tool for GitHub Issue Reports
by: Rahman, Md Nafiu, et al.
Published: (2026)
by: Rahman, Md Nafiu, et al.
Published: (2026)
A Systematic Mapping Study of Crowd Knowledge Enhanced Software Engineering Research Using Stack Overflow
by: Tanzil, Minaoar, et al.
Published: (2024)
by: Tanzil, Minaoar, et al.
Published: (2024)
"How do people decide?": A Model for Software Library Selection
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
A Large-Scale Empirical Study of COVID-19 Contact Tracing Mobile App Reviews
by: Parisa, Sifat Ishmam, et al.
Published: (2024)
by: Parisa, Sifat Ishmam, et al.
Published: (2024)
ChatGPT Inaccuracy Mitigation during Technical Report Understanding: Are We There Yet?
by: Tamanna, Salma Begum, et al.
Published: (2024)
by: Tamanna, Salma Begum, et al.
Published: (2024)
AutoMT: A Multi-Agent LLM Framework for Automated Metamorphic Testing of Autonomous Driving Systems
by: Liang, Linfeng, et al.
Published: (2025)
by: Liang, Linfeng, et al.
Published: (2025)
Stack Trace-Based Crash Deduplication with Transformer Adaptation
by: Mamun, Md Afif Al, et al.
Published: (2025)
by: Mamun, Md Afif Al, et al.
Published: (2025)
A Mixed Method Study of DevOps Challenges
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
ChatGPT Incorrectness Detection in Software Reviews
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
Metamorphic Coverage
by: Ba, Jinsheng, et al.
Published: (2025)
by: Ba, Jinsheng, et al.
Published: (2025)
Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations
by: Kulshreshtha, Ashir, et al.
Published: (2026)
by: Kulshreshtha, Ashir, et al.
Published: (2026)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
by: Xue, Pengyu, et al.
Published: (2024)
by: Xue, Pengyu, et al.
Published: (2024)
SWE-Shepherd: Advancing PRMs for Reinforcing Code Agents
by: Dihan, Mahir Labib, et al.
Published: (2026)
by: Dihan, Mahir Labib, et al.
Published: (2026)
METAMON: Finding Inconsistencies between Program Documentation and Behavior using Metamorphic LLM Queries
by: Lee, Hyeonseok, et al.
Published: (2025)
by: Lee, Hyeonseok, et al.
Published: (2025)
TriagerX: Dual Transformers for Bug Triaging Tasks with Content and Interaction Based Rankings
by: Mamun, Md Afif Al, et al.
Published: (2025)
by: Mamun, Md Afif Al, et al.
Published: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
by: Wang, Xiaoyin, et al.
Published: (2024)
by: Wang, Xiaoyin, et al.
Published: (2024)
Re-Evaluating Code LLM Benchmarks Under Semantic Mutation
by: Pan, Zhiyuan, et al.
Published: (2025)
by: Pan, Zhiyuan, et al.
Published: (2025)
You Can REST Now: Automated REST API Documentation and Testing via LLM-Assisted Request Mutations
by: Decrop, Alix, et al.
Published: (2024)
by: Decrop, Alix, et al.
Published: (2024)
Simulink Mutation Testing using CodeBERT
by: Zhang, Jingfan, et al.
Published: (2025)
by: Zhang, Jingfan, et al.
Published: (2025)
SGCR: A Specification-Grounded Framework for Trustworthy LLM Code Review
by: Wang, Kai, et al.
Published: (2025)
by: Wang, Kai, et al.
Published: (2025)
Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety
by: Uddin, S M Jamil
Published: (2026)
by: Uddin, S M Jamil
Published: (2026)
Hallucination Detection for LLM-based Text-to-SQL Generation via Two-Stage Metamorphic Testing
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
Specification and Detection of LLM Code Smells
by: Mahmoudi, Brahim, et al.
Published: (2025)
by: Mahmoudi, Brahim, et al.
Published: (2025)
Bootstrapping Coding Agents: The Specification Is the Program
by: Monperrus, Martin
Published: (2026)
by: Monperrus, Martin
Published: (2026)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
by: Asgari, Ali, et al.
Published: (2025)
by: Asgari, Ali, et al.
Published: (2025)
MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
by: Wang, Guanyu, et al.
Published: (2024)
by: Wang, Guanyu, et al.
Published: (2024)
Similar Items
-
LLM For Loop Invariant Generation and Fixing: How Far Are We?
by: Akhond, Mostafijur Rahman, et al.
Published: (2025) -
Applications and Challenges of Fairness APIs in Machine Learning Software
by: Das, Ajoy, et al.
Published: (2025) -
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
by: Salimian, Sina, et al.
Published: (2025) -
PAGENT: Learning to Patch Software Engineering Agents
by: Xue, Haoran, et al.
Published: (2025) -
Optimized Log Parsing with Syntactic Modifications
by: Enan, Nafid, et al.
Published: (2025)