Articulate but Wrong: Self-Review Failures in LLM-Based Code Modernization
Fuente:
arXiv
Guardado en:
| Autores principales: | Reddy, Gokul Chandra Purnachandra, Lolla, Aditya, Sanku, Harsha |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
por: Rabbi, Fazle, et al.
Publicado: (2026)
por: Rabbi, Fazle, et al.
Publicado: (2026)
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation
por: Sarker, Laboni, et al.
Publicado: (2024)
por: Sarker, Laboni, et al.
Publicado: (2024)
Modern Code Reviews -- Survey of Literature and Practice
por: Badampudi, Deepika, et al.
Publicado: (2024)
por: Badampudi, Deepika, et al.
Publicado: (2024)
A Roadmap on Modern Code Review: Challenges and Opportunities
por: Yang, Zezhou, et al.
Publicado: (2024)
por: Yang, Zezhou, et al.
Publicado: (2024)
Review of Tools for Zero-Code LLM Based Application Development
por: Pattnayak, Priyaranjan, et al.
Publicado: (2025)
por: Pattnayak, Priyaranjan, et al.
Publicado: (2025)
CodeCoR: An LLM-Based Self-Reflective Multi-Agent Framework for Code Generation
por: Pan, Ruwei, et al.
Publicado: (2025)
por: Pan, Ruwei, et al.
Publicado: (2025)
Same Weights, Different Words: Measuring Inter-Provider Divergence and Temperature-Zero Non-Determinism in Open-Weight LLM Inference
por: Gokul Chandra, Purnachandra Reddy
Publicado: (2026)
por: Gokul Chandra, Purnachandra Reddy
Publicado: (2026)
Deciphering Refactoring Branch Dynamics in Modern Code Review: An Empirical Study on Qt
por: AlOmar, Eman Abdullah
Publicado: (2024)
por: AlOmar, Eman Abdullah
Publicado: (2024)
AI-Assisted Assessment of Coding Practices in Modern Code Review
por: Vijayvergiya, Manushree, et al.
Publicado: (2024)
por: Vijayvergiya, Manushree, et al.
Publicado: (2024)
I Can't Share Code, but I need Translation -- An Empirical Study on Code Translation through Federated LLM
por: Kumar, Jahnavi, et al.
Publicado: (2025)
por: Kumar, Jahnavi, et al.
Publicado: (2025)
Benchmarking and Studying the LLM-based Code Review
por: Zeng, Zhengran, et al.
Publicado: (2025)
por: Zeng, Zhengran, et al.
Publicado: (2025)
Studying and Understanding the Effectiveness and Failures of Conversational LLM-Based Repair
por: Chen, Aolin, et al.
Publicado: (2025)
por: Chen, Aolin, et al.
Publicado: (2025)
ECO: An LLM-Driven Efficient Code Optimizer for Warehouse Scale Computers
por: Lin, Hannah, et al.
Publicado: (2025)
por: Lin, Hannah, et al.
Publicado: (2025)
LLM-Based Multi-Agent Systems for Code Generation: A Multi-Vocal Literature Review
por: Rasheeda, Zeeshan, et al.
Publicado: (2026)
por: Rasheeda, Zeeshan, et al.
Publicado: (2026)
Improving LLM-Based Go Code Review through Issue-List Generation and Context Augmentation
por: Sun, Kexin, et al.
Publicado: (2026)
por: Sun, Kexin, et al.
Publicado: (2026)
Failure-Aware Enhancements for Large Language Model (LLM) Code Generation: An Empirical Study on Decision Framework
por: Shen, Jianru, et al.
Publicado: (2026)
por: Shen, Jianru, et al.
Publicado: (2026)
R2Code: A Self-Reflective LLM Framework for Requirements-to-Code Traceability
por: Wang, Yifei, et al.
Publicado: (2026)
por: Wang, Yifei, et al.
Publicado: (2026)
AILINKPREVIEWER: Enhancing Code Reviews with LLM-Powered Link Previews
por: Trakoolgerntong, Panya, et al.
Publicado: (2025)
por: Trakoolgerntong, Panya, et al.
Publicado: (2025)
Rethinking Code Review Workflows with LLM Assistance: An Empirical Study
por: Aðalsteinsson, Fannar Steinn, et al.
Publicado: (2025)
por: Aðalsteinsson, Fannar Steinn, et al.
Publicado: (2025)
A Survey of Code Review Benchmarks and Evaluation Practices in Pre-LLM and LLM Era
por: Khan, Taufiqul Islam, et al.
Publicado: (2026)
por: Khan, Taufiqul Islam, et al.
Publicado: (2026)
VAPU: System for Autonomous Legacy Code Modernization
por: Ala-Salmi, Valtteri, et al.
Publicado: (2025)
por: Ala-Salmi, Valtteri, et al.
Publicado: (2025)
Dissecting Bug Triggers and Failure Modes in Modern Agentic Frameworks: An Empirical Study
por: Zhang, Xiaowen, et al.
Publicado: (2026)
por: Zhang, Xiaowen, et al.
Publicado: (2026)
Improving Code Reviewer Recommendation: Accuracy, Latency, Workload, and Bystanders
por: Rigby, Peter C., et al.
Publicado: (2023)
por: Rigby, Peter C., et al.
Publicado: (2023)
SGCR: A Specification-Grounded Framework for Trustworthy LLM Code Review
por: Wang, Kai, et al.
Publicado: (2025)
por: Wang, Kai, et al.
Publicado: (2025)
BitsAI-CR: Automated Code Review via LLM in Practice
por: Sun, Tao, et al.
Publicado: (2025)
por: Sun, Tao, et al.
Publicado: (2025)
Copilot Arena: A Platform for Code LLM Evaluation in the Wild
por: Chi, Wayne, et al.
Publicado: (2025)
por: Chi, Wayne, et al.
Publicado: (2025)
What's Wrong with Your Code Generated by Large Language Models? An Extensive Study
por: Dou, Shihan, et al.
Publicado: (2024)
por: Dou, Shihan, et al.
Publicado: (2024)
An Empirical Study of LLM-Based Code Clone Detection
por: Zhu, Wenqing, et al.
Publicado: (2025)
por: Zhu, Wenqing, et al.
Publicado: (2025)
Code vs Serialized AST Inputs for LLM-Based Code Summarization: An Empirical Study
por: Dong, Shijia, et al.
Publicado: (2026)
por: Dong, Shijia, et al.
Publicado: (2026)
Leveraging LLMs for Legacy Code Modernization: Challenges and Opportunities for LLM-Generated Documentation
por: Diggs, Colin, et al.
Publicado: (2024)
por: Diggs, Colin, et al.
Publicado: (2024)
Code Review Automation Via Multi-task Federated LLM -- An Empirical Study
por: Kumar, Jahnavi, et al.
Publicado: (2024)
por: Kumar, Jahnavi, et al.
Publicado: (2024)
LAURA: Enhancing Code Review Generation with Context-Enriched Retrieval-Augmented LLM
por: Zhang, Yuxin, et al.
Publicado: (2025)
por: Zhang, Yuxin, et al.
Publicado: (2025)
CRQBench: A Benchmark of Code Reasoning Questions
por: Dinella, Elizabeth, et al.
Publicado: (2024)
por: Dinella, Elizabeth, et al.
Publicado: (2024)
What Breaks When LLMs Code? Characterizing Operational Safety Failures of Agentic Code Assistants
por: Hasan, Alif Al, et al.
Publicado: (2026)
por: Hasan, Alif Al, et al.
Publicado: (2026)
LogSage: An LLM-Based Framework for CI/CD Failure Detection and Remediation with Industrial Validation
por: Xu, Weiyuan, et al.
Publicado: (2025)
por: Xu, Weiyuan, et al.
Publicado: (2025)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
por: Gong, Zhihao, et al.
Publicado: (2026)
por: Gong, Zhihao, et al.
Publicado: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
por: Gong, Zhihao, et al.
Publicado: (2025)
por: Gong, Zhihao, et al.
Publicado: (2025)
An Empirical Study of Bugs in Modern LLM Agent Frameworks
por: Zhu, Xinxue, et al.
Publicado: (2026)
por: Zhu, Xinxue, et al.
Publicado: (2026)
Quo Vadis, Code Review? Exploring the Future of Code Review
por: Dorner, Michael, et al.
Publicado: (2025)
por: Dorner, Michael, et al.
Publicado: (2025)
Engagement in Code Review: Emotional, Behavioral, and Cognitive Dimensions in Peer vs. LLM Interactions
por: Alami, Adam, et al.
Publicado: (2025)
por: Alami, Adam, et al.
Publicado: (2025)
Ejemplares similares
-
Beyond Translation Accuracy: Addressing False Failures in LLM-Based Code Translation
por: Rabbi, Fazle, et al.
Publicado: (2026) -
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation
por: Sarker, Laboni, et al.
Publicado: (2024) -
Modern Code Reviews -- Survey of Literature and Practice
por: Badampudi, Deepika, et al.
Publicado: (2024) -
A Roadmap on Modern Code Review: Challenges and Opportunities
por: Yang, Zezhou, et al.
Publicado: (2024) -
Review of Tools for Zero-Code LLM Based Application Development
por: Pattnayak, Priyaranjan, et al.
Publicado: (2025)