LLMs in Web Development: Evaluating LLM-Generated PHP Code Unveiling Vulnerabilities and Limitations
Fuente:
arXiv
Guardado en:
| Autores principales: | Tóth, Rebeka, Bisztray, Tamas, Erdodi, László |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
por: Tihanyi, Norbert, et al.
Publicado: (2025)
por: Tihanyi, Norbert, et al.
Publicado: (2025)
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
por: Toth, Rebeka, et al.
Publicado: (2025)
por: Toth, Rebeka, et al.
Publicado: (2025)
CASTLE: Benchmarking Dataset for Static Code Analyzers and LLMs towards CWE Detection
por: Dubniczky, Richard A., et al.
Publicado: (2025)
por: Dubniczky, Richard A., et al.
Publicado: (2025)
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
por: Bisztray, Tamas, et al.
Publicado: (2025)
por: Bisztray, Tamas, et al.
Publicado: (2025)
Reverse Browser: Vector-Image-to-Code Generator
por: Toth-Czifra, Zoltan
Publicado: (2025)
por: Toth-Czifra, Zoltan
Publicado: (2025)
WebDevJudge: Evaluating (M)LLMs as Critiques for Web Development Quality
por: Li, Chunyang, et al.
Publicado: (2025)
por: Li, Chunyang, et al.
Publicado: (2025)
WebApp1K: A Practical Code-Generation Benchmark for Web App Development
por: Cui, Yi
Publicado: (2024)
por: Cui, Yi
Publicado: (2024)
Evaluating the Energy-Efficiency of the Code Generated by LLMs
por: Islam, Md Arman, et al.
Publicado: (2025)
por: Islam, Md Arman, et al.
Publicado: (2025)
Unseen Horizons: Unveiling the Real Capability of LLM Code Generation Beyond the Familiar
por: Zhang, Yuanliang, et al.
Publicado: (2024)
por: Zhang, Yuanliang, et al.
Publicado: (2024)
WebCompass: Towards Multimodal Web Coding Evaluation for Code Language Models
por: Lei, Xinping, et al.
Publicado: (2026)
por: Lei, Xinping, et al.
Publicado: (2026)
Holistic Evaluation of State-of-the-Art LLMs for Code Generation
por: Zhang, Le, et al.
Publicado: (2025)
por: Zhang, Le, et al.
Publicado: (2025)
An Empirical Evaluation of LLM-Based Approaches for Code Vulnerability Detection: RAG, SFT, and Dual-Agent Systems
por: Saju, Md Hasan, et al.
Publicado: (2026)
por: Saju, Md Hasan, et al.
Publicado: (2026)
LLM-Powered Code Vulnerability Repair with Reinforcement Learning and Semantic Reward
por: Islam, Nafis Tanveer, et al.
Publicado: (2024)
por: Islam, Nafis Tanveer, et al.
Publicado: (2024)
The Code Whisperer: LLM and Graph-Based AI for Smell and Vulnerability Resolution
por: Baqar, Mohammad, et al.
Publicado: (2026)
por: Baqar, Mohammad, et al.
Publicado: (2026)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
por: Cui, Yi
Publicado: (2025)
por: Cui, Yi
Publicado: (2025)
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
por: Zhu, Yuecai, et al.
Publicado: (2026)
por: Zhu, Yuecai, et al.
Publicado: (2026)
Bridging LLM-Generated Code and Requirements: Reverse Generation technique and SBC Metric for Developer Insights
por: Ponnusamy, Ahilan Ayyachamy Nadar
Publicado: (2025)
por: Ponnusamy, Ahilan Ayyachamy Nadar
Publicado: (2025)
LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation
por: Xu, Dong, et al.
Publicado: (2026)
por: Xu, Dong, et al.
Publicado: (2026)
AccessGuru: Leveraging LLMs to Detect and Correct Web Accessibility Violations in HTML Code
por: Fathallah, Nadeen, et al.
Publicado: (2025)
por: Fathallah, Nadeen, et al.
Publicado: (2025)
Harnessing the Power of LLMs in Source Code Vulnerability Detection
por: Mahyari, Andrew A
Publicado: (2024)
por: Mahyari, Andrew A
Publicado: (2024)
Studying Vulnerable Code Entities in R
por: Zhao, Zixiao, et al.
Publicado: (2024)
por: Zhao, Zixiao, et al.
Publicado: (2024)
Understanding the Limits of Automated Evaluation for Code Review Bots in Practice
por: Karakaya, Veli, et al.
Publicado: (2026)
por: Karakaya, Veli, et al.
Publicado: (2026)
Correct Code, Vulnerable Dependencies: A Large Scale Measurement Study of LLM-Specified Library Versions
por: Wang, Chengjie, et al.
Publicado: (2026)
por: Wang, Chengjie, et al.
Publicado: (2026)
A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics
por: Katzy, Jonathan, et al.
Publicado: (2025)
por: Katzy, Jonathan, et al.
Publicado: (2025)
Investigating The Smells of LLM Generated Code
por: Paul, Debalina Ghosh, et al.
Publicado: (2025)
por: Paul, Debalina Ghosh, et al.
Publicado: (2025)
Evaluating the Use of LLMs for Automated DOM-Level Resolution of Web Performance Issues
por: Peters, Gideon, et al.
Publicado: (2026)
por: Peters, Gideon, et al.
Publicado: (2026)
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs' Vulnerability Reasoning
por: Sun, Yuqiang, et al.
Publicado: (2024)
por: Sun, Yuqiang, et al.
Publicado: (2024)
Detecting Data Poisoning in Code Generation LLMs via Black-Box, Vulnerability-Oriented Scanning
por: Yan, Shenao, et al.
Publicado: (2026)
por: Yan, Shenao, et al.
Publicado: (2026)
From PowerPoint UI Sketches to Web-Based Applications: Pattern-Driven Code Generation for GIS Dashboard Development Using Knowledge-Augmented LLMs, Context-Aware Visual Prompting, and the React Framework
por: Xu, Haowen, et al.
Publicado: (2025)
por: Xu, Haowen, et al.
Publicado: (2025)
WebVIA: A Web-based Vision-Language Agentic Framework for Interactive and Verifiable UI-to-Code Generation
por: Xu, Mingde, et al.
Publicado: (2025)
por: Xu, Mingde, et al.
Publicado: (2025)
Software Vulnerability and Functionality Assessment using LLMs
por: Jensen, Rasmus Ingemann Tuffveson, et al.
Publicado: (2024)
por: Jensen, Rasmus Ingemann Tuffveson, et al.
Publicado: (2024)
Reasoning with LLMs for Zero-Shot Vulnerability Detection
por: Zibaeirad, Arastoo, et al.
Publicado: (2025)
por: Zibaeirad, Arastoo, et al.
Publicado: (2025)
Test-Driven Development for Code Generation
por: Mathews, Noble Saji, et al.
Publicado: (2024)
por: Mathews, Noble Saji, et al.
Publicado: (2024)
WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics
por: Liu, Chenxu, et al.
Publicado: (2026)
por: Liu, Chenxu, et al.
Publicado: (2026)
Code Copycat Conundrum: Demystifying Repetition in LLM-based Code Generation
por: Liu, Mingwei, et al.
Publicado: (2025)
por: Liu, Mingwei, et al.
Publicado: (2025)
Coding in a Bubble? Evaluating LLMs in Resolving Context Adaptation Bugs During Code Adaptation
por: Zhang, Tanghaoran, et al.
Publicado: (2026)
por: Zhang, Tanghaoran, et al.
Publicado: (2026)
Verification Limits Code LLM Training
por: Gureja, Srishti, et al.
Publicado: (2025)
por: Gureja, Srishti, et al.
Publicado: (2025)
Uncertainty Quantification for LLM-based Code Generation
por: Xu, Senrong, et al.
Publicado: (2026)
por: Xu, Senrong, et al.
Publicado: (2026)
Insights from Benchmarking Frontier Language Models on Web App Code Generation
por: Cui, Yi
Publicado: (2024)
por: Cui, Yi
Publicado: (2024)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
por: Li, Ziyu, et al.
Publicado: (2024)
por: Li, Ziyu, et al.
Publicado: (2024)
Ejemplares similares
-
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
por: Tihanyi, Norbert, et al.
Publicado: (2025) -
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
por: Toth, Rebeka, et al.
Publicado: (2025) -
CASTLE: Benchmarking Dataset for Static Code Analyzers and LLMs towards CWE Detection
por: Dubniczky, Richard A., et al.
Publicado: (2025) -
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
por: Bisztray, Tamas, et al.
Publicado: (2025) -
Reverse Browser: Vector-Image-to-Code Generator
por: Toth-Czifra, Zoltan
Publicado: (2025)