Verbatim Data Transcription Failures in LLM Code Generation: A State-Tracking Stress Test
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Haque, Mohd Ariful, Gupta, Kishor Datta, Rahman, Mohammad Ashiqur, George, Roy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SOK: Exploring Hallucinations and Security Risks in AI-Assisted Software Development with Insights for LLM Deployment
von: Haque, Ariful, et al.
Veröffentlicht: (2025)
von: Haque, Ariful, et al.
Veröffentlicht: (2025)
LLM Security Guard for Code
von: Kavian, Arya, et al.
Veröffentlicht: (2024)
von: Kavian, Arya, et al.
Veröffentlicht: (2024)
QASecClaw: A Multi-Agent LLM Approach for False Positive Reduction in Static Application Security Testing
von: Ameen, Mohd Ruhul, et al.
Veröffentlicht: (2026)
von: Ameen, Mohd Ruhul, et al.
Veröffentlicht: (2026)
SAFuzz: Semantic-Guided Adaptive Fuzzing for LLM-Generated Code
von: Yang, Ziyi, et al.
Veröffentlicht: (2026)
von: Yang, Ziyi, et al.
Veröffentlicht: (2026)
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
von: Ge, Yifei, et al.
Veröffentlicht: (2026)
von: Ge, Yifei, et al.
Veröffentlicht: (2026)
How Secure is Secure Code Generation? Adversarial Prompts Put LLM Defenses to the Test
von: Tessa, Melissa, et al.
Veröffentlicht: (2026)
von: Tessa, Melissa, et al.
Veröffentlicht: (2026)
Teaching an Old LLM Secure Coding: Localized Preference Optimization on Distilled Preferences
von: Hasan, Mohammad Saqib, et al.
Veröffentlicht: (2025)
von: Hasan, Mohammad Saqib, et al.
Veröffentlicht: (2025)
SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
von: Hu, Qi, et al.
Veröffentlicht: (2026)
von: Hu, Qi, et al.
Veröffentlicht: (2026)
HYDRA: A Hybrid Heuristic-Guided Deep Representation Architecture for Predicting Latent Zero-Day Vulnerabilities in Patched Functions
von: Farhad, Mohammad, et al.
Veröffentlicht: (2025)
von: Farhad, Mohammad, et al.
Veröffentlicht: (2025)
How to Compare the Security of Code Written by Humans to LLM-generated Code
von: Balebako, Rebecca, et al.
Veröffentlicht: (2026)
von: Balebako, Rebecca, et al.
Veröffentlicht: (2026)
WildCode: An Empirical Analysis of Code Generated by ChatGPT
von: Khanmohammadi, Kobra, et al.
Veröffentlicht: (2025)
von: Khanmohammadi, Kobra, et al.
Veröffentlicht: (2025)
CodableLLM: Automating Decompiled and Source Code Mapping for LLM Dataset Generation
von: Manuel, Dylan, et al.
Veröffentlicht: (2025)
von: Manuel, Dylan, et al.
Veröffentlicht: (2025)
Fine-Tuning LLMs for Code Mutation: A New Era of Cyber Threats
von: Setak, Mohammad, et al.
Veröffentlicht: (2024)
von: Setak, Mohammad, et al.
Veröffentlicht: (2024)
An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2026)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2026)
Measuring the Permission Gate: A Stress-Test Evaluation of Claude Code's Auto Mode
von: Ji, Zimo, et al.
Veröffentlicht: (2026)
von: Ji, Zimo, et al.
Veröffentlicht: (2026)
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
von: Deng, Gelei, et al.
Veröffentlicht: (2023)
von: Deng, Gelei, et al.
Veröffentlicht: (2023)
Exploring Security Practices in Infrastructure as Code: An Empirical Study
von: Verdet, Alexandre, et al.
Veröffentlicht: (2023)
von: Verdet, Alexandre, et al.
Veröffentlicht: (2023)
Automated Testing of Broken Authentication Vulnerabilities in Web APIs with AuthREST
von: Corradini, Davide, et al.
Veröffentlicht: (2025)
von: Corradini, Davide, et al.
Veröffentlicht: (2025)
Synergistic Directed Execution and LLM-Driven Analysis for Zero-Day AI-Generated Malware Detection
von: Edwards, George, et al.
Veröffentlicht: (2026)
von: Edwards, George, et al.
Veröffentlicht: (2026)
Residual Risk Analysis in Benign Code: How Far Are We? A Multi-Model Semantic and Structural Similarity Approach
von: Farhad, Mohammad, et al.
Veröffentlicht: (2026)
von: Farhad, Mohammad, et al.
Veröffentlicht: (2026)
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
von: Ning, Kaiwen, et al.
Veröffentlicht: (2024)
von: Ning, Kaiwen, et al.
Veröffentlicht: (2024)
CKGFuzzer: LLM-Based Fuzz Driver Generation Enhanced By Code Knowledge Graph
von: Xu, Hanxiang, et al.
Veröffentlicht: (2024)
von: Xu, Hanxiang, et al.
Veröffentlicht: (2024)
What Makes a Good LLM Agent for Real-world Penetration Testing?
von: Deng, Gelei, et al.
Veröffentlicht: (2026)
von: Deng, Gelei, et al.
Veröffentlicht: (2026)
SCAFFOLD-CEGIS: Preventing Latent Security Degradation in LLM-Driven Iterative Code Refinement
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
Concerned with Data Contamination? Assessing Countermeasures in Code Language Model
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
Detecting Stealthy Data Poisoning Attacks in AI Code Generators
von: Improta, Cristina
Veröffentlicht: (2025)
von: Improta, Cristina
Veröffentlicht: (2025)
Integrating APK Image and Text Data for Enhanced Threat Detection: A Multimodal Deep Learning Approach to Android Malware
von: Arifin, Md Mashrur, et al.
Veröffentlicht: (2026)
von: Arifin, Md Mashrur, et al.
Veröffentlicht: (2026)
OpDiffer: LLM-Assisted Opcode-Level Differential Testing of Ethereum Virtual Machine
von: Ma, Jie, et al.
Veröffentlicht: (2025)
von: Ma, Jie, et al.
Veröffentlicht: (2025)
CleanVul: Automatic Function-Level Vulnerability Detection in Code Commits Using LLM Heuristics
von: Li, Yikun, et al.
Veröffentlicht: (2024)
von: Li, Yikun, et al.
Veröffentlicht: (2024)
Trim My View: An LLM-Based Code Query System for Module Retrieval in Robotic Firmware
von: Arasteh, Sima, et al.
Veröffentlicht: (2025)
von: Arasteh, Sima, et al.
Veröffentlicht: (2025)
Usability as a Weapon: Attacking the Safety of LLM-Based Code Generation via Usability Requirements
von: Li, Yue, et al.
Veröffentlicht: (2026)
von: Li, Yue, et al.
Veröffentlicht: (2026)
How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection
von: Chen, Maofei, et al.
Veröffentlicht: (2026)
von: Chen, Maofei, et al.
Veröffentlicht: (2026)
AutoDFBench 1.0: A Benchmarking Framework for Digital Forensic Tool Testing and Generated Code Evaluation
von: Wickramasekara, Akila, et al.
Veröffentlicht: (2025)
von: Wickramasekara, Akila, et al.
Veröffentlicht: (2025)
Execution-State-Aware LLM Reasoning for Automated Proof-of-Vulnerability Generation
von: Li, Haoyu, et al.
Veröffentlicht: (2026)
von: Li, Haoyu, et al.
Veröffentlicht: (2026)
Who Tests the Testers? Systematic Enumeration and Coverage Audit of LLM Agent Tool Call Safety
von: Chen, Xuan, et al.
Veröffentlicht: (2026)
von: Chen, Xuan, et al.
Veröffentlicht: (2026)
PrediQL: Automated Testing of GraphQL APIs with LLMs
von: Liu, Shaolun, et al.
Veröffentlicht: (2025)
von: Liu, Shaolun, et al.
Veröffentlicht: (2025)
D-LiFT: Improving LLM-based Decompiler Backend via Code Quality-driven Fine-tuning
von: Zou, Muqi, et al.
Veröffentlicht: (2025)
von: Zou, Muqi, et al.
Veröffentlicht: (2025)
Tracking Down Software Cluster Bombs: A Current State Analysis of the Free/Libre and Open Source Software (FLOSS) Ecosystem
von: Tatschner, Stefan, et al.
Veröffentlicht: (2025)
von: Tatschner, Stefan, et al.
Veröffentlicht: (2025)
RepoMark: A Data-Usage Auditing Framework for Code Large Language Models
von: Qu, Wenjie, et al.
Veröffentlicht: (2025)
von: Qu, Wenjie, et al.
Veröffentlicht: (2025)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
von: Yang, Hua, et al.
Veröffentlicht: (2025)
von: Yang, Hua, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SOK: Exploring Hallucinations and Security Risks in AI-Assisted Software Development with Insights for LLM Deployment
von: Haque, Ariful, et al.
Veröffentlicht: (2025) -
LLM Security Guard for Code
von: Kavian, Arya, et al.
Veröffentlicht: (2024) -
QASecClaw: A Multi-Agent LLM Approach for False Positive Reduction in Static Application Security Testing
von: Ameen, Mohd Ruhul, et al.
Veröffentlicht: (2026) -
SAFuzz: Semantic-Guided Adaptive Fuzzing for LLM-Generated Code
von: Yang, Ziyi, et al.
Veröffentlicht: (2026) -
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
von: Ge, Yifei, et al.
Veröffentlicht: (2026)