PoLLMgraph: Unraveling Hallucinations in Large Language Models via State Transition Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Derui, Chen, Dingfan, Li, Qing, Chen, Zongxiong, Ma, Lei, Grossklags, Jens, Fritz, Mario |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SmartPoC: Generating Executable and Validated PoCs for Smart Contract Bug Reports
by: Chen, Longfei, et al.
Published: (2025)
by: Chen, Longfei, et al.
Published: (2025)
PBFuzz: Agentic Directed Fuzzing for PoV Generation
by: Zeng, Haochen, et al.
Published: (2025)
by: Zeng, Haochen, et al.
Published: (2025)
EvoPoC: Automated Exploit Synthesis for DeFi Smart Contracts via Hierarchical Knowledge Graphs
by: Liang, Ruichao, et al.
Published: (2026)
by: Liang, Ruichao, et al.
Published: (2026)
PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages
by: Simsek, Deniz, et al.
Published: (2025)
by: Simsek, Deniz, et al.
Published: (2025)
HALURust: Exploiting Hallucinations of Large Language Models to Detect Vulnerabilities in Rust
by: Luo, Yu, et al.
Published: (2025)
by: Luo, Yu, et al.
Published: (2025)
From Transactions to Exploits: Automated PoC Synthesis for Real-World DeFi Attacks
by: Su, Xing, et al.
Published: (2026)
by: Su, Xing, et al.
Published: (2026)
Patch-to-PoC: A Systematic Study of Agentic LLM Systems for Linux Kernel N-Day Reproduction
by: Pu, Juefei, et al.
Published: (2026)
by: Pu, Juefei, et al.
Published: (2026)
TELSAFE: Security Gap Quantitative Risk Assessment Framework
by: Siddiqui, Sarah Ali, et al.
Published: (2025)
by: Siddiqui, Sarah Ali, et al.
Published: (2025)
Understanding the Supply Chain and Risks of Large Language Model Applications
by: Ma, Yujie, et al.
Published: (2025)
by: Ma, Yujie, et al.
Published: (2025)
AutoFirm: Automatically Identifying Reused Libraries inside IoT Firmware at Large-Scale
by: Chen, YongLe, et al.
Published: (2024)
by: Chen, YongLe, et al.
Published: (2024)
PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
by: Liu, Xiting, et al.
Published: (2026)
by: Liu, Xiting, et al.
Published: (2026)
Large Language Model Supply Chain: Open Problems From the Security Perspective
by: Hu, Qiang, et al.
Published: (2024)
by: Hu, Qiang, et al.
Published: (2024)
DeCoMa: Detecting and Purifying Code Dataset Watermarks through Dual Channel Code Abstraction
by: Xiao, Yuan, et al.
Published: (2025)
by: Xiao, Yuan, et al.
Published: (2025)
Large Language Models-Aided Program Debloating
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
LineBreaker: Finding Token-Inconsistency Bugs with Large Language Models
by: Chen, Hongbo, et al.
Published: (2024)
by: Chen, Hongbo, et al.
Published: (2024)
PoCo: Agentic Proof-of-Concept Exploit Generation for Smart Contracts
by: Andersson, Vivi, et al.
Published: (2025)
by: Andersson, Vivi, et al.
Published: (2025)
SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
by: Hu, Qi, et al.
Published: (2026)
by: Hu, Qi, et al.
Published: (2026)
ReposVul: A Repository-Level High-Quality Vulnerability Dataset
by: Wang, Xinchen, et al.
Published: (2024)
by: Wang, Xinchen, et al.
Published: (2024)
LUNA: A Model-Based Universal Analysis Framework for Large Language Models
by: Song, Da, et al.
Published: (2023)
by: Song, Da, et al.
Published: (2023)
A Large-scale Empirical Study on the Generalizability of Disclosed Java Library Vulnerability Exploits
by: Chen, Zirui, et al.
Published: (2026)
by: Chen, Zirui, et al.
Published: (2026)
How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection
by: Chen, Maofei, et al.
Published: (2026)
by: Chen, Maofei, et al.
Published: (2026)
An Empirical Study on the Effectiveness of Large Language Models for Binary Code Understanding
by: Shang, Xiuwei, et al.
Published: (2025)
by: Shang, Xiuwei, et al.
Published: (2025)
Out of Sight, Still at Risk: The Lifecycle of Transitive Vulnerabilities in Maven
by: Przymus, Piotr, et al.
Published: (2025)
by: Przymus, Piotr, et al.
Published: (2025)
Evaluating Large Language Models for Line-Level Vulnerability Localization
by: Zhang, Jian, et al.
Published: (2024)
by: Zhang, Jian, et al.
Published: (2024)
A Large Scale Study of AI-based Binary Function Similarity Detection Techniques for Security Researchers and Practitioners
by: Shi, Jingyi, et al.
Published: (2025)
by: Shi, Jingyi, et al.
Published: (2025)
Dissecting Payload-based Transaction Phishing on Ethereum
by: Chen, Zhuo, et al.
Published: (2024)
by: Chen, Zhuo, et al.
Published: (2024)
FuzzingBrain V2: A Multi-Agent LLM System for Automated Vulnerability Discovery and Reproduction
by: Sheng, Ze, et al.
Published: (2026)
by: Sheng, Ze, et al.
Published: (2026)
How Far Have We Gone in Binary Code Understanding Using Large Language Models
by: Shang, Xiuwei, et al.
Published: (2024)
by: Shang, Xiuwei, et al.
Published: (2024)
Toward Quantum-Safe Software Engineering: A Vision for Post-Quantum Cryptography Migration
by: Zhang, Lei
Published: (2026)
by: Zhang, Lei
Published: (2026)
CrossInspector: A Static Analysis Approach for Cross-Contract Vulnerability Detection
by: Chen, Xiao
Published: (2024)
by: Chen, Xiao
Published: (2024)
Vulnerability-Hunter: An Adaptive Feature Perception Attention Network for Smart Contract Vulnerabilities
by: Chen, Yizhou
Published: (2024)
by: Chen, Yizhou
Published: (2024)
SoK: Web3 RegTech for Cryptocurrency VASP AML/CFT Compliance
by: Mao, Qian'ang, et al.
Published: (2025)
by: Mao, Qian'ang, et al.
Published: (2025)
Is Stateful Fuzzing Really Challenging?
by: Daniele, Cristian
Published: (2024)
by: Daniele, Cristian
Published: (2024)
JSidentify-V2: Leveraging Dynamic Memory Fingerprinting for Mini-Game Plagiarism Detection
by: Li, Zhihao, et al.
Published: (2025)
by: Li, Zhihao, et al.
Published: (2025)
VulnRepairEval: An Exploit-Based Evaluation Framework for Assessing Large Language Model Vulnerability Repair Capabilities
by: Wang, Weizhe, et al.
Published: (2025)
by: Wang, Weizhe, et al.
Published: (2025)
Prompt Fuzzing for Fuzz Driver Generation
by: Lyu, Yunlong, et al.
Published: (2023)
by: Lyu, Yunlong, et al.
Published: (2023)
CNT: Safety-oriented Function Reuse across LLMs via Cross-Model Neuron Transfer
by: Zhao, Yue, et al.
Published: (2026)
by: Zhao, Yue, et al.
Published: (2026)
Train in Vain: Functionality-Preserving Poisoning to Prevent Unauthorized Use of Code Datasets
by: Xiao, Yuan, et al.
Published: (2026)
by: Xiao, Yuan, et al.
Published: (2026)
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
by: Peng, Yibo, et al.
Published: (2025)
by: Peng, Yibo, et al.
Published: (2025)
Demystifying and Detecting Cryptographic Defects in Ethereum Smart Contracts
by: Zhang, Jiashuo, et al.
Published: (2024)
by: Zhang, Jiashuo, et al.
Published: (2024)
Similar Items
-
SmartPoC: Generating Executable and Validated PoCs for Smart Contract Bug Reports
by: Chen, Longfei, et al.
Published: (2025) -
PBFuzz: Agentic Directed Fuzzing for PoV Generation
by: Zeng, Haochen, et al.
Published: (2025) -
EvoPoC: Automated Exploit Synthesis for DeFi Smart Contracts via Hierarchical Knowledge Graphs
by: Liang, Ruichao, et al.
Published: (2026) -
PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages
by: Simsek, Deniz, et al.
Published: (2025) -
HALURust: Exploiting Hallucinations of Large Language Models to Detect Vulnerabilities in Rust
by: Luo, Yu, et al.
Published: (2025)