Defining and Detecting the Defects of the Large Language Model-based Autonomous Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Ning, Kaiwen, Chen, Jiachi, Zhang, Jingwen, Li, Wei, Wang, Zexu, Feng, Yuming, Zhang, Weizhe, Zheng, Zibin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SSR: Safeguarding Staking Rewards by Defining and Detecting Logical Defects in DeFi Staking
by: Lin, Zewei, et al.
Published: (2026)
by: Lin, Zewei, et al.
Published: (2026)
Copy-and-Paste? Identifying EVM-Inequivalent Code Smells in Multi-chain Reuse Contracts
by: Wang, Zexu, et al.
Published: (2025)
by: Wang, Zexu, et al.
Published: (2025)
One Signature, Multiple Payments: Demystifying and Detecting Signature Replay Vulnerabilities in Smart Contracts
by: Wang, Zexu, et al.
Published: (2025)
by: Wang, Zexu, et al.
Published: (2025)
Definition and Detection of Centralization Defects in Smart Contracts
by: Lin, Zewei, et al.
Published: (2024)
by: Lin, Zewei, et al.
Published: (2024)
Efficiently Detecting Reentrancy Vulnerabilities in Complex Smart Contracts
by: Wang, Zexu, et al.
Published: (2024)
by: Wang, Zexu, et al.
Published: (2024)
A Survey of Large Language Models for Code: Evolution, Benchmarking, and Future Trends
by: Zheng, Zibin, et al.
Published: (2023)
by: Zheng, Zibin, et al.
Published: (2023)
Unity is Strength: Enhancing Precision in Reentrancy Vulnerability Detection of Smart Contract Analysis Tools
by: Wang, Zexu, et al.
Published: (2024)
by: Wang, Zexu, et al.
Published: (2024)
V2E: Validating Smart Contract Vulnerabilities through Profit-driven Exploit Generation and Execution
by: Zhang, Jingwen, et al.
Published: (2026)
by: Zhang, Jingwen, et al.
Published: (2026)
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
by: Ning, Kaiwen, et al.
Published: (2024)
by: Ning, Kaiwen, et al.
Published: (2024)
SmartReco: Detecting Read-Only Reentrancy via Fine-Grained Cross-DApp Analysis
by: Zhang, Jingwen, et al.
Published: (2024)
by: Zhang, Jingwen, et al.
Published: (2024)
Towards an Understanding of Large Language Models in Software Engineering Tasks
by: Zheng, Zibin, et al.
Published: (2023)
by: Zheng, Zibin, et al.
Published: (2023)
CRPWarner: Warning the Risk of Contract-related Rug Pull in DeFi Smart Contracts
by: Lin, Zewei, et al.
Published: (2024)
by: Lin, Zewei, et al.
Published: (2024)
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
An Empirical Study on Low-Code Programming using Traditional vs Large Language Model Support
by: Liu, Yongkun, et al.
Published: (2024)
by: Liu, Yongkun, et al.
Published: (2024)
NumScout: Unveiling Numerical Defects in Smart Contracts using LLM-Pruning Symbolic Execution
by: Chen, Jiachi, et al.
Published: (2025)
by: Chen, Jiachi, et al.
Published: (2025)
DAppSCAN: Building Large-Scale Datasets for Smart Contract Weaknesses in DApp Projects
by: Zheng, Zibin, et al.
Published: (2023)
by: Zheng, Zibin, et al.
Published: (2023)
An Empirical Study of Agent Developer Practices in AI Agent Frameworks
by: Wang, Yanlin, et al.
Published: (2025)
by: Wang, Yanlin, et al.
Published: (2025)
Uncover the Premeditated Attacks: Detecting Exploitable Reentrancy Vulnerabilities by Identifying Attacker Contracts
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Beyond Functional Correctness: Investigating Coding Style Inconsistencies in Large Language Models
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
RiskTagger: An LLM-based Agent for Automatic Annotation of Web3 Crypto Money Laundering Behaviors
by: Lin, Dan, et al.
Published: (2025)
by: Lin, Dan, et al.
Published: (2025)
RMCBench: Benchmarking Large Language Models' Resistance to Malicious Code
by: Chen, Jiachi, et al.
Published: (2024)
by: Chen, Jiachi, et al.
Published: (2024)
RLCoder: Reinforcement Learning for Repository-Level Code Completion
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
Comment Traps: How Defective Commented-out Code Augment Defects in AI-Assisted Code Generation
by: Huang, Yuan, et al.
Published: (2025)
by: Huang, Yuan, et al.
Published: (2025)
SmartOracle: Generating Smart Contract Oracle via Fine-Grained Invariant Detection
by: Su, Jianzhong, et al.
Published: (2024)
by: Su, Jianzhong, et al.
Published: (2024)
When Large Language Models Meet UAV Projects: An Empirical Study from Developers' Perspective
by: Chen, Yihua, et al.
Published: (2025)
by: Chen, Yihua, et al.
Published: (2025)
LLM Hallucinations in Practical Code Generation: Phenomena, Mechanism, and Mitigation
by: Zhang, Ziyao, et al.
Published: (2024)
by: Zhang, Ziyao, et al.
Published: (2024)
AgentRaft: Automated Detection of Data Over-Exposure in LLM Agents
by: Lin, Yixi, et al.
Published: (2026)
by: Lin, Yixi, et al.
Published: (2026)
Enhancing The Open Network: Definition and Automated Detection of Smart Contract Defects
by: Song, Hao, et al.
Published: (2025)
by: Song, Hao, et al.
Published: (2025)
Identifying Smart Contract Security Issues in Code Snippets from Stack Overflow
by: Chen, Jiachi, et al.
Published: (2024)
by: Chen, Jiachi, et al.
Published: (2024)
Demystifying and Detecting Cryptographic Defects in Ethereum Smart Contracts
by: Zhang, Jiashuo, et al.
Published: (2024)
by: Zhang, Jiashuo, et al.
Published: (2024)
You Augment Me: Exploring ChatGPT-based Data Augmentation for Semantic Code Search
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
When ChatGPT Meets Smart Contract Vulnerability Detection: How Far Are We?
by: Chen, Chong, et al.
Published: (2023)
by: Chen, Chong, et al.
Published: (2023)
When to Stop? Towards Efficient Code Generation in LLMs with Excess Token Prevention
by: Guo, Lianghong, et al.
Published: (2024)
by: Guo, Lianghong, et al.
Published: (2024)
An Empirical Study of Interaction Bugs in ROS-based Software
by: Chen, Zhixiang, et al.
Published: (2025)
by: Chen, Zhixiang, et al.
Published: (2025)
OmniGIRL: A Multilingual and Multimodal Benchmark for GitHub Issue Resolution
by: Guo, Lianghong, et al.
Published: (2025)
by: Guo, Lianghong, et al.
Published: (2025)
A Preliminary Study on the Robustness of Code Generation by Large Language Models
by: Li, Zike, et al.
Published: (2025)
by: Li, Zike, et al.
Published: (2025)
Hyperion: Unveiling DApp Inconsistencies using LLM and Dataflow-Guided Symbolic Execution
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Accelerating Automatic Program Repair with Dual Retrieval-Augmented Fine-Tuning and Patch Generation on Large Language Models
by: Guo, Hanyang, et al.
Published: (2025)
by: Guo, Hanyang, et al.
Published: (2025)
iJTyper: An Iterative Type Inference Framework for Java by Integrating Constraint- and Statistically-based Methods
by: Chen, Zhixiang, et al.
Published: (2024)
by: Chen, Zhixiang, et al.
Published: (2024)
FeedbackEval: A Benchmark for Evaluating Large Language Models in Feedback-Driven Code Repair Tasks
by: Dai, Dekun, et al.
Published: (2025)
by: Dai, Dekun, et al.
Published: (2025)
Similar Items
-
SSR: Safeguarding Staking Rewards by Defining and Detecting Logical Defects in DeFi Staking
by: Lin, Zewei, et al.
Published: (2026) -
Copy-and-Paste? Identifying EVM-Inequivalent Code Smells in Multi-chain Reuse Contracts
by: Wang, Zexu, et al.
Published: (2025) -
One Signature, Multiple Payments: Demystifying and Detecting Signature Replay Vulnerabilities in Smart Contracts
by: Wang, Zexu, et al.
Published: (2025) -
Definition and Detection of Centralization Defects in Smart Contracts
by: Lin, Zewei, et al.
Published: (2024) -
Efficiently Detecting Reentrancy Vulnerabilities in Complex Smart Contracts
by: Wang, Zexu, et al.
Published: (2024)