DeCon: Detecting Incorrect Assertions via Postconditions Generated by a Large Language Model
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Hao, Chen, Tianyu, Huang, Jiaming, Li, Zongyang, Ran, Dezhi, Wang, Xinyu, Li, Ying, Marron, Assaf, Harel, David, Xie, Yuan, Xie, Tao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Infrastructure Software Perspective Toward Computation Offloading between Executable Specifications and Foundation Models
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
Preparing for Super-Reactivity: Early Fault-Detection in the Development of Exceedingly Complex Reactive Systems
by: Harel, David, et al.
Published: (2024)
by: Harel, David, et al.
Published: (2024)
On Augmenting Scenario-Based Modeling with Generative AI
by: Harel, David, et al.
Published: (2024)
by: Harel, David, et al.
Published: (2024)
Beyond Pass or Fail: Multi-Dimensional Benchmarking of Foundation Models for Goal-based Mobile UI Navigation
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
Expecting the Unexpected: Developing Autonomous-System Design Principles for Reacting to Unpredicted Events and Conditions
by: Marron, Assaf, et al.
Published: (2020)
by: Marron, Assaf, et al.
Published: (2020)
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models
by: Yu, Hao, et al.
Published: (2023)
by: Yu, Hao, et al.
Published: (2023)
Skill-Adpative Imitation Learning for UI Test Reuse
by: Wu, Mengzhou, et al.
Published: (2024)
by: Wu, Mengzhou, et al.
Published: (2024)
Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software
by: Ran, Dezhi, et al.
Published: (2024)
by: Ran, Dezhi, et al.
Published: (2024)
POSTCONDBENCH: Benchmarking Correctness and Completeness in Formal Postcondition Inference
by: Zhang, Gehao, et al.
Published: (2026)
by: Zhang, Gehao, et al.
Published: (2026)
Assertion Messages with Large Language Models (LLMs) for Code
by: Aljohani, Ahmed, et al.
Published: (2025)
by: Aljohani, Ahmed, et al.
Published: (2025)
Beyond Code Generation: Assessing Code LLM Maturity with Postconditions
by: He, Fusen, et al.
Published: (2024)
by: He, Fusen, et al.
Published: (2024)
AssertionBench: A Benchmark to Evaluate Large-Language Models for Assertion Generation
by: Pulavarthi, Vaishnavi, et al.
Published: (2024)
by: Pulavarthi, Vaishnavi, et al.
Published: (2024)
Automatic Assertion Mining in Assertion-Based Verification: Techniques, Challenges, and Future Directions
by: Iman, Mohammad Reza Heidari, et al.
Published: (2026)
by: Iman, Mohammad Reza Heidari, et al.
Published: (2026)
Improving Retrieval-Augmented Deep Assertion Generation via Joint Training
by: Zhang, Quanjun, et al.
Published: (2025)
by: Zhang, Quanjun, et al.
Published: (2025)
SPADE: Synthesizing Data Quality Assertions for Large Language Model Pipelines
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
Studying the Impact of Early Test Termination Due to Assertion Failure on Code Coverage and Spectrum-based Fault Localization
by: Uddin, Md. Ashraf, et al.
Published: (2025)
by: Uddin, Md. Ashraf, et al.
Published: (2025)
Improving Deep Assertion Generation via Fine-Tuning Retrieval-Augmented Pre-trained Language Models
by: Zhang, Quanjun, et al.
Published: (2025)
by: Zhang, Quanjun, et al.
Published: (2025)
Breaking the Myth: Can Small Models Infer Postconditions Too?
by: Zhang, Gehao, et al.
Published: (2025)
by: Zhang, Gehao, et al.
Published: (2025)
Understanding and Characterizing Mock Assertions in Unit Tests
by: Zhu, Hengcheng, et al.
Published: (2025)
by: Zhu, Hengcheng, et al.
Published: (2025)
Fine-Grained Assertion-Based Test Selection
by: Gu, Sijia, et al.
Published: (2024)
by: Gu, Sijia, et al.
Published: (2024)
SimdBench: Benchmarking Large Language Models for SIMD-Intrinsic Code Generation
by: He, Yibo, et al.
Published: (2025)
by: He, Yibo, et al.
Published: (2025)
The First Prompt Counts the Most! An Evaluation of Large Language Models on Iterative Example-Based Code Generation
by: Fu, Yingjie, et al.
Published: (2024)
by: Fu, Yingjie, et al.
Published: (2024)
ChatGPT Incorrectness Detection in Software Reviews
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
by: Tanzil, Minaoar Hossain, et al.
Published: (2024)
Toward Programming Languages for Reasoning: Humans, Symbolic Systems, and AI Agents
by: Marron, Mark
Published: (2024)
by: Marron, Mark
Published: (2024)
A New Generation of Intelligent Development Environments
by: Marron, Mark
Published: (2024)
by: Marron, Mark
Published: (2024)
An Effectively $Ω(c)$ Language and Runtime
by: Marron, Mark
Published: (2024)
by: Marron, Mark
Published: (2024)
A Time Series Analysis of Assertions in the Linux Kernel
by: Ruohonen, Jukka
Published: (2024)
by: Ruohonen, Jukka
Published: (2024)
Enhancing LLM-based Fault Localization with a Functionality-Aware Retrieval-Augmented Generation Framework
by: Shi, Xinyu, et al.
Published: (2025)
by: Shi, Xinyu, et al.
Published: (2025)
VulStamp: Vulnerability Assessment using Large Language Model
by: Shen, Hao, et al.
Published: (2025)
by: Shen, Hao, et al.
Published: (2025)
Towards Large Language Model Guided Kernel Direct Fuzzing
by: Li, Xie, et al.
Published: (2025)
by: Li, Xie, et al.
Published: (2025)
AppForge: From Assistant to Independent Developer -- Are GPTs Ready for Software Development?
by: Ran, Dezhi, et al.
Published: (2025)
by: Ran, Dezhi, et al.
Published: (2025)
SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference
by: Le, Cuong Chi, et al.
Published: (2026)
by: Le, Cuong Chi, et al.
Published: (2026)
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
by: Endres, Madeline, et al.
Published: (2023)
by: Endres, Madeline, et al.
Published: (2023)
Assertion-Aware Test Code Summarization with Large Language Models
by: Mollah, Anamul Haque, et al.
Published: (2025)
by: Mollah, Anamul Haque, et al.
Published: (2025)
Measuring the Influence of Incorrect Code on Test Generation
by: Huang, Dong, et al.
Published: (2024)
by: Huang, Dong, et al.
Published: (2024)
Towards Automatic Translation of Machine Learning Visual Insights to Analytical Assertions
by: Shome, Arumoy, et al.
Published: (2024)
by: Shome, Arumoy, et al.
Published: (2024)
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
by: Richter, Cedric, et al.
Published: (2025)
by: Richter, Cedric, et al.
Published: (2025)
On the Rationale and Use of Assertion Messages in Test Code: Insights from Software Practitioners
by: Peruma, Anthony, et al.
Published: (2024)
by: Peruma, Anthony, et al.
Published: (2024)
A Regression Testing Framework with Automated Assertion Generation for Machine Learning Notebooks
by: Yao, Yingao Elaine, et al.
Published: (2025)
by: Yao, Yingao Elaine, et al.
Published: (2025)
SiblingRepair: Sibling-Based Multi-Hunk Repair with Large Language Models
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
Similar Items
-
An Infrastructure Software Perspective Toward Computation Offloading between Executable Specifications and Foundation Models
by: Ran, Dezhi, et al.
Published: (2025) -
Preparing for Super-Reactivity: Early Fault-Detection in the Development of Exceedingly Complex Reactive Systems
by: Harel, David, et al.
Published: (2024) -
On Augmenting Scenario-Based Modeling with Generative AI
by: Harel, David, et al.
Published: (2024) -
Beyond Pass or Fail: Multi-Dimensional Benchmarking of Foundation Models for Goal-based Mobile UI Navigation
by: Ran, Dezhi, et al.
Published: (2025) -
Expecting the Unexpected: Developing Autonomous-System Design Principles for Reacting to Unpredicted Events and Conditions
by: Marron, Assaf, et al.
Published: (2020)