Reinforcement Learning with Negative Tests as Completeness Signal for Formal Specification Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Zhechong, Zhang, Zhao, Sun, Zeyu, Sun, Huifeng, Xiong, Yingfei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
by: Huang, Zhechong, et al.
Published: (2025)
by: Huang, Zhechong, et al.
Published: (2025)
GramTrans: A Better Code Representation Approach in Code Generation
by: Zhang, Zhao, et al.
Published: (2025)
by: Zhang, Zhao, et al.
Published: (2025)
Effective Random Test Generation for Deep Learning Compilers
by: Ren, Luyao, et al.
Published: (2023)
by: Ren, Luyao, et al.
Published: (2023)
On Reasoning-Centric LLM-based Automated Theorem Proving
by: Sun, Yican, et al.
Published: (2026)
by: Sun, Yican, et al.
Published: (2026)
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning
by: Wang, Guoqing, et al.
Published: (2026)
by: Wang, Guoqing, et al.
Published: (2026)
Trustworthy Software Project Generation : a Case Study with an Interactive Theorem Prover
by: Fang, Jian, et al.
Published: (2026)
by: Fang, Jian, et al.
Published: (2026)
A Learning Method for Symbolic Systems Using Large Language Models
by: Fang, Jian, et al.
Published: (2026)
by: Fang, Jian, et al.
Published: (2026)
VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
by: Misu, Md Rakib Hossain, et al.
Published: (2026)
Lyra: A Benchmark for Turducken-Style Code Generation
by: Liang, Qingyuan, et al.
Published: (2021)
by: Liang, Qingyuan, et al.
Published: (2021)
Extracting Formal Specifications from Documents Using LLMs for Automated Testing
by: Li, Hui, et al.
Published: (2025)
by: Li, Hui, et al.
Published: (2025)
SemOpt: LLM-Driven Code Optimization via Rule-Based Analysis
by: Zhao, Yuwei, et al.
Published: (2025)
by: Zhao, Yuwei, et al.
Published: (2025)
Condor: A Code Discriminator Integrating General Semantics with Code Details
by: Liang, Qingyuan, et al.
Published: (2024)
by: Liang, Qingyuan, et al.
Published: (2024)
Automatically Learning a Precise Measurement for Fault Diagnosis Capability of Test Cases
by: Zhao, Yifan, et al.
Published: (2025)
by: Zhao, Yifan, et al.
Published: (2025)
Automatic Generation of Formal Specification and Verification Annotations Using LLMs and Test Oracles
by: Faria, João Pascoal, et al.
Published: (2026)
by: Faria, João Pascoal, et al.
Published: (2026)
POSTCONDBENCH: Benchmarking Correctness and Completeness in Formal Postcondition Inference
by: Zhang, Gehao, et al.
Published: (2026)
by: Zhang, Gehao, et al.
Published: (2026)
Interleaved Learning and Exploration: A Self-Adaptive Fuzz Testing Framework for MLIR
by: Sun, Zeyu, et al.
Published: (2025)
by: Sun, Zeyu, et al.
Published: (2025)
Generating Project-Specific Test Cases with Requirement Validation Intention
by: Qi, Binhang, et al.
Published: (2025)
by: Qi, Binhang, et al.
Published: (2025)
LLM-based Vulnerability Detection at Project Scale: An Empirical Study
by: Li, Fengjie, et al.
Published: (2026)
by: Li, Fengjie, et al.
Published: (2026)
Accelerating Patch Validation for Program Repair with Interception-Based Execution Scheduling
by: Xiao, Yuan-An, et al.
Published: (2023)
by: Xiao, Yuan-An, et al.
Published: (2023)
An Agile Formal Specification Language Design Based on K Framework
by: Zhang, Jianyu, et al.
Published: (2024)
by: Zhang, Jianyu, et al.
Published: (2024)
SpecSyn: LLM-based Synthesis and Refinement of Formal Specifications for Real-world Program Verification
by: Ma, Lezhi, et al.
Published: (2026)
by: Ma, Lezhi, et al.
Published: (2026)
Line-level Semantic Structure Learning for Code Vulnerability Detection
by: Wang, Ziliang, et al.
Published: (2024)
by: Wang, Ziliang, et al.
Published: (2024)
RLCoder: Reinforcement Learning for Repository-Level Code Completion
by: Wang, Yanlin, et al.
Published: (2024)
by: Wang, Yanlin, et al.
Published: (2024)
Event-B Agent: Towards LLM Agent for Formal Model Synthesis and Repair
by: Wang, Hongshu, et al.
Published: (2026)
by: Wang, Hongshu, et al.
Published: (2026)
IRCoCo: Immediate Rewards-Guided Deep Reinforcement Learning for Code Completion
by: Li, Bolun, et al.
Published: (2024)
by: Li, Bolun, et al.
Published: (2024)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
by: Guo, Liwei, et al.
Published: (2025)
by: Guo, Liwei, et al.
Published: (2025)
From Exploration to Specification: LLM-Based Property Generation for Mobile App Testing
by: Xiong, Yiheng, et al.
Published: (2026)
by: Xiong, Yiheng, et al.
Published: (2026)
Exploring and Evaluating Interplays of BPpy with Deep Reinforcement Learning and Formal Methods
by: Yaacov, Tom, et al.
Published: (2025)
by: Yaacov, Tom, et al.
Published: (2025)
Counterexample Classification against Signal Temporal Logic Specifications
by: Zhang, Zhenya, et al.
Published: (2026)
by: Zhang, Zhenya, et al.
Published: (2026)
PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion Models
by: Huang, Haocheng, et al.
Published: (2026)
by: Huang, Haocheng, et al.
Published: (2026)
FormalRTL: Verified RTL Synthesis at Scale
by: Li, Kezhi, et al.
Published: (2026)
by: Li, Kezhi, et al.
Published: (2026)
Formal Synthesis of Uncertainty Reduction Controllers
by: Carwehl, Marc, et al.
Published: (2024)
by: Carwehl, Marc, et al.
Published: (2024)
Testing for Fault Diversity in Reinforcement Learning
by: Mazouni, Quentin, et al.
Published: (2024)
by: Mazouni, Quentin, et al.
Published: (2024)
Reusable Test Suites for Reinforcement Learning
by: Betten, Jørn Eirik, et al.
Published: (2025)
by: Betten, Jørn Eirik, et al.
Published: (2025)
ATGen: Adversarial Reinforcement Learning for Test Case Generation
by: Li, Qingyao, et al.
Published: (2025)
by: Li, Qingyao, et al.
Published: (2025)
Deep Reinforcement Learning for Automated Web GUI Testing
by: Gu, Zhiyu, et al.
Published: (2025)
by: Gu, Zhiyu, et al.
Published: (2025)
Validity-Preserving Delta Debugging via Generator Trace Reduction
by: Ren, Luyao, et al.
Published: (2024)
by: Ren, Luyao, et al.
Published: (2024)
PredicateFix: Repairing Static Analysis Alerts with Bridging Predicates
by: Xiao, Yuan-An, et al.
Published: (2025)
by: Xiao, Yuan-An, et al.
Published: (2025)
Reducing Cost of LLM Agents with Trajectory Reduction
by: Xiao, Yuan-An, et al.
Published: (2025)
by: Xiao, Yuan-An, et al.
Published: (2025)
Test Adequacy for Metamorphic Testing: Criteria, Measurement, and Implication
by: Fu, An, et al.
Published: (2024)
by: Fu, An, et al.
Published: (2024)
Similar Items
-
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
by: Huang, Zhechong, et al.
Published: (2025) -
GramTrans: A Better Code Representation Approach in Code Generation
by: Zhang, Zhao, et al.
Published: (2025) -
Effective Random Test Generation for Deep Learning Compilers
by: Ren, Luyao, et al.
Published: (2023) -
On Reasoning-Centric LLM-based Automated Theorem Proving
by: Sun, Yican, et al.
Published: (2026) -
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning
by: Wang, Guoqing, et al.
Published: (2026)