FormalSpecCpp: A Dataset of C++ Formal Specifications created using LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chakraborty, Madhurima, Pirkelbauer, Peter, Yi, Qing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
von: Xia, Shihao, et al.
Veröffentlicht: (2026)
von: Xia, Shihao, et al.
Veröffentlicht: (2026)
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024)
Understanding Tool-Augmented Agents for Lean Formalization: A Factorial Analysis
von: Zhang, Ke, et al.
Veröffentlicht: (2026)
von: Zhang, Ke, et al.
Veröffentlicht: (2026)
Intent-aligned Formal Specification Synthesis via Traceable Refinement
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
Descriptor: C++ Self-Admitted Technical Debt Dataset (CppSATD)
von: Pham, Phuoc, et al.
Veröffentlicht: (2025)
von: Pham, Phuoc, et al.
Veröffentlicht: (2025)
Understanding Formal Reasoning Failures in LLMs as Abstract Interpreters
von: Mitchell, Jacqueline L., et al.
Veröffentlicht: (2025)
von: Mitchell, Jacqueline L., et al.
Veröffentlicht: (2025)
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
von: Endres, Madeline, et al.
Veröffentlicht: (2023)
von: Endres, Madeline, et al.
Veröffentlicht: (2023)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization
von: Agarwal, Anmol, et al.
Veröffentlicht: (2026)
von: Agarwal, Anmol, et al.
Veröffentlicht: (2026)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
von: Kadosh, Tal, et al.
Veröffentlicht: (2023)
von: Kadosh, Tal, et al.
Veröffentlicht: (2023)
Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
von: Yu, Simon, et al.
Veröffentlicht: (2026)
von: Yu, Simon, et al.
Veröffentlicht: (2026)
Is Programming by Example solved by LLMs?
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
von: Li, Wen-Ding, et al.
Veröffentlicht: (2024)
Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
QEDCartographer: Automating Formal Verification Using Reward-Free Reinforcement Learning
von: Sanchez-Stern, Alex, et al.
Veröffentlicht: (2024)
von: Sanchez-Stern, Alex, et al.
Veröffentlicht: (2024)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
Beyond Postconditions: Can Large Language Models infer Formal Contracts for Automatic Software Verification?
von: Richter, Cedric, et al.
Veröffentlicht: (2025)
von: Richter, Cedric, et al.
Veröffentlicht: (2025)
VeriSoftBench: Repository-Scale Formal Verification Benchmarks for Lean
von: Xin, Yutong, et al.
Veröffentlicht: (2026)
von: Xin, Yutong, et al.
Veröffentlicht: (2026)
A Joint Learning Model with Variational Interaction for Multilingual Program Translation
von: Du, Yali, et al.
Veröffentlicht: (2024)
von: Du, Yali, et al.
Veröffentlicht: (2024)
PerfRL: A Small Language Model Framework for Efficient Code Optimization
von: Duan, Shukai, et al.
Veröffentlicht: (2023)
von: Duan, Shukai, et al.
Veröffentlicht: (2023)
A Multi-Expert Large Language Model Architecture for Verilog Code Generation
von: Nadimi, Bardia, et al.
Veröffentlicht: (2024)
von: Nadimi, Bardia, et al.
Veröffentlicht: (2024)
JavaBench: A Benchmark of Object-Oriented Code Generation for Evaluating Large Language Models
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
von: Cao, Jialun, et al.
Veröffentlicht: (2024)
Incoherence as Oracle-less Measure of Error in LLM-Based Code Generation
von: Valentin, Thomas, et al.
Veröffentlicht: (2025)
von: Valentin, Thomas, et al.
Veröffentlicht: (2025)
ChatDBG: Augmenting Debugging with Large Language Models
von: Levin, Kyla H., et al.
Veröffentlicht: (2024)
von: Levin, Kyla H., et al.
Veröffentlicht: (2024)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
von: Hooda, Ashish, et al.
Veröffentlicht: (2024)
von: Hooda, Ashish, et al.
Veröffentlicht: (2024)
On the Effectiveness of Machine Learning-based Call Graph Pruning: An Empirical Study
von: Mir, Amir M., et al.
Veröffentlicht: (2024)
von: Mir, Amir M., et al.
Veröffentlicht: (2024)
ScenicNL: Generating Probabilistic Scenario Programs from Natural Language
von: Elmaaroufi, Karim, et al.
Veröffentlicht: (2024)
von: Elmaaroufi, Karim, et al.
Veröffentlicht: (2024)
Large Language Models for Code Summarization
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
Representing Prompting Patterns with PDL: Compliance Agent Case Study
von: Vaziri, Mandana, et al.
Veröffentlicht: (2025)
von: Vaziri, Mandana, et al.
Veröffentlicht: (2025)
FlakyGuard: Automatically Fixing Flaky Tests at Industry Scale
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
APRIL: API Synthesis with Automatic Prompt Optimization and Reinforcement Learning
von: Zhong, Hua, et al.
Veröffentlicht: (2025)
von: Zhong, Hua, et al.
Veröffentlicht: (2025)
Encoding architecture algebra
von: Bersier, Stephane, et al.
Veröffentlicht: (2024)
von: Bersier, Stephane, et al.
Veröffentlicht: (2024)
Large Language Models Synergize with Automated Machine Learning
von: Xu, Jinglue, et al.
Veröffentlicht: (2024)
von: Xu, Jinglue, et al.
Veröffentlicht: (2024)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
Linguacodus: A Synergistic Framework for Transformative Code Generation in Machine Learning Pipelines
von: Trofimova, Ekaterina, et al.
Veröffentlicht: (2024)
von: Trofimova, Ekaterina, et al.
Veröffentlicht: (2024)
The 4/$δ$ Bound: Designing Predictable LLM-Verifier Systems for Formal Method Guarantee
von: Dantas, PIerre, et al.
Veröffentlicht: (2025)
von: Dantas, PIerre, et al.
Veröffentlicht: (2025)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
ReGAL: Refactoring Programs to Discover Generalizable Abstractions
von: Stengel-Eskin, Elias, et al.
Veröffentlicht: (2024)
von: Stengel-Eskin, Elias, et al.
Veröffentlicht: (2024)
EffiPair: Improving the Efficiency of LLM-generated Code with Relative Contrastive Feedback
von: Hajizadeh, Samira, et al.
Veröffentlicht: (2026)
von: Hajizadeh, Samira, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Doc2Spec: Synthesizing Formal Programming Specifications from Natural Language via Grammar Induction
von: Xia, Shihao, et al.
Veröffentlicht: (2026) -
DafnyBench: A Benchmark for Formal Software Verification
von: Loughridge, Chloe, et al.
Veröffentlicht: (2024) -
Understanding Tool-Augmented Agents for Lean Formalization: A Factorial Analysis
von: Zhang, Ke, et al.
Veröffentlicht: (2026) -
Intent-aligned Formal Specification Synthesis via Traceable Refinement
von: Ye, Zhe, et al.
Veröffentlicht: (2026) -
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)