ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Ahuja, Riyaz, Rowney, Tate, Avigad, Jeremy, Welleck, Sean
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866913154005467136
author Ahuja, Riyaz
Rowney, Tate
Avigad, Jeremy
Welleck, Sean
author_facet Ahuja, Riyaz
Rowney, Tate
Avigad, Jeremy
Welleck, Sean
contents Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data quality for neural provers. However, scalable proof optimization is hindered by heterogeneous and heuristically specified objectives, scarce data, and high training and inference costs. To overcome these challenges, we introduce ImProver 2, a neurosymbolic framework for automated proof optimization in Lean 4. ImProver 2 combines a data-efficient expert-iteration pipeline with a scaffold that exposes formal structure alongside lightweight informal abstractions. We further introduce a suite of metrics capturing structural proof properties. Using ImProver 2, we train a 7B-parameter model that outperforms orders-of-magnitude larger models within the same model family, and is competitive with mid-tier frontier models across metrics. We additionally demonstrate that our neurosymbolic scaffold significantly improves performance across both small and frontier models. We show that with proper scaffolding and training, small models can effectively restructure research-level proofs over complex and varied metrics, matching substantially larger systems and establishing proof optimization as a scalable, learnable task.
format Preprint
id arxiv_https___arxiv_org_abs_2605_22885
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
Ahuja, Riyaz
Rowney, Tate
Avigad, Jeremy
Welleck, Sean
Artificial Intelligence
Computation and Language
Machine Learning
Logic in Computer Science
Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data quality for neural provers. However, scalable proof optimization is hindered by heterogeneous and heuristically specified objectives, scarce data, and high training and inference costs. To overcome these challenges, we introduce ImProver 2, a neurosymbolic framework for automated proof optimization in Lean 4. ImProver 2 combines a data-efficient expert-iteration pipeline with a scaffold that exposes formal structure alongside lightweight informal abstractions. We further introduce a suite of metrics capturing structural proof properties. Using ImProver 2, we train a 7B-parameter model that outperforms orders-of-magnitude larger models within the same model family, and is competitive with mid-tier frontier models across metrics. We additionally demonstrate that our neurosymbolic scaffold significantly improves performance across both small and frontier models. We show that with proper scaffolding and training, small models can effectively restructure research-level proofs over complex and varied metrics, matching substantially larger systems and establishing proof optimization as a scalable, learnable task.
title ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
topic Artificial Intelligence
Computation and Language
Machine Learning
Logic in Computer Science
url https://arxiv.org/abs/2605.22885