ProofOptimizer: Training Language Models to Simplify Proofs without Human Demonstrations
Fuente:
arXiv
Guardado en:
| Autores principales: | Gu, Alex, Piotrowski, Bartosz, Gloeckle, Fabian, Yang, Kaiyu, Markosyan, Aram H. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FormalProofBench: Can Models Write Graduate Level Math Proofs That Are Formally Verified?
por: Ravi, Nikil, et al.
Publicado: (2026)
por: Ravi, Nikil, et al.
Publicado: (2026)
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
por: Cao, Jialun, et al.
Publicado: (2025)
por: Cao, Jialun, et al.
Publicado: (2025)
A Case Study on the Effectiveness of LLMs in Verification with Proof Assistants
por: Bayazıt, Barış, et al.
Publicado: (2025)
por: Bayazıt, Barış, et al.
Publicado: (2025)
Solving Inequality Proofs with Large Language Models
por: Lu, Pan, et al.
Publicado: (2025)
por: Lu, Pan, et al.
Publicado: (2025)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
por: Thakur, Amitayush, et al.
Publicado: (2025)
por: Thakur, Amitayush, et al.
Publicado: (2025)
What I cannot execute, I do not understand: Training and Evaluating LLMs on Program Execution Traces
por: Armengol-Estapé, Jordi, et al.
Publicado: (2025)
por: Armengol-Estapé, Jordi, et al.
Publicado: (2025)
Towards Neural Synthesis for SMT-Assisted Proof-Oriented Programming
por: Chakraborty, Saikat, et al.
Publicado: (2024)
por: Chakraborty, Saikat, et al.
Publicado: (2024)
MUSTARD: Mastering Uniform Synthesis of Theorem and Proof Data
por: Huang, Yinya, et al.
Publicado: (2024)
por: Huang, Yinya, et al.
Publicado: (2024)
A Certified Proof Checker for Deep Neural Network Verification in Imandra
por: Desmartin, Remi, et al.
Publicado: (2024)
por: Desmartin, Remi, et al.
Publicado: (2024)
Detecting Prefix Bias in LLM-based Reward Models
por: Kumar, Ashwin, et al.
Publicado: (2025)
por: Kumar, Ashwin, et al.
Publicado: (2025)
Beyond Pass-by-Pass Optimization: Intent-Driven IR Optimization with Large Language Models
por: Qiu, Lei, et al.
Publicado: (2026)
por: Qiu, Lei, et al.
Publicado: (2026)
Proof Automation with Large Language Models
por: Lu, Minghai, et al.
Publicado: (2024)
por: Lu, Minghai, et al.
Publicado: (2024)
Can Proof Assistants Verify Multi-Agent Systems?
por: Mendez, Julian Alfredo, et al.
Publicado: (2025)
por: Mendez, Julian Alfredo, et al.
Publicado: (2025)
Proof2Hybrid: Automatic Mathematical Benchmark Synthesis for Proof-Centric Problems
por: Peng, Yebo, et al.
Publicado: (2025)
por: Peng, Yebo, et al.
Publicado: (2025)
Meta Large Language Model Compiler: Foundation Models of Compiler Optimization
por: Cummins, Chris, et al.
Publicado: (2024)
por: Cummins, Chris, et al.
Publicado: (2024)
SLaDe: A Portable Small Language Model Decompiler for Optimized Assembly
por: Armengol-Estapé, Jordi, et al.
Publicado: (2023)
por: Armengol-Estapé, Jordi, et al.
Publicado: (2023)
Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search
por: Lu, Jialin, et al.
Publicado: (2026)
por: Lu, Jialin, et al.
Publicado: (2026)
The Proof is in the Almond Cookies
por: van Trijp, Remi, et al.
Publicado: (2025)
por: van Trijp, Remi, et al.
Publicado: (2025)
The Open Proof Corpus: A Large-Scale Study of LLM-Generated Mathematical Proofs
por: Dekoninck, Jasper, et al.
Publicado: (2025)
por: Dekoninck, Jasper, et al.
Publicado: (2025)
Emergent Representations of Program Semantics in Language Models Trained on Programs
por: Jin, Charles, et al.
Publicado: (2023)
por: Jin, Charles, et al.
Publicado: (2023)
Reliable Fine-Grained Evaluation of Natural Language Math Proofs
por: Ma, Wenjie, et al.
Publicado: (2025)
por: Ma, Wenjie, et al.
Publicado: (2025)
ProofSketch: Efficient Verified Reasoning for Large Language Models
por: Sheshanarayana, Disha, et al.
Publicado: (2025)
por: Sheshanarayana, Disha, et al.
Publicado: (2025)
Exploring the Role of Reasoning Structures for Constructing Proofs in Multi-Step Natural Language Reasoning with Large Language Models
por: Zheng, Zi'ou, et al.
Publicado: (2024)
por: Zheng, Zi'ou, et al.
Publicado: (2024)
Next-Token Prediction Task Assumes Optimal Data Ordering for LLM Training in Proof Generation
por: An, Chenyang, et al.
Publicado: (2024)
por: An, Chenyang, et al.
Publicado: (2024)
SAC-Opt: Semantic Anchors for Iterative Correction in Optimization Modeling
por: Zhang, Yansen, et al.
Publicado: (2025)
por: Zhang, Yansen, et al.
Publicado: (2025)
Neuro-Symbolic Integration Brings Causal and Reliable Reasoning Proofs
por: Yang, Sen, et al.
Publicado: (2023)
por: Yang, Sen, et al.
Publicado: (2023)
SGLang: Efficient Execution of Structured Language Model Programs
por: Zheng, Lianmin, et al.
Publicado: (2023)
por: Zheng, Lianmin, et al.
Publicado: (2023)
How Do Humans Write Code? Large Models Do It the Same Way Too
por: Li, Long, et al.
Publicado: (2024)
por: Li, Long, et al.
Publicado: (2024)
Algorithmic Language Models with Neurally Compiled Libraries
por: Saldyt, Lucas, et al.
Publicado: (2024)
por: Saldyt, Lucas, et al.
Publicado: (2024)
Can Language Models Solve Olympiad Programming?
por: Shi, Quan, et al.
Publicado: (2024)
por: Shi, Quan, et al.
Publicado: (2024)
EnvTrace: Simulation-Based Semantic Evaluation of LLM Code via Execution Trace Alignment -- Demonstrated at Synchrotron Beamlines
por: van der Vleuten, Noah, et al.
Publicado: (2025)
por: van der Vleuten, Noah, et al.
Publicado: (2025)
A Natural Formalized Proof Language
por: Xie, Lihan, et al.
Publicado: (2024)
por: Xie, Lihan, et al.
Publicado: (2024)
From Reasoning to Learning: A Survey on Hypothesis Discovery and Rule Learning with Large Language Models
por: He, Kaiyu, et al.
Publicado: (2025)
por: He, Kaiyu, et al.
Publicado: (2025)
Aletheia tackles FirstProof autonomously
por: Feng, Tony, et al.
Publicado: (2026)
por: Feng, Tony, et al.
Publicado: (2026)
Autograding Mathematical Induction Proofs with Natural Language Processing
por: Zhao, Chenyan, et al.
Publicado: (2024)
por: Zhao, Chenyan, et al.
Publicado: (2024)
Assessing the Interpretability of Programmatic Policies with Large Language Models
por: Bashir, Zahra, et al.
Publicado: (2023)
por: Bashir, Zahra, et al.
Publicado: (2023)
XGrammar: Flexible and Efficient Structured Generation Engine for Large Language Models
por: Dong, Yixin, et al.
Publicado: (2024)
por: Dong, Yixin, et al.
Publicado: (2024)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
por: Singhvi, Arnav, et al.
Publicado: (2023)
por: Singhvi, Arnav, et al.
Publicado: (2023)
SuperCoder: Assembly Program Superoptimization with Large Language Models
por: Wei, Anjiang, et al.
Publicado: (2025)
por: Wei, Anjiang, et al.
Publicado: (2025)
CodeMind: Evaluating Large Language Models for Code Reasoning
por: Liu, Changshu, et al.
Publicado: (2024)
por: Liu, Changshu, et al.
Publicado: (2024)
Ejemplares similares
-
FormalProofBench: Can Models Write Graduate Level Math Proofs That Are Formally Verified?
por: Ravi, Nikil, et al.
Publicado: (2026) -
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
por: Cao, Jialun, et al.
Publicado: (2025) -
A Case Study on the Effectiveness of LLMs in Verification with Proof Assistants
por: Bayazıt, Barış, et al.
Publicado: (2025) -
Solving Inequality Proofs with Large Language Models
por: Lu, Pan, et al.
Publicado: (2025) -
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
por: Thakur, Amitayush, et al.
Publicado: (2025)