From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Dylan, Wang, Justin, Charton, Francois |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Instruction Diversity Drives Generalization To Unseen Tasks
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
NLP Verification: Towards a General Methodology for Certifying Robustness
by: Casadio, Marco, et al.
Published: (2024)
by: Casadio, Marco, et al.
Published: (2024)
Hierarchical Attention Generates Better Proofs
by: Chen, Jianlong, et al.
Published: (2025)
by: Chen, Jianlong, et al.
Published: (2025)
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
by: Luo, Ziyan, et al.
Published: (2023)
by: Luo, Ziyan, et al.
Published: (2023)
VERINA: Benchmarking Verifiable Code Generation
by: Ye, Zhe, et al.
Published: (2025)
by: Ye, Zhe, et al.
Published: (2025)
PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition
by: Tsoukalas, George, et al.
Published: (2024)
by: Tsoukalas, George, et al.
Published: (2024)
$\textbf{Only-IF}$:Revealing the Decisive Effect of Instruction Diversity on Generalization
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
by: Thakur, Amitayush, et al.
Published: (2025)
by: Thakur, Amitayush, et al.
Published: (2025)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
by: Li, Guchan, et al.
Published: (2026)
by: Li, Guchan, et al.
Published: (2026)
RLSF: Fine-tuning LLMs via Symbolic Feedback
by: Jha, Piyush, et al.
Published: (2024)
by: Jha, Piyush, et al.
Published: (2024)
Lattice Annotated Temporal (LAT) Logic for Non-Markovian Reasoning
by: Mukherji, Kaustuv, et al.
Published: (2025)
by: Mukherji, Kaustuv, et al.
Published: (2025)
ScenicProver: A Framework for Compositional Probabilistic Verification of Learning-Enabled Systems
by: Vin, Eric, et al.
Published: (2025)
by: Vin, Eric, et al.
Published: (2025)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
by: Thakur, Amitayush, et al.
Published: (2025)
by: Thakur, Amitayush, et al.
Published: (2025)
An In-Context Learning Agent for Formal Theorem-Proving
by: Thakur, Amitayush, et al.
Published: (2023)
by: Thakur, Amitayush, et al.
Published: (2023)
Decidable By Construction: Design-Time Verification for Trustworthy AI
by: Haynes, Houston
Published: (2026)
by: Haynes, Houston
Published: (2026)
VeriThoughts: Enabling Automated Verilog Code Generation using Reasoning and Formal Verification
by: Yubeaton, Patrick, et al.
Published: (2025)
by: Yubeaton, Patrick, et al.
Published: (2025)
Dafny as Verification-Aware Intermediate Language for Code Generation
by: Li, Yue Chen, et al.
Published: (2025)
by: Li, Yue Chen, et al.
Published: (2025)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Transformer-Based Models Are Not Yet Perfect At Learning to Emulate Structural Recursion
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
The Geometry of Reasoning: Flowing Logics in Representation Space
by: Zhou, Yufa, et al.
Published: (2025)
by: Zhou, Yufa, et al.
Published: (2025)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
by: Zhang, Terry Jingchen, et al.
Published: (2025)
by: Zhang, Terry Jingchen, et al.
Published: (2025)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
by: Chen, Michael K., et al.
Published: (2025)
by: Chen, Michael K., et al.
Published: (2025)
VerMCTS: Synthesizing Multi-Step Programs using a Verifier, a Large Language Model, and Tree Search
by: Brandfonbrener, David, et al.
Published: (2024)
by: Brandfonbrener, David, et al.
Published: (2024)
Towards Regulated Deep Learning
by: García-Camino, Andrés
Published: (2019)
by: García-Camino, Andrés
Published: (2019)
Intent-aligned Formal Specification Synthesis via Traceable Refinement
by: Ye, Zhe, et al.
Published: (2026)
by: Ye, Zhe, et al.
Published: (2026)
Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Herald: A Natural Language Annotated Lean 4 Dataset
by: Gao, Guoxiong, et al.
Published: (2024)
by: Gao, Guoxiong, et al.
Published: (2024)
A Neurosymbolic Approach to Loop Invariant Generation via Weakest Precondition Reasoning
by: King, Daragh, et al.
Published: (2025)
by: King, Daragh, et al.
Published: (2025)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
by: Shen, Ziju, et al.
Published: (2025)
by: Shen, Ziju, et al.
Published: (2025)
Bridging the Knowledge Void: Inference-time Acquisition of Unfamiliar Programming Languages for Coding Tasks
by: Shen, Chen, et al.
Published: (2026)
by: Shen, Chen, et al.
Published: (2026)
MiniF2F in Rocq: Automatic Translation Between Proof Assistants -- A Case Study
by: Viennot, Jules, et al.
Published: (2025)
by: Viennot, Jules, et al.
Published: (2025)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
by: Zhao, Xueliang, et al.
Published: (2024)
by: Zhao, Xueliang, et al.
Published: (2024)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
by: Lalwani, Abhinav, et al.
Published: (2024)
by: Lalwani, Abhinav, et al.
Published: (2024)
Consistent Joint Decision-Making with Heterogeneous Learning Models
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
Guiding Word Equation Solving using Graph Neural Networks (Extended Technical Report)
by: Abdulla, Parosh Aziz, et al.
Published: (2024)
by: Abdulla, Parosh Aziz, et al.
Published: (2024)
FLARE: Faithful Logic-Aided Reasoning and Exploration
by: Arakelyan, Erik, et al.
Published: (2024)
by: Arakelyan, Erik, et al.
Published: (2024)
Quantifying artificial intelligence through algorithmic generalization
by: Ito, Takuya, et al.
Published: (2024)
by: Ito, Takuya, et al.
Published: (2024)
Harnessing the Power of Semi-Structured Knowledge and LLMs with Triplet-Based Prefiltering for Question Answering
by: Boer, Derian, et al.
Published: (2024)
by: Boer, Derian, et al.
Published: (2024)
Transformers Can Learn Connectivity in Some Graphs but Not Others
by: Roy, Amit, et al.
Published: (2025)
by: Roy, Amit, et al.
Published: (2025)
Similar Items
-
Instruction Diversity Drives Generalization To Unseen Tasks
by: Zhang, Dylan, et al.
Published: (2024) -
NLP Verification: Towards a General Methodology for Certifying Robustness
by: Casadio, Marco, et al.
Published: (2024) -
Hierarchical Attention Generates Better Proofs
by: Chen, Jianlong, et al.
Published: (2025) -
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
by: Luo, Ziyan, et al.
Published: (2023) -
VERINA: Benchmarking Verifiable Code Generation
by: Ye, Zhe, et al.
Published: (2025)