From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Dylan, Wang, Justin, Charton, Francois |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Instruction Diversity Drives Generalization To Unseen Tasks
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
NLP Verification: Towards a General Methodology for Certifying Robustness
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
Hierarchical Attention Generates Better Proofs
von: Chen, Jianlong, et al.
Veröffentlicht: (2025)
von: Chen, Jianlong, et al.
Veröffentlicht: (2025)
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
von: Luo, Ziyan, et al.
Veröffentlicht: (2023)
von: Luo, Ziyan, et al.
Veröffentlicht: (2023)
VERINA: Benchmarking Verifiable Code Generation
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
von: Ye, Zhe, et al.
Veröffentlicht: (2025)
PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition
von: Tsoukalas, George, et al.
Veröffentlicht: (2024)
von: Tsoukalas, George, et al.
Veröffentlicht: (2024)
$\textbf{Only-IF}$:Revealing the Decisive Effect of Instruction Diversity on Generalization
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
CLEVER: A Curated Benchmark for Formally Verified Code Generation
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
von: Li, Guchan, et al.
Veröffentlicht: (2026)
von: Li, Guchan, et al.
Veröffentlicht: (2026)
RLSF: Fine-tuning LLMs via Symbolic Feedback
von: Jha, Piyush, et al.
Veröffentlicht: (2024)
von: Jha, Piyush, et al.
Veröffentlicht: (2024)
Lattice Annotated Temporal (LAT) Logic for Non-Markovian Reasoning
von: Mukherji, Kaustuv, et al.
Veröffentlicht: (2025)
von: Mukherji, Kaustuv, et al.
Veröffentlicht: (2025)
ScenicProver: A Framework for Compositional Probabilistic Verification of Learning-Enabled Systems
von: Vin, Eric, et al.
Veröffentlicht: (2025)
von: Vin, Eric, et al.
Veröffentlicht: (2025)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
An In-Context Learning Agent for Formal Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2023)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2023)
Decidable By Construction: Design-Time Verification for Trustworthy AI
von: Haynes, Houston
Veröffentlicht: (2026)
von: Haynes, Houston
Veröffentlicht: (2026)
VeriThoughts: Enabling Automated Verilog Code Generation using Reasoning and Formal Verification
von: Yubeaton, Patrick, et al.
Veröffentlicht: (2025)
von: Yubeaton, Patrick, et al.
Veröffentlicht: (2025)
Dafny as Verification-Aware Intermediate Language for Code Generation
von: Li, Yue Chen, et al.
Veröffentlicht: (2025)
von: Li, Yue Chen, et al.
Veröffentlicht: (2025)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Transformer-Based Models Are Not Yet Perfect At Learning to Emulate Structural Recursion
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
The Geometry of Reasoning: Flowing Logics in Representation Space
von: Zhou, Yufa, et al.
Veröffentlicht: (2025)
von: Zhou, Yufa, et al.
Veröffentlicht: (2025)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
von: Chen, Michael K., et al.
Veröffentlicht: (2025)
VerMCTS: Synthesizing Multi-Step Programs using a Verifier, a Large Language Model, and Tree Search
von: Brandfonbrener, David, et al.
Veröffentlicht: (2024)
von: Brandfonbrener, David, et al.
Veröffentlicht: (2024)
Towards Regulated Deep Learning
von: García-Camino, Andrés
Veröffentlicht: (2019)
von: García-Camino, Andrés
Veröffentlicht: (2019)
Intent-aligned Formal Specification Synthesis via Traceable Refinement
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
Herald: A Natural Language Annotated Lean 4 Dataset
von: Gao, Guoxiong, et al.
Veröffentlicht: (2024)
von: Gao, Guoxiong, et al.
Veröffentlicht: (2024)
A Neurosymbolic Approach to Loop Invariant Generation via Weakest Precondition Reasoning
von: King, Daragh, et al.
Veröffentlicht: (2025)
von: King, Daragh, et al.
Veröffentlicht: (2025)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
von: Shen, Ziju, et al.
Veröffentlicht: (2025)
von: Shen, Ziju, et al.
Veröffentlicht: (2025)
Bridging the Knowledge Void: Inference-time Acquisition of Unfamiliar Programming Languages for Coding Tasks
von: Shen, Chen, et al.
Veröffentlicht: (2026)
von: Shen, Chen, et al.
Veröffentlicht: (2026)
MiniF2F in Rocq: Automatic Translation Between Proof Assistants -- A Case Study
von: Viennot, Jules, et al.
Veröffentlicht: (2025)
von: Viennot, Jules, et al.
Veröffentlicht: (2025)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
von: Lalwani, Abhinav, et al.
Veröffentlicht: (2024)
Consistent Joint Decision-Making with Heterogeneous Learning Models
von: Faghihi, Hossein Rajaby, et al.
Veröffentlicht: (2024)
von: Faghihi, Hossein Rajaby, et al.
Veröffentlicht: (2024)
Guiding Word Equation Solving using Graph Neural Networks (Extended Technical Report)
von: Abdulla, Parosh Aziz, et al.
Veröffentlicht: (2024)
von: Abdulla, Parosh Aziz, et al.
Veröffentlicht: (2024)
FLARE: Faithful Logic-Aided Reasoning and Exploration
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
Quantifying artificial intelligence through algorithmic generalization
von: Ito, Takuya, et al.
Veröffentlicht: (2024)
von: Ito, Takuya, et al.
Veröffentlicht: (2024)
Harnessing the Power of Semi-Structured Knowledge and LLMs with Triplet-Based Prefiltering for Question Answering
von: Boer, Derian, et al.
Veröffentlicht: (2024)
von: Boer, Derian, et al.
Veröffentlicht: (2024)
Transformers Can Learn Connectivity in Some Graphs but Not Others
von: Roy, Amit, et al.
Veröffentlicht: (2025)
von: Roy, Amit, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Instruction Diversity Drives Generalization To Unseen Tasks
von: Zhang, Dylan, et al.
Veröffentlicht: (2024) -
NLP Verification: Towards a General Methodology for Certifying Robustness
von: Casadio, Marco, et al.
Veröffentlicht: (2024) -
Hierarchical Attention Generates Better Proofs
von: Chen, Jianlong, et al.
Veröffentlicht: (2025) -
Chronosymbolic Learning: Efficient CHC Solving with Symbolic Reasoning and Inductive Learning
von: Luo, Ziyan, et al.
Veröffentlicht: (2023) -
VERINA: Benchmarking Verifiable Code Generation
von: Ye, Zhe, et al.
Veröffentlicht: (2025)