miniCTX: Neural Theorem Proving with (Long-)Contexts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Jiewen, Zhu, Thomas, Welleck, Sean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
miniCodeProps: a Minimal Benchmark for Proving Code Properties
von: Lohn, Evan, et al.
Veröffentlicht: (2024)
von: Lohn, Evan, et al.
Veröffentlicht: (2024)
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
An In-Context Learning Agent for Formal Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2023)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2023)
Steering LLMs for Formal Theorem Proving
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
Optimizing Temperature for Language Models with Multi-Sample Inference
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
Learning to Reason with Insight for Informal Theorem Proving
von: Li, Yunhe, et al.
Veröffentlicht: (2026)
von: Li, Yunhe, et al.
Veröffentlicht: (2026)
Agentic-R1: Distilled Dual-Strategy Reasoning
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
von: Petrov, Ivo, et al.
Veröffentlicht: (2025)
von: Petrov, Ivo, et al.
Veröffentlicht: (2025)
QED-Nano: Teaching a Tiny Model to Prove Hard Theorems
von: LM-Provers, et al.
Veröffentlicht: (2026)
von: LM-Provers, et al.
Veröffentlicht: (2026)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2024)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
Ineq-Comp: Benchmarking Human-Intuitive Compositional Reasoning in Automated Theorem Proving on Inequalities
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
von: Ahuja, Riyaz, et al.
Veröffentlicht: (2026)
von: Ahuja, Riyaz, et al.
Veröffentlicht: (2026)
Neural Theorem Proving: Generating and Structuring Proofs for Formal Verification
von: Rao, Balaji, et al.
Veröffentlicht: (2025)
von: Rao, Balaji, et al.
Veröffentlicht: (2025)
Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
von: Rajaee, Sara, et al.
Veröffentlicht: (2025)
Training Proactive and Personalized LLM Agents
von: Sun, Weiwei, et al.
Veröffentlicht: (2025)
von: Sun, Weiwei, et al.
Veröffentlicht: (2025)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
Proving that Cryptic Crossword Clue Answers are Correct
von: Andrews, Martin, et al.
Veröffentlicht: (2024)
von: Andrews, Martin, et al.
Veröffentlicht: (2024)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
BAIT: Benchmarking (Embedding) Architectures for Interactive Theorem-Proving
von: Lamont, Sean, et al.
Veröffentlicht: (2024)
von: Lamont, Sean, et al.
Veröffentlicht: (2024)
100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?
von: Yang, Wang, et al.
Veröffentlicht: (2025)
von: Yang, Wang, et al.
Veröffentlicht: (2025)
Revisiting In-Context Learning with Long Context Language Models
von: Baek, Jinheon, et al.
Veröffentlicht: (2024)
von: Baek, Jinheon, et al.
Veröffentlicht: (2024)
LongSafety: Enhance Safety for Long-Context LLMs
von: Huang, Mianqiu, et al.
Veröffentlicht: (2024)
von: Huang, Mianqiu, et al.
Veröffentlicht: (2024)
MOM: Memory-Efficient Offloaded Mini-Sequence Inference for Long Context Language Models
von: Zhang, Junyang, et al.
Veröffentlicht: (2025)
von: Zhang, Junyang, et al.
Veröffentlicht: (2025)
The Impossibility Triangle of Long-Context Modeling
von: Zhou, Yan
Veröffentlicht: (2026)
von: Zhou, Yan
Veröffentlicht: (2026)
PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents
von: Gu, Zhuohan, et al.
Veröffentlicht: (2026)
von: Gu, Zhuohan, et al.
Veröffentlicht: (2026)
Latent Context Compilation: Distilling Long Context into Compact Portable Memory
von: Li, Zeju, et al.
Veröffentlicht: (2026)
von: Li, Zeju, et al.
Veröffentlicht: (2026)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
von: Fu, Qichen, et al.
Veröffentlicht: (2024)
von: Fu, Qichen, et al.
Veröffentlicht: (2024)
Every Attention Matters: An Efficient Hybrid Architecture for Long-Context Reasoning
von: Ling Team, et al.
Veröffentlicht: (2025)
von: Ling Team, et al.
Veröffentlicht: (2025)
LycheeCluster: Efficient Long-Context Inference with Structure-Aware Chunking and Hierarchical KV Indexing
von: Li, Dongfang, et al.
Veröffentlicht: (2026)
von: Li, Dongfang, et al.
Veröffentlicht: (2026)
Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning
von: Yang, Wang, et al.
Veröffentlicht: (2025)
von: Yang, Wang, et al.
Veröffentlicht: (2025)
LLoCO: Learning Long Contexts Offline
von: Tan, Sijun, et al.
Veröffentlicht: (2024)
von: Tan, Sijun, et al.
Veröffentlicht: (2024)
Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models
von: Qiao, Boyu, et al.
Veröffentlicht: (2026)
von: Qiao, Boyu, et al.
Veröffentlicht: (2026)
Tactic: Adaptive Sparse Attention with Clustering and Distribution Fitting for Long-Context LLMs
von: Zhu, Kan, et al.
Veröffentlicht: (2025)
von: Zhu, Kan, et al.
Veröffentlicht: (2025)
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
von: Jin, Bowen, et al.
Veröffentlicht: (2024)
LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
von: Jin, Hongye, et al.
Veröffentlicht: (2024)
von: Jin, Hongye, et al.
Veröffentlicht: (2024)
Too Long, Didn't Model: Decomposing LLM Long-Context Understanding With Novels
von: Hamilton, Sil, et al.
Veröffentlicht: (2025)
von: Hamilton, Sil, et al.
Veröffentlicht: (2025)
LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification
von: Yang, Penghui, et al.
Veröffentlicht: (2025)
von: Yang, Penghui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
miniCodeProps: a Minimal Benchmark for Proving Code Properties
von: Lohn, Evan, et al.
Veröffentlicht: (2024) -
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025) -
An In-Context Learning Agent for Formal Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2023) -
Steering LLMs for Formal Theorem Proving
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025) -
Optimizing Temperature for Language Models with Multi-Sample Inference
von: Du, Weihua, et al.
Veröffentlicht: (2025)