The Proof is in the Almond Cookies
Fuente:
arXiv
Guardado en:
| Autores principales: | van Trijp, Remi, Beuls, Katrien, Van Eecke, Paul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PyFCG: Fluid Construction Grammar in Python
por: Van Eecke, Paul, et al.
Publicado: (2025)
por: Van Eecke, Paul, et al.
Publicado: (2025)
The Computational Learning of Construction Grammars: State of the Art and Prospective Roadmap
por: Doumen, Jonas, et al.
Publicado: (2024)
por: Doumen, Jonas, et al.
Publicado: (2024)
A Method for Learning Large-Scale Computational Construction Grammars from Semantically Annotated Corpora
por: Van Eecke, Paul, et al.
Publicado: (2026)
por: Van Eecke, Paul, et al.
Publicado: (2026)
Decentralised Emergence of Robust and Adaptive Linguistic Conventions in Populations of Autonomous Agents Grounded in Continuous Worlds
por: Ekila, Jérôme Botoko, et al.
Publicado: (2024)
por: Ekila, Jérôme Botoko, et al.
Publicado: (2024)
The evolution of case grammar
por: van Trijp, Remi
Publicado: (2016)
por: van Trijp, Remi
Publicado: (2016)
A Certified Proof Checker for Deep Neural Network Verification in Imandra
por: Desmartin, Remi, et al.
Publicado: (2024)
por: Desmartin, Remi, et al.
Publicado: (2024)
Proof2Hybrid: Automatic Mathematical Benchmark Synthesis for Proof-Centric Problems
por: Peng, Yebo, et al.
Publicado: (2025)
por: Peng, Yebo, et al.
Publicado: (2025)
The Open Proof Corpus: A Large-Scale Study of LLM-Generated Mathematical Proofs
por: Dekoninck, Jasper, et al.
Publicado: (2025)
por: Dekoninck, Jasper, et al.
Publicado: (2025)
Reliable Fine-Grained Evaluation of Natural Language Math Proofs
por: Ma, Wenjie, et al.
Publicado: (2025)
por: Ma, Wenjie, et al.
Publicado: (2025)
Proof of Time: A Benchmark for Evaluating Scientific Idea Judgments
por: Ye, Bingyang, et al.
Publicado: (2026)
por: Ye, Bingyang, et al.
Publicado: (2026)
Neuro-Symbolic Integration Brings Causal and Reliable Reasoning Proofs
por: Yang, Sen, et al.
Publicado: (2023)
por: Yang, Sen, et al.
Publicado: (2023)
FormalProofBench: Can Models Write Graduate Level Math Proofs That Are Formally Verified?
por: Ravi, Nikil, et al.
Publicado: (2026)
por: Ravi, Nikil, et al.
Publicado: (2026)
Prove Your Point!: Bringing Proof-Enhancement Principles to Argumentative Essay Generation
por: Xiao, Ruiyu, et al.
Publicado: (2024)
por: Xiao, Ruiyu, et al.
Publicado: (2024)
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
por: Cao, Jialun, et al.
Publicado: (2025)
por: Cao, Jialun, et al.
Publicado: (2025)
Positional Encoding via Token-Aware Phase Attention
por: Wang, Yu, et al.
Publicado: (2025)
por: Wang, Yu, et al.
Publicado: (2025)
Next-Token Prediction Task Assumes Optimal Data Ordering for LLM Training in Proof Generation
por: An, Chenyang, et al.
Publicado: (2024)
por: An, Chenyang, et al.
Publicado: (2024)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
por: Nguyen, Van Bach, et al.
Publicado: (2024)
por: Nguyen, Van Bach, et al.
Publicado: (2024)
Aletheia tackles FirstProof autonomously
por: Feng, Tony, et al.
Publicado: (2026)
por: Feng, Tony, et al.
Publicado: (2026)
Exploring the Role of Reasoning Structures for Constructing Proofs in Multi-Step Natural Language Reasoning with Large Language Models
por: Zheng, Zi'ou, et al.
Publicado: (2024)
por: Zheng, Zi'ou, et al.
Publicado: (2024)
Solving Inequality Proofs with Large Language Models
por: Lu, Pan, et al.
Publicado: (2025)
por: Lu, Pan, et al.
Publicado: (2025)
Autograding Mathematical Induction Proofs with Natural Language Processing
por: Zhao, Chenyan, et al.
Publicado: (2024)
por: Zhao, Chenyan, et al.
Publicado: (2024)
Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study
por: Das, Sudeshna, et al.
Publicado: (2024)
por: Das, Sudeshna, et al.
Publicado: (2024)
Proof-of-Guardrail in AI Agents and What (Not) to Trust from It
por: Jin, Xisen, et al.
Publicado: (2026)
por: Jin, Xisen, et al.
Publicado: (2026)
Creativity in AI: Progresses and Challenges
por: Ismayilzada, Mete, et al.
Publicado: (2024)
por: Ismayilzada, Mete, et al.
Publicado: (2024)
Emergence of Phonemic, Syntactic, and Semantic Representations in Artificial Neural Networks
por: Orhan, Pierre, et al.
Publicado: (2026)
por: Orhan, Pierre, et al.
Publicado: (2026)
ProofSketch: Efficient Verified Reasoning for Large Language Models
por: Sheshanarayana, Disha, et al.
Publicado: (2025)
por: Sheshanarayana, Disha, et al.
Publicado: (2025)
A Proof of Learning Rate Transfer under $μ$P
por: Hayou, Soufiane
Publicado: (2025)
por: Hayou, Soufiane
Publicado: (2025)
Automatic Image Annotation (AIA) of AlmondNet-20 Method for Almond Detection by Improved CNN-based Model
por: Ilani, Mohsen Asghari, et al.
Publicado: (2024)
por: Ilani, Mohsen Asghari, et al.
Publicado: (2024)
Do We Need Frontier Models to Verify Mathematical Proofs?
por: Naik, Aaditya, et al.
Publicado: (2026)
por: Naik, Aaditya, et al.
Publicado: (2026)
New Benchmark Dataset and Fine-Grained Cross-Modal Fusion Framework for Vietnamese Multimodal Aspect-Category Sentiment Analysis
por: Nguyen, Quy Hoang, et al.
Publicado: (2024)
por: Nguyen, Quy Hoang, et al.
Publicado: (2024)
MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
por: Opedal, Andreas, et al.
Publicado: (2024)
por: Opedal, Andreas, et al.
Publicado: (2024)
Compactor: Calibrated Query-Agnostic KV Cache Compression with Approximate Leverage Scores
por: Chari, Vivek, et al.
Publicado: (2025)
por: Chari, Vivek, et al.
Publicado: (2025)
Refining Transcripts With TV Subtitles by Prompt-Based Weakly Supervised Training of ASR
por: Zhao, Xinnian, et al.
Publicado: (2025)
por: Zhao, Xinnian, et al.
Publicado: (2025)
Hierarchical Attention Generates Better Proofs
por: Chen, Jianlong, et al.
Publicado: (2025)
por: Chen, Jianlong, et al.
Publicado: (2025)
From Eigenmodes to Proofs: Integrating Graph Spectral Operators with Symbolic Interpretable Reasoning
por: Kiruluta, Andrew, et al.
Publicado: (2025)
por: Kiruluta, Andrew, et al.
Publicado: (2025)
BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards
por: Dorn, Diego, et al.
Publicado: (2024)
por: Dorn, Diego, et al.
Publicado: (2024)
My Life in Artificial Intelligence: People, anecdotes, and some lessons learnt
por: van Deemter, Kees
Publicado: (2025)
por: van Deemter, Kees
Publicado: (2025)
LiveMathematicianBench: A Live Benchmark for Mathematician-Level Reasoning with Proof Sketches
por: He, Linyang, et al.
Publicado: (2026)
por: He, Linyang, et al.
Publicado: (2026)
Can we Evaluate RAGs with Synthetic Data?
por: van Elburg, Jonas, et al.
Publicado: (2025)
por: van Elburg, Jonas, et al.
Publicado: (2025)
LITERA: An LLM Based Approach to Latin-to-English Translation
por: Rosu, Paul
Publicado: (2025)
por: Rosu, Paul
Publicado: (2025)
Ejemplares similares
-
PyFCG: Fluid Construction Grammar in Python
por: Van Eecke, Paul, et al.
Publicado: (2025) -
The Computational Learning of Construction Grammars: State of the Art and Prospective Roadmap
por: Doumen, Jonas, et al.
Publicado: (2024) -
A Method for Learning Large-Scale Computational Construction Grammars from Semantically Annotated Corpora
por: Van Eecke, Paul, et al.
Publicado: (2026) -
Decentralised Emergence of Robust and Adaptive Linguistic Conventions in Populations of Autonomous Agents Grounded in Continuous Worlds
por: Ekila, Jérôme Botoko, et al.
Publicado: (2024) -
The evolution of case grammar
por: van Trijp, Remi
Publicado: (2016)