Quantifying artificial intelligence through algorithmic generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Ito, Takuya, Campbell, Murray, Horesh, Lior, Klinger, Tim, Ram, Parikshit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Multi-Step Deductive Reasoning Over Natural Language: An Empirical Study on Out-of-Distribution Generalisation
by: Bao, Qiming, et al.
Published: (2022)
by: Bao, Qiming, et al.
Published: (2022)
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
by: Zhao, Xueliang, et al.
Published: (2024)
by: Zhao, Xueliang, et al.
Published: (2024)
RLSF: Fine-tuning LLMs via Symbolic Feedback
by: Jha, Piyush, et al.
Published: (2024)
by: Jha, Piyush, et al.
Published: (2024)
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
by: Lalwani, Abhinav, et al.
Published: (2024)
by: Lalwani, Abhinav, et al.
Published: (2024)
Consistent Joint Decision-Making with Heterogeneous Learning Models
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
Guiding Word Equation Solving using Graph Neural Networks (Extended Technical Report)
by: Abdulla, Parosh Aziz, et al.
Published: (2024)
by: Abdulla, Parosh Aziz, et al.
Published: (2024)
FLARE: Faithful Logic-Aided Reasoning and Exploration
by: Arakelyan, Erik, et al.
Published: (2024)
by: Arakelyan, Erik, et al.
Published: (2024)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Harnessing the Power of Semi-Structured Knowledge and LLMs with Triplet-Based Prefiltering for Question Answering
by: Boer, Derian, et al.
Published: (2024)
by: Boer, Derian, et al.
Published: (2024)
Herald: A Natural Language Annotated Lean 4 Dataset
by: Gao, Guoxiong, et al.
Published: (2024)
by: Gao, Guoxiong, et al.
Published: (2024)
JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models
by: Chen, Michael K., et al.
Published: (2025)
by: Chen, Michael K., et al.
Published: (2025)
SATBench: Benchmarking LLMs' Logical Reasoning via Automated Puzzle Generation from SAT Formulas
by: Wei, Anjiang, et al.
Published: (2025)
by: Wei, Anjiang, et al.
Published: (2025)
Transformers Can Learn Connectivity in Some Graphs but Not Others
by: Roy, Amit, et al.
Published: (2025)
by: Roy, Amit, et al.
Published: (2025)
Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors
by: Bao, Qiming, et al.
Published: (2025)
by: Bao, Qiming, et al.
Published: (2025)
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
by: Rajaee, Sara, et al.
Published: (2025)
by: Rajaee, Sara, et al.
Published: (2025)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
by: Zhang, Terry Jingchen, et al.
Published: (2025)
by: Zhang, Terry Jingchen, et al.
Published: (2025)
Reasoning Inconsistencies and How to Mitigate Them in Deep Learning
by: Arakelyan, Erik
Published: (2025)
by: Arakelyan, Erik
Published: (2025)
Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2
by: Miyashita, Hisashi, et al.
Published: (2026)
by: Miyashita, Hisashi, et al.
Published: (2026)
Mathematics with large language models as provers and verifiers
by: Duc, Hieu Le, et al.
Published: (2025)
by: Duc, Hieu Le, et al.
Published: (2025)
Combining Textual and Structural Information for Premise Selection in Lean
by: Petrovčič, Job, et al.
Published: (2025)
by: Petrovčič, Job, et al.
Published: (2025)
Are Language Models Efficient Reasoners? A Perspective from Logic Programming
by: Opedal, Andreas, et al.
Published: (2025)
by: Opedal, Andreas, et al.
Published: (2025)
Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
by: Su, DiJia, et al.
Published: (2025)
by: Su, DiJia, et al.
Published: (2025)
A Neurosymbolic Approach to Natural Language Formalization and Verification
by: Bayless, Sam, et al.
Published: (2025)
by: Bayless, Sam, et al.
Published: (2025)
Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning
by: Noël, Valentin
Published: (2026)
by: Noël, Valentin
Published: (2026)
The Geometry of Reasoning: Flowing Logics in Representation Space
by: Zhou, Yufa, et al.
Published: (2025)
by: Zhou, Yufa, et al.
Published: (2025)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
by: Shen, Ziju, et al.
Published: (2025)
by: Shen, Ziju, et al.
Published: (2025)
Self-Supervised Transformers as Iterative Solution Improvers for Constraint Satisfaction
by: Xu, Yudong W., et al.
Published: (2025)
by: Xu, Yudong W., et al.
Published: (2025)
ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
by: Ahuja, Riyaz, et al.
Published: (2026)
by: Ahuja, Riyaz, et al.
Published: (2026)
Hierarchical Attention Generates Better Proofs
by: Chen, Jianlong, et al.
Published: (2025)
by: Chen, Jianlong, et al.
Published: (2025)
DeepOnto: A Python Package for Ontology Engineering with Deep Learning
by: He, Yuan, et al.
Published: (2023)
by: He, Yuan, et al.
Published: (2023)
Bolzano: Case Studies in LLM-Assisted Mathematical Research
by: Balko, Martin, et al.
Published: (2026)
by: Balko, Martin, et al.
Published: (2026)
Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
by: Qasim, Kaleem Ullah, et al.
Published: (2025)
Machine Learning for Quantifier Selection in cvc5
by: Jakubův, Jan, et al.
Published: (2024)
by: Jakubův, Jan, et al.
Published: (2024)
Finding Clustering Algorithms in the Transformer Architecture
by: Clarkson, Kenneth L., et al.
Published: (2025)
by: Clarkson, Kenneth L., et al.
Published: (2025)
PutnamBench: Evaluating Neural Theorem-Provers on the Putnam Mathematical Competition
by: Tsoukalas, George, et al.
Published: (2024)
by: Tsoukalas, George, et al.
Published: (2024)
From Symbolic Tasks to Code Generation: Diversification Yields Better Task Performers
by: Zhang, Dylan, et al.
Published: (2024)
by: Zhang, Dylan, et al.
Published: (2024)
NLP Verification: Towards a General Methodology for Certifying Robustness
by: Casadio, Marco, et al.
Published: (2024)
by: Casadio, Marco, et al.
Published: (2024)
Lattice Annotated Temporal (LAT) Logic for Non-Markovian Reasoning
by: Mukherji, Kaustuv, et al.
Published: (2025)
by: Mukherji, Kaustuv, et al.
Published: (2025)
ScenicProver: A Framework for Compositional Probabilistic Verification of Learning-Enabled Systems
by: Vin, Eric, et al.
Published: (2025)
by: Vin, Eric, et al.
Published: (2025)
Similar Items
-
EXPLORER: Exploration-guided Reasoning for Textual Reinforcement Learning
by: Basu, Kinjal, et al.
Published: (2024) -
Multi-Step Deductive Reasoning Over Natural Language: An Empirical Study on Out-of-Distribution Generalisation
by: Bao, Qiming, et al.
Published: (2022) -
SubgoalXL: Subgoal-based Expert Learning for Theorem Proving
by: Zhao, Xueliang, et al.
Published: (2024) -
RLSF: Fine-tuning LLMs via Symbolic Feedback
by: Jha, Piyush, et al.
Published: (2024) -
Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
by: Lalwani, Abhinav, et al.
Published: (2024)