Learning to Reason via Program Generation, Emulation, and Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Weir, Nathaniel, Khalifa, Muhammad, Qiu, Linlu, Weller, Orion, Clark, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
von: Jiang, Dongwei, et al.
Veröffentlicht: (2024)
von: Jiang, Dongwei, et al.
Veröffentlicht: (2024)
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
von: Weller, Orion, et al.
Veröffentlicht: (2023)
von: Weller, Orion, et al.
Veröffentlicht: (2023)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
von: Weir, Nathaniel, et al.
Veröffentlicht: (2024)
von: Weir, Nathaniel, et al.
Veröffentlicht: (2024)
Generating Data-Driven Reasoning Rubrics for Domain-Adaptive Reward Modeling
von: Sanders, Kate, et al.
Veröffentlicht: (2026)
von: Sanders, Kate, et al.
Veröffentlicht: (2026)
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning
von: Singh, Vikash, et al.
Veröffentlicht: (2026)
von: Singh, Vikash, et al.
Veröffentlicht: (2026)
Reframing Tax Law Entailment as Analogical Reasoning
von: Zou, Xinrui, et al.
Veröffentlicht: (2024)
von: Zou, Xinrui, et al.
Veröffentlicht: (2024)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning
von: Corallo, Giulio, et al.
Veröffentlicht: (2025)
von: Corallo, Giulio, et al.
Veröffentlicht: (2025)
GRACE: Discriminator-Guided Chain-of-Thought Reasoning
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2023)
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2023)
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2023)
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2023)
Bias Amplification in Language Model Evolution: An Iterated Learning Perspective
von: Ren, Yi, et al.
Veröffentlicht: (2024)
von: Ren, Yi, et al.
Veröffentlicht: (2024)
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
von: Sanders, Kate, et al.
Veröffentlicht: (2024)
von: Sanders, Kate, et al.
Veröffentlicht: (2024)
NELLIE: A Neuro-Symbolic Inference Engine for Grounded, Compositional, and Explainable Reasoning
von: Weir, Nathaniel, et al.
Veröffentlicht: (2022)
von: Weir, Nathaniel, et al.
Veröffentlicht: (2022)
The Qiyas Benchmark: Measuring ChatGPT Mathematical and Language Understanding in Arabic
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2024)
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2024)
Reason2Decide: Rationale-Driven Multi-Task Learning
von: Hasan, H M Quamran, et al.
Veröffentlicht: (2025)
von: Hasan, H M Quamran, et al.
Veröffentlicht: (2025)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
von: Qiu, Linlu, et al.
Veröffentlicht: (2023)
von: Qiu, Linlu, et al.
Veröffentlicht: (2023)
VeriCoT: Neuro-symbolic Chain-of-Thought Validation via Logical Consistency Checks
von: Feng, Yu, et al.
Veröffentlicht: (2025)
von: Feng, Yu, et al.
Veröffentlicht: (2025)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
von: Clark, Peter, et al.
Veröffentlicht: (2023)
von: Clark, Peter, et al.
Veröffentlicht: (2023)
A Survey of Large Language Models for Arabic Language and its Dialects
von: Mashaabi, Malak, et al.
Veröffentlicht: (2024)
von: Mashaabi, Malak, et al.
Veröffentlicht: (2024)
Logit Arithmetic Elicits Long Reasoning Capabilities Without Training
von: Zhang, Yunxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Yunxiang, et al.
Veröffentlicht: (2025)
Defending Against Disinformation Attacks in Open-Domain Question Answering
von: Weller, Orion, et al.
Veröffentlicht: (2022)
von: Weller, Orion, et al.
Veröffentlicht: (2022)
When do Generative Query and Document Expansions Fail? A Comprehensive Study Across Methods, Retrievers, and Datasets
von: Weller, Orion, et al.
Veröffentlicht: (2023)
von: Weller, Orion, et al.
Veröffentlicht: (2023)
Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2026)
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2026)
The Surprising Effectiveness of Test-Time Training for Few-Shot Learning
von: Akyürek, Ekin, et al.
Veröffentlicht: (2024)
von: Akyürek, Ekin, et al.
Veröffentlicht: (2024)
Search-R2: Enhancing Search-Integrated Reasoning via Actor-Refiner Collaboration
von: He, Bowei, et al.
Veröffentlicht: (2026)
von: He, Bowei, et al.
Veröffentlicht: (2026)
DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search
von: Yue, Murong, et al.
Veröffentlicht: (2024)
von: Yue, Murong, et al.
Veröffentlicht: (2024)
Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
von: Shen, Maohao, et al.
Veröffentlicht: (2025)
von: Shen, Maohao, et al.
Veröffentlicht: (2025)
Accelerating Large Language Model Reasoning via Speculative Search
von: Wang, Zhihai, et al.
Veröffentlicht: (2025)
von: Wang, Zhihai, et al.
Veröffentlicht: (2025)
Generating Literature-Driven Scientific Theories at Scale
von: Jansen, Peter, et al.
Veröffentlicht: (2026)
von: Jansen, Peter, et al.
Veröffentlicht: (2026)
MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning
von: Yuan, Qianhao, et al.
Veröffentlicht: (2025)
von: Yuan, Qianhao, et al.
Veröffentlicht: (2025)
Emulating Clinician Cognition via Self-Evolving Deep Clinical Research
von: Ren, Ruiyang, et al.
Veröffentlicht: (2026)
von: Ren, Ruiyang, et al.
Veröffentlicht: (2026)
CoSearch: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Search
von: Zeng, Hansi, et al.
Veröffentlicht: (2026)
von: Zeng, Hansi, et al.
Veröffentlicht: (2026)
ARise: Towards Knowledge-Augmented Reasoning via Risk-Adaptive Search
von: Zhang, Yize, et al.
Veröffentlicht: (2025)
von: Zhang, Yize, et al.
Veröffentlicht: (2025)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
From Models to Microtheories: Distilling a Model's Topical Knowledge for Grounded Question Answering
von: Weir, Nathaniel, et al.
Veröffentlicht: (2024)
von: Weir, Nathaniel, et al.
Veröffentlicht: (2024)
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting
von: Getachew, Nathaniel, et al.
Veröffentlicht: (2025)
von: Getachew, Nathaniel, et al.
Veröffentlicht: (2025)
LLMs versus the Halting Problem: Characterizing Program Termination Reasoning
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
von: Sultan, Oren, et al.
Veröffentlicht: (2026)
Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search Trajectories
von: Liu, Peiyang, et al.
Veröffentlicht: (2026)
von: Liu, Peiyang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
von: Jiang, Dongwei, et al.
Veröffentlicht: (2024) -
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
von: Weller, Orion, et al.
Veröffentlicht: (2023) -
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
von: Weir, Nathaniel, et al.
Veröffentlicht: (2024) -
Generating Data-Driven Reasoning Rubrics for Domain-Adaptive Reward Modeling
von: Sanders, Kate, et al.
Veröffentlicht: (2026) -
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning
von: Singh, Vikash, et al.
Veröffentlicht: (2026)