Language models can learn implicit multi-hop reasoning, but only if they have lots of training data
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Yuekun, Du, Yupei, Zhu, Dawei, Hahn, Michael, Koller, Alexander |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simple and effective data augmentation for compositional generalization
by: Yao, Yuekun, et al.
Published: (2024)
by: Yao, Yuekun, et al.
Published: (2024)
Predicting generalization performance with correctness discriminators
by: Yao, Yuekun, et al.
Published: (2023)
by: Yao, Yuekun, et al.
Published: (2023)
Barriers to Universal Reasoning With Transformers (And How to Overcome Them)
by: Kraus, Oliver, et al.
Published: (2026)
by: Kraus, Oliver, et al.
Published: (2026)
Distributional reasoning in LLMs: Parallel reasoning processes in multi-hop reasoning
by: Shalev, Yuval, et al.
Published: (2024)
by: Shalev, Yuval, et al.
Published: (2024)
On the Ability of Transformers to Verify Plans
by: Sarrof, Yash, et al.
Published: (2026)
by: Sarrof, Yash, et al.
Published: (2026)
Reason to Rote: Rethinking Memorization in Reasoning
by: Du, Yupei, et al.
Published: (2025)
by: Du, Yupei, et al.
Published: (2025)
Learning without training: The implicit dynamics of in-context learning
by: Dherin, Benoit, et al.
Published: (2025)
by: Dherin, Benoit, et al.
Published: (2025)
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
by: Lindemann, Matthias, et al.
Published: (2024)
by: Lindemann, Matthias, et al.
Published: (2024)
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
Can language models learn analogical reasoning? Investigating training objectives and comparisons to human performance
by: Petersen, Molly R., et al.
Published: (2023)
by: Petersen, Molly R., et al.
Published: (2023)
Investigating the translation capabilities of Large Language Models trained on parallel data only
by: Gilabert, Javier García, et al.
Published: (2024)
by: Gilabert, Javier García, et al.
Published: (2024)
Anything Goes? A Crosslinguistic Study of (Im)possible Language Learning in LMs
by: Yang, Xiulin, et al.
Published: (2025)
by: Yang, Xiulin, et al.
Published: (2025)
GenDec: A robust generative Question-decomposition method for Multi-hop reasoning
by: Wu, Jian, et al.
Published: (2024)
by: Wu, Jian, et al.
Published: (2024)
Code-enabled language models can outperform reasoning models on diverse tasks
by: Zhang, Cedegao E., et al.
Published: (2025)
by: Zhang, Cedegao E., et al.
Published: (2025)
Fine-grained Controllable Text Generation through In-context Learning with Feedback
by: Thillainathan, Sarubi, et al.
Published: (2024)
by: Thillainathan, Sarubi, et al.
Published: (2024)
A Survey on Complex Tasks for Goal-Directed Interactive Agents
by: Hartmann, Mareike, et al.
Published: (2024)
by: Hartmann, Mareike, et al.
Published: (2024)
Markovian ODE-guided scoring can assess the quality of offline reasoning traces in language models
by: Nandi, Arghodeep, et al.
Published: (2026)
by: Nandi, Arghodeep, et al.
Published: (2026)
On Leveraging Encoder-only Pre-trained Language Models for Effective Keyphrase Generation
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Do we really have to filter out random noise in pre-training data for language models?
by: Ru, Jinghan, et al.
Published: (2025)
by: Ru, Jinghan, et al.
Published: (2025)
Large language models have learned to use language
by: Lupyan, Gary
Published: (2025)
by: Lupyan, Gary
Published: (2025)
Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?
by: Song, Yingjin, et al.
Published: (2025)
by: Song, Yingjin, et al.
Published: (2025)
When can transformers reason with abstract symbols?
by: Boix-Adsera, Enric, et al.
Published: (2023)
by: Boix-Adsera, Enric, et al.
Published: (2023)
Evaluating Spatiotemporal Consistency in Automatically Generated Sewing Instructions
by: Geiger, Luisa, et al.
Published: (2025)
by: Geiger, Luisa, et al.
Published: (2025)
FTFT: Efficient and Robust Fine-Tuning by Transferring Training Dynamics
by: Du, Yupei, et al.
Published: (2023)
by: Du, Yupei, et al.
Published: (2023)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
by: Lindemann, Matthias, et al.
Published: (2023)
by: Lindemann, Matthias, et al.
Published: (2023)
LLMs syntactically adapt their language use to their conversational partner
by: Kandra, Florian, et al.
Published: (2025)
by: Kandra, Florian, et al.
Published: (2025)
Collaborative Problem-Solving in an Optimization Game
by: Jeknic, Isidora, et al.
Published: (2025)
by: Jeknic, Isidora, et al.
Published: (2025)
A Dialogue Game for Eliciting Balanced Collaboration
by: Jeknić, Isidora, et al.
Published: (2024)
by: Jeknić, Isidora, et al.
Published: (2024)
ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation
by: Zhao, Raoyuan, et al.
Published: (2026)
by: Zhao, Raoyuan, et al.
Published: (2026)
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
by: Han, Yifu, et al.
Published: (2025)
by: Han, Yifu, et al.
Published: (2025)
Same evaluation, more tokens: On the effect of input length for machine translation evaluation using Large Language Models
by: Domhan, Tobias, et al.
Published: (2025)
by: Domhan, Tobias, et al.
Published: (2025)
BRIDGE: Benchmark for multi-hop Reasoning In long multimodal Documents with Grounded Evidence
by: Xiang, Biao, et al.
Published: (2026)
by: Xiang, Biao, et al.
Published: (2026)
Slm-mux: Orchestrating small language models for reasoning
by: Wang, Chenyu, et al.
Published: (2025)
by: Wang, Chenyu, et al.
Published: (2025)
Exploring the limits of decoder-only models trained on public speech recognition corpora
by: Gupta, Ankit, et al.
Published: (2024)
by: Gupta, Ankit, et al.
Published: (2024)
Manipulating language models' training data to study syntactic constraint learning: the case of English passivization
by: Leong, Cara Su-Yi, et al.
Published: (2024)
by: Leong, Cara Su-Yi, et al.
Published: (2024)
Large Language Models for Depression Recognition in Spoken Language Integrating Psychological Knowledge
by: Li, Yupei, et al.
Published: (2025)
by: Li, Yupei, et al.
Published: (2025)
Positional Biases Shift as Inputs Approach Context Window Limits
by: Veseli, Blerta, et al.
Published: (2025)
by: Veseli, Blerta, et al.
Published: (2025)
Scope-enhanced Compositional Semantic Parsing for DRT
by: Yang, Xiulin, et al.
Published: (2024)
by: Yang, Xiulin, et al.
Published: (2024)
AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing
by: Thillainathan, Sarubi, et al.
Published: (2026)
by: Thillainathan, Sarubi, et al.
Published: (2026)
World model inspired sarcasm reasoning with large language model agents
by: Inoshita, Keito, et al.
Published: (2025)
by: Inoshita, Keito, et al.
Published: (2025)
Similar Items
-
Simple and effective data augmentation for compositional generalization
by: Yao, Yuekun, et al.
Published: (2024) -
Predicting generalization performance with correctness discriminators
by: Yao, Yuekun, et al.
Published: (2023) -
Barriers to Universal Reasoning With Transformers (And How to Overcome Them)
by: Kraus, Oliver, et al.
Published: (2026) -
Distributional reasoning in LLMs: Parallel reasoning processes in multi-hop reasoning
by: Shalev, Yuval, et al.
Published: (2024) -
On the Ability of Transformers to Verify Plans
by: Sarrof, Yash, et al.
Published: (2026)