RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Qu, Yuxiao, Singh, Anikait, Lee, Yoonho, Setlur, Amrith, Salakhutdinov, Ruslan, Finn, Chelsea, Kumar, Aviral
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908573702815744
author Qu, Yuxiao
Singh, Anikait
Lee, Yoonho
Setlur, Amrith
Salakhutdinov, Ruslan
Finn, Chelsea
Kumar, Aviral
author_facet Qu, Yuxiao
Singh, Anikait
Lee, Yoonho
Setlur, Amrith
Salakhutdinov, Ruslan
Finn, Chelsea
Kumar, Aviral
contents Reasoning requires going beyond pattern matching or memorization of solutions to identify and implement "algorithmic procedures" that can be used to deduce answers to hard problems. Doing so requires realizing the most relevant primitives, intermediate results, or shared procedures, and building upon them. While RL post-training on long chains of thought ultimately aims to uncover this kind of algorithmic behavior, most reasoning traces learned by large models fail to consistently capture or reuse procedures, instead drifting into verbose and degenerate exploration. To address more effective reasoning, we introduce reasoning abstractions: concise natural language descriptions of procedural and factual knowledge that guide the model toward learning successful reasoning. We train models to be capable of proposing multiple abstractions given a problem, followed by RL that incentivizes building a solution while using the information provided by these abstractions. This results in a two-player RL training paradigm, abbreviated as RLAD, that jointly trains an abstraction generator and a solution generator. This setup effectively enables structured exploration, decouples learning signals of abstraction proposal and solution generation, and improves generalization to harder problems. We also show that allocating more test-time compute to generating abstractions is more beneficial for performance than generating more solutions at large test budgets, illustrating the role of abstractions in guiding meaningful exploration.
format Preprint
id arxiv_https___arxiv_org_abs_2510_02263
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
Qu, Yuxiao
Singh, Anikait
Lee, Yoonho
Setlur, Amrith
Salakhutdinov, Ruslan
Finn, Chelsea
Kumar, Aviral
Artificial Intelligence
Computation and Language
Machine Learning
Reasoning requires going beyond pattern matching or memorization of solutions to identify and implement "algorithmic procedures" that can be used to deduce answers to hard problems. Doing so requires realizing the most relevant primitives, intermediate results, or shared procedures, and building upon them. While RL post-training on long chains of thought ultimately aims to uncover this kind of algorithmic behavior, most reasoning traces learned by large models fail to consistently capture or reuse procedures, instead drifting into verbose and degenerate exploration. To address more effective reasoning, we introduce reasoning abstractions: concise natural language descriptions of procedural and factual knowledge that guide the model toward learning successful reasoning. We train models to be capable of proposing multiple abstractions given a problem, followed by RL that incentivizes building a solution while using the information provided by these abstractions. This results in a two-player RL training paradigm, abbreviated as RLAD, that jointly trains an abstraction generator and a solution generator. This setup effectively enables structured exploration, decouples learning signals of abstraction proposal and solution generation, and improves generalization to harder problems. We also show that allocating more test-time compute to generating abstractions is more beneficial for performance than generating more solutions at large test budgets, illustrating the role of abstractions in guiding meaningful exploration.
title RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
topic Artificial Intelligence
Computation and Language
Machine Learning
url https://arxiv.org/abs/2510.02263