Beyond Inference-Time Search: Reinforcement Learning Synthesizes Reusable Solvers
Fuente:
arXiv
Saved in:
| Main Authors: | Massoudi, Soheyl, Apaza, Gabriel, Habibi, Milad, Fuge, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agentic Large Language Models for Conceptual Systems Engineering and Design
by: Massoudi, Soheyl, et al.
Published: (2025)
by: Massoudi, Soheyl, et al.
Published: (2025)
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
by: Molinari, Gioele, et al.
Published: (2026)
by: Molinari, Gioele, et al.
Published: (2026)
GLUE: Coordinating Pre-Trained Generative Models for System-Level Design
by: Aebersold, Tim, et al.
Published: (2025)
by: Aebersold, Tim, et al.
Published: (2025)
Inverse design with conditional cascaded diffusion models
by: Habibi, Milad, et al.
Published: (2024)
by: Habibi, Milad, et al.
Published: (2024)
EngiBench: A Framework for Data-Driven Engineering Design Research
by: Felten, Florian, et al.
Published: (2025)
by: Felten, Florian, et al.
Published: (2025)
Active Inference with Reusable State-Dependent Value Profiles
by: Poschl, Jacob
Published: (2025)
by: Poschl, Jacob
Published: (2025)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
by: Liu, Max, et al.
Published: (2024)
by: Liu, Max, et al.
Published: (2024)
DiffOPF: Diffusion Solver for Optimal Power Flow
by: Hoseinpour, Milad, et al.
Published: (2025)
by: Hoseinpour, Milad, et al.
Published: (2025)
Inference-Time Alignment Control for Diffusion Models with Reinforcement Learning Guidance
by: Jin, Luozhijie, et al.
Published: (2025)
by: Jin, Luozhijie, et al.
Published: (2025)
Action-List Reinforcement Learning Syndrome Decoding for Binary Linear Block Codes
by: Taghipour, Milad, et al.
Published: (2025)
by: Taghipour, Milad, et al.
Published: (2025)
Beyond The Rainbow: High Performance Deep Reinforcement Learning on a Desktop PC
by: Clark, Tyler, et al.
Published: (2024)
by: Clark, Tyler, et al.
Published: (2024)
Boosting Cross-problem Generalization in Diffusion-Based Neural Combinatorial Solver via Inference Time Adaptation
by: Lei, Haoyu, et al.
Published: (2025)
by: Lei, Haoyu, et al.
Published: (2025)
Dynamic Search for Inference-Time Alignment in Diffusion Models
by: Li, Xiner, et al.
Published: (2025)
by: Li, Xiner, et al.
Published: (2025)
Edge AI-Powered Real-Time Decision-Making for Autonomous Vehicles in Adverse Weather Conditions
by: Rahmati, Milad
Published: (2025)
by: Rahmati, Milad
Published: (2025)
Flash Inference: Near Linear Time Inference for Long Convolution Sequence Models and Beyond
by: Oncescu, Costin-Andrei, et al.
Published: (2024)
by: Oncescu, Costin-Andrei, et al.
Published: (2024)
Learning Time-Series Representations by Hierarchical Uniformity-Tolerance Latent Balancing
by: Jalali, Amin, et al.
Published: (2025)
by: Jalali, Amin, et al.
Published: (2025)
UNSAT Solver Synthesis via Monte Carlo Forest Search
by: Cameron, Chris, et al.
Published: (2022)
by: Cameron, Chris, et al.
Published: (2022)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
by: Sokota, Samuel, et al.
Published: (2025)
by: Sokota, Samuel, et al.
Published: (2025)
Thermodynamic Focusing for Inference-Time Search: Practical Methods for Target-Conditioned Sampling and Prompted Inference
by: Zhang, Zhan
Published: (2025)
by: Zhang, Zhan
Published: (2025)
Towards the Reusability and Compositionality of Causal Representations
by: Talon, Davide, et al.
Published: (2024)
by: Talon, Davide, et al.
Published: (2024)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
by: Zhao, Eric, et al.
Published: (2025)
by: Zhao, Eric, et al.
Published: (2025)
RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design
by: Kshirsagar, Meghana, et al.
Published: (2026)
by: Kshirsagar, Meghana, et al.
Published: (2026)
Tree Search for LLM Agent Reinforcement Learning
by: Ji, Yuxiang, et al.
Published: (2025)
by: Ji, Yuxiang, et al.
Published: (2025)
Beyond Rewards in Reinforcement Learning for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2026)
by: Bates, Elizabeth, et al.
Published: (2026)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
by: Mu, Tongzhou, et al.
Published: (2024)
by: Mu, Tongzhou, et al.
Published: (2024)
STEER: Inference-Time Risk Control via Constrained Quality-Diversity Search
by: Yang, Eric, et al.
Published: (2026)
by: Yang, Eric, et al.
Published: (2026)
Learning to Reduce Search Space for Generalizable Neural Routing Solver
by: Zhou, Changliang, et al.
Published: (2025)
by: Zhou, Changliang, et al.
Published: (2025)
Beyond Accuracy: EcoL2 Metric for Sustainable Neural PDE Solvers
by: Kapoor, Taniya, et al.
Published: (2025)
by: Kapoor, Taniya, et al.
Published: (2025)
Beyond Model Collapse: Scaling Up with Synthesized Data Requires Verification
by: Feng, Yunzhen, et al.
Published: (2024)
by: Feng, Yunzhen, et al.
Published: (2024)
Reinforcement Learning Methods for Neighborhood Selection in Local Search
by: Molinghen, Yannick, et al.
Published: (2026)
by: Molinghen, Yannick, et al.
Published: (2026)
Log-Augmented Generation: Scaling Test-Time Reasoning with Reusable Computation
by: Chen, Peter Baile, et al.
Published: (2025)
by: Chen, Peter Baile, et al.
Published: (2025)
Learning to Guide Local Search for MPE Inference in Probabilistic Graphical Models
by: Malhotra, Brij, et al.
Published: (2026)
by: Malhotra, Brij, et al.
Published: (2026)
Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
by: Riemer, Matthew, et al.
Published: (2024)
by: Riemer, Matthew, et al.
Published: (2024)
Archon: An Architecture Search Framework for Inference-Time Techniques
by: Saad-Falcon, Jon, et al.
Published: (2024)
by: Saad-Falcon, Jon, et al.
Published: (2024)
In Search for Architectures and Loss Functions in Multi-Objective Reinforcement Learning
by: Terekhov, Mikhail, et al.
Published: (2024)
by: Terekhov, Mikhail, et al.
Published: (2024)
Enhancing Reinforcement Learning for the Floorplanning of Analog ICs with Beam Search
by: Della Rovere, Sandro Junior, et al.
Published: (2025)
by: Della Rovere, Sandro Junior, et al.
Published: (2025)
A Survey on Neural Architecture Search Based on Reinforcement Learning
by: Shao, Wenzhu
Published: (2024)
by: Shao, Wenzhu
Published: (2024)
Beyond Prediction: Reinforcement Learning as the Defining Leap in Healthcare AI
by: Perera, Dilruk, et al.
Published: (2025)
by: Perera, Dilruk, et al.
Published: (2025)
Inverse Reinforcement Learning by Estimating Expertise of Demonstrators
by: Beliaev, Mark, et al.
Published: (2024)
by: Beliaev, Mark, et al.
Published: (2024)
Similar Items
-
Agentic Large Language Models for Conceptual Systems Engineering and Design
by: Massoudi, Soheyl, et al.
Published: (2025) -
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
by: Molinari, Gioele, et al.
Published: (2026) -
GLUE: Coordinating Pre-Trained Generative Models for System-Level Design
by: Aebersold, Tim, et al.
Published: (2025) -
Inverse design with conditional cascaded diffusion models
by: Habibi, Milad, et al.
Published: (2024) -
EngiBench: A Framework for Data-Driven Engineering Design Research
by: Felten, Florian, et al.
Published: (2025)