Saved in:
| Main Authors: | Rentschler, Micah, Roberts, Jesse |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.14176 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025)
by: Rentschler, Micah, et al.
Published: (2025)
The Ends Justify the Thoughts: RL-Induced Motivated Reasoning in LLM CoTs
by: Howe, Nikolaus, et al.
Published: (2025)
by: Howe, Nikolaus, et al.
Published: (2025)
Memory-Enhanced Neural Solvers for Routing Problems
by: Chalumeau, Felix, et al.
Published: (2024)
by: Chalumeau, Felix, et al.
Published: (2024)
General-Purpose In-Context Learning by Meta-Learning Transformers
by: Kirsch, Louis, et al.
Published: (2022)
by: Kirsch, Louis, et al.
Published: (2022)
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
by: Xiao, Yubin, et al.
Published: (2024)
by: Xiao, Yubin, et al.
Published: (2024)
Benchmarking General-Purpose In-Context Learning
by: Wang, Fan, et al.
Published: (2024)
by: Wang, Fan, et al.
Published: (2024)
Beyond Accuracy: EcoL2 Metric for Sustainable Neural PDE Solvers
by: Kapoor, Taniya, et al.
Published: (2025)
by: Kapoor, Taniya, et al.
Published: (2025)
General Flexible $f$-divergence for Challenging Offline RL Datasets with Low Stochasticity and Diverse Behavior Policies
by: Wang, Jianxun, et al.
Published: (2026)
by: Wang, Jianxun, et al.
Published: (2026)
Rethinking Light Decoder-based Solvers for Vehicle Routing Problems
by: Huang, Ziwei, et al.
Published: (2025)
by: Huang, Ziwei, et al.
Published: (2025)
Solver-Free Decision-Focused Learning for Linear Optimization Problems
by: Berden, Senne, et al.
Published: (2025)
by: Berden, Senne, et al.
Published: (2025)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
by: Liang, Anthony, et al.
Published: (2026)
by: Liang, Anthony, et al.
Published: (2026)
Distance-aware Attention Reshaping: Enhance Generalization of Neural Solver for Large-scale Vehicle Routing Problems
by: Wang, Yang, et al.
Published: (2024)
by: Wang, Yang, et al.
Published: (2024)
Loss Landscape Degeneracy and Stagewise Development in Transformers
by: Hoogland, Jesse, et al.
Published: (2024)
by: Hoogland, Jesse, et al.
Published: (2024)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
AttentionSmithy: A Modular Framework for Rapid Transformer Development and Customization
by: Cranney, Caleb, et al.
Published: (2025)
by: Cranney, Caleb, et al.
Published: (2025)
IVP-VAE: Modeling EHR Time Series with Initial Value Problem Solvers
by: Xiao, Jingge, et al.
Published: (2023)
by: Xiao, Jingge, et al.
Published: (2023)
The Autonomy-Alignment Problem in Open-Ended Learning Robots: Formalising the Purpose Framework
by: Baldassarre, Gianluca, et al.
Published: (2024)
by: Baldassarre, Gianluca, et al.
Published: (2024)
Reinforcement Learning from Meta-Evaluation: Aligning Language Models Without Ground-Truth Labels
by: Rentschler, Micah, et al.
Published: (2026)
by: Rentschler, Micah, et al.
Published: (2026)
Out-of-Distribution Generalization for Neural Physics Solvers
by: Wei, Zhao, et al.
Published: (2026)
by: Wei, Zhao, et al.
Published: (2026)
Improving Transformer World Models for Data-Efficient RL
by: Dedieu, Antoine, et al.
Published: (2025)
by: Dedieu, Antoine, et al.
Published: (2025)
Online Finetuning Decision Transformers with Pure RL Gradients
by: Luo, Junkai, et al.
Published: (2026)
by: Luo, Junkai, et al.
Published: (2026)
Lifelong Learner: Discovering Versatile Neural Solvers for Vehicle Routing Problems
by: Feng, Shaodi, et al.
Published: (2025)
by: Feng, Shaodi, et al.
Published: (2025)
DPN: Decoupling Partition and Navigation for Neural Solvers of Min-max Vehicle Routing Problems
by: Zheng, Zhi, et al.
Published: (2024)
by: Zheng, Zhi, et al.
Published: (2024)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Towards Efficient Constraint Handling in Neural Solvers for Routing Problems
by: Bi, Jieyi, et al.
Published: (2026)
by: Bi, Jieyi, et al.
Published: (2026)
Efficient Neural Combinatorial Optimization Solver for the Min-max Heterogeneous Capacitated Vehicle Routing Problem
by: Wu, Xuan, et al.
Published: (2025)
by: Wu, Xuan, et al.
Published: (2025)
Is PRM Necessary? Problem-Solving RL Implicitly Induces PRM Capability in LLMs
by: Feng, Zhangying, et al.
Published: (2025)
by: Feng, Zhangying, et al.
Published: (2025)
Generative Latent Neural PDE Solver using Flow Matching
by: Li, Zijie, et al.
Published: (2025)
by: Li, Zijie, et al.
Published: (2025)
Latent Generative Solvers for Generalizable Long-Term Physics Simulation
by: Chen, Zituo, et al.
Published: (2026)
by: Chen, Zituo, et al.
Published: (2026)
FLIP: Flow-Centric Generative Planning as General-Purpose Manipulation World Model
by: Gao, Chongkai, et al.
Published: (2024)
by: Gao, Chongkai, et al.
Published: (2024)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
by: Wang, Yufei, et al.
Published: (2024)
by: Wang, Yufei, et al.
Published: (2024)
Toward Explainable Offline RL: Analyzing Representations in Intrinsically Motivated Decision Transformers
by: Guiducci, Leonardo, et al.
Published: (2025)
by: Guiducci, Leonardo, et al.
Published: (2025)
Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data
by: Ran-Milo, Yuval, et al.
Published: (2026)
by: Ran-Milo, Yuval, et al.
Published: (2026)
Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramér Surrogate
by: C., Simo Alami, et al.
Published: (2025)
by: C., Simo Alami, et al.
Published: (2025)
An Efficient Learning-based Solver Comparable to Metaheuristics for the Capacitated Arc Routing Problem
by: Guo, Runze, et al.
Published: (2024)
by: Guo, Runze, et al.
Published: (2024)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
by: Bhatia, Abhinav, et al.
Published: (2023)
by: Bhatia, Abhinav, et al.
Published: (2023)
Stop Regressing: Training Value Functions via Classification for Scalable Deep RL
by: Farebrother, Jesse, et al.
Published: (2024)
by: Farebrother, Jesse, et al.
Published: (2024)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE Solvers
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Similar Items
-
Exploitation Is All You Need... for Exploration
by: Rentschler, Micah, et al.
Published: (2025) -
The Ends Justify the Thoughts: RL-Induced Motivated Reasoning in LLM CoTs
by: Howe, Nikolaus, et al.
Published: (2025) -
Memory-Enhanced Neural Solvers for Routing Problems
by: Chalumeau, Felix, et al.
Published: (2024) -
General-Purpose In-Context Learning by Meta-Learning Transformers
by: Kirsch, Louis, et al.
Published: (2022) -
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
by: Xiao, Yubin, et al.
Published: (2024)