Learning from Less: SINDy Surrogates in RL
Fuente:
arXiv
Saved in:
| Main Authors: | Dixit, Aniket, Khan, Muhammad Ibrahim, Ahmed, Faizan, Brusey, James |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Higher-Order Action Regularization in Deep Reinforcement Learning: From Continuous Control to Building Energy Management
by: Ahmed, Faizan, et al.
Published: (2026)
by: Ahmed, Faizan, et al.
Published: (2026)
Deadline-Aware, Energy-Efficient Control of Domestic Immersion Hot Water Heater
by: Khan, Muhammad Ibrahim, et al.
Published: (2026)
by: Khan, Muhammad Ibrahim, et al.
Published: (2026)
LIMR: Less is More for RL Scaling
by: Li, Xuefeng, et al.
Published: (2025)
by: Li, Xuefeng, et al.
Published: (2025)
Less is more? Rewards in RL for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2025)
by: Bates, Elizabeth, et al.
Published: (2025)
Smart Sampling: Self-Attention and Bootstrapping for Improved Ensembled Q-Learning
by: Khan, Muhammad Junaid, et al.
Published: (2024)
by: Khan, Muhammad Junaid, et al.
Published: (2024)
Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing
by: Khan, Azal Ahmad, et al.
Published: (2026)
by: Khan, Azal Ahmad, et al.
Published: (2026)
Paying Less Generalization Tax: A Cross-Domain Generalization Study of RL Training for LLM Agents
by: Liu, Zhihan, et al.
Published: (2026)
by: Liu, Zhihan, et al.
Published: (2026)
Surrogate models for Rock-Fluid Interaction: A Grid-Size-Invariant Approach
by: Pinheiro, Nathalie C., et al.
Published: (2026)
by: Pinheiro, Nathalie C., et al.
Published: (2026)
When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation
by: Faisal, Faizan
Published: (2026)
by: Faisal, Faizan
Published: (2026)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
by: Bhatia, Abhinav, et al.
Published: (2023)
by: Bhatia, Abhinav, et al.
Published: (2023)
Adaptive Online Learning with LSTM Networks for Energy Price Prediction
by: Salihoglu, Salih, et al.
Published: (2025)
by: Salihoglu, Salih, et al.
Published: (2025)
Artificial Intelligence and Deep Learning Algorithms for Epigenetic Sequence Analysis: A Review for Epigeneticists and AI Experts
by: Tahir, Muhammad, et al.
Published: (2025)
by: Tahir, Muhammad, et al.
Published: (2025)
LoRA Learns Less and Forgets Less
by: Biderman, Dan, et al.
Published: (2024)
by: Biderman, Dan, et al.
Published: (2024)
Surrogate Fitness Metrics for Interpretable Reinforcement Learning
by: Altmann, Philipp, et al.
Published: (2025)
by: Altmann, Philipp, et al.
Published: (2025)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization
by: Khan, Zaid, et al.
Published: (2026)
by: Khan, Zaid, et al.
Published: (2026)
Algorithmic Fairness in AI Surrogates for End-of-Life Decision-Making
by: Ahmad, Muhammad Aurangzeb
Published: (2025)
by: Ahmad, Muhammad Aurangzeb
Published: (2025)
The Road Less Scheduled
by: Defazio, Aaron, et al.
Published: (2024)
by: Defazio, Aaron, et al.
Published: (2024)
Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramér Surrogate
by: C., Simo Alami, et al.
Published: (2025)
by: C., Simo Alami, et al.
Published: (2025)
Is Monotonic Sampling Necessary in Diffusion Models?
by: Khan, Muhammad Haris
Published: (2026)
by: Khan, Muhammad Haris
Published: (2026)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
An Explainable Disease Surveillance System for Early Prediction of Multiple Chronic Diseases
by: Khan, Shaheer Ahmad, et al.
Published: (2025)
by: Khan, Shaheer Ahmad, et al.
Published: (2025)
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
by: Bauer, Justin, et al.
Published: (2026)
by: Bauer, Justin, et al.
Published: (2026)
Watch Less, Feel More: Sim-to-Real RL for Generalizable Articulated Object Manipulation via Motion Adaptation and Impedance Control
by: Do, Tan-Dzung, et al.
Published: (2025)
by: Do, Tan-Dzung, et al.
Published: (2025)
DeepFeatIoT: Unifying Deep Learned, Randomized, and LLM Features for Enhanced IoT Time Series Sensor Data Classification in Smart Industries
by: Inan, Muhammad Sakib Khan, et al.
Published: (2025)
by: Inan, Muhammad Sakib Khan, et al.
Published: (2025)
Scalable Decision Focused Learning via Online Trainable Surrogates
by: Signorelli, Gaetano, et al.
Published: (2025)
by: Signorelli, Gaetano, et al.
Published: (2025)
GCA Framework: A GCC Countries-Grounded Dataset and Agentic Pipeline for Climate Decision Support
by: Sheikh, Muhammad Umer, et al.
Published: (2026)
by: Sheikh, Muhammad Umer, et al.
Published: (2026)
SparseRL-Sync: Lossless Weight Synchronization with ~100x Less Communication
by: Hu, Lucas, et al.
Published: (2026)
by: Hu, Lucas, et al.
Published: (2026)
Transfer Learning of Surrogate Models: Integrating Domain Warping and Affine Transformations
by: Pan, Shuaiqun, et al.
Published: (2025)
by: Pan, Shuaiqun, et al.
Published: (2025)
Minimizing Surrogate Losses for Decision-Focused Learning using Differentiable Optimization
by: Mandi, Jayanta, et al.
Published: (2025)
by: Mandi, Jayanta, et al.
Published: (2025)
Learning Surrogates for Offline Black-Box Optimization via Gradient Matching
by: Hoang, Minh, et al.
Published: (2025)
by: Hoang, Minh, et al.
Published: (2025)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
by: Mark, Max Sobol, et al.
Published: (2024)
by: Mark, Max Sobol, et al.
Published: (2024)
Transitive RL: Value Learning via Divide and Conquer
by: Park, Seohong, et al.
Published: (2025)
by: Park, Seohong, et al.
Published: (2025)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
by: Bortkiewicz, Michał, et al.
Published: (2025)
by: Bortkiewicz, Michał, et al.
Published: (2025)
Policy Learning for Off-Dynamics RL with Deficient Support
by: Van, Linh Le Pham, et al.
Published: (2024)
by: Van, Linh Le Pham, et al.
Published: (2024)
EchoRL: Reinforcement Learning via Rollout Echoing
by: Bi, Jinhe, et al.
Published: (2026)
by: Bi, Jinhe, et al.
Published: (2026)
RL-GPT: Integrating Reinforcement Learning and Code-as-policy
by: Liu, Shaoteng, et al.
Published: (2024)
by: Liu, Shaoteng, et al.
Published: (2024)
Is Value Learning Really the Main Bottleneck in Offline RL?
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
The Role of Deep Learning Regularizations on Actors in Offline RL
by: Tarasov, Denis, et al.
Published: (2024)
by: Tarasov, Denis, et al.
Published: (2024)
Unsupervised Surrogate Anomaly Detection
by: Klüttermann, Simon, et al.
Published: (2025)
by: Klüttermann, Simon, et al.
Published: (2025)
Similar Items
-
Higher-Order Action Regularization in Deep Reinforcement Learning: From Continuous Control to Building Energy Management
by: Ahmed, Faizan, et al.
Published: (2026) -
Deadline-Aware, Energy-Efficient Control of Domestic Immersion Hot Water Heater
by: Khan, Muhammad Ibrahim, et al.
Published: (2026) -
LIMR: Less is More for RL Scaling
by: Li, Xuefeng, et al.
Published: (2025) -
Less is more? Rewards in RL for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2025) -
Smart Sampling: Self-Attention and Bootstrapping for Improved Ensembled Q-Learning
by: Khan, Muhammad Junaid, et al.
Published: (2024)