Leveraging weights signals -- Predicting and improving generalizability in reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Moulin, Olivier, Francois-lavet, Vincent, Elbers, Paul, Hoogendoorn, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Disentangled (Un)Controllable Features
by: Kooi, Jacob E., et al.
Published: (2022)
by: Kooi, Jacob E., et al.
Published: (2022)
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
by: Mirbakhsh, Shahin, et al.
Published: (2024)
by: Mirbakhsh, Shahin, et al.
Published: (2024)
Interactive incremental learning of generalizable skills with local trajectory modulation
by: Knauer, Markus, et al.
Published: (2024)
by: Knauer, Markus, et al.
Published: (2024)
Leveraging LLMs for reward function design in reinforcement learning control tasks
by: Cardenoso, Franklin, et al.
Published: (2025)
by: Cardenoso, Franklin, et al.
Published: (2025)
Exploring the impact of traffic signal control and connected and automated vehicles on intersections safety: A deep reinforcement learning approach
by: Karbasi, Amir Hossein, et al.
Published: (2024)
by: Karbasi, Amir Hossein, et al.
Published: (2024)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Inhibitory normalization of error signals improves learning in neural circuits
by: Eyono, Roy Henha, et al.
Published: (2026)
by: Eyono, Roy Henha, et al.
Published: (2026)
Bayesian Modelling in Practice: Using Uncertainty to Improve Trustworthiness in Medical Applications
by: Ruhe, David, et al.
Published: (2019)
by: Ruhe, David, et al.
Published: (2019)
A meshfree exterior calculus for generalizable and data-efficient learning of physics from point clouds
by: Shaffer, Benjamin D., et al.
Published: (2026)
by: Shaffer, Benjamin D., et al.
Published: (2026)
Curriculum reinforcement learning with measurable task representation learning
by: Wen, Yongyan, et al.
Published: (2026)
by: Wen, Yongyan, et al.
Published: (2026)
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)
by: Lee, Sunbowen, et al.
Published: (2025)
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025)
by: Kadurha, David Krame, et al.
Published: (2025)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Deep reinforcement learning with time-scale invariant memory
by: Kabir, Md Rysul, et al.
Published: (2024)
by: Kabir, Md Rysul, et al.
Published: (2024)
Delayed homomorphic reinforcement learning for environments with delayed feedback
by: Lee, Jongsoo, et al.
Published: (2026)
by: Lee, Jongsoo, et al.
Published: (2026)
Offline reinforcement learning for job-shop scheduling problems
by: Echeverria, Imanol, et al.
Published: (2024)
by: Echeverria, Imanol, et al.
Published: (2024)
Q-Net: Queue Length Estimation via Kalman-based Neural Networks
by: Gao, Ting, et al.
Published: (2025)
by: Gao, Ting, et al.
Published: (2025)
Driving pattern interpretation based on action phases clustering
by: Yao, Xue, et al.
Published: (2024)
by: Yao, Xue, et al.
Published: (2024)
Improving the accuracy and generalizability of molecular property regression models with a substructure-substitution-rule-informed framework
by: Fan, Xiaoyu, et al.
Published: (2025)
by: Fan, Xiaoyu, et al.
Published: (2025)
IGUANe: a 3D generalizable CycleGAN for multicenter harmonization of brain MR images
by: Roca, Vincent, et al.
Published: (2024)
by: Roca, Vincent, et al.
Published: (2024)
Dynamic feature selection in medical predictive monitoring by reinforcement learning
by: Chen, Yutong, et al.
Published: (2024)
by: Chen, Yutong, et al.
Published: (2024)
Economic span selection of bridge based on deep reinforcement learning
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Not all tokens are needed(NAT): token efficient reinforcement learning
by: Sang, Hejian, et al.
Published: (2026)
by: Sang, Hejian, et al.
Published: (2026)
Performance Comparison of Deep RL Algorithms for Mixed Traffic Cooperative Lane-Changing
by: Yao, Xue, et al.
Published: (2024)
by: Yao, Xue, et al.
Published: (2024)
Discovering highly efficient low-weight quantum error-correcting codes with reinforcement learning
by: He, Austin Yubo, et al.
Published: (2025)
by: He, Austin Yubo, et al.
Published: (2025)
An efficient deep reinforcement learning environment for flexible job-shop scheduling
by: Wu, Xinquan, et al.
Published: (2025)
by: Wu, Xinquan, et al.
Published: (2025)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
by: Kobayashi, Seijin, et al.
Published: (2025)
by: Kobayashi, Seijin, et al.
Published: (2025)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
by: Saghafian, Armin, et al.
Published: (2024)
by: Saghafian, Armin, et al.
Published: (2024)
Task diversity produces systematic transfer but inhibits continual reinforcement learning
by: Seth, Purab, et al.
Published: (2026)
by: Seth, Purab, et al.
Published: (2026)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
by: Hutson, Miles, et al.
Published: (2024)
by: Hutson, Miles, et al.
Published: (2024)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving
by: Qu, Yansong, et al.
Published: (2026)
by: Qu, Yansong, et al.
Published: (2026)
An advantage based policy transfer algorithm for reinforcement learning with measures of transferability
by: Alam, Md Ferdous, et al.
Published: (2023)
by: Alam, Md Ferdous, et al.
Published: (2023)
Survey on reinforcement learning for language processing
by: Uc-Cetina, Victor, et al.
Published: (2021)
by: Uc-Cetina, Victor, et al.
Published: (2021)
Novel RL approach for efficient Elevator Group Control Systems
by: Vaartjes, Nathan, et al.
Published: (2025)
by: Vaartjes, Nathan, et al.
Published: (2025)
Does learning the right latent variables necessarily improve in-context learning?
by: Mittal, Sarthak, et al.
Published: (2024)
by: Mittal, Sarthak, et al.
Published: (2024)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Designing an efficient and equitable humanitarian supply chain dynamically via reinforcement learning
by: Jin, Weijia
Published: (2025)
by: Jin, Weijia
Published: (2025)
Learning to summarize user information for personalized reinforcement learning from human feedback
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs
by: Gan, Xingwei, et al.
Published: (2026)
by: Gan, Xingwei, et al.
Published: (2026)
Similar Items
-
Disentangled (Un)Controllable Features
by: Kooi, Jacob E., et al.
Published: (2022) -
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
by: Mirbakhsh, Shahin, et al.
Published: (2024) -
Interactive incremental learning of generalizable skills with local trajectory modulation
by: Knauer, Markus, et al.
Published: (2024) -
Leveraging LLMs for reward function design in reinforcement learning control tasks
by: Cardenoso, Franklin, et al.
Published: (2025) -
Exploring the impact of traffic signal control and connected and automated vehicles on intersections safety: A deep reinforcement learning approach
by: Karbasi, Amir Hossein, et al.
Published: (2024)